Method and device for controlling a recording volume and storage medium

DE602020055362T2Active Publication Date: 2025-07-30BEIJING XIAOMI PINECONE ELECTRONICS CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
DE602020055362
Authority / Receiving Office
DE · DE
Patent Type
Patents
Current Assignee / Owner
Priority Date
2020-03-10
Filing Date
2020-08-28
Publication Date
2025-07-30
Estimated Expiration
2040-08-28

AI Technical Summary

Technical Problem

Existing voice interaction apparatuses with different acoustic hardware architectures struggle to generate consistent recording volumes, affecting the accuracy of nearby wakeup mechanisms based on energy or signal-to-noise ratios.

Method used

A method and device for controlling recording volume by determining adjustment gains based on microphone sensitivities and analog-to-digital conversion quantified reference voltages to calibrate voice interaction apparatuses, ensuring they produce the same recording volume under the same sound pressure level.

Benefits of technology

This approach enables accurate running of nearby wakeup mechanisms by ensuring all voice interaction apparatuses generate consistent recording volumes, improving the precision of response determination based on energy or signal-to-noise ratios.

✦ Generated by Eureka AI based on patent content.
Patent Text Reader
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] The present disclosure relates to the technical field of voice interaction apparatuses, and more particularly, to a method and device for controlling a recording volume, and a storage medium.BACKGROUND

[0002] With the development of artificial intelligent technologies, voice control has become an important application in smart home scenes. US 6 914 989 B2 relates to microphone calibration in a conference system with a plurality of microphones. US 2012 / 106749 A1 relates to the calibration of a plurality of microphones in a vehicle environment.

[0003] There may be a plurality of voice interaction apparatuses based on voice control in a user's house. In order to avoid concurrent responding to one instruction by the plurality of voice interaction apparatuses, a function of waking up a nearby voice interaction apparatus is typically implemented based on voice energy received or picked up by a microphone or a signal-to-noise ratio.

[0004] Therefore, how to guarantee accurate running of a nearby wakeup mechanism based on energy or the signal-to-noise ratio needs to be addressed.SUMMARY

[0005] In order to overcome problems in relevant technologies, the present disclosure provides a method and device for controlling a recording volume, and a storage medium.

[0006] According to a first aspect of the embodiments of the disclosure, a method for controlling a recording volume according to claim 1 is provided.

[0007] In an implementation, the operation of determining the adjustment gain based on the first microphone sensitivity, the second microphone sensitivity of the voice interaction apparatus and the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus may include: determining a difference between a sum of the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus plus the first microphone sensitivity and the second microphone sensitivity as the adjustment gain.

[0008] In an implementation, after determining the adjustment gain, the method may further include: calibrating the voice interaction apparatus with the adjustment gain determined at present; testing an actual recording volume of the calibrated voice interaction apparatus under a preset sound pressure level; determining a reference recording volume of the reference apparatus under the preset sound pressure level; and determining a calibration value of the adjustment gain determined at present based on the reference recording volume and the actual recording volume, and determining the adjustment gain obtained after calibration as a final adjustment gain for calibrating the voice interaction apparatus.

[0009] In an implementation, determining the calibration value of the adjustment gain determined at present based on the reference recording volume and the actual recording volume may include: determining a difference between the reference recording volume and the actual recording volume; and determining the difference as the calibration value of the adjustment gain determined at present.

[0010] In an implementation, the reference apparatus may be a voice interaction apparatus provided with a digital microphone; and / or the reference apparatus may be a voice interaction apparatus provided with a direct-sound microphone sound chamber.

[0011] According to a second aspect of the embodiments of the present disclosure, a device for controlling a recording volume according to claim 6 is provided.

[0012] In an implementation, the determination module may be configured to: determine a difference between a sum of the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus plus the first microphone sensitivity and the second microphone sensitivity as the adjustment gain.

[0013] In an implementation, the determination module may be further configured to: calibrate the voice interaction apparatus with the adjustment gain determined at present; test an actual recording volume of the calibrated voice interaction apparatus under a preset sound pressure level; determine a reference recording volume of the reference apparatus under the preset sound pressure level; and determine a calibration value of the adjustment gain determined at present based on the reference recording volume and the actual recording volume, and determine an adjustment gain obtained after calibration as a final adjustment gain for calibrating the voice interaction apparatus.

[0014] In an implementation, the determination module may be configured to: determine a difference between the reference recording volume and the actual recording volume; and determine the difference as the calibration value of the adjustment gain determined at present.

[0015] In an implementation, the reference apparatus may be a voice interaction apparatus provided with a digital microphone; and / or the reference apparatus may be a voice interaction apparatus provided with a direct-sound microphone sound chamber

[0016] According to another aspect of the embodiments of the present disclosure, a non-transitory computer readable storage medium according to claim 9 is provided. Computer executable instructions are stored in the non-transitory computer readable storage medium. When the computer readable instructions are executed by a processor, the method for controlling a recording volume in the first aspect or any one of the implementations of the first aspect is implemented.

[0017] The technical solution provided by the embodiments of the present disclosure may include the following beneficial effects. Through calibration of the voice interaction apparatuses with the adjustment gain, the voice interaction apparatuses with different acoustic hardware architectures can generate a same recording volume after picking up a user voice instruction under a same sound pressure level. Thus, accurate running of a nearby wakeup mechanism is achieved based on energy or a signal-to-noise ratio.

[0018] It should be understood that the above general descriptions and the following detailed descriptions are exemplary and explanatory only, and are not intended to limit the disclosure.BRIEF DESCRIPTION OF THE DRAWINGS

[0019] The accompanying drawings, which are incorporated in and constitute a part of this specification, illustrate embodiments consistent with the disclosure and, together with the specification, serve to explain the principles of the disclosure. FIG. 1 is a flowchart of a method for controlling a recording volume according to an exemplary embodiment of the disclosure; FIG. 2 is a flowchart of operations of determining an adjustment gain in a method for controlling a recording volume according to an exemplary embodiment of the disclosure; FIG. 3 is a schematic structural diagram of a voice interaction apparatus provided with a direct-sound microphone sound chamber; FIG. 4 is a schematic diagram of a pickup path of a voice interaction apparatus provided with a digital microphone; FIG. 5 is a schematic diagram of a pickup path of a voice interaction apparatus provided with an analog microphone and an analog-to-digital converter; FIG. 6 is a flowchart of another method for controlling a recording volume according to an exemplary embodiment of the disclosure; FIG. 7 is a flowchart of operations of determining a calibration value of a present adjustment gain in a method for controlling a recording volume according to an exemplary embodiment of the disclosure; FIG. 8 is a schematic diagram of a nearby wakeup mechanism based on energy or a signal-to-noise ratio; FIG. 9 is a block diagram of a device for controlling a recording volume according to an exemplary embodiment of the disclosure; and FIG. 10 is a block diagram of a device for controlling a recording volume according to an exemplary embodiment of the disclosure. DETAILED DESCRIPTION

[0020] Reference will now be made in detail to exemplary embodiments, examples of which are illustrated in the accompanying drawings. The following description refers to the accompanying drawings in which the same numbers in different drawings represent the same or similar elements unless otherwise indicated. The implementations set forth in the following description of exemplary embodiments do not represent all implementations consistent with the disclosure. Instead, they are merely examples of apparatuses and methods consistent with aspects related to the disclosure as recited in detail in the appended claims.

[0021] Based on a function of nearby wakeup of voice interaction apparatuses, respective initial recording volumes of the voice interaction apparatuses can be collected, and such information can be saved on a decision making apparatus. After the voice interaction apparatuses pick up a voice instruction of a user, volume information picked up respectively can be uploaded to the decision making apparatus.

[0022] The decision making apparatus, based on the inquired initial recording volumes of each voice interaction apparatus, can make corresponding compensation to volume information picked up by the voice interaction apparatuses. Based on compensated volume information, the decision making apparatus can determine a voice interaction apparatus with maximum signal energy or signal-to-noise ratio and confirm that the voice interaction apparatus is closest to the user. The determined voice interaction apparatus needs to make a response to operations of the user.

[0023] The above working mode, on one side, results in large workloads of the decision making apparatus and slows down confirmation of the decision making apparatus; and on the other side, due to limited storage space of the decision making apparatus, initial recording volumes of new voice interaction apparatuses may hardly be uploaded to the decision making apparatus in time. Hence, confirmation accuracy of the decision making apparatus is degraded and thus, accurate running of a nearby wakeup mechanism based on energy or a signal-to-noise ratio is affected.

[0024] A method for controlling a recording volume provided by the embodiments of the disclosure may guarantee accurate running of the nearby wakeup mechanism based on energy or the signal-to-noise ratio.

[0025] FIG. 1 is a flowchart of a method for controlling a recording volume according to an exemplary embodiment of the disclosure.

[0026] As shown in FIG. 1, in an exemplary embodiment of the disclosure, the method for controlling a recording volume includes operation S101 and operation S102. Operations S101 and S102 will be introduced as follows.

[0027] In operation S101, an adjustment gain is determined according to a preset reference index of a reference apparatus.

[0028] An output signal can be amplified by a gain. Through setting of the gain for a voice interaction apparatus, signals output by the voice interaction apparatus can be amplified.

[0029] The adjustment gain may be a hardware gain or a software gain.

[0030] In operation S102, at least one voice interaction apparatus is calibrated based on the adjustment gain, so that the at least one voice interaction apparatus has a same recording volume under a same sound pressure level.

[0031] The recording volume herein may be understood as a level or amplitude of a digital signal. For example, when recording is conducted by a device under a sound pressure level of 94 dB, a digital signal level generated by a 1 kHz single-frequency acoustic signal is taken as a characterization value of the recording volume.

[0032] The voice interaction apparatuses may have different acoustic hardware architectures. Hence, under the same sound pressure level, original volumes picked up by the voice interaction apparatuses may be different.

[0033] By setting corresponding adjustment gains for different voice interaction apparatuses, the voice interaction apparatuses can have a same recording volume under the same sound pressure level.

[0034] For an instance, based on the nearby wakeup mechanism, among the voice interaction apparatuses calibrated with the adjustment gain, the voice interaction apparatus picking up maximum volume information may be determined as a voice interaction apparatus needing to respond to a user instruction according to volume information of the user picked up by the voice interaction apparatus.

[0035] In an embodiment, a reference apparatus may be a voice interaction apparatus provided with a Digital Microphone (DMIC) or may also be a voice interaction apparatus provided with an Analog Microphone (AMIC) and an Analog-to-Digital Converter (ADC).

[0036] Adjustment gains of different voice interaction apparatuses need to be determined based on the reference apparatus, so that the stability and generality of the reference apparatus need to be achieved.

[0037] The voice interaction apparatuses may include those provided with a DMIC and those provided with an AMIC and an ADC. Compared to the voice interaction apparatuses with an AMIC and an ADC, the voice interaction apparatuses with a DMIC do not have a hardware gain that serves as a variable and thus, have higher stability.

[0038] FIG. 3 is a schematic structural diagram of a voice interaction apparatus provided with a direct-sound microphone sound chamber.

[0039] In an embodiment, as shown in FIG. 3, a reference apparatus may be a voice interaction apparatus provided with a direct-sound microphone sound chamber.

[0040] The direct-sound microphone sound chamber may be a direct sound chamber with a straight-line microphone sound chamber structure. In other words, a microphone picking up voice right faces a sound bore. Hence, when being propagated to the microphone, voice can be not shielded by anything, and there is no loss of voice energy.

[0041] According to the method for controlling a recording volume provided by the disclosure, the voice interaction apparatuses can be calibrated with an adjustment gain, so the voice interaction apparatuses with different acoustic hardware architectures can generate a same recording volume while picking up a user voice instruction under a same sound pressure level. Through the disclosure, accurate running of the nearby wakeup mechanism based on energy or the signal-to-noise ratio can be achieved.

[0042] FIG. 2 is a flowchart of operations of determining an adjustment gain in a method for controlling a recording volume according to an exemplary embodiment of the disclosure.

[0043] As shown in FIG. 2, the operations of determining an adjustment gain include operation S201 and operation S202. The operations S201 and S202 are introduced as follows. In the present embodiment, a preset reference index may include microphone sensitivity.

[0044] In operation S201, first microphone sensitivity of a reference apparatus is determined.

[0045] The preset reference index may be also other indexes. The microphone sensitivity is taken as the example here. Same theories are applicable to other indexes.

[0046] The first microphone sensitivity of the reference apparatus may be acquired through inquiry of specifications of the apparatus.

[0047] In an embodiment, first microphone sensitivity of a reference apparatus provided with a direct-sound microphone sound chamber and a DMIC can be determined.

[0048] Reference apparatuses of different models, provided with direct-sound microphone sound chambers and DMICs may have different first microphone sensitivity.

[0049] The first microphone sensitivity may be a preset value, such as -26 dB.

[0050] In operation S202, an adjustment gain is determined based on the first microphone sensitivity and second microphone sensitivity of a voice interaction apparatus.

[0051] According to the present invention, the adjustment gain is determined based on the first microphone sensitivity and the second microphone sensitivity. Based on the adjustment gain, recording volumes of a plurality of voice interaction apparatuses may be compensated. Hence, after picking up a user voice instruction under a same sound pressure level, the plurality of voice interaction apparatuses can generate a same recording volume.

[0052] In an example not being part of the present invention, the voice interaction apparatus may include a DMIC.

[0053] The adjustment gain of the voice interaction apparatus with the DMIC may be determined according to a difference between the first microphone sensitivity and the second microphone sensitivity.

[0054] For clearer illustration, as assigned, the adjustment gain is Gain*, the first microphone sensitivity is Smic 1 and the second microphone sensitivity is Smic 2 .

[0055] As for the voice interaction apparatus with the DMIC, the adjustment gain Gain*, the first microphone sensitivity Smic 1 and the second microphone sensitivity Smic 2 Gain * = Smic 1 - Smic 2 . can satisfy the following relationship:

[0056] The second microphone sensitivity Smic 2 of the voice interaction apparatus can be acquired through inquiry of specifications of the apparatus.

[0057] For example, the first microphone sensitivity Smic 1 may be -26 dB. If the second microphone sensitivity Smic 2 of the voice interaction apparatus with the DMIC is -22 dB, then Gain*=-26-(-22)=-4 dB.

[0058] FIG. 4 is a schematic diagram of a pickup path of a voice interaction apparatus provided with a DMIC.

[0059] As shown in FIG. 4, the adjustment gain Gain* of the voice interaction apparatus provided with the DMIC is a software gain. The adjustment gain Gain* may be implemented in a processor of the voice interaction apparatus.

[0060] In an exemplary embodiment of the disclosure, the voice interaction apparatus may include a voice interaction apparatus provided with an AMIC and an ADC.

[0061] According to the present invention, an adjustment gain of the voice interaction apparatus provided with the AMIC and the ADC is determined based on first microphone sensitivity, second microphone sensitivity of the voice interaction apparatus and an analog-to-digital conversion quantified reference voltage of the voice interaction apparatus.

[0062] In an embodiment, as for the voice interaction apparatus provided with the AMIC and the ADC, a difference between a sum of the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus plus the first microphone sensitivity and the second microphone sensitivity may be determined as the adjustment gain.

[0063] The adjustment gain is Gain*. The first microphone sensitivity is Smic 1 . The second microphone sensitivity is Smic 2 . For clearer illustration, the analog-to-digital conversion quantified reference voltage is assigned as Vref.

[0064] As for the voice interaction apparatus provided with the AMIC and the ADC, the adjustment gain Gain*, the first microphone sensitivity Smic 1 , the second microphone sensitivity Smic 2 and the analog-to-digital conversion quantified reference voltage Vref can satisfy the following relationship: Gain * = Vref + Smic 1 - Smic 2 .

[0065] The analog-to-digital conversion quantified reference Vref may be acquired based on specifications of the ADC. For example, the analog-to-digital conversion quantified reference voltage Vref of a model tlv320adc3101 from Texas Instruments (see "TLV320ADC3101 Low-Power Stereo ADC With Embedded miniDSP for Wireless Handsets and Portable Audio", http: / / www.ti.com / lit / ds / slas553b / slas553b.pdf) may be -3 dB.

[0066] The second microphone sensitivity Smic 2 of the voice interaction apparatus may be acquired according to specifications of the apparatus. For example, the second sensitivity Smic 2 of an AMIC of Goertek S18OB381-015 may be -38 dB.

[0067] For example, when the first microphone sensitivity Smic 1 is -26 dB, Gain*=-3-26-(-38)=+9 dB.

[0068] FIG. 5 is a schematic diagram of a pickup path of a voice interaction apparatus provided with an AMIC and an ADC.

[0069] As shown in FIG. 5, the adjustment gain Gain* of the voice interaction apparatus provided with the AMIC and the ADC is a hardware gain.

[0070] The adjustment gain Gain* may be acquired through adjustment of a gain module of a pre-amplifier Pre-AMP or a programmable gain amplifier PGA integrated with an ADC module ADC.

[0071] FIG. 6 is a flowchart of a method for controlling a recording volume according to an exemplary embodiment of the disclosure.

[0072] As shown in FIG. 6, in an exemplary embodiment of the disclosure, the method for controlling a recording volume includes operations S301-306.

[0073] In operation S301, the adjustment gain is determined according to a preset reference index of a reference apparatus. In operation S306, based on the adjustment gain, a recording volume of at least one voice interaction apparatus is calibrated, so that the at least one voice interaction apparatus has a same recording volume under a same sound pressure level. During determination of an adjustment gain according to a preset reference index of the reference apparatus, an index value of a used preset reference index is a theoretical index value. For example, the index value of the preset reference index may be acquired from a product instruction, and the adjustment gain is determined based on the theoretical index value. However, in actual production and design of a product, there may be a small error between an index value of a finished product in actual use and an index value in a product instruction. In order to make the acquired adjustment gain more accurate, the adjustment gain can be calibrated for an actual product after being determined based on the theoretical index value of the preset reference index. The adjustment gain obtained after calibration is taken as an adjustment gain finally used for calibrating the recording volume of the voice interaction apparatus in operation S306. See operations S302-S305 for operations of calibration.

[0074] Operation S301 and operation S306 are introduced above in detail, so no unnecessary details will be given here. Operations S302-S305 are introduced as follows.

[0075] In operation S302, the adjustment gain currently determined is used to calibrate the voice interaction apparatus.

[0076] In operation S303, an actual recording volume of the calibrated voice interaction apparatus is tested under a preset sound pressure level.

[0077] In an embodiment, the preset sound pressure level may be a sound pressure level of 94 dB. The preset sound pressure level may be adjusted according to actual situations. In the disclosure, the preset sound pressure level is not specifically defined.

[0078] Under the sound pressure level of 94 dB, the actual recording volume of the voice interaction apparatus which is already calibrated with the adjustment gain determined at present is S1 as tested.

[0079] In operation S304, a reference recording volume of the reference apparatus under the preset sound pressure level is determined.

[0080] In an embodiment, as determined, the reference recording volume of the voice interaction apparatus with a direct-sound microphone sound chamber, the sensitivity of x dB and a DMIC is S0 under the sound pressure level of y dB, where x and y are integers, x may be -26, and y may be 94. Then, as determined, the reference recording volume of the reference apparatus with the direct-sound microphone sound chamber, the sensitivity of -26 dB and the DMIC is S0 under the sound pressure level of 94 dB.

[0081] In operation S305, based on the reference recording volume and the actual recording volume, a calibration value of the adjustment gain determined at present is determined; and the calibrated adjustment gain is determined as the final adjustment gain for calibrating the voice interaction apparatus.

[0082] As for the voice interaction apparatus provided with the DMIC, the adjustment gain Gain* determined at present is Smic 1 - Smic 2 .

[0083] As for the voice interaction apparatus provided with the AMIC and the ADC, the adjustment gain Gain*determined at present is Vref + Smic 1 - Smic 2 .

[0084] Based on the reference recording volume S0 and the actual recording volume S1, the calibration value of the adjustment gain Gain* determined at present can be determined. The adjustment gain Gain* determined at present can be calibrated with the calibration value. The recording volume of the voice interaction apparatus may be calibrated based on the calibrated adjustment gain Gain*. Thus, the accuracy for the plurality of voice interaction apparatuses to generate the same recording volume after picking up a user instruction under the same sound pressure level is further increased, even the difference in the recording volumes generated by the plurality of voice interaction apparatuses is zero or very small.

[0085] FIG. 7 is a flowchart of operations of determining a calibration value of a present adjustment gain in a method for controlling a recording volume according to an exemplary embodiment of the disclosure.

[0086] As shown in FIG. 7, in an exemplary embodiment of the disclosure, the operation of determining the calibration value of the adjustment gain determined at present based on the a reference recording volume and an actual recording volume include operation S401 and operation S402. The operations S401 and S402 are introduced as follows.

[0087] In operation S401, a difference between the reference recording volume and the actual recording volume is determined.

[0088] In operation S402, the difference is determined as the calibration value of the adjustment gain determined at present.

[0089] The calibration value of the adjustment gain Gain* determined at present is (S0-S1).

[0090] For clearer illustration, as assigned, the calibrated adjustment gain is Gain*'. The calibrated adjustment gain Gain*', the adjustment gain Gain* determined at present and the calibration value (S0-S1) can satisfy the following relationship: Gain * ' = Gain * + S 0 - S 1 .

[0091] For example, S0 and S1 may be represented by peak levels of recording signals at the frequent point of 1 kHz.

[0092] After being processed with the method for controlling a recording volume in the embodiments of the disclosure, the voice volumes picked up by each of the voice interaction apparatuses under a user instruction with a same sound pressure level can be the same.

[0093] A distance between a user and a voice interaction apparatus is in positive correlation with a sound pressure level of a user instruction picked up from the user. Furthermore, the distance between the user and the voice interaction apparatus is in positive correlation with a voice volume picked up by the voice interaction apparatus.

[0094] Thus, when distances between a user and voice interaction apparatuses are different, the voice interaction apparatus nearest to the user may pick up a user voice instruction with a maximum sound pressure level in comparison with other voice interaction apparatuses. Based on the nearby wakeup mechanism, the voice interaction apparatus picking up the user voice instruction with the maximum sound pressure level may be determined as the voice interaction apparatus needing to respond to the user instruction.

[0095] FIG. 8 is a schematic diagram of a nearby wakeup mechanism based on energy or a signal-to-noise ratio.

[0096] As shown in FIG. 8, in an embodiment, the voice interaction apparatuses may include a speaker A, a television B, a speaker C, an air conditioner D and a voice switch E.

[0097] A center node may be a voice interaction apparatus or may be a network center apparatus such as a gateway or a router, serving as a judgment apparatus.

[0098] After each of the voice interaction apparatuses uploads a picked voice volume of a user instruction to the center node (judgment apparatus), the center node may judge which voice interaction apparatus picks up the user instruction with a maximum sound pressure level (or voice volume) and determine the voice interaction apparatus as the voice interaction apparatus needing to respond to the user instruction and determine other apparatuses as apparatuses for continuous dormancy.

[0099] Furthermore, the center node may feed back a command instruction to the corresponding voice interaction apparatus, to enable the voice interaction apparatus picking up the used instruction with the maximum sound pressure level (or voice volume) to respond to the user instruction, and make other voice interaction apparatuses continue dormancy.

[0100] Based on the same ideas, the embodiments of the disclosure also provide a device for controlling a recording volume.

[0101] It is understandable that, in order to realize above functions, the device for controlling a recording volume provided by the embodiments of the disclosure includes a corresponding hardware structure and / or a software module configured to execute each of the functions. In combination with units and algorithmic steps in the examples in the embodiments of the disclosure, the embodiments of the disclosure may be realized in a hardware form or a hardware-computer software combined form. Whether one of the functions is executed by hardware or by driving the hardware with computer software is decided by specific applications and design constraint conditions of the technical solution. Those skilled in the art may use different methods to realize the described functions aiming at each specific application, but such implementation is not deemed to exceed the scope of the technical solution in the embodiments of the disclosure.

[0102] FIG. 9 is a block diagram of a device for controlling a recording volume according to an exemplary embodiment of the disclosure.

[0103] As shown in FIG. 9, in an exemplary embodiment of the disclosure, the device for controlling a recording volume includes a determination module 201 and a calibration module 202. The determination module 201 and the calibration module 202 are introduced as follows.

[0104] The determination module 201 is configured to determine an adjustment gain according to a preset reference index of a reference apparatus.

[0105] The calibration module 202 is configured to, based on the adjustment gain, calibrate a recording volume of at least one voice interaction apparatus to make the at least one voice interaction apparatus have a same recording volume under a same sound pressure level.

[0106] In an exemplary embodiment of the disclosure, the determination module 201 is configured to determine first microphone sensitivity of the reference apparatus, the preset reference index including microphone sensitivity; and determine the adjustment gain based on the first microphone sensitivity and second microphone sensitivity of the voice interaction apparatus.

[0107] In an example not being part of the present invention, the voice interaction apparatus may include a DMIC. The determination module 201 is configured to determine a difference between the first microphone sensitivity and the second microphone sensitivity as the adjustment gain.

[0108] In an exemplary embodiment of the disclosure, the voice interaction apparatus may include a voice interaction apparatus provided with an AMIC and an ADC. The determination module 201 is configured to determine the adjustment gain based on the first microphone sensitivity, the second microphone sensitivity of the voice interaction apparatus and an analog-to-digital conversion quantified reference voltage of the voice interaction apparatus.

[0109] In an exemplary embodiment of the disclosure, the determination module 201 is configured to determine a difference between a sum of the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus plus the first microphone sensitivity and the second microphone sensitivity as the adjustment gain.

[0110] In an exemplary embodiment of the disclosure, the determination module 201 is configured to: calibrate the voice interaction apparatus with the adjustment gain determined at present; test an actual recording volume of the calibrated voice interaction apparatus under a preset sound pressure level; determine a reference recording volume of the reference apparatus under the preset sound pressure level; and determine a calibration value of the adjustment gain determined at present based on the reference recording volume and the actual recording volume, and determine the adjustment gain obtained after calibration as a final adjustment gain for calibrating the voice interaction apparatus.

[0111] In an exemplary embodiment of the disclosure, the determination module 201 is configured to determine a difference between the reference recording volume and the actual recording volume, and determine the difference as the calibration value of the adjustment gain determined at present.

[0112] In an exemplary embodiment of the disclosure, the reference apparatus may be a voice interaction apparatus provided with a DMIC; and / or the reference apparatus may be a voice interaction apparatus provided with a direct-sound microphone sound chamber.

[0113] With respect to the device in the above embodiments, the specific manners for performing operations for individual modules therein have been described in detail in the embodiments regarding the method for controlling a recording volume, which will not be elaborated herein.

[0114] FIG. 10 is a block diagram of a device for controlling a recording volume according to an exemplary embodiment of the disclosure. For example, the device for controlling a recording volume may be a mobile phone, a computer, a digital broadcast terminal, a messaging device, a gaming console, a tablet, a medical device, exercise equipment, a personal digital assistant, and the like.

[0115] Referring to FIG. 10, the device for controlling a recording volume may include one or more of the following components: a processing component 302, a memory 304, a power component 306, a multimedia component 308, an audio component 310, an input / output (I / O) interface 312, a sensor component 314, and a communication component 316.

[0116] The processing component 302 typically controls overall operations of the device, such as the operations associated with display, telephone calls, data communications, camera operations, and recording operations. The processing component 302 may include one or more processors 320 to execute instructions to perform all or part of the operations in the above described methods. Moreover, the processing component 302 may include one or more modules which facilitate the interaction between the processing component 302 and other components. For instance, the processing component 302 may include a multimedia module to facilitate the interaction between the multimedia component 308 and the processing component 302.

[0117] The memory 304 is configured to store various types of data to support the operation of the device. Examples of such data include instructions for any applications or methods operated on the device, contact data, phonebook data, messages, pictures, video, etc. The memory 304 may be implemented using any type of volatile or non-volatile memory devices, or a combination thereof, such as a static random access memory (SRAM), an electrically erasable programmable read-only memory (EEPROM), an erasable programmable read-only memory (EPROM), a programmable read-only memory (PROM), a read-only memory (ROM), a magnetic memory, a flash memory, a magnetic or optical disk.

[0118] The power component 306 provides power to various components of the device. The power component 306 may include a power management system, one or more power sources, and any other components associated with the generation, management, and distribution of power in the device.

[0119] The multimedia component 308 includes a screen providing an output interface between the device and the user. In some embodiments, the screen may include a liquid crystal display (LCD) and a touch panel (TP). If the screen includes the touch panel, the screen may be implemented as a touch screen to receive input signals from the user. The touch panel includes one or more touch sensors to sense touches, swipes, and gestures on the touch panel. The touch sensors may not only sense a boundary of a touch or swipe action, but also sense a period of time and a pressure associated with the touch or swipe action. In some embodiments, the multimedia component 308 includes a front camera and / or a rear camera. The front camera and the rear camera may receive an external multimedia datum while the device is in an operation mode, such as a photographing mode or a video mode. Each of the front camera and the rear camera may be a fixed optical lens system or have focus and optical zoom capability.

[0120] The audio component 310 is configured to output and / or input audio signals. For example, the audio component 310 includes a microphone ("MIC") configured to receive an external audio signal when the device is in an operation mode, such as a call mode, a recording mode, and a voice recognition mode. The received audio signal may be further stored in the memory 304 or transmitted via the communication component 316. In some embodiments, the audio component 310 further includes a speaker to output audio signals.

[0121] The I / O interface 312 provides an interface between the processing component 302 and peripheral interface modules, such as a keyboard, a click wheel, buttons, and the like. The buttons may include, but are not limited to, a home button, a volume button, a starting button, and a locking button.

[0122] The sensor component 314 includes one or more sensors to provide status assessments of various aspects of the device. For instance, the sensor component 314 may detect an open / closed status of the device, relative positioning of components, e.g., the display and the keypad, of the device, a change in position of the device or a component of the device, a presence or absence of user contact with the device, an orientation or an acceleration / deceleration of the device, and a change in temperature of the device. The sensor component 314 may include a proximity sensor configured to detect the presence of nearby objects without any physical contact. The sensor component 314 may also include a light sensor, such as a CMOS or CCD image sensor, for use in imaging applications. In some embodiments, the sensor component 314 may also include an accelerometer sensor, a gyroscope sensor, a magnetic sensor, a pressure sensor, or a temperature sensor.

[0123] The communication component 316 is configured to facilitate communication, wired or wirelessly, between the device and other devices. The device can access a wireless network based on a communication standard, such as WiFi, 2G, or 3G, or a combination thereof. In one exemplary embodiment, the communication component 316 receives a broadcast signal or broadcast associated information from an external broadcast management system via a broadcast channel. In one exemplary embodiment, the communication component 316 further includes a near field communication (NFC) module to facilitate short-range communications. For example, the NFC module may be implemented based on a radio frequency identification (RFID) technology, an infrared data association (IrDA) technology, an ultra-wideband (UWB) technology, a Bluetooth (BT) technology, and other technologies.

[0124] In exemplary embodiments, the device for controlling a recording volume may be implemented with one or more application specific integrated circuits (ASICs), digital signal processors (DSPs), digital signal processing devices (DSPDs), programmable logic devices (PLDs), field programmable gate arrays (FPGAs), controllers, micro-controllers, microprocessors, or other electronic components, for performing the above described methods.

[0125] In exemplary embodiments, there is also provided a non-transitory computer-readable storage medium including instructions, such as included in the memory 304, executable by the processor 320 in the device, for performing the above-described methods. For example, the non-transitory computer-readable storage medium may be a ROM, a CD-ROM, a magnetic tape, a floppy disc, an optical data storage device, and the like.

[0126] It is further understandable that, "a plurality of" in the disclosure refers to two or more than two. The principle is also applied to other measure words. "An / or" describes a correlative relationship between correlated objects, involving three relations. For example, A and / or B may denote three situations: sole existence of A, coexistence of A and B and sole existence of B. The character " / " generally denotes former and later correlated objects are correlated by a relationship of "or". In a singular form, "a", "the said" and "the" refer to existence of a plurality of forms, unless other connotations are specified clearly in the context.

[0127] It is further understandable that, operations are described in specific sequences as shown in diagrams of the embodiments of the disclosure, but it does not mean that the operations must be executed according to the displayed sequences or serial sequences, or all the displayed operations need to be executed for realization of an expected result. Under specific conditions, a plurality of tasks and concurrent processing may be beneficial.

[0128] Other embodiments of the disclosure will be apparent to those skilled in the art from consideration of the specification and practice of the disclosure disclosed here. This application is intended to cover any variations, uses, or adaptations of the disclosure following the general principles thereof and including such departures from the present disclosure as come within known or customary practice in the art. It is intended that the specification and examples be considered as exemplary only, with the scope of the invention being indicated by the following claims.

Claims

1. A method for controlling a recording volume, comprising: determining (S101) an adjustment gain according to a preset reference index of a reference apparatus; and calibrating (S102) a recording volume of at least one voice interaction apparatus based on the adjustment gain to make the at least one voice interaction apparatus have a same recording volume under a same sound pressure level, wherein the preset reference index comprises microphone sensitivity; and the voice interaction apparatus comprises a voice interaction apparatus provided with an analog microphone and an analog-to-digital converter, wherein determining the adjustment gain according to the preset reference index of the reference apparatus comprises: determining first microphone sensitivity of the reference apparatus; and determining the adjustment gain based on the first microphone sensitivity, the second microphone sensitivity of the voice interaction apparatus and an analog-to-digital conversion quantified reference voltage of the voice interaction apparatus.

2. The method of claim 1, wherein determining the adjustment gain based on the first microphone sensitivity, the second microphone sensitivity of the voice interaction apparatus and the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus comprises: determining a difference between a sum of the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus plus the first microphone sensitivity and the second microphone sensitivity as the adjustment gain, wherein the analog-to-digital conversion quantified reference voltage is a parameter for the analog-to-digital converter.

3. The method of any one of claims 1-2, after determining the adjustment gain, the method further comprising: calibrating (S302) the voice interaction apparatus with the adjustment gain determined at present; testing (S303) an actual recording volume of the calibrated voice interaction apparatus under a preset sound pressure level; determining (S304) a reference recording volume of the reference apparatus under the preset sound pressure level; determining (S305) a calibration value of the adjustment gain determined at present based on the reference recording volume and the actual recording volume; and determining an adjustment gain obtained after calibration as a final adjustment gain for calibrating the voice interaction apparatus.

4. The method of claim 3, wherein determining the calibration value of the adjustment gain determined at present based on the reference recording volume and the actual recording volume comprises: determining (S401) a difference between the reference recording volume and the actual recording volume; and determining (S402) the difference as the calibration value of the adjustment gain determined at present.

5. The method of claim 1, wherein the reference apparatus is a voice interaction apparatus provided with a digital microphone; and / or the reference apparatus is a voice interaction apparatus provided with a direct-sound microphone sound chamber.

6. A device for controlling a recording volume, comprising: a memory (304), configured to store instructions; and a processor (320), configured to call the instructions to implement operations of: determining an adjustment gain according to a preset reference index of a reference apparatus; and calibrating a recording volume of at least one voice interaction apparatus based on the adjustment gain to make the at least one voice interaction apparatus have a same recording volume under a same sound pressure level, wherein the preset reference index comprises microphone sensitivity; and the voice interaction apparatus comprises a voice interaction apparatus provided with an analog microphone and an analog-to-digital converter, wherein determining the adjustment gain according to the preset reference index of the reference apparatus comprises: determining first microphone sensitivity of the reference apparatus; and determining the adjustment gain based on the first microphone sensitivity, the second microphone sensitivity of the voice interaction apparatus and an analog-to-digital conversion quantified reference voltage of the voice interaction apparatus.

7. The device of claim 6, wherein the processor is further configured to determine a difference between a sum of the analog-to-digital conversion quantified reference voltage of the voice interaction apparatus plus the first microphone sensitivity and the second microphone sensitivity as the adjustment gain, wherein the analog-to-digital conversion quantified reference voltage is a parameter for the analog-to-digital converter.

8. The device of any one of claims 6-7, wherein the processor is further configured to: calibrate the voice interaction apparatus with the adjustment gain determined at present; test an actual recording volume of the calibrated voice interaction apparatus under a preset sound pressure level; determine a reference recording volume of the reference apparatus under the preset sound pressure level; determine a calibration value of the adjustment gain determined at present based on the reference recording volume and the actual recording volume; and determine the adjustment gain obtained after calibration as a final adjustment gain for calibrating the voice interaction apparatus.

9. A non-transitory computer readable storage medium, having stored computer executable instructions thereon that, when executed by a processor, implement the method for controlling a recording volume of any one of claims 1-5.