A noise suppression system selects a gain value based on sub-band and time-frame features to optimize automatic speech recognition.
A dialog system refines proper name lists using dynamic contextual weights and user models to improve recognition accuracy.
A voice processing apparatus clusters unknown words by phoneme similarity to update the recognition dictionary automatically.
Electronic apparatus assigns priority levels to devices so only the highest priority unit executes voice commands, preventing simultaneous activation confusion.
A voice input system segments recognition using parameter-based sifting conditions to refine candidate lists.
A low-power event generator analyzes audio streams to activate a digital signal processor only upon keyword detection, reducing mobile device power consumption.
Dynamic grammar generation narrows the search scope to resolve the contradiction between large vocabulary coverage and reliable voice recognition accuracy.
A speech recognition correction system converts results to pinyin for candidate text generation.
A hybrid speech recognition system combines embedded and network-based engines to process audio signals.
A customizable speech recognition neural network adapts to target domains using frozen layers and aligned attention weights.
An automatic evaluation system extracts nine rhythm and fluency features from speech signals to generate consistent pronunciation proficiency scores.
An analysis circuit modifies voice detection thresholds based on false positive rates, reducing unnecessary activation of the speech recognition module.
An adaptive voice activity detection circuit selects optimal algorithms based on detected environmental noise characteristics.
An electronic device transmits user utterances to an external server for processing and displays sample utterances for selection.
A digital assistant initiates natural language processing and task flow evaluation during speech end-point detection to reduce response delays.
Flexible substrate layer interposed between fabric and housing diffuses light from embedded LEDs for clear visual feedback.
A system corrects semantic unit sets by matching improvement phonetic sounds against reference templates to replace inaccurate units.
Random confirmation mechanism samples uncertain speech results to assess system performance and reduce call routing errors.
A server processes unrecognized voice inputs to map user intent to actions, reducing cognitive load from predefined command constraints.
System filters available actions based on identified spoken entities to resolve interface complexity and usability issues.
A speech signal processing method combines noise reduction filtering with spectral envelope reconstruction to enhance audio output.
An enhanced MMSE determiner warps speech presence probability using a sigmoid function responsive to external signal-to-noise ratio.
A haptic output device translates audible speech into tactile vibrations for hearing impaired users.
A virtual assistant system builds dynamic speaker profiles by continuously monitoring audio inputs and updating voice models in real time.
Replacing adaptive codebooks with a glottal-shape codebook reduces error propagation during frame erasures while maintaining coding efficiency.
Server matches voice commands to target devices using pre-registered profiles, eliminating manual device naming requirements.
A transcription correction system matches extracted features against domain-specific indexes to replace mistranscribed entries.
A voice device generates sound profiles for incoming audio and compares them against known media patterns to stop processing commands from media broadcasts.
Iterative recognition extracts multiple semantic items from one utterance using prior constraints.
A multimodal browser tracks hyperlink visibility to resolve ambiguous speech recognition inputs.
A voice processing system segments users by analyzing interaction patterns to customize command responses.