A Transformation Method from Lip Image Features to Speech Coding Parameters
A technology of speech coding and image features, which is applied in speech analysis, speech synthesis, neural learning methods, etc., can solve the problem of complex conversion process and achieve the effect of facilitating construction training
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment 1
[0048] The following is a specific implementation method, but the method and principle of the present invention are not limited to the specific numbers given therein.
[0049] (1) The predictor can be implemented using artificial neural networks. Other machine learning techniques can also be used to construct the predictor. In the following process, the predictor uses an artificial neural network, that is, the predictor is equivalent to an artificial neural network.
[0050] In this embodiment, the neural network is composed of 3 LSTM layers + 2 fully connected layers Dense connected in sequence. A Dropout layer is added between every two layers and between the internal feedback layer of LSTM. For clarity of the architecture, these are not shown in the figure. Such as image 3 Shown:
[0051] Among them, the three layers of LSTM each have 80 neurons, and the first two layers use the "return_sequences" mode. The two dense layers have 100 neurons and 14 neurons respectively.
[005...
PUM
Abstract
Description
Claims
Application Information
- R&D Engineer
- R&D Manager
- IP Professional
- Industry Leading Data Capabilities
- Powerful AI technology
- Patent DNA Extraction
Browse by: Latest US Patents, China's latest patents, Technical Efficacy Thesaurus, Application Domain, Technology Topic, Popular Technical Reports.
© 2024 PatSnap. All rights reserved.Legal|Privacy policy|Modern Slavery Act Transparency Statement|Sitemap|About US| Contact US: help@patsnap.com