Method for converting lip image sequence into voice coding parameters
A speech coding and image sequence technology, applied in speech analysis, speech synthesis, computer components, etc., can solve the problem of complex conversion process and achieve the effect of facilitating construction training
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment 1
[0050] The following is a specific implementation method, but the methods and principles described in the present invention are not limited to the specific numbers given therein.
[0051] (1) Predictor, which can be realized by artificial neural network. Predictors can also be built using other machine learning techniques. In the following process, the predictor uses a deep artificial neural network, that is, the predictor is equivalent to a deep artificial neural network;
[0052] like image 3 As shown, the artificial neural network is mainly composed of 3 convolutional LSTM network layers (ConvLSTM2D) and 2 fully connected layers (Dense) connected in turn, as shown in the following figure. Each ConvLSTM2D is followed by a pooling layer (MaxPooling2D), and the two Dense layers are preceded by a dropout layer (Dropout), for the clarity of the structure, these are in image 3 not drawn in.
[0053] Among them, each of the three layers of convolutional LSTM has 80 neurons, ...
PUM
Abstract
Description
Claims
Application Information
- R&D Engineer
- R&D Manager
- IP Professional
- Industry Leading Data Capabilities
- Powerful AI technology
- Patent DNA Extraction
Browse by: Latest US Patents, China's latest patents, Technical Efficacy Thesaurus, Application Domain, Technology Topic, Popular Technical Reports.
© 2024 PatSnap. All rights reserved.Legal|Privacy policy|Modern Slavery Act Transparency Statement|Sitemap|About US| Contact US: help@patsnap.com