Visual question and answer method based on multi-modal decomposition model
A multi-modal, model technology, applied in character and pattern recognition, instruments, computer parts, etc., can solve the problems of low performance and low accuracy, and achieve the effect of improving performance
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment Construction
[0035] specific implementation plan
[0036] It should be noted that, if there is no conflict, the experimental examples and the features in the experiments herein can be combined with each other. The present invention will be further described in detail below in conjunction with the drawings and specific embodiments.
[0037] figure 1 It is a flow chart of a system for performing visual question answering based on a multimodal decomposition model in the present invention. It mainly includes multi-modal decomposition bilinear pooling (MFB), multi-modal decomposition high-order pooling (MFH), and assisted attention model.
[0038] figure 2 It is an MFB flow chart for visual question answering based on a multimodal decomposition model in the present invention. Multimodal decomposition bilinear pooling (MFB), different modalities have two feature vectors, where the visual features of the image are Visual features of question text The formula for multimodal decomposition bil...
PUM
Abstract
Description
Claims
Application Information
- R&D Engineer
- R&D Manager
- IP Professional
- Industry Leading Data Capabilities
- Powerful AI technology
- Patent DNA Extraction
Browse by: Latest US Patents, China's latest patents, Technical Efficacy Thesaurus, Application Domain, Technology Topic, Popular Technical Reports.
© 2024 PatSnap. All rights reserved.Legal|Privacy policy|Modern Slavery Act Transparency Statement|Sitemap|About US| Contact US: help@patsnap.com