Training method and system for language model
A technology of language model and training method, which is applied in the field of language model training method and system, can solve the problems that the language model is easy to lose the original statistical distribution of big data, the language recognition rate is reduced, and the amount of computing resources is large, so as to improve the speech recognition rate , reduce the amount of computing resources, and the effect of reasonable parameters
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment 1
[0054] refer to figure 2 , which shows a flow chart of the steps of Embodiment 1 of a language model training method of the present invention, which may specifically include the following steps:
[0055] Step 201, obtaining seed corpus in various fields;
[0056] In the embodiment of the present invention, the domain may refer to the application scenarios of the data, such as news, place names, website addresses, people's names, map navigation, chatting, short messages, Q&A, Weibo, etc. are common domains. In practical applications, the corresponding seed corpus can be obtained through professional crawling and cooperation for specific fields. The cooperation can be with the website operator to obtain the corresponding seed corpus through the log files of the website, such as through The corresponding seed corpus is obtained from the log file of the microblog website, and the embodiment of the present invention does not limit the specific method for obtaining the seed corpus...
Embodiment 2
[0121] refer to image 3 , which shows a flow chart of the steps of Embodiment 2 of an information search method of the present invention, which may specifically include the following steps:
[0122] Step 301. Obtain seed corpus in each field, and train a seed model in a corresponding field according to the seed corpus in each field;
[0123] Step 302: Screen the big data corpus according to the vector space model of the seed corpus in each field, and obtain the seed screening corpus in the corresponding field;
[0124] Step 303, respectively use the seed screening corpus training in each field to obtain the screening model in the corresponding field;
[0125] Step 304, fusing the screening models of all domains to obtain a corresponding screening fusion model.
[0126] Step 305, fuse the seed models of all domains to obtain the corresponding seed fusion model;
[0127] Step 306: Fusion the screening fusion module and the seed fusion model to obtain a corresponding general ...
PUM
Abstract
Description
Claims
Application Information
- R&D Engineer
- R&D Manager
- IP Professional
- Industry Leading Data Capabilities
- Powerful AI technology
- Patent DNA Extraction
Browse by: Latest US Patents, China's latest patents, Technical Efficacy Thesaurus, Application Domain, Technology Topic, Popular Technical Reports.
© 2024 PatSnap. All rights reserved.Legal|Privacy policy|Modern Slavery Act Transparency Statement|Sitemap|About US| Contact US: help@patsnap.com