Deep learning operator automatic optimization system and method based on Shenwei processor
A deep learning and automatic optimization technology, which is applied in neural learning methods, neural architectures, biological neural network models, etc., can solve problems such as difficult transplantation, high optimization time overhead, and low optimization performance
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment 1
[0034] Embodiment 1 of the present invention is based on Shenwei processor (SW26010) to automatically optimize convolution (including three calculation methods of im2col method, implicit convolution and Winograd method) and full connection two calculation-intensive operators. Such as image 3 , the optimization implementation of the operator is split into two parts, the tensor assembly primitive that makes full use of the hardware characteristics and the optimization scheduling that can be automatically tuned, so as to separate the hardware-related and hardware-independent optimization strategies. The tensor assembly primitive is used as the construction unit, combined with multiple loop scheduling, to complete the calculation tasks of the convolution operator and the fully connected operator.
[0035] Based on the above discussion, as figure 1 As shown, the Shenwei processor-based deep learning operator automatic optimization system provided by Embodiment 1 of the present in...
Embodiment 2
[0041] Such as figure 2 As shown, Embodiment 2 of the present invention provides an optimization method based on the Shenwei processor-based deep learning operator automatic optimization system provided in Embodiment 1, including:
[0042] S101: Acquiring a dedicated description language to define a computing task and a description of an optimization space;
[0043] S102: Construct the optimization space according to the description of the optimization space, generate several different calculation realizations for the description and scheduling of the calculation tasks according to different optimization methods in the optimization space, and output the calculation realization expressed by the intermediate representation;
[0044] S103: Perform optimization on the intermediate representation, and output the optimized intermediate representation;
[0045] S104: Search for an optimal calculation implementation from the optimized intermediate representation;
[0046] S105: Tra...
PUM
Abstract
Description
Claims
Application Information
- R&D Engineer
- R&D Manager
- IP Professional
- Industry Leading Data Capabilities
- Powerful AI technology
- Patent DNA Extraction
Browse by: Latest US Patents, China's latest patents, Technical Efficacy Thesaurus, Application Domain, Technology Topic, Popular Technical Reports.
© 2024 PatSnap. All rights reserved.Legal|Privacy policy|Modern Slavery Act Transparency Statement|Sitemap|About US| Contact US: help@patsnap.com