An end-side large model weight protection method and system based on decryption inference card private memory
By decrypting the private memory system of the inference card, the problems of high leakage risk, low efficiency, or poor versatility in edge-side large model weight protection are solved, achieving efficient and secure edge-side large model weight protection, applicable to various hardware architectures, and improving user experience.
Patent Information
- Application Number
- CN202610016010.7
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2026-01-07
- Publication Date
- 2026-06-05
AI Technical Summary
Existing mid-side large model weight protection schemes suffer from high leakage risk, low inference efficiency, or poor versatility, failing to simultaneously meet the requirements of secure and reliable operation, efficient inference, and broad applicability.
An edge-side large model weight protection system based on the private memory of the decryption inference card is adopted. Through the hardware isolation mechanism between the decryption inference card and the host, it is ensured that the model weight is decrypted and inferred only in the private memory of the decryption inference card, avoiding plaintext data leakage. Data interaction is carried out through the PCIe interface to reduce computational overhead.
It achieves efficient and secure protection of large model weights on the edge, reduces the risk of model weight leakage, improves inference efficiency, is applicable to various hardware architectures, requires no environment switching, and enhances the user experience.
Smart Images

Figure FT_1 
Figure FT_2 
Figure FT_3
Abstract
Citation Information
Patent Citations
End-side large model parameter protection method and system based on trusted execution environment
CN120744908A