Long text information processing method and system based on large model, medium, product and terminal
By performing preset threshold segmentation and differential processing on long texts, compressed semantic vectors and text semantic vector representations are generated, solving the problems of computational power consumption and context integrity in long text processing of large language models, and realizing efficient and lossless long text information processing.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- SHANGHAI GUANGYU XINCHEN TECHNOLOGY CO LTD
- Filing Date
- 2026-01-20
- Publication Date
- 2026-06-02
AI Technical Summary
Existing large language models suffer from high computational costs and easy loss of contextual integrity when processing long text information. Furthermore, solutions that modify the model architecture or rely on external knowledge bases have high development costs, poor compatibility, or the risk of information loss.
By dividing a long text into a first sub-text segment and a second sub-text segment using a preset threshold, image conversion and encoding compression and lexicalization are applied respectively to generate compressed semantic vectors and text semantic vector representations, which are then input into a large language model through splicing and adaptation.
Without modifying the large language model architecture or relying on external knowledge bases, it achieves efficient processing of long texts while preserving the complete semantic logic, reducing computational complexity and resource consumption, and ensuring processing efficiency and accuracy.
Smart Images

Figure CN121562565B_ABST
Abstract
Citation Information
Patent Citations
Method and device for training large language model based on long text
CN119004107A
Fault-tolerant processing method and system for super-long text in large model service, and storage medium
CN121144496A