A method for processing the interactive context of a language model, an electronic device, and a storage medium.

By monitoring and offloading infrequently accessed data segments, combined with compression and reassembly processing, the context bloat problem of LLM in long dialogue scenarios is solved, optimizing computational resource consumption and decision quality, and improving response efficiency and system robustness.

CN122088461APending Publication Date: 2026-05-26ZHEJIANG MEIRI HUDONG NETWORK TECH CO LTD +1
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
ZHEJIANG MEIRI HUDONG NETWORK TECH CO LTD
Filing Date
2026-04-15
Publication Date
2026-05-26

Smart Images

  • Figure CN122088461A_ABST
    Figure CN122088461A_ABST
Patent Text Reader

Abstract

This invention discloses a method, electronic device, and storage medium for processing the interactive context of a language model. The method includes: real-time monitoring of the length of the dialogue context and the access frequency of each data fragment in the session context cache; if there are data fragments with an access frequency below a threshold, they are migrated to external storage, and step transformation data containing semantic summaries, key findings, and file indexes is generated; if there are no such data fragments and the context length exceeds the limit, historical content is summarized and compressed based on preset anchor points, retaining recent content; then, the processed content is reorganized into a structured context with type labels according to a preset structure; finally, a context state snapshot associated with a version identifier is generated and stored. This method achieves adaptive optimization management of the context, significantly reduces computational load and cost, improves response speed and decision quality, and enhances the traceability of the state.
Need to check novelty before this filing date? Find Prior Art