Token marking method for large language model input data and cognitive system
By using a four-dimensional discrete spatiotemporal system as a token for large language models, the problems of confusion and logical jumps in multi-turn dialogues and IoT control are solved, enabling more accurate dialogue understanding and device control, and constructing a unified cognitive system.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- 黄宝明
- Filing Date
- 2026-04-06
- Publication Date
- 2026-07-03
AI Technical Summary
Existing large language models lack explicit context differentiation, logical progression indicators, physical world associations, and a unified tokenization framework when handling multi-turn dialogues and IoT control. This leads to model confusion, logical jumps, and inaccurate device control in complex tasks.
The token is a four-dimensional discrete spatiotemporal system used as a large language model. It includes coordinates of four dimensions: time (T), context (C), logic (L), and device (D). These coordinates explicitly identify the token's time, dialogue order, logical position, and device identity, thus constructing a structured cognitive system.
It improves the model's dialogue understanding ability, enhances the coherence of logical reasoning, supports time synchronization and device control in the Internet of Things scenario, is compatible with existing model architecture, and has low deployment cost and good scalability.
Smart Images

Figure FT_1 
Figure FT_2 
Figure FT_3