This invention discloses a distributed
large model token counting method based on trace linking, relating to the field of
large model management technology. The method involves collecting token consumption data from the previous day's data via a distributed cloud
large model server, uploading this data to Kafka, and receiving and
parsing the token consumption data from Kafka via a large
model management center. Data from different nodes is integrated based on the trace ID to form complete request chain data. Following a hierarchical and domain-based approach, the large
model management center stores user token consumption data in
object storage. Finally, based on business needs, the large
model management center displays relevant statistical data of message chains within a filtered region.