The invention discloses a petition
data set construction method based on a large
language model, and relates to the technical field of
natural language processing, and the method comprises the steps: collecting petition case information, and obtaining a standard petition text object through preprocessing; based on a large
language model, performing semantic structuring
processing on the standard petition text object to generate structured case features; according to the structured case features, a case
similarity relation graph is constructed through similarity
correlation analysis, comprehensive evaluation is carried out, and difficulty hierarchical case objects are obtained; selecting an alternative standard path through a multi-target
path search algorithm based on the difficulty hierarchical case object, and generating an anti-fact path candidate set; and according to the anti-fact path candidate set and the case
similarity relation graph, performing hierarchical sample arrangement on the difficulty hierarchical case objects, and constructing a petition
data set. According to the method, the alternative standard path is selected by executing the multi-target
path search algorithm, so that the decision reference capability and generalization robustness of a large
language model in a complex petition scene are improved.