The invention discloses a method for evaluating the
code generation capability of a fine-tuned credential
large model, and relates to the technical field of
artificial intelligence, and the method comprises the steps: S1, constructing an
evaluation data set which is constructed according to a half-and-half rule, S2, executing a model generation test, and based on the
evaluation data set, carrying out a model generation test; the method comprises the following steps of S1, adjusting a code output result, S2, calling the adjusted credential
large model to generate a corresponding code output result, S3, carrying out multi-dimensional evaluation analysis, S4, forming a comprehensive
evaluation result, and calculating the credibility 7 degree of the
large model code generation capability for guiding model performance optimization and subsequent iteration improvement; according to the fine-tuned evaluation method for the
code generation capability of the credential large model, evaluation deviation caused by a single index is effectively reduced, and the obtained quality credibility is more practical and explanatory and can be used as a decision basis for supporting large model optimization strategy formulation and parameter
fine tuning.