Embodiments of the present specification provide a task
processing model training method, a task
processing method and a
data labeling model training method. The task
processing model training method comprises: obtaining training data of a target training task; inputting the training data into a
data labeling model to obtain a training labeling result of the training data, wherein the
data labeling model is trained based on sample data, sample auxiliary data of the sample data and a sample labeling result, the sample auxiliary data is obtained through multi-round argumentation of multiple argumentation roles based on the sample data, and the argumentation roles have different argumentation
viewpoints; and training an initial processing model according to the training data and the training labeling result to obtain a task processing model. The labeling capability of the data labeling model and the robustness and generalization capability of the data labeling model when facing new data are improved through the multi-round argumentation process, so that the data labeling model can generate reliable training labeling results consistent with human intentions, and the accuracy of the task processing model is improved.