Embodiments of the invention provide an MPP cluster
exception handling method and apparatus, and a storage medium. The method comprises the steps of obtaining a to-be-executed job; when the to-be-executed jobs are dominant faults, whether the queuing time of the to-be-executed jobs exceeds a first threshold value or not is judged, and if the queuing time of the to-be-executed jobs exceeds the first threshold value, the jobs are submitted and removed out of a to-be-distributed
job queue; if the first threshold value is not exceeded, the to-be-executed job is queued continuously; when the to-be-executed job is a hidden fault, whether the queuing time of the to-be-executed job exceeds a second threshold value or not is judged, and if the queuing time of the to-be-executed job exceeds the second threshold value, the to-be-executed job is submitted and removed out of the queuing
queue; if the number does not exceed a second threshold value, judging the number of job instances which are associated with the
metadata cluster in the to-be-executed job and are in operation, and if the number of the job instances does not exceed a third threshold value, submitting the to-be-executed job and removing the to-be-distributed
job queue; and if so, continuing to
queue the to-be-executed jobs. According to the method, invalid scheduling submission is reduced, and the
automation level is improved.