The application discloses a method and equipment for automatically generating a training
data set in a mowing
robot vision
system, and the method comprises the following steps: S01, acquiring a scene condition
control signal and generating a geometric constraint
signal; S02, constructing a structured prompt word and inputting the same into a text condition of a text
encoder to obtain text encoding; S03, inputting initial
noise, the geometric constraint
signal and the text encoding into a
diffusion generation model, and converting the initial
noise, the geometric constraint
signal and the text encoding into a preliminary image
data set through image decoding; S04, preliminarily screening the preliminary image
data set according to
semantic consistency and visual quality, performing instance segmentation on the preliminarily screened image data set to obtain an instance
mask, and finely screening the instance
mask according to a quality feature of the instance
mask to obtain a finally generated training image data set. The application can automatically generate a high-quality
visual training data set meeting training requirements without relying on real scene collection and manual labeling.