The invention provides a smart glasses AI voice
interaction method, which comprises the following steps: simultaneously acquiring a user gesture image and a voice
signal, acquiring a user historical interaction
record, extracting key point coordinates and pointing
direction information of the gesture image, acquiring hand distance information, generating a gesture motion path, simultaneously extracting
frequency spectrum information and an intonation
peak value of the voice
signal, and generating a gesture motion path; forming a voice beat sequence; extracting a spatial semantic mode of the gesture and voice fusion data, and recognizing a
core object of a user pointing instruction according to pointing coordinates of a gesture motion path and an intonation
peak value of a voice beat sequence; after the
core object pointing to the instruction is recognized, the user intention is determined in combination with the spatial semantic mode and the historical interaction
record of the user; and the complete intention analysis result is output to the intelligent glasses display module to execute corresponding operation, and is fed back to the acquisition module to adjust the next capture parameters including the acquisition frequency and the recognition sensitivity, so that the response speed and the accuracy are improved.