The application relates to the technical field of intelligent generation, in particular to a unified autoregressive three-dimensional model automatic skeleton binding method and device, wherein the method comprises the following steps: converting an input three-dimensional model into
point cloud data; performing sampling and normalization
processing on the
point cloud data to extract local and global geometric features of the three-dimensional model, and determining feature embedding of the
point cloud data according to the local and global geometric features; based on the feature embedding and a preset autoregressive model, generating a skeleton tree token sequence and a discrete
skin token sequence, to generate a skeleton tree and
skin weights, and fusing the skeleton tree and the
skin weights to obtain an automatic three-dimensional grid binding result suitable for an
animation driving condition. Therefore, the problem that
animation effect presentation is affected due to the fact that related technologies cannot uniformly model the skeleton and the skin weights and cannot process sparse skin weights in a more stable manner is solved.