The disclosed method for generating virtual objects includes generating, based on object data, compressed object data, performing, based on the object data and scales, operations to
train a first untrained
machine learning model to generate a first trained
machine learning model comprising a trained
codebook and a trained decoder, wherein the first trained
machine learning model is trained to generate a reconstruction of the compressed object data, generating, based on the compressed object data and the scales and using the first trained
machine learning model, token maps data, performing, based on the token maps data and conditions, operations to
train a second untrained
machine learning model to generate a second trained
machine learning model comprising a trained autoregressive model, wherein the second trained machine learning model is trained to generate predicted token maps, and generating, based on the scales, conditions, and using both trained models, a virtual object.