Embodiments of the present specification provide a GPU
resource management method and device, the method is applied to a cloud side node, comprising: receiving a
processing request sent by a user, analyzing the
processing request, determining an
object identifier and GPU resource demand information corresponding to the
processing request; in the case of determining that the container responds to the processing request according to the
object identifier, determining a target
edge node from the
edge node cluster based on the idle container GPU
resource information of each
edge node in the edge node group according to the GPU resource demand information, and in the case of determining that the
virtual machine responds to the processing request according to the
object identifier, determining a target edge node from the edge node
group based on the idle
virtual machine GPU
resource information of each edge node according to the GPU resource demand information, wherein the edge node group contains at least two edge nodes, and each edge node is configured with at least two GPUs; sending the object identifier and the GPU resource demand information to the target edge node.