A video
signal processing method comprises: obtaining a first
syntax element indicating whether combined prediction is applied to the current block, wherein the combined prediction is a prediction mode that combines inter-prediction and intra-prediction; and if the first
syntax element indicates that the combined prediction is applied to the current block, reconstructing the current block based on a combined prediction block, characterised in that the method further comprises: if the first
syntax element indicates that the combined prediction is not applied to the current block, obtaining a second syntax element indicating whether sub-
block transform is applied to the current block, wherein the sub-
block transform indicates a transform mode that applies transform to one of sub-blocks of the current block divided in a horizontal direction or in a vertical direction; and if the second syntax element indicates that the sub-
block transform is applied to the current block, reconstructing the current block based on the sub-block transform; wherein the combined prediction block is obtained by performing a weighted-sum of an inter-prediction block and an intra-prediction block, wherein the inter-prediction block is obtained by using the inter-prediction for the current block, and the intra-prediction block is obtained by using a planar mode for the current block, wherein a weight for the weighted-sum is determined based on a prediction mode of each neighbouring locations, and wherein the neighbouring locations are (xCb-1, yCb-1+cbHeight) and (xCb-1+cbWidth, yCb-1), wherein (xCb, yCb) is a coordinate of the top-left of the current block, the cbHeight is a height of the current block and the cbWidth is a width of the current block.