Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

65 results about "Video Media" patented technology

AI automatic editing method and related equipment

The invention discloses an AI automatic editing method and related equipment, and the method comprises the steps: obtaining a video prototype template selected by a user, prompt information inputted by the user, and original materials uploaded by the user through a user interaction interface, and calling a language large model to generate a structured video script; performing association degree matching on the structured script and the original material through a target detection model, and screening the original material according to a matching result and a quality scoring model to obtain an original video clip; performing background music matching according to the video script to obtain background music, generating voice dubbing and a flower-character chartlet according to the video script, generating video subtitles according to the voice dubbing, performing time axis alignment on each video media element, and synthesizing into a finished video. According to the embodiment of the invention, full-process automatic processing from script generation to video synthesis can be realized, and a high-quality finished video meeting user requirements can be efficiently output. The method can be widely applied to the technical field of automatic editing.
Owner:GUANGZHOU FAISCO INFORMATON TECH

AI engine docking decision-making method based on 5G new call

The invention discloses an AI engine docking decision-making method based on a 5G new call, and the method comprises the steps: carrying out the call service, and analyzing the call information of a user; the AI engine provides an intelligent interaction analysis capability; the media docking module is used for receiving a function instruction of calling service logic, providing basic media capability and making and displaying media content according to user operation; the BM module is used for storing user services and setting data; wherein the call service synchronizes a call instruction, audio and video media stream information and user behavior data to the media docking module, and receives a media instruction. According to the method, real-time multi-mode interactive perception and dynamic AI engine docking can be realized, so that the bottleneck in the prior art is broken through.
Owner:CHINA UNICOM WO MUSIC & CULTURE CO LTD

Video decoder initialization information signaling

A mechanism for processing video data is disclosed. A parent block is partitioned with an Extended Ternary-Tree (ETT) partition to create three sub-blocks. At least one of the sub-blocks includes a side with a measurement that is not a power of two. A conversion is performed between a visual media data and a bitstream based on the sub-blocks.
Owner:DOUYIN VISION CO LTD +1

Video recommendation method fusing social information

A video recommendation method fusing social information comprises the steps that firstly, a user-video interaction graph and a social graph are constructed through video media platform data, different combination strategies of a diffusion model and graph convolution operation are applied in a denoising priority path and a structure priority path respectively, and collaborative perception and fusion of user embedding and structure information are achieved; secondly, aligning user embedding output by the double-track denoising path by adopting comparative learning, and enhancing robustness; secondly, performing deep coding on the user-video interaction graph and the social graph by applying a double-graph neural network, comprehensively capturing user preferences and embedding videos; and finally, for noise possibly introduced by multi-module fusion, deep denoising is executed to purify user embedding again. According to the method, guidance of structural information is introduced in the diffusion denoising process, and more robust user embedding is obtained by using a comparative learning strategy, so that a structural perception high-fidelity social denoising task is completed, and personalized video recommendation is realized on a video media platform.
Owner:ZHEJIANG UNIV OF TECH

Generating large scale rewarded video advertisements with high impression and relevance

In one embodiment, a method includes receiving an advertisement exchange requesting a rewarded video advertisement from an external server, tokenizing the advertisement exchange based on coding schemes, unwrapping wrappers associated with the advertisement exchange to fetch an original video media file, generating a re-encoded video media file based on re-encoding the original video media file, wherein generating the re-encoded video media file comprises integrating visual elements into the original video media file, wherein the visual elements are configured for enabling the rewarded video advertisement, generating the rewarded video advertisement based on replacing the original video media file with the re-encoded video media file, and sending the rewarded video advertisement to the external server.
Owner:CONSUMABLE INC

Video media platform recommendation method based on distributed robust graph neural network

The invention discloses a video media platform recommendation method based on a distributed robust graph neural network, and the method comprises the steps: constructing a user-video interaction graph based on behavior data, including like, collection and comment, of a user watching a video in a video platform, and carrying out the modeling of the interaction graph through a variational graph auto-encoder, obtaining a low-dimensional embedding vector of the user and the video; thirdly, modeling is carried out on the potential environmental factors; and then, gradually disturbing and reconstructing an embedded vector in combination with a diffusion model, learning a structural relationship embedded in a noise evolution process through a graph neural network, and introducing a distribution robust optimization mechanism to realize personalized video recommendation. According to the method, through fusion graph representation learning, environment modeling and diffusion denoising mechanisms, the influence of data noise and distribution offset on recommendation performance is effectively relieved, and the robustness and recommendation accuracy of a video recommendation system in a complex dynamic environment are improved.
Owner:ZHEJIANG UNIV OF TECH

Neural network-based post filter for video coding

A method of processing video data. The method includes determining that a supplemental enhancement information (SEI) message of a bitstream includes indicators specifying one or more neural network (NN) filter model candidates or selections for a video unit or samples within the video unit, and converting between a video media file comprising the video unit and the bitstream based on the indicators. A corresponding video coding apparatus and non-transitory computer readable medium are also disclosed.
Owner:BYTEDANCE INC

A screen projection control method of a video medium and a related device

This application discloses a method and related apparatus for controlling the projection of video media, relating to the field of software technology. It acquires the vehicle's gear position and the video playback configuration under that gear position, including playback mode and screen configuration. When the gear position changes from P to non-P, the playback mode is video mode, and the screen configuration is multi-screen synchronous playback, a first screen for multi-screen synchronous playback is determined. If the central control screen is the initiating screen and other screens are receiving screens in the first screen, the central control screen is controlled to exit video mode, and a new initiating screen is determined among the other screens in the first screen (excluding the central control screen) to maintain video mode on the other screens. This application effectively solves the experience problem caused by changes in playback mode due to changes in vehicle gear position by shifting the initiating screen. The user's viewing behavior in the car is not disturbed by gear changes, improving the continuity and integrity of the viewing experience.
Owner:BEIJING CO WHEELS TECH CO LTD

Using scalability dimension information

A method of processing video data includes using a scalability dimension information (SDI) supplemental enhancement information (SEI) message to indicate which primary layers are associated with an auxiliary layer when auxiliary information is present in a bitstream, and converting between a video media file and the bitstream based on the SDI SEI message. A corresponding video coding apparatus and non-transitory computer readable medium are also disclosed.
Owner:DOUYIN VISION CO LTD +1

On Processing Video Of Different Colour Formats For In-loop Filtering In Video Coding

PendingUS20260255000A1Pattern recognitionLoop filter
A mechanism for processing video data is disclosed. The mechanism determines to pad at least one unavailable sample of a video before feeding the video into a process. A conversion is performed between a visual media data and a bitstream based on the padded samples.
Owner:BYTEDANCE INC

Summary generation based on trip

Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing a program and a method for generating a summary based on trip information. The program and method include operations for: determining that one or more criteria associated with a user correspond to a trip taken by the user during a given time interval; retrieving a plurality of visual media items generated by a client device of the user during the given time interval; determining location information for the plurality of visual media items; automatically generating a trip graphic to represent the trip based on the plurality of visual media items generated by the user during the given time interval and the determined location information; and causing the trip graphic to be displayed on the client device.
Owner:SNAP INC

Cartoon video medium-based poster batch generation method and system

The invention discloses a poster batch generation method and system based on a cartoon video medium, and the method comprises the steps: obtaining the configuration parameters of a one-card video, and the configuration parameters comprise role information and text information; performing image conversion processing on the cartoon video to obtain an image set; identifying the role information of each image in the image set by using a preset model to obtain a role image set; acquiring a poster image set of the cartoon video by utilizing a role position in the role image set and an adding position of the text information; and outputting the poster image set. According to the poster batch generation method and system based on the cartoon video medium, the poster pictures of the cartoon video can be automatically generated, the forms of the poster pictures are various, the production efficiency is greatly improved, the method and system can automatically adapt to the identification process of animation images, and the success rate of the poster is improved.
Owner:SHANGHAI SENYU MEDIA HLDG CO LTD

System and method for video / audio comprehension and automated clipping

Systems and Methods for Video / Audio Comprehension and Automated Clipping includes providing at least one media clip (MC) within an event for display or listening on a user device including receiving audio or video media data indicative of the event, transcribing the media data into timestamped text, identifying entities within the text, creating text segments having a begin timestamp and end timestamp and having a minimum number of entity mentions in the text segments, clipping from the media data the at least one media clip having a begin timestamp and end timestamp corresponding to the begin timestamp and end timestamp of a corresponding one of the text segments, and providing the at least one media clip to the user device for viewing or listening by a user. Feedback may also be provided to adjust the logic that identifies MCs. MC Alerts may also be sent to users autonomously or based on user-set parameters.
Owner:DISNEY ENTERPRISES INC

Padding methods for adaptive loop filter in video coding

A mechanism for processing video data is disclosed. A padding process is determined for application to derive out of boundary samples for use by a filter applied to in boundary samples within a video unit. A conversion is performed between a visual media data and a bitstream based on the filter.
Owner:BYTEDANCE INC +1

Synchronizing audio and video data

Methods, systems, and storage media for synchronizing audio and video data are disclosed. Exemplary implementations may: receive at least one visual media file; receive audio data; determine a beat pattern associated with the audio data; generate a visual effect sequence comprising the at least one visual media file and at least one three-dimensional feature; and generate a resultant video file comprising the visual effect sequence and a transition sequence associated with the beat pattern.
Owner:META PLATFORMS INC

User interface for capturing and managing visual media

To provide a user interface that stops display of control for adjusting a capture duration according to determination that a low light condition is not satisfied.SOLUTION: A user interface displays control for adjusting a capture duration for capturing a medium according to a request for capturing the medium, simultaneously with representation of visual fields of one or more cameras according to determination that a low light condition is satisfied which includes a condition that is satisfied when the rays of ambient light within the visual fields of the one or more cameras fall below respective thresholds. The capture duration is a period during which different images used for creating a composite image are captured. The user interface stops display of the control for adjusting the capture duration according to determination that the low light condition is not satisfied.SELECTED DRAWING: Figure 6B
Owner:APPLE INC

Event source content and remote content synchronization

A method and apparatus for synchronizing event media content including remote audio and video content recorded by spectator or fan users at an event performance, wherein the remote audio content is recorded from the speakers of the event performance and directly from source audio content recorded at the performance by originators, clubs, etc. The source audio content has higher quality than the remote audio content recorded by the spectators. The higher quality audio source content replaces the lower quality audio content recorded by the spectators. The resulting source audio / remote video media content provides a pure recording clear audio quality for the user's personalized event memento.
Owner:VAUDIO LTD

Using side information for sample adaptive offset in video coding

A mechanism for processing video data is disclosed. The mechanism includes determining to employ side information as input to a sample adaptive offset (SAO) filter. The SAO filter can also be a cross component SAO (CCSAO) filter. A conversion is performed between a visual media data and a bitstream based on the SAO filter.
Owner:DOUYIN VISION CO LTD +1

Dynamic adaptation of volumetric content component qbitstreams in streaming services

A media content processing device can decode visual volumetric content based on one or more messages that can indicate which of one or more attribute sub-bitstreams indicated in a parameter set is active. The parameter set can comprise a parameter set based on visual volumetric video. The messages indicating one or more active attribute sub-bitstreams can be received by a decoder. The decoder can perform decoding based on the one or more messages, such as determining which attribute sub-bitstream to use to decode visual media content. The one or more messages can be generated and sent to a decoder, for example, to indicate deactivation of the one or more attribute sub-bitstreams. The decoder can determine, based on the one or more messages, an inactive attribute sub-bitstream and skip that inactive attribute sub-bitstream to decode the visual media content.
Owner:INTERDIGITAL VC HOLDINGS INC

Bilateral filter in video coding

A mechanism for processing video data is disclosed. The mechanism determines to classify picture data into groups based on statistical information related to the picture data and apply a bilateral filter to filter samples in each of the groups. A conversion is performed between a visual media data and a bitstream based on the bilateral filter.
Owner:DOUYIN VISION CO LTD +1

Video Enhancement

The media application receives a request for enhanced video from the user. The media application records the input video of the scene, and the input video has a first format. The media application converts the input video to a second format. The media application converts the input video to the second format by performing front-end processing and conversion from the red, green, and blue (RGB) color space to the YUV color space using an image signal processor, and the file size of the input video in the second format is smaller than that of the input video in the first format. The media application sends the input video in the second format to the server for cloud processing. The media application receives the enhanced video from the server.
Owner:GOOGLE LLC

Video media platform sequence recommendation method fusing frequency domain information and cross-user enhancement

A video media platform sequence recommendation method fusing frequency domain information and cross-user enhancement comprises the steps that firstly, a user behavior sequence is segmented through a sliding window, a sub-sequence set is constructed, the sub-sequence set is classified according to a target video, and a plurality of video category subsets are formed; secondly, Fourier transform is carried out on sub-sequence embedded vectors, high-frequency noise is dynamically filtered based on frequency domain energy, and de-noised intention representation is reconstructed through inverse transform so as to consider both long-term preference and instantaneous interest; thirdly, clustering intention representations of similar video subsequences to select positive sample pairs, clustering intention representations of all the subsequences to extract intention prototypes of all clusters, and constructing a sequence-level and prototype-level cross-granularity contrast learning framework; and finally, obtaining user and video representation to realize a personalized video recommendation result. According to the method, hidden dynamic association of long-term and short-term intentions in a potential user behavior sequence is fully tapped, so that the accuracy of sequence recommendation videos in a video media platform is improved.
Owner:ZHEJIANG UNIV OF TECH

Neural network-based post-filter for video coding and decoding

A neural network-based post-filter for video coding is described, and a method, apparatus, and medium for processing video data are provided. The method comprises: deriving, for an encoder and a decoder, one or more neural network (NN) filter model indices and / or availability of one or more NN filter models from supplemental enhancement information (SEI) messages of video units of a video media file in the same manner; and performing a conversion between the video media file and the bitstream based on the one or more NN filter model indices and / or the availability of the one or more NN filter models.
Owner:FACE CUTE CO LTD

IMS DC session switching method, system, and apparatus, terminal device, and storage medium

The present application relates to an IMS DC session switching method, system, and apparatus, a terminal device, and a storage medium. The method comprises: sending an SDP offer message, wherein the SDP offer message comprises information for changing an IMS DC session; and receiving an SDP answer message sent by a second terminal, wherein the changing the IMS DC session includes adding audio and / or video media information to an established standalone IMS DC session, and deleting audio and / or video media information in an established IMS DC session.
Owner:CHINA TELECOM CORP LTD TECHNOLOGY INNOVATION CENTER +1

Using boundary strength for adaptive loop filter in video coding

A mechanism for processing video data is disclosed. The mechanism includes determining to employ a boundary strength of a deblocking filter (DBF-BS) as side information input into an adaptive loop filter (ALF) or a cross component ALF (CC-ALF). A conversion can then be performed between a visual media data and a bitstream based on the ALF or CC-ALF.
Owner:DOUYIN VISION CO LTD +1