Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

26results about "Using detectable carrier information" patented technology

Image processing method and apparatus, device, and medium

An image processing method is provided. In the method, a target video frame set is acquired from video data of a plurality of video frames. The target video frame set includes a subset of the video frames that is selected based on characteristics of the subset of the video frames. A global color feature of a reference video frame is acquired. An image semantic feature of the reference video frame is acquired. An enhancement parameter of the reference video frame is acquired for each of at least one image information dimension according to the global color feature and the image semantic feature. Image enhancement is separately performed on the video frames in the target video frame set according to each enhancement parameter of the reference video frame to obtain target image data for each of the video frames in the target video frame set.
Owner:TENCENT CLOUD COMPUTING (BEIJING) CO LTD

Systems and methods for intelligent playback

PendingUS20260155154A1Speech analysisUsing detectable carrier informationSpeech rateSyllable
Systems and methods for intelligent playback of media content may include an intelligent media playback system that, in response to determining the speech tempo in audio content by measuring syllable density of speech in the audio content, automatically adjusts a playback speed of the audio content as the audio content is being played based on the determined speech tempo. In some embodiments, the system may automatically and dynamically adjust the playback speed to result in a desired target speech tempo. In addition, the system may determine whether to automatically adjust playback speed of the audio content, as the media is being played, based on the detected speech tempo of the speech in the audio content and the determined type of content of media. Such automatic adjustments in playback speed result in more efficient playback of the audio content.
Owner:DISH NETWORK TECHNOLOGIES INDIA PTE LTD +1

Methods for serving interactive content to a user

ActiveUS12664352B2Electronic editing digitised analogue information signalsUsing non-detectable carrier informationDigital videoInteractive content
One variation of a method for serving interactive content to a user includes, at a visual element inserted into a document accessed by a computing device: loading a first frame from a digital video; in response to a scroll-down event that moves the visual element upward from a bottom of a window rendered on the computing device toward a top of the window, seeking from the first frame through a subset of frames in the digital video in a first direction at a rate corresponding to a scroll rate of the scroll-down event, the subset of frames spanning a duration of the digital video corresponding to a length of the scroll-down event; and, in response to termination of the scroll-down event with the visual element remaining in view within the window, playing the digital video forward from a last frame in the subset of frames in the digital video.
Owner:YIELDMO

COMMERCIAL AUTOMATIC PLAYBACK SYSTEM

ActiveDE602013087547T2Television system detailsUsing detectable carrier information
Owner:ADEIA MEDIA SOLUTIONS INC

Multimedia scene break detection

ActiveUS12659555B2Using detectable carrier informationSelective content distribution
A system and method for multimedia scene break detection, including: a computer processor, a scene break detection service executing on the computer processor and comprising functionality to: (i) receive a request for scene break detection on a media item, (ii) identify a set of candidate scene break timestamps selected based on analysis of an audio component of the media item and a video component of the media item, (iii) execute a computer vision scoring model for each candidate scene break timestamp to generate a score representing a visual distance between a first set of proximal frames and a second set of proximal frames of the candidate scene break timestamp, and (iv) select, based at least on the score of each of the set of candidate scene break timestamps, a final set of scene break timestamps for the media item.
Owner:TUBI INC

Intelligent video editor for creating non-linear editing timeline

PendingUS20260148754A1Electronic editing digitised analogue information signalsUsing detectable carrier informationPersonalizationSocial media
An automated video editing system facilitates the creation of non-linear editing (NLE) timelines using a single prompt from a user. The system automatically ingests digital media, including video, audio, text, and images, and processes them to generate a proxy version with extracted features such as speech transcription, shot detection, facial recognition, and text recognition. A prompt-driven editing engine interprets user input and generates an edit decision list (EDL) using a large language model, which guides the assembly of an edited video timeline. The system also applies advanced editing features-such as captioning, animated title cards, font and color styling, sound effects, and transitions-based on learned user preferences. Additionally, it enables contextual overlays, chapter cards, and hierarchical timelines, while continuously learning user preferences to personalize editing results. The system may be integrated with social media platforms and existing media libraries for content sourcing, customization, and automated publishing.
Owner:OPEN VIDEO EXPLORATION INC

Method and system for seamless media synchronization and switching

ActiveCN114257324BBroadcast-related systemsUsing detectable carrier informationMediaFLOMicrophone signal
The invention is entitled "Method and system for seamless media synchronization and switching." A method performed by a portable media player device is disclosed. The method receives a microphone signal that includes audio content output by an audio playback device via a loudspeaker. The method determines identifying information about the audio content, where the identifying information is determined through acoustic signal analysis of the microphone signal. In response to determining that the audio playback device has ceased outputting the audio content, the method retrieves an audio signal corresponding to the audio content from a local memory of the portable media player device or a remote device communicatively coupled thereto, and uses the audio signal to drive a speaker that is not part of the audio playback device to continue outputting the audio content and any additional audio content related to the audio content.
Owner:APPLE INC

Device for recording data for generating a local street panorama image and method for same

The invention relates to a device for recording data to generate a georeferenced street panorama image, comprising a camera, a satellite-based positioning and time-determination device, and a storage unit. According to the invention, a processing unit is provided which encodes the time data from the satellite-based positioning and time-determination device into a format recordable by the camera and forwards it to the camera, the camera being designed to simultaneously record this data and to record a continuous video. The storage unit stores the video containing the time data as well as the position and time data from the satellite-based positioning and time-determination device. The invention further relates to a method for this.
Owner:PARKLING GMBH

Video playback device and video playback program

The challenge is to enable video playback by easily customizing the video playback state. [Solution] The video playback device includes a setting unit that sets parameters relating to the playback state of a first video, and a playback control unit that controls the playback state of a second video in a display mode that includes a mode in which the first video and a second video different from the first video are displayed simultaneously, based on the parameter settings.
Owner:ALPHATHETA CORP

Advertisement break in video within a messaging system

Aspects of the disclosure relate to systems and methods for setting advertisement breakpoints in a video, the system comprising a computer-readable storage medium storing a program. The program and method cause accessing a video; determining a plurality of shot boundaries of the video, each shot boundary defining a shot corresponding to a contiguous series of video frames without a cut or transition; and for each shot boundary of the plurality of shot boundaries, performing a set of breakpoint tests on the shot boundary, each breakpoint test configured to return a respective score indicating whether the shot boundary corresponds to a breakpoint for potentially inserting an advertisement during playback of the video, computing a combined score for the shot boundary based on combining each of the respective scores, and setting the shot boundary as a breakpoint if the combined score satisfies a threshold.
Owner:SNAP INC

Apparatus and method for controlling a camera

An apparatus and method for controlling a camera are provided, the apparatus having an input unit configured to acquire video data from the camera, an image processing unit configured to determine from the acquired video data whether a particular person is alone in a first area monitored by the camera, a control unit configured to generate a control signal for controlling the camera to operate in a first monitoring mode or a second monitoring mode based on a determination by the image processing unit of whether the particular person is alone in the first area monitored by the camera, and an output unit configured to output the control signal to the camera.
Owner:KONINKLIJKE PHILIPS NV

Method performed by electronic apparatus, electronic apparatus and storage medium

A method performed by an electronic apparatus, an electronic apparatus and a storage medium, which involves the field of artificial intelligence are provided. The method includes obtaining target sound masks of a target in a first video at respective moments, based on image-related information of the target, a first audio signal corresponding to the first video, and direction information of the first audio signal, and obtaining a second audio signal in which a sound related to the target is excluded, based on the target sound masks of the target at the respective moments and the first audio signal.
Owner:SAMSUNG ELECTRONICS CO LTD

Method, apparatus, device and product for adding effect

PendingUS20260148753A1Electronic editing digitised analogue information signalsUsing non-detectable carrier informationSimulationMechanical engineering
The disclosure relates to a method, an apparatus, a device and a product for adding an effect. The method comprises obtaining beat marker information of music of a video, the beat marker information indicating a change in a rhythm of the music. The method further comprises obtaining, variable-speed information for the video, the variable-speed information indicating a playback speed of each video segment of the video. In addition, the method further comprises adding an effect to the video based on the variable-speed information.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD +1

Methods and apparatus for streaming content

ActiveUS12641209B2Static indicating devicesUsing detectable carrier informationComputer graphics (images)3d image
Methods and apparatus for streaming content corresponding to a 360 degree field of view are described. The methods and apparatus of the present invention are well suited for use with 3D immersion systems and / or head mounted display which allow a user to turn his or her head and see a corresponding scene portion. The methods and apparatus can support real or near real time streaming of 3D image content corresponding to a 360 degree field of view.
Owner:NEVERMIND CAPITAL LLC

Automatic commercial playback system

ActiveEP4250749B1Television system detailsUsing detectable carrier information
While a multimedia device is fast-forwarding content, the multimedia device reads "jump back" tags expressed in or derived from a closed-caption stream. When the multimedia device detects the presence of a "jump back" tag while fast-forwarding, the multimedia device enters a special state. While in this special state, if the multimedia device detects that the user has instructed the multimedia device to stop fast-forwarding, the multimedia device locates a specified temporal location in a recorded commercial break. This specified temporal location may be specified by the particular tag, for example. The multimedia device stops performing whatever activity in which the multimedia device was engaged, "jumps back" to the specified temporal location in the recorded commercial break, and resumes playing the recorded content stream at normal speed from the specified temporal location.
Owner:ADEIA MEDIA SOLUTIONS INC

Integration of Video Language Models with AI for Filmmaking

A computer-implemented method includes receiving metadata related to filmmaking techniques and Lidar data, processing the data to produce adapted training data, and applying transfer learning techniques to a video model to generate an adapted model for simulating filmmaking techniques based on user input. A computing system includes processors and memory having stored instructions to receive filmmaking metadata and Lidar data, process the data into adapted training data, and apply transfer learning to generate an adapted model for simulating filmmaking techniques based on user input. A computer-readable medium includes instructions that, when executed, perform receiving filmmaking metadata and Lidar data, processing the data into adapted training data, and applying transfer learning to generate an adapted model for simulating filmmaking techniques based on user input.
Owner:INTERPOSITIVE LLC

Video assembly using generative artificial intelligence

Embodiments of the present invention provide systems, methods, and computer storage media for identifying the relevant segments that effectively summarize the larger input video and / or form a rough cut, and assembling them into one or more smaller trimmed videos. For example, visual scenes and corresponding scene captions are extracted from the input video and associated with an extracted diarized and timestamped transcript to generate an augmented transcript. The augmented transcript is applied to a large language model to extract sentences that characterize a trimmed version of the input video (e.g., a natural language summary, a representation of identified sentences from the transcript). As such, corresponding video segments are identified (e.g., using similarity to match each sentence in a generated summary with a corresponding transcript sentence) and assembled into one or more trimmed videos. In some embodiments, the trimmed video is generated based on a user's query and / or desired length.
Owner:ADOBE INC

Method executed by electronic equipment, electronic equipment and storage medium

The embodiment of the invention provides a method executed by electronic equipment, the electronic equipment and a storage medium, and relates to the field of artificial intelligence. The method comprises the following steps: obtaining a target sound mask of a target at each moment based on image-related information of the target in a first video, a first audio signal corresponding to the first video and direction information of the first audio signal; and based on the target sound mask of the target at each moment and the first audio signal, a second audio signal is obtained, and the second audio signal does not include sound related to the target. Optionally, the method performed by the electronic device may be performed using an artificial intelligence model.
Owner:BEIJING SAMSUNG TELECOM R&D CENT +1

Processing monocular videos using three-dimensional gaussian splatting

The present disclosure describes techniques for processing monocular videos using three-dimensional gaussian splatting (3DGS). Spatial decomposition and temporal decomposition are performed on a monocular video to generate a plurality of clips. A first set of 3DGS representing foreground objects in each of the plurality of clips are initialized and optimized. A second set of 3DGS representing background in each of the plurality of clips are initialized and optimized. Two images are generated for each frame comprised in each of the plurality of clips based on the first set of 3DGS and the second set of 3DGS, respectively. Two images are merged to generate a resulting image for each frame in each of the plurality of clips. The resulting image accurately represents a corresponding frame in the monocular video.
Owner:LEMON INC(GB)

Method, device, equipment and product for adding special effects

The invention relates to a method, a device, equipment and a product for adding special effects. The method includes acquiring sticking point information of music of a video, wherein the sticking point information indicates a change in rhythm of the music. The method further includes obtaining speed change information for the video based on the sticking point information, where the speed change information indicates a playback rate for each video segment of the video. In addition, the method also includes adding a special effect to the video based on the speed change information.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD +1

Artificial Intelligence-Based Video Content Creation with Predetermined Styles

A computer-implemented method includes capturing test images of a scene using different film formats, constructing a training dataset with a variety of shots captured under varied lighting conditions, training an artificial intelligence model to learn specific visual signatures, assessing the AI model to determine authenticity, and optimizing learning cycles to enhance training. A computing system comprises one or more processors and one or more memories storing instructions that, when executed, cause the system to perform these steps. A non-transitory computer-readable medium stores instructions that, when executed by one or more processors, cause a system to capture test images using different film formats, construct a training dataset with varied lighting conditions, train an AI model to learn specific visual signatures, assess authenticity, and optimize learning cycles.
Owner:INTERPOSITIVE LLC