Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

81results about "Multimedia data retrieval" patented technology

Audio content playback management

An example implementation involves a playback device receiving a request to add continuous automated streaming audio content to a playback queue, the request indicating a playback start time, and the playback queue indicating a plurality of audio content. The implementation further involves the playback device adding the continuous automated streaming audio content to the playback queue. The example implementation also involves the playback device determining that a duration until the playback start time is less than a duration of the given audio content before playing a given audio content in the playback queue. The example implementation involves the playback device responsively, playing the continuous automated streaming audio content.
Owner:SONOS INC

Content archive model

An archive model can be used for managing networked storage of recorded content, such as network DVR (digital video recorder) content. Content may be initially recorded to an active storage device, with individual duplicate copies recorded for each requesting user, and subsequently archived to an archive storage device. For playback, the content can be reconstituted into the active storage device prior to delivery to the requesting user. Content can be predictively reconstituted in anticipation of user needs, and the reconstitution capacity of the system can be dynamically reallocated for load balancing.
Owner:COMCAST CABLE COMM LLC

System and method for managing a file

A computer-implemented method, performed by a system, for creating a video representation of a file is disclosed. The method comprises receiving, by the client device a request to start a synchronization addon with a file, whose content is requested to be shared by a user of the client device, wherein synchronization of data, collaborated on in the collaboration session, is performed by a synchronization framework used by the synchronization addon, and receiving, by the recording function, a message, originating from the client device instructing the central server to capture image frames of the content displayed by an instance of the synchronization addon. The method further comprises listening, by the recording function, to event messages from the synchronization addon executing in the recording function, and determining, by the recording function, a respective event time stamp for each respective event message of the event messages to obtain metadata associated with the content.
Owner:LIVEARENA TECH AB

Digital media asset management method and system

The invention relates to the technical field of digital media asset management, in particular to a digital media asset management method and system, which comprises the processes of metadata compliance tag injection, policy awareness mixed query and privacy affinity cache management. A cross-cloud query path is dynamically decided through an intelligent metadata gateway, double cloud caches are deployed in a hierarchical manner, and label consistency and copyright transaction security are ensured by using a block chain evidence storage module. According to the method, the global retrieval efficiency can be improved, the sensitive data cross-border transmission risk is reduced, illegal access and data leakage are effectively avoided, meanwhile, the high-frequency sensitive data retrieval performance is optimized by dynamically adjusting the cache strategy, and efficient cooperation of cross-layer events is achieved. The invention provides a digital media asset management method and a digital media asset management system, and aims to solve or at least alleviate the problems of low metadata retrieval efficiency and sensitive data cross-border transmission compliance risk coexistence in a mixed cloud environment.
Owner:大河传媒有限公司

Contextual digital media processing systems and methods

Systems and methods for contextual digital media processing are disclosed herein. An example method includes receiving content from a source as digital media that are being displayed to a user, processing the digital media to determine contextual information within the content, searching at least one network for supplementary content based on the determined contextual information, and transmitting the supplementary content for use with at least one of the source or a receiving device.
Owner:SONY INTERACTIVE ENTERTAINMENT LLC

Interactive multimedia content processing method and apparatus, device, medium and product

The present invention relates to the technical field of computers, and relates to an interactive multimedia content processing method and apparatus, a device, a medium and a product. The interactive multimedia content processing method of the present invention comprises: on the basis of content text of an original story, generating content text of one or more story branches; on the basis of the content text of the original story and the content text of the one or more story branches, generating images, wherein the images include images of characters and images of scenes; and on the basis of the content text of the original story, the content text of the one or more story branches and the images, generating interactive multimedia content.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Methods and system for distributing information via multiple forms of delivery services

A content distribution facilitation system is described comprising configured servers and a network interface configured to interface with a plurality of terminals in a client server relationship and optionally with a cloud-based storage system. A request from a first source for content comprising content criteria is received, the content criteria comprising content subject matter. At least a portion of the content request content criteria is transmitted to a selected content contributor. If recorded content is received from the first content contributor, the first source is provided with access to the received recorded content. The recorded content may be transmitted via one or more networks to one or more destination devices. Optionally, a voice analysis and / or facial recognition engine are utilized to determine if the recorded content is from the first content contributor.
Owner:GREENFLY

Generating image scene based on events

Methods and systems are disclosed for suggesting a scene of an image using one or more machine learning models based on detected events. The methods and systems detect an event associated with a second user through an interaction system associated with a first user, and generate a prompt including the event and a request for a plurality of scenarios associated with the event. The method and system processes the cues through a large language model (LLM) to generate a plurality of scenes related to the event, and presents individual content items corresponding to individual ones of the plurality of scenes.
Owner:SNAP INC

Methods, systems, and media for presenting media content items with aggregated timed reactions

In one example, a computing system comprises processing circuitry that executes instructions to determine whether a media content item is eligible for identification of key moments, retrieve a plurality of reactions associated with the media content item, wherein each reaction comprises a reaction timestamp that corresponds to a time at which a reaction was received and a graphical icon that corresponds to one of the reactions, group the reactions based on the reaction timestamp and the selected graphical icon, identify a plurality of key moments within the media content item based on the groupings, wherein each of the key moments is associated with a representative graphical icon that corresponds to an aggregated viewer reaction, and cause an indication of the media content item to be presented in a user interface while presenting the aggregated viewer reaction at each time window associated with each of the key moments.
Owner:GOOGLE LLC

Multimedia autonomous electronic survival device

The present invention relates to a multimedia autonomous electronic survival device characterized in that it consists of at least a hardware base (21) formed by a motherboard (1) equipped with an electronic microprocessor, and an internal memory (2). A series of input and output periphery elements including a touch screen (3), a battery (4) and an electric power source (S) of the set are connected to the motherboard, all of which are contained in a mechanical casing (6). The internal memory (2) contains a database containing ordered and structured knowledge on survival and its techniques. The device is equipped with software having a visual and acoustic interface that allows access to the information of the database through the input and output peripherals, incorporating a radiofrequency receiver and emitter (10) for communication and a GPS positioner (9) for positioning that are connected to the motherboard (1), and the power source (5) incorporating autonomous and renewable power supply elements that can additionally be connected to the electrical grid.
Owner:ALEGRE MASO BENJAMI

Mobile terminal and method for controlling same

PendingEP4765792A2Devices with GPS signal receiverSubstation equipment
The present disclosure relates to a mobile terminal having an artificial intelligence unit, and a control method therefor. A mobile terminal according to the present disclosure includes: GPS configured to receive location information of the mobile terminal, a wireless communication unit configured to perform wireless communication with an external device, and a controller configured to store communication information when a communication event with the external device is generated. The controller matches the location information received from the GPS with the communication information and stores the matched information when the communication event is generated, and outputs at least one information corresponding to a search query for retrieving communication information from pre-stored communication information by using location information matched to the pre-stored communication information and the search query, when the search query is input by a user.
Owner:LG ELECTRONICS INC

Method for processing at least one grid of data representative of a scene captured by a set of sensors, device and corresponding program.

Method for processing at least one data grid representing a scene captured by a set of sensors, device and corresponding program. The invention relates to a method for processing data grids representing a scene captured by sensors, a grid being divided into cells (C1, C2, C3, Ci) each associated with a probability distribution of states (PD) of an area of ​​the scene covered by the cell at a current time step, determined as a function of data reported by at least one of the sensors and / or data predicted by a prediction model as a function of data reported at a previous time step by said sensor.The process includes determining, for at least one cell of the grid, a data point of interest (DP) representative of the degree of interest of said cell with regard to predetermined criteria related to a scene capture context, a characteristic of the captured scene, and / or an application associated with said capture, said determination delivering an augmented data grid (ADG). Abbreviated figure: Figure 1.
Owner:INRIA INSTITUT NATIONAL DE RECHERCHE EN INFORMATIQUE ET EN AUTOMATIQUE

System and method for changing the size of a user group to which media items are to be presented

A system and method are disclosed for: identifying a media item to be provided to a user group of a content sharing platform, wherein the media item is associated with a category, and wherein each user in the user group is associated with a weight indicating a probability of correspondence between the respective user and the category associated with the media item; receiving a request to change the size of the user group from a first level to a second level, the first level corresponding to a first weight threshold for the user group; and calculating a second weight threshold for the user group corresponding to the second level based on a value indicating a difference between the first level and the second level, the second weight threshold to be subsequently used to determine whether the media item is to be provided to a user requesting content from the content sharing platform.
Owner:GOOGLE LLC

Method and system for reproducing contents, and computer-readable recording medium thereof

A content reproducing method and system for performing seamless playback of contents between devices is provided. The contents reproducing system includes a portable device (100) which, when a short distance communication with a remote control (110) which is configured to control an electronic device (120) occurs during reproducing of contents, generates data required by the electronic device (120) for reproducing the contents that are being reproduced, and which transmits the generated data to the remote control (110); the remote control (110) which receives the data from the portable device (100) and which transmits the received data to the electronic device (120), in conjunction with the occurrence of the short distance communication with the portable device (100); and the electronic device (120) for receiving the contents from a contents provider and reproducing the contents.
Owner:SAMSUNG ELECTRONICS CO LTD

Methods and systems for meme generation with multi-modal input and planning

The disclosure relates generally to methods and systems for meme generation with multi-modal input and planning. Conventional AI-based techniques either rely on input text prompt or the user-provided image as an input to generate the meme. Such input specification styles result in restrictive for clearly specifying an intent with a single modality. The present disclosure explores a multi-modal input specification style where a user can provide the input through a text prompt along with a widely popular meme template image. According to the present disclosure, the meme generation task is defined as a combination of two sub-tasks. In the first sub-task, a meme image template is retrieved from a dataset of existing meme templates using a template planning strategy. In the second sub-task, the text caption is generated for the retrieved template conditioned on the multi-modal input provided by the user through the caption planning strategy.
Owner:TATA CONSULTANCY SERVICES LTD

Methods and systems for meme generation with multi-modal input and planning

The disclosure relates generally to methods and systems for meme generation with multi-modal input and planning. Conventional AI-based techniques either rely on input text prompt or the user-provided image as an input to generate the meme. Such input specification styles result in restrictive for clearly specifying an intent with a single modality. The present disclosure explores a multi-modal input specification style where a user can provide the input through a text prompt along with a widely popular meme template image. According to the present disclosure, the meme generation task is defined as a combination of two sub-tasks. In the first sub-task, a meme image template is retrieved from a dataset of existing meme templates using a template planning strategy. In the second sub-task, the text caption is generated for the retrieved template conditioned on the multi-modal input provided by the user through the caption planning strategy.
Owner:TATA CONSULTANCY SERVICES LTD

Methods and apparatus for displaying, compressing and / or indexing information relating to a meeting

A method of visualising a meeting between one or more participants on a display includes, in an electronic processing device, the steps of: determining a plurality of signals, each of the plurality of signals being at least partially indicative of the meeting; generating a plurality of features using the plurality of signals, the features being at least partially indicative of the signals; generating at least one of: at least one phase indicator associated with the plurality of features, the at least one phase indicator being indicative of a temporal segmentation of at least part of the meeting; and at least one event indicator associated with the plurality of features, the at least one event indicator being indicative of an event during the meeting. The method also includes the step of causing a representation indicative of the at least one phase indicator and / or the at least one event indicator to be displayed on the display to thereby provide visualisation of the meeting.
Owner:PINCH LABS PTY LTD

System and methods for changing a size of a group of users to be presented with a media item

A media item to be provided to a group of users of a content sharing platform is identified. Each user is associated with one or more weights, indicating the probability of correspondence with a category associated with the media item. A request to dynamically change the group size from a first to a second level is received. A second weight threshold for the group of users corresponding to the second level is obtained. Upon receiving a content request from a client device associated with a user, it is determined whether the user's weight meets the second threshold. If so, the media item is provided to the client device.
Owner:GOOGLE LLC

Generative model for suggesting image modifications

Methods and systems for suggesting image modifications using one or more machine learning models are disclosed. The methods and systems select, by an interactive application, individual content items from a plurality of previously captured content items that match one or more criteria corresponding to sharable content, and generate prompts that include the individual content items and a request for a plurality of suggested modifications of the individual content items. The methods and systems process cues through a large language model (LLM) to generate a plurality of suggested modifications to individual content items, and generate modified individual content items corresponding to individual suggested modifications of the plurality of suggested modifications.
Owner:SNAP INC

Providing guidance regarding content viewed via augmented reality devices

Methods, systems, and storage media for providing feedback regarding augmented reality device content are disclosed. Exemplary implementations may: detect, via a sensor of an augmented reality device having an outwardly facing camera, that a user of the augmented reality device appears within content presented in a view area thereof; responsive to detecting that the user of the augmented reality device appears within the content presented in the view area of the augmented reality device, determine an action being performed by the user; and responsive to determining the action being performed by the user, supplement the content presented in the view area of the augmented reality device.
Owner:META PLATFORMS TECHNOLOGIES LLC

Method for controlling dissemination of instructional content to operators performing procedures within a facility

A method for modifying a procedure includes: accessing an instructional block library containing a set of verified instructional blocks associated with approved digital procedures performed within a facility; accessing an unverified instructional block, authored by an operator, for a new digital procedure at the facility; detecting a set of language signals in the unverified instructional block; correlating an equipment unit language signal, in the set of language signals, with an equipment unit located within the facility; correlating an action language signal, in the set of language signals, with an action prompt related to the equipment unit; identifying a verified instructional block, in the set of verified instructional blocks in the instructional block library, as analogous to the unverified instructional block in response to the verified instructional block including language signals associated with the equipment unit and the action prompt; and inserting the verified instructional block in the new digital procedure.
Owner:APPRENTICE FS INC

A speaker tracking method based on audio-visual dual modality

This invention discloses a speaker tracking method based on audio-visual dual-modal fusion. The method includes: acquiring speaker location information through a camera and a microphone array; fusing audio and image information to locate the speaker based on the number of people in the image and the presence of voice input; deriving different motion control instructions based on the different positioning results, planning a tracking path in real time, and driving the motion control system to track the speaker. In cases where the visual module fails to detect a person, a method is provided to actively search for the speaker using voice positioning information, making the system more intelligent. Ultimately, the invention enables real-time speaker tracking by a mobile robot, facilitating human-machine interaction.
Owner:CHINA UNIV OF GEOSCIENCES (WUHAN)

Content delivery system, content delivery method, and program

To easier provide a content and a comment on it.SOLUTION: The present invention provides a content provision system comprising a control part in which a script formed by at least text information in regard to identification information of a content and an advertisement is stored in a predetermined storage medium so as to enable a browsing by a user; in accordance with the script selected by the user, a reading of a content indicated by content identification information contained in their script is controlled so that the content is executed and provided to the user by utilizing the rights that the user has already acquired via a contract with a specific service; and the text information contained in the script is formed by a voice synthesis, and is provided to the user at least either before or after a provision of the content.SELECTED DRAWING: Figure 1
Owner:SONY GROUP CORP