Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

146results about "Video data browsing/visualisation" patented technology

Landscape interaction method and apparatus, electronic device, and storage medium

Provided are a landscape interaction method and apparatus, an electronic device, and a storage medium. The method includes: receiving a first trigger operation acting on a first display area of a landscape playback page, where the first display area includes a display area of author information of a current video author, and the current video author is a publisher of a first target video in the landscape playback page; and in response to the first trigger operation, displaying a target-video list of the current video author in a second display area of the landscape playback page, where the target-video list includes at least one video item of at least one target video published by the current video author.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

System and method for analyzing videos

The present invention relates to a system and method for analyzing videos. The system comprises a video section and a tabs section, wherein the video section comprises a video and the tabs section comprises at least one tab with items or information related to the video. In most embodiments one of the tabs is a products tab which lists items in the video which can be purchased. These items are associated with specific time points in the video, generally when the item is visible in the video. When a user watches the video, the individual products will be highlighted as their time points are viewed. The user can additionally scroll through the list of products without effecting video playback. The method comprises the gathering and uploading of information associated with the video.
Owner:VIKTRS LTD

Camera quick calling and visualization method based on natural language understanding

The invention discloses a camera quick calling and visualization method based on natural language understanding, and relates to the technical field of video monitoring and artificial intelligence crossing. Comprising the steps that S1, a full-amount camera database is constructed, S2, a camera space retrieval interface is provided, and S3, a map visualization front end is constructed; s4, providing an address-coordinate conversion service; s5, local privatization deployment: deploying a camera space retrieval interface, a front-end static resource and a large language model mirror image in an intranet GPU server; s6, constructing a large model knowledge newly establishing a knowledge base table Knowledgeprompt in the model mirror image for writing data, and constructing the large model knowledge base, and S7, receiving a query demand of a user for inputting a natural language in a dialog box; s8, performing semantic understanding, and calling an interface to return data; step S9, automatically dotting the map; and S10, playing the video stream as required.
Owner:浪潮智慧城市科技有限公司

Structured video documents

A method includes receiving a content feed that includes audio data corresponding to speech utterances and processing the content feed to generate a semantically-rich, structured document. The structured document includes a transcription of the speech utterances and includes a plurality of words each aligned with a corresponding audio segment of the audio data that indicates a time when the word was recognized in the audio data. During playback of the content feed, the method also includes receiving a query from a user requesting information contained in the content feed and processing, by a large language model, the query and the structured document to generate a response to the query. The response conveys the requested information contained in the content feed. The method also includes providing, for output from a user device associated with the user, the response to the query.
Owner:GOOGLE LLC

Transformation of database entries for improved association with related content items

A content analysis system includes processor and memory hardware storing data analyzed content items and instructions for execution by the processor hardware. The instructions include, in response to a first intermediate content item being analyzed to generate a first text description, receiving the first intermediate content item and analyzing the first text description to generate a first reduced text description. The instructions include identifying a first set of tags by applying a tag model to the first text description and generating a first analyzed content item. The instructions include adding the first analyzed content item to the analyzed content database and, in response to a displayed content item being associated with at least one tag of the first set of tags, displaying a first user-selectable link corresponding to the first analyzed content item on a portion of a user interface of a user device displaying the displayed content item.
Owner:CHARLES SCHWAB & CO INC

Human-machine interaction method and apparatus, and electronic device

Embodiments of this application provide a human-machine interaction method and apparatus, and an electronic device. In the method, a first interface is displayed in a gallery application in response to a first operation of a user on a thumbnail that is of a video and that is displayed in a first album or a photo tab. A picture of the video is displayed in a display region of the first interface. The first interface further includes an association region and an operation region. The association region includes an identifier of the video and an identifier of at least one photo associated with the video. The at least one photo is generated based on the video. The operation region includes a share control. The video is shared with a target terminal in response to the user operating the share control. According to embodiments of this application, a video can be shared in a scenario in which an associated photo is generated based on the video.
Owner:HONOR DEVICE CO LTD

A method, device and system for storing and synchronously triggering augmented reality events in a panoramic video

This invention provides a method for storing augmented reality events in panoramic videos, a method for synchronously triggering events, and an apparatus. The method uses time-triggered special effects or footage in the panoramic video as trigger events. It uses the ID of the trigger event as the key and the content of the trigger event as the value to create an index for a linked list. Each trigger event is stored in the linked list in chronological order. Simultaneously, a balanced tree is built using the trigger time of each event as the key and the ID of each trigger event as the value. To insert or delete a trigger event, the method searches the balanced tree to obtain the IDs of time-adjacent trigger events, then finds the corresponding position in the linked list based on that ID for insertion or deletion. It also allows for timed triggering of each event at specific times. This invention enables convenient and quick insertion, deletion, and synchronous triggering of events in panoramic videos. By using a balanced tree to store events, it greatly reduces the computational requirements during event lookup and improves the synchronization during panoramic video playback.
Owner:BEIJING UNIV OF POSTS & TELECOMM

Retrieving images for video scrubbing at a client device

Methods, software, devices and systems for video scrubbing enable a client device to retrieve images for scrubbing based on a user-requested time along a video timeline of a video stored in a server. The client device checks if a cached image meets specified conditions, including a timestamp within a precision margin around the requested time. The precision margin scales with the timeline length, providing a smaller margin for shorter timelines and a larger margin for longer timelines. If a relevant cached image is found, it is retrieved; if not, an image with a highest relevance score within the precision margin is fetched from the server and stored in memory.
Owner:AXIS

Retrieving images for video scrubbing at a client device

The present disclosure relates to retrieving images for video scrubbing at a client device, in particular to a method 400, software, apparatus and system for video scrubbing enabling a client device to retrieve images for scrubbing based on a user requested time along a video timeline of a video stored in a server. The client device checks S404 if a cached image meets a specified condition, which is a timestamp included within a precision margin around the requested time. The precision margin scales with the timeline length, providing a smaller margin for shorter timelines and a larger margin for longer timelines. If a relevant cached image is found, the image is retrieved S416; if not, the image with the highest relevance score within the precision margin is fetched S418 from the server and stored in memory.
Owner:AXIS

Video processing method and apparatus

The present application discloses a video processing method. In an example, the method can be performed by a first client. After posting a plurality of videos, the first client displays first video identifiers corresponding to the plurality of videos on a first page. After displaying the first video identifiers corresponding to the plurality of videos, if a preset condition is satisfied, the first client can switch, in response to satisfying the preset condition, the first video identifiers corresponding to the plurality of videos to second video identifiers. Therefore, by using the solution, according to the second video identifiers, a user can determine that the plurality of videos is a series of videos. Correspondingly, a viewing experience of a user is enhanced.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Multimedia focalization

Example implementations are directed to methods and systems for individualized multimedia navigation and control including receiving metadata for a piece of digital content, where the metadata comprises a primary image and text that is used to describes the digital content; analyzing the primary image to detect one or more objects; selecting one or more secondary images corresponding to each detected object; and generating a data structure for the digital content comprising the one or more secondary images, where the digital content is described by a preferred secondary image.
Owner:OPEN TV INC

Character dynamic effect data generation method and device, storage medium and electronic equipment

The invention relates to a text dynamic effect data generation method and device, a storage medium and electronic equipment. The method comprises the following steps: constructing a copywriting library and a background library; according to the character dynamic effect template, randomly selecting target character content from a copywriting library, and generating a corresponding black-matrix character dynamic effect video; fusing the black-matrix character dynamic effect video with a target background image randomly selected from a background library by adopting a grid division strategy to generate a target dynamic effect video; and extracting a first frame of mask image of the target dynamic effect video to obtain a character mask marking image of the target dynamic effect video, and generating a training data set of the target character content in combination with the copywriting index information of the target dynamic effect video. The technical problems that a large amount of diversified dynamic character and background pairing data in a real scene is lacked, and high-quality character dynamic effect training samples are difficult to automatically generate are solved.
Owner:BEIJING QIYI CENTURY SCI & TECH CO LTD

Human-computer interaction method and device and electronic equipment

The embodiment of the invention provides a man-machine interaction method and device and electronic device.In the method, in a picture library application program, in response to a first operation of a user on a thumbnail of a video displayed in a first photo album or a photo tab, a first interface is displayed, and a display area of the first interface displays a picture of the video; the first interface further comprises an association area and an operation area, the association area comprises an identifier of the video and an identifier of at least one photo associated with the video, the at least one photo is generated based on the video, and the operation area comprises a sharing control; and in response to a user operating the sharing control, sharing the video to the target terminal. According to the embodiment of the invention, video sharing can be realized in a scene in which the video generates the associated photo.
Owner:HONOR DEVICE CO LTD

Media item and product pairing

A method includes providing, for presentation on a user device, a user interface (UI) comprising one or more graphical representations of one or more content items, wherein each graphical representation of a respective content item is selectable to initiate presentation of the respective content item, and is displayed with a UI element in a collapsed state, the UI element comprising information identifying a plurality of products related to the respective video. The method further includes responsive to receiving an indication of a user interaction with the UI element in the collapsed state, causing a presentation of the UI element to be modified from the collapsed state to an expanded state, wherein the UI element in the expanded state comprises a plurality of visual components each associated with one of the plurality of products, and responsive to receiving an indication of a user selection of one of the plurality of visual components of the UI element in the expanded state, facilitating presentation of a first content item.
Owner:GOOGLE LLC

Intelligent management and control platform task closed-loop management system, method, equipment and medium

The invention discloses an intelligent management and control platform task closed-loop management system, method and device and a medium, and belongs to the technical field of intelligent management and control system integration, and the method comprises the steps: directly extracting sample data through an ETL tool after data mapping, recognizing and processing an outlier through a K-means clustering algorithm, and processing missing data through multiple interpolation; an extraction rule is established according to the mode and the structure of the data, the accuracy of the extraction rule is optimized by using a machine learning method, and different extraction processes are entered according to the type of the data; through text topic classification and data labeling, a data analysis model based on a deep learning model is established for data analysis, and then data extraction is performed. By switching the first display mode and the second display mode and displaying the preview elements, layered display and quick retrieval of tasks are realized, and integrated processing of multi-source data is realized by extracting and covering digitized, semi-structured and non-structured information through multi-modal data.
Owner:SANXIA JINSHAJIANG YUNCHUAN HYDROPOWER DEV CO LTD

Unlocking sharing destinations in an interaction system

A third-party user input content item is presented. A content sharing function is invoked responsive to determining user selection of a content sharing graphical element. The content sharing function comprises presentation of a destination graphical element identifying a first content sharing destination. The first content sharing destination is locked. A combination graphical element is user-selectable to invoke a combination function. Responsive to determining user selection of the combination graphical element, the combination function is invoked to access a second user input content item and combine the third-party user input content item with the second user input content item to create a combined user input content item. The first content sharing destination is unlocked and the user is enabled to share the combined user input content item to the unlocked first content sharing destination.
Owner:SNAP INC

Electronic device and operation method thereof

The present disclosure relates to an artificial intelligence (AI) system and application thereof, which use a machine learning algorithm. An electronic device according to the present disclosure may include memory storing one or more instructions, and one or more processors configured to execute the one or more instructions stored in the memory, wherein the one or more processors are configured to transmit, to a server, request information that is obtained from at least one of situation information and metadata corresponding to content, the request information including input conversational text information, and receive, from the server, a recommendation result based on the request information.
Owner:SAMSUNG ELECTRONICS CO LTD

Work display method and apparatus, electronic device, storage medium, and program product

Embodiments of the present disclosure provide a work display method and apparatus, an electronic device, a storage medium, and a program product. The method comprises: displaying a target work in a work display page; in response to a page switching operation acting in the work display page, displaying a personal homepage of a target publisher, and displaying a position control in the personal homepage, the target publisher being a publisher of the target work, and the personal homepage being configured to display work items of works published by the target publisher; and in response to a first trigger operation acting on the position control, displaying a work item of the target work in the personal homepage, before the work item of the target work is displayed, the display of the position control being maintained in the personal homepage.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Artificial intelligence-based copy data generation method, device, equipment and medium

This application belongs to the fields of artificial intelligence and fintech, and relates to a method, apparatus, computer device, and storage medium for generating text data based on artificial intelligence. The method includes: acquiring a video to be processed; extracting an image sequence from the video and generating a comprehensive image vector corresponding to the image sequence based on an image encoder; extracting target audio from the video and generating a comprehensive audio vector corresponding to the target audio based on an audio encoder; extracting target text from the video and generating a comprehensive text vector corresponding to the target text based on a text encoder; combining the comprehensive image vector, comprehensive audio vector, and comprehensive text vector to obtain a target vector; decoding the target vector based on a decoder to obtain word data; and generating target text corresponding to the video based on the word data. Furthermore, the target text can be stored in a blockchain. This application improves the efficiency and accuracy of text generation based on a three-modal encoder and decoder.
Owner:PING AN TECH (SHENZHEN) CO LTD

Traffic accident scene analysis system

A traffic accident scene analysis system comprises a photographing unit, an input unit, a user interface and an arithmetic unit electrically connected with the photographing unit, the input unit and the user interface, the photographing unit photographs a scene image of an accident to output a scene picture, the input unit is used for transmitting record information, and the arithmetic unit is used for transmitting the record information to the user interface. The arithmetic unit is used for carrying out image segmentation on the scene picture to generate at least one accident vehicle information and one road information, carrying out word segmentation on the record information to generate a plurality of pieces of keyword information, and executing program data of a generative artificial intelligent model; and generating a preliminary analysis report containing an accident animation according to the at least one piece of accident vehicle information, the road information and the plurality of pieces of keyword information, and displaying the preliminary analysis report through the user interface.
Owner:FENG CHIA UNIVERSITY

A method and computer program product for flight travel booking

PendingCN122154981AImprove purchasing efficiencyImprove turnover rateVideo data browsing/visualisationReservationsSimulationComputer program
Embodiments of the present application provide a flight travel ticket booking method and a computer program product, applied to a travel service platform, comprising: determining a flight ticket booking scheme of a user, the flight ticket booking scheme comprising a departure place, a destination and at least one transfer place; generating a corresponding video according to the flight ticket booking scheme; the video at least comprising a flight trajectory corresponding to the flight ticket booking scheme dynamically displayed on a map; the flight trajectory at least comprising a flight trajectory of a trip and / or a flight trajectory of a return trip, the flight trajectory comprising flight trajectories of at least three segments of a trip displayed in sequence; obtaining the video corresponding to the flight ticket booking scheme, and displaying. The trip is presented to the user through multi-dimensional information of pictures and sounds, attracting the user while improving the purchase efficiency of the user for the travel vehicle ticket, and improving the transaction rate of the travel service platform.
Owner:浙江飞猪网络技术有限公司

Double camera streams

Image augmentation effects are provided on a device that includes a display and a camera. A first stream of image frames captured by the camera is received and an augmented reality effect is applied thereto to generate an augmented stream of image frames. The augmented stream of image frames is displayed on the display in real time. A second stream of image frames, corresponding to the first stream of image frames, is concurrently saved to an initial video file. The second stream of image frames can later be retrieved from the initial video file and the augmented reality effects applied thereto independently of the first stream of image frames.
Owner:SNAP INC

Material tag recommendation method and device, electronic equipment and storage medium

This disclosure relates to a method, apparatus, electronic device, and storage medium for recommending content tags. The method includes: in response to a content tag recommendation request from a target terminal, acquiring and sending a list of content tag combinations to the target terminal; the list of content tag combinations is obtained based on a tag combination recommendation model that performs tag combination prediction processing on the content tags and display association information corresponding to multiple historical content; the display indicator information corresponding to the content tag combinations in the content tag combination list meets the display indicator conditions; in response to a filtering request for the content tag combination list, extracting target object information from the filtering request; acquiring a sub-list of content tag combinations matching the target object information from the content tag combination list; and sending the sub-list of content tag combinations to the target terminal. According to the technical solution provided by this disclosure, accurate content tags can be recommended for content creation, improving content creation efficiency and resource utilization.
Owner:BEIJING DAJIA INTERNET INFORMATION TECH CO LTD

Electronic transaction activated augmented reality experience

Aspects of the disclosure relate to systems and methods for performing operations comprising: receiving, by a client device implementing a messaging application, a request to access a display of a plurality of augmented reality experiences; retrieving a plurality of identifiers for each of the plurality of augmented reality experiences; determining that a given augmented reality experience of the plurality of augmented reality experiences is associated with an access restriction; in response to determining that the given augmented reality experience is associated with the access restriction, modifying a given identifier of the plurality of identifiers that is associated with the given augmented reality experience; and generating a graphical user interface to be displayed on the client device, the graphical user interface comprising the plurality of identifiers including the modified given identifier.
Owner:SNAP INC

Interaction methods, devices, electronic devices, storage media, and computer programs

An interaction method, apparatus, electronic device, and storage medium. The method includes receiving a text display operation on a first media content including video content (S101), and in response to the text display operation, displaying a text sentence list of the first media content in a predetermined area, where the text sentence list includes at least two text sentences, and for each of the at least two text sentences, there is a corresponding audio sentence in the target audio data of the first media content (S102).
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

A knowledge graph-based cultural digital video content display method and system

The application discloses a culture digital video content display method and system based on a knowledge graph, processes cultural entities through a named entity recognition method, generates triples knowledge by combining visual feature coding and audio semantic alignment, forms a cultural knowledge graph, further analyzes the relationship between entities by using relationship extraction, fuses time sequence to organize custom order, determines the video content structure, automatically links video segments from the graph, constructs a main video plus knowledge graph panel double-view interface, supports synchronous display and inheritance relationship jump by clicking entity attributes, and integrates audio alignment to enhance interaction; according to the user navigation path, the graph is incrementally updated, the visual coding feedback is fused to adjust the entity priority, and an optimized propagation sequence is obtained. The application significantly improves the semantic association depth and interactive continuity of cultural data, realizes personalized content recommendation and dynamic propagation optimization, and finally promotes the immersive experience and efficient propagation of cultural inheritance.
Owner:HUNAN LABORATORY CULTURE MEDIA CO LTD

Selecting and providing digital components during content display

This disclosure relates to selecting and providing digital components during content display. Methods, systems, and apparatus for selecting, providing, and displaying one or more digital components during content display include computer programs encoded on a computer storage medium. The method may include identifying a plurality of digital components that can be presented on a client device. A maximum number of digital components that can be presented in a time slot of content and the duration of said time slot are determined. For each digital component, a score is generated based on the duration, location requirements, and the number of times said digital component can be provided within said time slot. A first set of digital components is selected based on the score and provided to the client device.
Owner:GOOGLE LLC

Film and television work network propagation public opinion analysis method

The invention relates to the technical field of network public opinion monitoring, and particularly discloses a film and television work network propagation public opinion analysis method, which comprises the steps of multi-modal data acquisition, theme tracking, sentiment analysis and public opinion situation awareness. According to the scheme, a mode of combining LDA topic modeling and hierarchical clustering is adopted, a tree hierarchical structure and a vocabulary evolution relation of a topic are constructed, potential topics of film and television works in network propagation public opinions are learned from word frequency co-occurrence, topic popularity is quantized and mapped to a timeline through dynamic popularity calculation, and therefore the topic popularity is obtained. The interpretability and operability of the result are ensured; the method comprises the following steps: quantifying themes of film and television works by synthesizing popularity fluctuation and emotion dispute, automatically screening key nodes in network propagation public opinions based on a theme tree, quantifying emotion intensity by fusing confidence, and aggregating the emotion intensity to a macroscopic theme and a time window, so as to construct a macroscopic and microscopic integrated dynamic public opinion map; and the sentiment analysis accuracy is greatly improved.
Owner:BEIJING INST OF CLOTHING TECH

Page processing method, apparatus and device, and storage medium

The present disclosure provides a page processing method, apparatus and device, and a storage medium. The method comprises: first, displaying a target interactive control on a first page, wherein a target video is played on the first page, and the target interactive control is used for triggering switching from the first page to a second page for display; and then, in response to a preset first operation for the target interactive control, controlling the target interactive control on the first page to switch from a first display state to a second display state.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Matching video content with podcast episodes

A system and method are provided for matching videos and podcast episodes. A data store comprising podcast episode identifiers is accessed. The podcast episode identifiers are associated with one or more podcast episode attributes. A video content item is identified. The video content item includes one or more video content item attributes. A matching podcast episode identifier that matches the video content item is determined based on the one or more podcast episode attributes and the one or more video content item attributes. A ranking of one of the video content items or the matching podcast episode identifiers is adjusted to reflect a correspondence between the video content item and the matching podcast episode identifier. Information associated with the matching podcast episode identifier is provided to a first user device.
Owner:GOOGLE LLC