Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

386results about "Video data browsing/visualisation" patented technology

Intelligent tool warehouse management system

The invention provides an intelligent tool warehouse management system, which belongs to the field of tool management and comprises a multi-mode sensing module, an intelligent access control module, an edge calculation module, an anomaly analysis module and a visualization module. The multi-mode sensing module realizes identity recognition, real-time positioning and integrity verification of the tool through RFID and UWB dual-mode positioning, visual recognition and sensor fusion technologies; the intelligent access control module is based on dynamic authority control and gravity sensing goods shelves, and tool storing and taking compliance is ensured; the edge computing module adopts localized data processing and low-power-consumption communication, and supports network disconnection disaster recovery; the abnormity analysis module identifies violation operation through an LSTM model and triggers grading alarm; the visualization module displays the tool state and the warehouse thermodynamic diagram in real time based on the digital twinning technology. The problems that traditional warehouse management depends on manpower, efficiency is low and errors are prone to occurring are solved, and high-precision tracking, intelligent early warning and optimal scheduling of the tool in the whole life cycle are achieved.
Owner:TAIYUAN LONGWAY ELECTRONICS SCI & TECH

Video generation method, electronic device, and computer readable storage medium

Provided are a video generation method, an electronic device, and a computer readable storage medium, relating to the technical fields of computers and video processing. The video generation method comprises: acquiring a target text, wherein the target text is used for describing video content to be generated (S21); and using a target video generation model to perform video generation processing on the target text to obtain a target video, wherein the target video generation model is a model obtained by performing model alignment on an initial video generation model and a preset reward model in a fine-tuning manner (S22). The present invention solves the technical problems in the related art that a video generated by a video generation model obtained by training based on network data has poor quality and does not meet user expectations.
Owner:ALIBABA (CHINA) CO LTD

Visual display method and system applied to security and protection and storage medium

The invention relates to the technical field of security and protection visualization, in particular to a security and protection visualization display method and system and a storage medium. The method comprises the following steps: acquiring a data stream acquired by security and protection equipment, and compressing the data stream into a unified acquisition frame set; identifying a road boundary point set in the unified collection frame set, and detecting a dynamic target and a static target at each moment in the unified collection frame set; performing semantic annotation according to the motion state change condition of the dynamic target at each moment in the unified collection frame set to obtain a road boundary semantic distribution map; according to the road boundary semantic distribution map, identifying and predicting the motion state change condition of the static target; and predicting an overlapping time point of the motion state of the target based on the motion state change condition of the static target and the motion state change condition of the dynamic target, and transmitting the overlapping time point to communication equipment of the static target through the Internet of Things so as to execute a voice prompt task. According to the invention, the display precision and response efficiency of security and protection visualization can be improved.
Owner:JINGGANGSHAN YUJIE FIRE SCI & TECH

Methods and systems for automated content generation

An aspect of the disclosure related to methods and systems configured to distribute streaming content and to enable users to discover content. A user interface is rendered on a user device, comprising thumbnail representations of longform content items arranged in rows, wherein a first row comprises a first set of thumbnail representations of longform content items identified as a first category and a second row comprises a second set of thumbnail representations of longform content items identified as a second category. A plurality of shortform video previews corresponding to at least a portion of the longform content items identified as the first category, A first shortform video automatically begins playing. A second shortform video begins playing in response to a first event. In response to a user interaction while the second shortform video is displayed, the second longform content item is streamed to and displayed by the user device.
Owner:PLUTO INC

Information display method and device, electronic equipment, readable storage medium and program product

The invention relates to an information display method and device, electronic equipment, a computer readable storage medium and a computer program product, and relates to the technical field of Internet of Things. The information display efficiency of the time axis of the information display page to the event can be improved. The method comprises the steps that an information display page is displayed, the information display page comprises a time axis corresponding to monitoring data, the monitoring data comprises monitoring data obtained through monitoring of target equipment and / or associated equipment, and the associated equipment is equipment associated with the target equipment; and if the target event exists, according to the event time and / or the event type of the target event, an event identifier and event information corresponding to the target event are displayed on the time axis according to the target style, and the target event is obtained by identifying the monitoring data.
Owner:SHENZHEN LUMIUNITED TECH CO LTD

Unmanned aerial vehicle video abstract semantic description method and system based on multi-modal large model

The invention relates to the technical field of unmanned aerial vehicle video data interpretation, in particular to an unmanned aerial vehicle video abstract semantic description method and system based on a multi-modal large model, and the method comprises the steps: obtaining a plurality of segmented video frame images of unmanned aerial vehicle video data; extracting image features by using a multi-modal large model, wherein the multi-modal large model adopts an image encoder in a visual language basic model to encode the input segmented video frame images and extract the corresponding image features; performing adaptive clustering on the image features to obtain a clustering center of each segmented video, and generating an unmanned aerial vehicle video abstract by taking a frame position where the clustering center is located as a frame position where the video abstract is located; and acquiring scene semantic description of the unmanned aerial vehicle video abstract by utilizing a semantic description model, wherein the semantic description model is obtained by finely adjusting a large model by utilizing an unmanned aerial vehicle image semantic description data set. The core intelligence information can be accurately and efficiently extracted from the unmanned aerial vehicle video data, and the utilization efficiency of the unmanned aerial vehicle video data is improved.
Owner:Chinese People's Liberation Army Cyberspace Force Information Engineering University

Short video retrieval method combining pre-training model and dependency syntax tree

The invention discloses a short video retrieval method combining a pre-training model and a dependency syntax tree, and relates to the technical field of multi-modal learning, computer vision and information retrieval cross fusion. The semantic information in the context is captured by using the advantages of BERT, and the deep semantic representation is generated, so that the meaning of the sentence is more accurately understood. A clear syntactic relationship is provided by using the dependency syntactic tree, and the structured dependency relationship between words can be revealed. The relationship can help the model to better understand logic semantics in sentences and extract text features related to tasks in a more targeted manner. Through the dependency syntax tree and the model, primary and secondary information in the long sentences can be distinguished more accurately, the problems of word order ambiguity and the like are eliminated, feature extraction is more stable, and robustness is improved. The dependency syntactic tree is explicit grammatical representation, and structured information of the dependency syntactic tree can assist in analyzing BERT implicit features, so that more explainable basis is provided for the decision process of the model, and the interpretability is remarkably enhanced.
Owner:NORTHEASTERN UNIV CHINA

Emotion distribution display method and device of video

The embodiment of the invention provides a video emotion distribution display method and device, and relates to the technical field of computers. The method comprises the following steps: acquiring sentiment analysis data of a plurality of bullet screens corresponding to a target video, wherein the sentiment analysis data comprises sentiment tags of the bullet screens and creation time of the bullet screens; determining emotion distribution data of the target video according to the emotion analysis data of each bullet screen; and pushing the emotion distribution data to a user side, so that the user side displays the emotion distribution data. According to the technical scheme provided by the embodiment of the invention, the user can intuitively know the overall atmosphere and emotion trend of the story content of the target video currently watched by the user.
Owner:SHANGHAI BILIBILI TECH CO LTD

Landscape interaction method and apparatus, electronic device, and storage medium

Provided are a landscape interaction method and apparatus, an electronic device, and a storage medium. The method includes: receiving a first trigger operation acting on a first display area of a landscape playback page, where the first display area includes a display area of author information of a current video author, and the current video author is a publisher of a first target video in the landscape playback page; and in response to the first trigger operation, displaying a target-video list of the current video author in a second display area of the landscape playback page, where the target-video list includes at least one video item of at least one target video published by the current video author.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Method for generating information, method for displaying information, device, and storage medium

The present application provides a method for generating information, a method for displaying information, a device, and a storage medium. Description data corresponding to a first video is acquired, the description data representing historical comments on the first video; the description data is processed by means of a language model, and a virtual text is generated, the content of the virtual text comprising comments on target content in the first video based on a virtual character identity; and a virtual object for voice reading the virtual text is created, a second video is generated on the basis of the virtual object, and the second video is played back on one side of a terminal device. Historical comments on a first video are converted, on the basis of a language model, into a second video and said video is played back, so that the historical comments on the first video are displayed in the form of a video after being understood and sorted by means of the language model, thereby making the video content more concise and effective, and improving the efficiency and effect of displaying comment information.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Method and system for three-dimensional modeling of chemical industry park

The invention discloses a three-dimensional modeling method and system for a chemical industrial park, and the method comprises the steps: carrying a multi-lens inclined camera and a monitoring sensor through employing an unmanned plane, and obtaining the multi-view image data, positioning and attitude determination data of the chemical industrial park, and the real-time operation data of chemical equipment; performing data processing on the multi-view image data to generate sparse point clouds, and generating three-dimensional point clouds through dense matching and depth estimation; performing block processing on the three-dimensional point cloud, and performing splicing by adopting a point cloud registration algorithm to obtain global point cloud data; performing surface reconstruction on the global point cloud data to generate a three-dimensional grid model, and completing model texture mapping through a texture mapping algorithm; a three-dimensional visual management platform is constructed, a three-dimensional grid model and real-time monitoring data are integrated, and park global visualization, risk visualization and intelligent management of chemical engineering devices, equipment and facilities are achieved.
Owner:BEIJING UNIV OF CHEM TECH

Video generation method, device and equipment, computer readable storage medium and product

The embodiment of the invention provides a video generation method and device, equipment, a computer readable storage medium and a product, and the method comprises the steps: obtaining material contents determined by a user in response to a video generation operation triggered by the user; in response to a preview operation triggered by a user, displaying a story text content used for generating the target video in the first display interface, the story text content being generated by the large language model based on the material content; in response to a split picture generation operation triggered by a user, displaying a preset number of split pictures generated based on the story text content and preset first prompt content in the first display interface; and in response to a video generation operation triggered by a user, generating a target video based on the preset number of split pictures. Therefore, the user only needs to select the material content, the high-quality target video meeting the actual requirements of the user can be generated, and the generation efficiency of the target video is effectively improved.
Owner:BYTEDANCE TECHNOLOGY CO LTD +1

Video picture search split-screen interaction method and terminal based on multi-mode large model

The invention discloses a multi-mode large model-based video picture search split-screen interaction method and a terminal, and belongs to the technical field of intelligent terminal video interaction, and the method comprises the following steps: when a question search function is triggered and started, controlling to obtain a question search instruction, and simultaneously controlling to intercept video picture information of a front and back predetermined frame when the question search function is triggered and started; identifying the question search instruction, and identifying the intention of a user to search for video picture information; performing element identification on the intercepted predetermined frame of video picture information through a multi-mode large model, and finding out elements matched with the intention of a user to search for the video picture information; according to the found elements matched with the intention of the user and the identified intention of the user for searching the video picture information, automatically searching a search result through a multi-mode large model; and displaying through a preset split-screen interaction interface. According to the method and the device, a video picture can be deeply understood, generation type dialogue interaction can be performed on a user, and convenience is provided for use of the user.
Owner:NANJING KUKAI SMART SCREEN TECH CO LTD

Tax risk display method, system, equipment and medium

The invention provides a tax risk display method, system and device and a medium, and belongs to the field of tax risk analysis. The method comprises the following steps: extracting multi-modal data contained in cargo flow, contract flow, fund flow and invoice flow in a business process, constructing a transaction relation knowledge graph, and associating the multi-modal data with entities in the graph through a mapping rule base; performing consistency detection on the cargo flow, the contract flow, the fund flow and the invoice flow, determining a risk level in combination with historical risk events in the risk knowledge base, matching countermeasures from the risk knowledge base, associating tax regulations from the tax knowledge base, and generating a risk report; and generating video clips by adopting a text video technology, and splicing and merging the video clips to generate a target video. According to the method, the transaction relation knowledge graph is constructed through multi-modal data extraction, the risk indexes are quantified, the risk report is generated, and finally the result is converted into the video through the text video technology, so that the visual display of the tax risk is realized, and the accuracy and response efficiency of risk identification are improved.
Owner:INSPUR GENERSOFT CO LTD

Video retrieval method and device

The embodiment of the invention provides a video retrieval method and device, computer equipment, a computer readable storage medium and a computer program product, and relates to the technical field of video processing. The video retrieval method comprises the following steps: displaying a video editing interface, wherein a search entry is configured on the video editing interface; receiving target sub-mirror description information through the search entry; according to the sub-mirror description information, retrieving a corresponding target video slice from a preset local material retrieval table; wherein the local material retrieval table comprises video retrieval information of each video slice output based on the large language model. According to the technical scheme provided by the embodiment of the invention, the required video slice can be accurately retrieved from the local video material.
Owner:SHANGHAI BILIBILI TECH CO LTD

Information processing device, information processing method, and program

An information processing apparatus includes a control unit that acquires predetermined information from a specific storage that commonly manages the predetermined information used in a plurality of services each of which is capable of browsing information via application software or a website, and executes information presentation using the predetermined information by each of the plurality of services.
Owner:SONY GROUP CORP

Video recommendation method and device, electronic equipment, storage medium and program product

The invention relates to a video recommendation method and device, electronic equipment, a storage medium and a program product, and the method comprises the steps: displaying video recommendation information corresponding to at least one first video in a preset recommendation page corresponding to a plurality of first videos, and carrying out the preview playing of a first target video in the preset recommendation page based on a first preset magnification; the first preset magnification is greater than one, the first target video is any video in the plurality of first videos, and the at least one first video is a video except the first target video in the plurality of first videos. According to the embodiment of the invention, the video recommendation efficiency can be better improved while the video browsing efficiency of the user is improved, the user is helped to quickly find the interested video, the situation of mistakenly clicking the uninterested video is effectively reduced, the system resource waste caused by unnecessary page jump can be effectively reduced, and the system performance is improved.
Owner:BEIJING DAJIA INTERNET INFORMATION TECH CO LTD

Contextual digital media processing systems and methods

Systems and methods for contextual digital media processing are disclosed herein. An example method includes receiving content from a source as digital media that are being displayed to a user, processing the digital media to determine contextual information within the content, searching at least one network for supplementary content based on the determined contextual information, and transmitting the supplementary content for use with at least one of the source or a receiving device.
Owner:SONY INTERACTIVE ENTERTAINMENT LLC

Interaction method and apparatus, electronic device, and storage medium

The present disclosure relates to an interaction method, an apparatus, an electronic device and a storage medium. The method comprises: receiving (S101) a text display operation for a first media content, wherein the first media content includes a video content; in response to the text display operation, displaying (S102) in a preset region a list of text sentences of the first media content, wherein the list of text sentences contains at least two text sentences, and the at least two text sentences each have a corresponding audio sentence in target audio data of the first media content.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

System and method for analyzing videos

The present invention relates to a system and method for analyzing videos. The system comprises a video section and a tabs section, wherein the video section comprises a video and the tabs section comprises at least one tab with items or information related to the video. In most embodiments one of the tabs is a products tab which lists items in the video which can be purchased. These items are associated with specific time points in the video, generally when the item is visible in the video. When a user watches the video, the individual products will be highlighted as their time points are viewed. The user can additionally scroll through the list of products without effecting video playback. The method comprises the gathering and uploading of information associated with the video.
Owner:VIKTRS LTD

Camera quick calling and visualization method based on natural language understanding

The invention discloses a camera quick calling and visualization method based on natural language understanding, and relates to the technical field of video monitoring and artificial intelligence crossing. Comprising the steps that S1, a full-amount camera database is constructed, S2, a camera space retrieval interface is provided, and S3, a map visualization front end is constructed; s4, providing an address-coordinate conversion service; s5, local privatization deployment: deploying a camera space retrieval interface, a front-end static resource and a large language model mirror image in an intranet GPU server; s6, constructing a large model knowledge newly establishing a knowledge base table Knowledgeprompt in the model mirror image for writing data, and constructing the large model knowledge base, and S7, receiving a query demand of a user for inputting a natural language in a dialog box; s8, performing semantic understanding, and calling an interface to return data; step S9, automatically dotting the map; and S10, playing the video stream as required.
Owner:浪潮智慧城市科技有限公司

Intelligent teaching visualization system and method for college Chinese

The invention discloses an intelligent teaching visualization system and method for college Chinese, and relates to the technical field of intelligent teaching. Comprising a video preprocessing module used for acquiring teaching data including teaching videos, teaching outlines and teaching plan data; performing comprehensive analysis on the teaching data to obtain a segmentation evaluation value, and performing segmentation processing on the teaching video through the segmentation evaluation value to obtain teaching video segments; the progress tracking module is used for collecting operation data and video data when each student watches the teaching video, and comprehensively analyzing the operation data and the video data to obtain an operation evaluation value; and the behavior judgment module is used for setting a judgment threshold value, judging the operation evaluation value according to the judgment threshold value to obtain a judgment result, and transmitting the judgment result to the feedback module. According to the method, the problems that the learning progress record does not accord with the reality due to non-standard operation of students and teachers are difficult to accurately grasp learning conditions are solved, and accurate analysis of the learning conditions and personalized teaching are realized.
Owner:HEILONGJIANG COMM POLYTECHNIC

Structured video documents

A method includes receiving a content feed that includes audio data corresponding to speech utterances and processing the content feed to generate a semantically-rich, structured document. The structured document includes a transcription of the speech utterances and includes a plurality of words each aligned with a corresponding audio segment of the audio data that indicates a time when the word was recognized in the audio data. During playback of the content feed, the method also includes receiving a query from a user requesting information contained in the content feed and processing, by a large language model, the query and the structured document to generate a response to the query. The response conveys the requested information contained in the content feed. The method also includes providing, for output from a user device associated with the user, the response to the query.
Owner:GOOGLE LLC

Method, computer device, and storage medium for generating video cover

A method for generating a video cover is performed by a computer device and includes acquiring a video title and a candidate video cover of a target video; determining highlighted characters of the video title; determining typesetting parameters of the highlighted characters based on the highlighted characters and a cover parameter of the candidate video cover; and generating a target video cover of the target video by rendering the highlighted characters to the candidate video cover based on the typesetting parameters.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Transformation of database entries for improved association with related content items

A content analysis system includes processor and memory hardware storing data analyzed content items and instructions for execution by the processor hardware. The instructions include, in response to a first intermediate content item being analyzed to generate a first text description, receiving the first intermediate content item and analyzing the first text description to generate a first reduced text description. The instructions include identifying a first set of tags by applying a tag model to the first text description and generating a first analyzed content item. The instructions include adding the first analyzed content item to the analyzed content database and, in response to a displayed content item being associated with at least one tag of the first set of tags, displaying a first user-selectable link corresponding to the first analyzed content item on a portion of a user interface of a user device displaying the displayed content item.
Owner:CHARLES SCHWAB & CO INC

Human-machine interaction method and apparatus, and electronic device

Embodiments of this application provide a human-machine interaction method and apparatus, and an electronic device. In the method, a first interface is displayed in a gallery application in response to a first operation of a user on a thumbnail that is of a video and that is displayed in a first album or a photo tab. A picture of the video is displayed in a display region of the first interface. The first interface further includes an association region and an operation region. The association region includes an identifier of the video and an identifier of at least one photo associated with the video. The at least one photo is generated based on the video. The operation region includes a share control. The video is shared with a target terminal in response to the user operating the share control. According to embodiments of this application, a video can be shared in a scenario in which an associated photo is generated based on the video.
Owner:HONOR DEVICE CO LTD

A method, device and system for storing and synchronously triggering augmented reality events in a panoramic video

This invention provides a method for storing augmented reality events in panoramic videos, a method for synchronously triggering events, and an apparatus. The method uses time-triggered special effects or footage in the panoramic video as trigger events. It uses the ID of the trigger event as the key and the content of the trigger event as the value to create an index for a linked list. Each trigger event is stored in the linked list in chronological order. Simultaneously, a balanced tree is built using the trigger time of each event as the key and the ID of each trigger event as the value. To insert or delete a trigger event, the method searches the balanced tree to obtain the IDs of time-adjacent trigger events, then finds the corresponding position in the linked list based on that ID for insertion or deletion. It also allows for timed triggering of each event at specific times. This invention enables convenient and quick insertion, deletion, and synchronous triggering of events in panoramic videos. By using a balanced tree to store events, it greatly reduces the computational requirements during event lookup and improves the synchronization during panoramic video playback.
Owner:BEIJING UNIV OF POSTS & TELECOMM

Retrieving images for video scrubbing at a client device

Methods, software, devices and systems for video scrubbing enable a client device to retrieve images for scrubbing based on a user-requested time along a video timeline of a video stored in a server. The client device checks if a cached image meets specified conditions, including a timestamp within a precision margin around the requested time. The precision margin scales with the timeline length, providing a smaller margin for shorter timelines and a larger margin for longer timelines. If a relevant cached image is found, it is retrieved; if not, an image with a highest relevance score within the precision margin is fetched from the server and stored in memory.
Owner:AXIS

Methods, devices, equipment, and storage media for finding videos

This application discloses a method, apparatus, device, and storage medium for finding videos, belonging to the field of human-computer interaction. The method includes: displaying a video list interface showing video covers of at least two videos; responding to a trigger operation, displaying a scrolling selector on the video list interface, the scrolling selector displaying list options for the at least two videos, the list options being used to display relevant text about the videos; and responding to a selection operation of a first list option among the at least two video list options, displaying the target video. This application addresses the problem of users finding it difficult to quickly and accurately find target videos on a video list interface when the video covers are default icons or do not carry sufficient valid information, allowing users to quickly find the videos they wish to watch using the text content of each video.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD