Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

455 results about "Live video" patented technology

Live broadcast behavior tracking system based on deep learning

The invention relates to the technical field of live broadcast behavior monitoring, in particular to a live broadcast behavior tracking system based on deep learning, which obtains high-quality multi-source information and improves the accuracy of feature analysis by synchronously extracting image and audio data from a live broadcast video stream and combining frame extraction, image enhancement and voice recognition. According to the method, image features are extracted through a pre-trained convolutional neural network, audio features are extracted through a deep learning model, voice transliteration texts are fused, the weight of each modal feature is dynamically adjusted based on an attention mechanism, precise recognition of complex scenes and hidden violation behaviors is achieved, and image camouflage and latent language expression risks are effectively coped with. And the illegal type and confidence are output in real time, once suspected illegal behaviors are detected, alarm, interruption or shielding operation is triggered immediately, and related evidences are uploaded to an auditing database. And efficient, accurate and full-process management and control of the live broadcast violation behaviors are realized.
Owner:GUANGZHOU QUNGE INFORMATION TECHNOLOGY CO LTD

AR glasses blogger live video anti-shake processing method and processing system

The invention discloses an AR glasses blogger live video anti-shake processing method and processing system, and the method specifically comprises the steps: carrying out the calculation based on an original video stream and equipment motion data, and generating a preliminary stable frame sequence; based on the preliminary stable frame sequence, marking a key semantic region as a strong stability demand region, and segmenting a foreground main body and a background region; performing fine stabilization on the strong stability demand area and the foreground main body by applying non-rigid transformation, performing stabilization on the background area by applying rigid transformation and retaining the depth of the scene, and obtaining a frame sequence after partition correction; predicting a future multi-frame equipment motion track based on the frame sequence after partition correction, the equipment motion data and the segmentation information of the foreground main body / background region; and calculating and executing prospective canvas motion compensation in advance based on the future multi-frame equipment motion trail and the segmentation information of the foreground main body / background region. According to the invention, efficient anti-shake processing of the live video of the AR glasses is realized, and the stability of the live picture and the user experience are effectively improved.
Owner:东莞市三奕电子科技股份有限公司

Video coding method and device, electronic equipment and storage medium

The invention provides a video coding method and device, electronic equipment and a storage medium, relates to the technical field of image processing, in particular to the field of video coding and the like, and can be applied to application scenes such as video live broadcast and the like. The specific implementation scheme is as follows: acquiring a reference frame which is adjacent to a current frame and has established initial ROI hierarchical division; multiplexing a motion vector generated by the reference frame in a motion compensation time domain filtering process, and determining an ROI position predicted value of the current frame; extracting feature points of the five sense organs of the current frame, and correcting an ROI position predicted value based on a feature point spatial topological relation in the initial ROI; respectively configuring differential quantization parameters for the corrected main face region, the corrected secondary face region and the corrected background region; dynamically allocating a three-layer area code rate according to a real-time network bandwidth; and outputting the coded frame of the current frame and the associated ROI level metadata. According to the scheme, the coding efficiency and quality can be improved.
Owner:BEIJING BAIDU NETCOM SCI & TECH CO LTD

Intelligent global cloud live broadcast system based on AI and multi-operator and multi-cloud optimization

The invention relates to the technical field of network communication and video processing, in particular to an intelligent global cloud live broadcast system based on AI and multi-operator and multi-cloud optimization, and the system comprises a data collection module which is used for collecting live broadcast video streams; the video compression module adopts H.265 coding; the content review module is used for performing multi-modal review; the RPMTS packaging module is used for accessing three operator cloud private line POP nodes through a public network after packaging; the decoding module is used for decoding POP nodes in the middle area; the translation module is used for generating multi-language subtitles or dubbing; the multi-operator and multi-cloud adaptability matrix module calculates an optimal path in combination with an AI algorithm and a multi-cloud cost adaptability matrix; the multi-branch access module is used for realizing branch node cooperative forwarding; and the multi-language distribution module is used for generating multi-language subtitles and dubbles and pushing the multi-language subtitles and dubbles to each global live broadcast platform, so that synchronous distribution is realized, and the system can reduce global live broadcast delay, improve multi-language coverage capability and realize high reliability and cost optimization.
Owner:HAIJIAO CLOUD (SHENZHEN) INFORMATION TECHNOLOGY CO LTD

User Authentication, Spoofing and Replay Attack Prevention, Liveness Detection, and User-and-Document Verification using a Live Video Stream with Spatial Challenges

User authentication, spoofing and replay attack prevention, liveness detection, and user-and-document verification using a live video stream with spatial challenges. A camera of an electronic device captures and transmit a live selfie user-facing video, as part of a user registration process. The user is instructed to spatially move his body or face, such that his face would appear within a first particular on-screen shape; and to also, concurrently or simultaneously, spatially hold in his hand or move a particular an identification document such that it would appear within a second on-screen shape. Optionally, the on-screen shape moves on the screen, and the user is required to spatially move the relevant item to keep it within the boundaries of the moving on-screen shape. The system then analyzes the video via computerized vision, to determine whether the user complied with the spatial manipulation challenges.
Owner:IRONVEST INC

Instant check conversion

A computer implemented method, system, and non-transitory computer-readable device that may be used in a remote deposit environment. Upon receiving a user request, based on interactions with the UI, the method implements an electronic deposit of a financial instrument by activating a camera on the client device to generate a live video stream of image data of a field of view of at least one camera, wherein the live video stream includes imagery of at least a portion of each side of the financial instrument. The method continues by extracting data fields based on the formation of image objects on one or more sides of the financial instrument from the live video stream of image data. An EFT conversion of extracted data fields may be processed during or subsequent to the extraction process. A message is sent from a payee to a payor requesting the EFT. Upon acceptance, an EFT to the payee occurs. Upon denial, the remote deposit process is completed.
Owner:CAPITAL ONE SERVICES LLC

CDN offload via hybrid delivery over ATSC 3.0 for live video streaming

One of the promises of ATSC 3.0 has been the potential for data offload. Separately, advances have been made in hybrid and IP channel rollouts on ATSC 3.0 over the last year. The two have been combined to architect and implement a data offload system. This application explores a practical hybrid delivery model for streaming video services, allowing the distribution of video over CDNs and simultaneously over 3.0, drastically decreasing the bandwidth needs of these CDNs in 3.0 markets, all while integrating into third-party streaming apps to enable seamless streaming with no change to the viewer experience. Topics covered include the methodology for synchronization, encoding needs for the CDN and airchain, signaling design, integration into a streaming application, and the results of the real-world testing of this system.
Owner:ONE MEDIA LLC

Method and system for extracting inherent user feature using artificial intelligence

Disclosed is a computer-implemented method and system for training a subject-specific machine learning model to infer inherent subject features from recorded or live video data. The system preprocesses the visual and audio channels, converting audio to text, and employs multiple pre-trained extraction models to generate feature embeddings. Ground truth data is obtained to guide training, where weights are assigned to produce and combine predicted feature values. Model performance is optimized by minimizing error. The trained feature extraction models are deployed on an edge device, while the subject-specific model resides in the cloud. A lightweight edge model, derived via knowledge distillation and model compression, supports local inferencing with reduced reliance on cloud resources. Synchronization ensures iterative updates for sustained accuracy.
Owner:MOODMETRICS AI

Method and system for text search capability of live or recorded video content streamed over a distributed communication network

A server receives and rebroadcasts live streaming video content from a video capture device, such as a mobile phone or unmanned surveillance vehicle. The server includes a media server configured to stream selected video content to a client device, a video analysis system configured to analyze the live video content and generate object detection data, a storage system configured to store the generated object detection data and an identifier of the associated live video content, and a search engine configured to receive a text-based search request, search the object detection data stored in the storage system for relevant search results, and generate a list of live and stored video content associated with the relevant search results.
Owner:AERYON LABS

Information display method and device, equipment, storage medium and product

The invention provides an information display method and device, equipment, a storage medium and a product, and relates to the technical field of computers. The method comprises the steps that in a first device, a first object information set is displayed in a first preset page, the first object information set comprises first object information of a preset object in a preset object set, the preset object set is associated with a target live video interface, and target prompt information is displayed at the position adjacent to the first object information of the first target preset object, the business data of the first target preset object meets a preset requirement, the target prompt information is used for prompting that the first target preset object is set as a preset object with a target object attribute, and second object information of the preset object with the target object attribute contains a target label and is displayed at a target position; the second object information set is displayed on a second preset page associated with the target live video interface in the second device. By adopting the technical scheme, the setting efficiency of the target object attribute of the preset object can be improved.
Owner:BEIJING YOUZHUJU NETWORK TECH CO LTD

Live broadcast playback video downloading method and device and electronic equipment

The invention provides a live broadcast playback video downloading method and device and electronic equipment, and relates to the technical field of communication, and the method comprises the steps: carrying out the fragmentation task division of a live broadcast playback video based on a downloading task and a dynamic fragmentation strategy, and obtaining a plurality of target fragmentation tasks; each target fragment task corresponds to one to-be-downloaded video fragment and a serial number of the to-be-downloaded video fragment; executing a plurality of target fragmentation tasks based on a concurrency control mechanism and carrying out downloading detection; if it is detected that the video fragment corresponding to the target fragment task fails to download, retry processing is performed on the failed target fragment task based on a retry mechanism; merging and format conversion are carried out on the downloaded multiple video fragments to generate the target playback video file, the technical problems of low network bandwidth utilization rate and low downloading efficiency are solved through the method, and efficient and stable live video downloading is realized.
Owner:SHENZHEN HAIGUI NETWORK TECH CO LTD

Enhanced interactive features for a video presentation system

A system, related operating methods, and computer readable storage media are disclosed here. The disclosed subject matter relates to methods of providing augmented reality features in connection with presentation of video content. A disclosed method involves: causing presentation of a video program at a user device associated with a viewing user; obtaining metadata associated with the video program; processing a live video feed that includes the viewing user; generating augmented reality overlay content for the live video feed, wherein visual appearance of at least some of the augmented reality overlay content is influenced by the obtained metadata; and causing presentation of an enhanced version of the live video feed at the user device associated with the viewing user, the enhanced version of the live video feed comprising the augmented reality overlay content.
Owner:DISH NETWORK TECHNOLOGIES INDIA PTE LTD

Information processing method and system based on unmanned aerial vehicle video spatialization

The invention provides an information processing method and system based on unmanned aerial vehicle video spatialization, and relates to the technical field of unmanned aerial vehicle video processing. Firstly, camera optical parameters, camera pose information and sensor type identification are collected and synchronously coded to a non-display data segment of an unmanned aerial vehicle video frame structure to form a coded video with multi-dimensional space information, and then a coded video stream is formed through streaming packaging. And performing scene-based transcoding adaptation on the coded video stream to obtain an adaptive video stream, performing hierarchical decoding and splitting, constructing a dynamic space mapping model in combination with an image correction model, and calculating the geographic range of a video frame to obtain a live video with a multi-reference geographic range identifier. Based on the live video, geographic data with ground feature attributes and recognition confidence are generated and transmitted in a multi-link mode, a feedback modification track is received and layered updating processing is executed, bidirectional dynamic synchronization and conflict resolution of the geographic data and the live video are achieved, and the quality and reliability of geographic information are improved.
Owner:JILIN PROVINCIAL PUBLIC SECURITY BUREAU

Video coding method, live video coding method, video processing system, computing device, computer readable storage medium and computer program product

Embodiments of the present specification provide a video coding method, a live video coding method, a video processing system, a computing device, a computer readable storage medium and a computer program product, the video coding method comprising: acquiring video attribute information and target video quality information of a to-be-coded video, the to-be-coded video comprising a plurality of video frames; under the condition that the video frames are forward prediction coding frames, determining coding information corresponding to the plurality of video frames according to the video attribute information; determining quantization parameters respectively corresponding to the plurality of video frames according to the target video quality information and the coding information; and coding the plurality of video frames according to the quantization parameters to obtain a video coding result of the video to be coded. By modeling the relationship between the target video quality information and the quantization parameter, the quality information of the video coding result is as close as possible to the target video quality information, the video coding precision is improved, and the code control precision is ensured by calculating the quantization parameter frame by frame.
Owner:HANGZHOU ALICLOUD FEITIAN INFORMATION TECH CO LTD

Method and device for evaluating live broadcast content quality, medium and program product

The invention provides a method and device for evaluating live broadcast content quality, a medium and a program product. The method comprises the following steps: in response to a live broadcast content quality evaluation instruction, obtaining a target live broadcast video and attribute related information of the target live broadcast video; obtaining a quality scoring rule matched with the target live video according to the attribute related information; performing feature extraction processing to obtain one or more evaluation feature data corresponding to the plurality of evaluation dimensions; and based on the obtained evaluation feature data, according to the quality scoring rule, performing calculation to obtain scores of a plurality of evaluation dimensions and a comprehensive quality score, so as to evaluate the live broadcast content quality based on the comprehensive quality score. The live broadcast content quality is evaluated through multiple evaluation dimensions, the overall quality condition of live broadcast can be reflected more accurately, the victory or defeat judgment of the live broadcast PK is carried out based on the multi-dimensional comprehensive score, the PK content quality can be guided and improved, and the participation enthusiasm of the user is improved.
Owner:SHANGHAI BILIBILI TECH CO LTD

Touch controlled, low-latency, laser pet toy

A touch screen controlled light emitting pet toy that has a light emitting device equipped to direct a point of light on a surface to provide entertainment to a pet and includes movement about multiple axis and is operated by a touch screen computer based device communicably connected to a controller of the toy. The toy is equipped to receive user input from an application with a touch-screen directly or indirectly via a camera that provides live video data to the application and relays user input to the toy further allowingdikect control and pre-recorded pattern control of the light emitting device through low latency touch screen interactions.
Owner:HIEBER AARON M

A display method, apparatus, electronic device, computer readable medium

This application discloses a display method, apparatus, electronic device, and computer-readable medium. The method includes: when a client is displaying a live video page and a first live video is displayed on the live video page, after the client receives a first operation triggered on the live video page, displaying an object aggregation interface on the live video page, and displaying a first candidate object corresponding to first category description information on the object aggregation interface, the first candidate object being used to describe the first candidate live video; then, after the client receives a trigger operation on a first control on the object aggregation interface, displaying at least one second category description information on the object aggregation interface, so that the user can view live videos under multiple categories through the object aggregation interface, thereby better meeting the user's live video viewing needs and effectively improving the user's live video viewing experience.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Audio noise reduction method, system and device based on AI model and storage medium

The invention discloses an audio noise reduction method, system and device based on an AI model and a storage medium, and relates to the technical field of voice noise reduction, and the method comprises the steps: receiving a live video stream, and segmenting the live video stream into image frame data and audio stream data; the image frame data and the audio stream data are input to a preset scene recognition model, the recognition result of the current scene is obtained, and the preset scene recognition model comprises a character recognition sub-model, a voice recognition sub-model and a classification output module; and determining and executing a current noise reduction strategy according to the identification result to obtain an audio output signal, the current noise reduction strategy being a human voice enhancement strategy, an ambient sound enhancement strategy or a fusion enhancement strategy. According to the invention, scene recognition is carried out in a multi-modal fusion mode, different noise reduction strategies are adaptively executed according to the scene recognition result, noise optimization can be carried out for different scenes, and the audio output effect is improved.
Owner:NANJING PUTIAN TELEGE INTELLIGENT BUILDING

Systems and methods for managing and displaying handheld objects in video capture

The present application provides for a video system that displays handheld objects in video capture by detecting objects being held by the user within a segmentation mask and adding the object to an allowed list of permissibly visible objects. The system may be configured, during live video capture, to show or hide selected objects. As an added advantage, such approaches are not dependent on whether the hand is holding the object and moving synchronously. For example, if a video application on a video system included a pen in the allowed list, and the user waves and rotates the pen such that the pen moves out of sync with the body of the user, the video system will still make the pen visible. Moreover, the video application provides an efficient object detection mechanism that detects the handheld object once the object is placed within the segmentation mask.
Owner:ADEIA GUIDES INC

Live broadcast background replacement real-time optimization method and system based on portrait tracking

The invention provides a live broadcast background replacement real-time optimization method and system based on portrait tracking, and relates to the technical field of image processing, and the method comprises the steps: obtaining a real-time picture frame in a live broadcast video stream, constructing a feature pyramid, carrying out the portrait detection, obtaining an initial portrait region, inputting the initial portrait region into a dual-path tracking network, and obtaining a target tracking frame; inputting the RGB image of the corresponding area of the target tracking frame and the corresponding depth map into a bimodal segmentation network to generate a foreground mask, and correcting the foreground mask to obtain a portrait segmentation result; and fusing the portrait segmentation result with a preset background, and executing time sequence smoothing processing to obtain a background replacement result.
Owner:HANGZHOU ZERO ONEBIT TECHNOLOGY CO LTD

Video live broadcast content translation method and device, electronic equipment and storage medium

The invention provides a video live broadcast content translation method and device, electronic equipment and a storage medium, and relates to the technical field of video processing. The method is applied to a terminal, the terminal comprises a voice recognition module, a translation module and a text-to-voice module, and the method comprises the following steps: acquiring a live video stream, and splitting an audio track and a video track of the live video stream to obtain an audio and a video; a voice recognition module is adopted to recognize the audio, so that the audio is transferred into a text, and the text is the text of the first language; translating the text of the first language into a text of a second language by adopting a translation module; a text-to-language module is adopted to convert the text of the second language into voice; and performing video synthesis on the voice and the video to obtain a target video. According to the method and the device, the function module is deployed locally, so that the delay generated by transmitting the audio to the server is avoided, the processing efficiency is improved, and the real-time rebroadcasting of the live broadcast content in other languages is realized.
Owner:TERMINUSBEIJING TECH CO LTD

Instant check remembrance

A computer implemented method, system, and non-transitory computer-readable device that may be used in a remote deposit environment. Upon receiving a user request, based on interactions with the UI, the method implements an electronic deposit of a financial instrument by activating a camera on the client device to generate a live video stream of image data of a field of view of at least one camera, wherein the live video stream includes imagery of at least a portion of each side of the financial instrument. The method continues by extracting data fields based on the formation of image objects from one or both sides of the financial instrument from the live video stream of image data. The extracted data fields are converted, based on a payor agreement, into a recurring electronic funds transfer (EFT) schedule for future payments similar to the check.
Owner:CAPITAL ONE SERVICES LLC

Interaction method and device, equipment, storage medium and product

The invention provides an interaction method and device, equipment, a storage medium and a product, and relates to the technical field of computers. The method comprises the steps that a comment information set and a comment summary information set corresponding to a target live video interface are displayed in a preset page, and a single piece of comment summary information is determined based on multiple pieces of comment information belonging to the same preset category in the comment information set; the multiple pieces of comment information are released by audience users in the target live video interface; displaying a corresponding preset control at a first association position of the comment summary information; in response to a triggering operation for a target preset control, a preset event for first target comment summary information is triggered, the target preset control belongs to the preset control, and the first target comment summary information is the comment summary information corresponding to the target preset control. By adopting the technical scheme, the processing efficiency for the comment information of the preset category can be improved.
Owner:BEIJING YOUZHUJU NETWORK TECH CO LTD

Light control method of lamp and storage medium

The invention provides a light control method for lamps, a storage medium and a lamp control system applied to a live video scene, and the lamp control system is connected with at least one camera device and at least two lamps. Obtaining current display parameters sent by the camera equipment in the connected state, model information of the lamp and a face preference instruction and a skin color selection instruction input by a user; based on the model information, screening a target lamp meeting a first preset condition from the at least two lamps; performing matching in a plurality of target combination modes corresponding to the face preference instruction to obtain a target combination result corresponding to a target lamp meeting a second preset condition; matching the skin color selection instruction with a preset mapping relation to obtain a target display parameter of a target lamp in a target combination result; matching the target display parameter with the current display parameter to obtain a control instruction; and adjusting the target lamp based on the control instruction.
Owner:APUTURE IMAGING IND CO LTD

Methods and systems for privacy protecting a live video stream with an archived video stream

Methods, systems, and computer-readable media for producing a privacy-protected video stream are described herein. A request to display a live video stream of a camera is received. The live video stream is received in real-time and comprises a plurality of live image frames from the camera. An archived video stream of the camera is accessed in a data repository. A plurality of archived image frames of the archived video stream is processed to generate a background model comprising imagery common to multiple ones of the plurality of archived image frames. The plurality of archived image frames occurring in time prior to the request to display the live video stream. The privacy-protected video stream is processed in real-time. The privacy-protected video stream is output for display.
Owner:GENETEC

Lidar managed image generation

A computer implemented method, system, and non-transitory computer-readable device that may be used in a remote deposit environment. Upon receiving a user request, based on interactions with the UI, the method implements an electronic deposit of a financial instrument by activating a camera on the client device to generate a LIDAR managed live video stream of image data of a field of view of at least one camera, wherein the live video stream includes high quality confidence scored imagery of at least a portion of each side of the financial instrument. The method continues by extracting data fields based on the formation of image objects of each side of the financial instrument from the live video stream of image data. The extracted data fields are communicated to a remote deposit server to complete the remote deposit.
Owner:CAPITAL ONE SERVICES LLC

Video Processing Models with Streaming Feature Bank

One example aspect of the present disclosure is directed to a streaming model for video processing tasks, such as, for example, dense video captioning. Thanks to a memory mechanism, the proposed streaming model does not require access to all input frames concurrently in order to process the video. Moreover, thanks to a new streaming decoding algorithm, the proposed model can produce outputs causally without processing the entire input sequence. The streaming model is inherently suited to processing long videos—as it ingests frames sequentially (e.g., one at a time or in small batches). Moreover, as the output is streamed, intermediate predictions can be produced before processing the full video. This property means that the streaming model can be applied to process live video streams, as required for applications such as video conferencing, security and continuous monitoring among others.
Owner:GOOGLE LLC

Generating summary videos for user-selected portions of fixed-field videos using machine-learned classifiers

A media detection system receives a video corresponding to a fixed field of view. The media detection system may receive user input indicating one or more object types to identify or a subset of the video within which to identify objects. The media detection system applies one or more machine-learned classifiers to frames of the video and creates a summary video that includes the background of the video and identified instances for simultaneous playback within the fixed field of view. The media detection system may also identify instances of objects in a live video stream and use the identified instances to respond to user questions. The media detection system applies a language model to questions to identify the subject matter of the questions, identifies content within the live video stream associated with the subject matter, and uses the identified content to respond to the user's question.
Owner:MATROID INC

Method and system for extracting inherent user feature using artificial intelligence

Disclosed is a computer-implemented method and system for training a subject-specific machine learning model to infer inherent subject features from recorded or live video data. The system preprocesses the visual and audio channels, converting audio to text, and employs multiple pre-trained extraction models to generate feature embeddings. Ground truth data is obtained to guide training, where weights are assigned to produce and combine predicted feature values. Model performance is optimized by minimizing error. The trained feature extraction models are deployed on an edge device, while the subject-specific model resides in the cloud. A lightweight edge model, derived via knowledge distillation and model compression, supports local inferencing with reduced reliance on cloud resources. Synchronization ensures iterative updates for sustained accuracy.
Owner:MOODMETRICS AI