Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

16144results about "Television system details" patented technology

Method and apparatus for reducing the number of control messages transmitted by a set top terminal in an SDV system

A method is provided by which a subscriber accesses an SDV channel using a set top terminal. The method begins when the set top terminal receives a user request to tune to a first SDV channel. An active services list is also received over an access network. The active services list includes an entry for each currently available SDV program and a time-to-live (TTL) associated therewith. Tuning information is identified for the first SDV channel from its entry in the active services list. The set top terminal tunes to the first SDV channel using the identified tuning information. The channel change information associated with the user request is locally stored in set top terminal for transmission over the access network at a later time.
Owner:GENERAL INSTR CORP

Power transmission line multi-mode warning system and expelling method

The invention discloses a power transmission line multi-mode warning system and a power transmission line multi-mode expelling method, belongs to the technical field of power transmission line safety monitoring, and aims at solving the problems that power transmission line invasion target monitoring is not accurate, and the warning and expelling effect is poor. An environment image is acquired through an acquisition unit, and redundancy compensation is carried out on a fault unit. In the aspect of image splicing, a database is constructed by using structural feature points of a power transmission line, rapid projection transformation of a fixed area is realized, incremental feature matching is adopted for a dynamic area, and an environment panoramic image is generated. The method comprises the following steps: establishing a background coordinate system based on an environment panoramic image by aligning a fixed structure region, extracting multi-modal features of an intrusion target, constructing a dynamic trajectory parameter set, completing species classification and behavior recognition, constructing a multi-dimensional evaluation index system, dividing threat levels, and generating a thermodynamic diagram. And finally, according to data such as threat levels, a multi-mode grading warning system is constructed, warning equipment is dynamically adjusted, a target track is tracked, a warning effect is evaluated, and accurate and efficient invasion target expelling is realized.
Owner:SHENZHEN EVERBRIGHT LIGHTING CO LTD +1

Video plot generation and scene synthesis method and system based on natural language processing

The invention provides a video plot generation and scene synthesis method and system based on natural language processing, and relates to the technical field of video generation, and the method comprises the steps: receiving a script text to construct a multilayer scene map, extracting keywords to calculate semantic relevancy, executing feature decomposition and reconstruction to obtain a scene synthesis vector, and generating a video scene based on distance measurement. And extracting spatio-temporal features by using the feature pyramid, executing adaptive feature fusion, segmenting the video by applying a self-attention mechanism, and after a transition effect is inserted, performing style migration to output a finished video.
Owner:SHANDONG FOREIGN LANGUAGES VOCATIONAL AND TECH UNIV +1

Cloud gateway storage

A video gateway system at a worksite is coupled to multiple cameras on a network, and backs-up video streams generated by the cameras to a backend cloud backup video storage system and a frontend (cache) video storage system. The video gateway system generates an aggregated video asset from a plurality of streams of video from the multiple cameras, and generates metadata and a backup report associated with the video asset. The video asset, metadata, and backup report are stored on the backend cloud backup video storage system in an file system, and are also stored on the frontend (cache) video storage system such that for at least a period of time, the video asset and the associated the metadata and backup report are stored on both the cloud backup video storage system, facilitating quick access and retrieval to the stored video for retrieval for streaming, activity detection, and other uses.
Owner:SAMSARA INC

Multi-language cross-culture communication auxiliary method and system based on large model

The invention provides a multi-language cross-culture communication assisting method and system based on a large model. The method comprises the following steps: receiving a source language audio stream during a call, calling a multi-language sound frequency harmonic modulation feature library to extract fundamental frequency harmonic intensity distribution and tone turning features, and generating a cultural acoustic fingerprint vector; based on the vector, controlling a microphone array phase difference, directionally enhancing a fundamental frequency harmonic component of a speaker and suppressing noise, and outputting a high signal-to-noise ratio spectrogram; analyzing the pronunciation rhythm and tone turning characteristics of the spectrogram, capturing the pitch jump and duration of the syllable boundary, and generating an acoustic culture label; associating the spectrogram with a target semantic library, matching harmonic distribution and a cultural context rule based on a large model, and outputting a cultural interpretation prompt containing an ambiguity resolution suggestion; and generating a calibration result according to the acoustic tag and the semantic prompt, and overlapping the dynamic floating subtitles to the face area of the speaker in the video conference picture. According to the invention, cultural tone ambiguity in multi-language communication is eliminated.
Owner:LUSTER LIGHTWAVE CO LTD

Dual-stream video management

An Internet of Things (IoT) or vehicle dash cam may store both a high-resolution and low-resolution video stream on a device. The video streams are selectively accessible by remote devices. Because of the relatively smaller storage requirements of low-resolution video files, retaining of additional video data on the vehicle device (beyond what would be possible with only high-resolution video) is possible. The user may be provided an option to adjust the amount of low-resolution and high-resolution video to store on the device. A combined media file may be generated by a device to include time-synced high-resolution video, low-resolution video, and / or metadata for a particular time period.
Owner:SAMSARA INC

Broadcast receiver capable of enlarging broadcast-related information on screen and method of controlling the broadcast receiver

A method of controlling a broadcast receiver and which includes displaying a broadcast program on a screen of the broadcast receiver, determining if a broadcast-related information enlargement mode is selected, and enlarging the broadcast-related information on the screen of the broadcast receiver when the determining step determines the broadcast-related information enlargement mode is selected.
Owner:LG ELECTRONICS INC

Monitoring video enhancement method for farm

The invention belongs to the technical field of video processing, and particularly relates to a monitoring video enhancement method for a farm, which aims to solve the technical problem of low quality of an enhanced video in the prior art, and comprises the following steps: S1, processing each frame of image in a monitoring video sequence frame by frame; s2, distinguishing a target animal area from a background area, and identifying and generating an artifact mask; s3, aiming at the background area, carrying out key smoothing processing on the artifact position to inhibit the artifact; s4, for the image of the target animal area, performing adaptive nonlinear enhancement on the brightness component, and performing color correction on the chrominance component; s5, performing pixel-level fusion on the enhanced target animal area and the background area; and S6, spreading the information of the previous frame to the current frame by using the forward optical flow field, and carrying out weighted fusion on the information of the previous frame and the current frame. According to the method, the target bred animals, the background areas and the artifacts are accurately distinguished, so that refined and differentiated processing of pictures is realized.
Owner:EGG NO 1 FOOD CO LTD

Method and apparatus for encoding, decoding, or displaying picture-in-picture

Various embodiments provide example apparatus, method, and computer program product. An example apparatus includes: receiving or generating a first encoded bitstream comprising independently encoded subpictures; receiving or generating a second encoded bitstream comprising independently encoded subpictures; wherein resolution of subpictures in the second encoded bitstream is same or substantially same as resolution of corresponding subpictures in the first encoded bitstream; generating an encapsulated file with a first track and second track, the first track comprises the first encoded bitstream comprising the independently encoded subpicture, the second track comprises the second encoded bitstream comprising the independently encoded subpictures; and wherein generating the encapsulated file comprises including following information in the encapsulated file: a picture-in-picture relationship between the first track and second track; and data units independently encoded subpictures in the first encoded bitstream that are to be replaced by data units of the independently encoded subpictures of the second encoded bitstream.
Owner:NOKIA TECHNOLOGIES OY

Object monitoring method and device based on multi-source data and storage medium

The invention relates to the field of image processing, and discloses an object monitoring method and device based on multi-source data and a storage medium, and the method comprises the steps: synchronously collecting multi-channel video streams, environment parameters and equipment position information, and carrying out the time alignment and data association processing, and forming associated data; performing feature matching on the multiple paths of video streams, fusing position information and environment parameters, and establishing a mapping relation between video feature points and a unified space coordinate system; splicing the multiple paths of video streams in real time according to the mapping relation to generate a panoramic video stream; performing dynamic scene analysis based on the panoramic video stream and the environmental parameters, and identifying a target type and a state to obtain an analysis result; and generating an equipment regulation and control strategy based on the analysis result, and generating and issuing a regulation and control instruction for adjusting the working parameters of the multi-view image acquisition device based on the equipment regulation and control strategy. According to the invention, the splicing precision, the real-time performance and the target identification accuracy of panoramic monitoring can be improved, and the dynamic adaptive regulation and control of the equipment can be realized.
Owner:SHENZHEN STARCAM TECH

Intelligent construction site face recognition attendance supervision and management system and management method

The invention relates to the technical field of face recognition, in particular to an intelligent construction site face recognition attendance supervision and management system and management method. Comprising a binocular living body detection camera, an identity card reader, an IC / ID card reader, an infrared monitoring sensor and a gate acceleration sensor; the dynamic weight calculation module is used for dynamically adjusting the weight parameter of the face matching degree according to the living body detection result, the human-certificate comparison data and the environment illumination condition; the malicious gate rushing judgment module is used for generating a gate rushing risk value based on gate acceleration sensor data, infrared monitoring border crossing behaviors and continuous face recognition failure times; the attendance data optimization module is used for analyzing and correcting the confidence coefficient of attendance records through a historical recognition success rate and a time sequence; according to the invention, safety management requirements in complex scenes can be met.
Owner:HUIYUE (GUANGZHOU) TECHNOLOGY CO LTD

Multi-target fire positioning method and device based on multispectral dynamic fusion

The invention provides a multi-target fire positioning method and device based on multispectral dynamic fusion, and relates to the technical field of unmanned aerial vehicle inspection, and the method comprises the steps: obtaining a multispectral image sequence collected by an unmanned aerial vehicle in an inspection task execution process; according to the visible light image and the infrared image contained in the image frame group, smoke identification processing, fire source identification processing and result dynamic fusion processing based on confidence are carried out, and a target pixel position identified as a target area in the image frame group is obtained; and performing ray inversion by taking the pose data of the unmanned aerial vehicle as positioning compensation based on the target pixel position identified as the target area in the image frame group, and determining the target longitude and latitude position identified as the target area in the image frame group. According to the invention, the method can achieve the recognition and positioning of a plurality of target regions during the cruise task execution process of the unmanned aerial vehicle, facilitates the improvement of the positioning efficiency of the target regions, and also can remarkably improve the positioning precision of the target regions.
Owner:TIANJIN YUNSHENG INTELLIGENT TECH CO LTD

Text-to-video full link generation method and system based on multi-modal large model

The invention discloses a text-to-video full-link generation method and system based on a multi-modal large model, and belongs to the technical field of artificial intelligence content generation. Through cooperative work of multiple agents, analysis of a text input by a user and construction of a cross-modal memory library, unification of videos and audios for generating a sub-mirror is ensured based on memory library content, and the generation efficiency of the sub-mirror is improved. The full-process automatic generation from the text to the video is realized; the method is implemented by the following steps of: acquiring text input of a user; text analysis: dynamically extracting, analyzing, generating, associating and storing image-text-tone multi-modal information from an input text through an Agent working cooperatively, and constructing a multi-modal memory library; generating a split mirror, and generating a split mirror video and an audio according to the memory bank; and performing audio and video synthesis and audio and picture synchronous alignment to form a final video. According to the method, the narrative continuity of long video generation can be realized, the feature consistency of the split image is improved, the consistency of cross-modal emotion is enhanced, manual intervention is reduced, and the video production efficiency is improved.
Owner:INSPUR QILU SOFTWARE IND

Preparation method of MEMS micromirror driven by vertical comb teeth

The invention discloses a preparation method of a vertical comb-driven MEMS (Micro Electro Mechanical System) micromirror, which comprises the following steps of: forming a through alignment mark and a pre-etched comb structure on a device layer by using a first composite mask, and forming the through alignment mark and the pre-etched comb structure on the back surface of the device layer on a first substrate through a composite mask process, after the first substrate and the second substrate are bonded, the self-alignment composite mask of the comb tooth structure is aligned with the back face pre-etched comb tooth structure through the through alignment mark, high-precision alignment of the comb tooth structure is achieved, accumulated errors are avoided, and the yield of device preparation is improved.
Owner:NORTHWESTERN POLYTECHNICAL UNIV

Virtual fitting video generation method and system based on multi-view face fixation

The invention discloses a virtual fitting video generation method and system based on multi-view face fixation, and relates to the field of generative artificial intelligence and computer vision, and the virtual fitting video generation method based on explicit geometric constraints comprises the following steps: S1, obtaining an original fitting image and an action cue word, and constructing a multi-source input data set; s2, segmenting the original fitting image to obtain multi-view modeling, and generating a multi-angle face image and a clothing texture feature parameter; s3, analyzing the action cue word to generate a target posture sequence, and generating head and tail frame virtual fitting images; s4, performing pairing analysis on the virtual fitting images of the head frame and the tail frame to generate attitude transition parameters; s5, performing dynamic texture repair on the image to generate a transition frame image sequence; and S6, generating a virtual fitting video and outputting the virtual fitting video. According to the method, accurate segmentation and multi-angle modeling of the face area and the clothing area are realized through image segmentation and a rigid geometric transformation algorithm.
Owner:QINGDAO UNIV

MEMS device manufacturing method and MEMS device

The invention provides a manufacturing method of an MEMS device and the MEMS device, and the manufacturing method comprises the steps: providing a first wafer which is provided with a groove; a protective layer is arranged in the groove; providing a second wafer, and bonding the second wafer and the first wafer to seal the groove to obtain a first cavity; the second wafer is etched to form a comb tooth structure, and the protection layer is used for protecting the comb tooth structure. According to the embodiment of the invention, the protection layer is arranged in the first cavity of the first wafer as a buffer layer, so that when the second wafer is etched to form the comb tooth structure, the plasma backwash in the over-etching stage can be reduced, the comb tooth structure is further protected, the damage to the bottom of the comb tooth caused by the backwash of etching particles is reduced and solved, and the service life of the comb tooth structure is prolonged. Therefore, the mechanical strength and the anti-failure capability of the device are enhanced.
Owner:NINGBO SEMICON INT CORP

Smart park resource scheduling method and system

The invention relates to the technical field of resource scheduling, in particular to a smart park resource scheduling method and system, and the method comprises the steps: constructing a heterogeneous information network representing the current state of a park; performing feature extraction on the heterogeneous information network by adopting a graph attention convolutional network, and generating a state embedding vector fused with high-order neighborhood information; inputting the state embedding vector into a scheduling strategy model, and predicting the expected variation of the candidate scheduling action to future park energy consumption, security index and traffic efficiency; according to the state embedding vector and an external management instruction, determining an optimization target of a current scheduling period from a preset operation normal form, and allocating a dynamic weight; and calculating a comprehensive utility score of each candidate scheduling action, and issuing and executing the candidate scheduling action with the highest score as a final scheduling instruction. According to the invention, the refinement level and the overall operation benefit of park management are improved.
Owner:SHANDONG ZHENGTU INFORMATION POLYTRON TECH INC

Generating A Unified Virtual Background Image For Multiple Video Conference Participants

A unified virtual background image is generated for multiple participants of a video conference to create an immersive conference experience based on its use within video streams of those multiple participants. Generative artificial intelligence software associated with a conferencing system obtains input associated with a video conference. The generative artificial intelligence software generates a virtual background image based on the input. The virtual background image is then for use within multiple participant video streams during the video conference
Owner:ZOOM COMMUNICATIONS INC

Generating digital images utilizing a diffusion-based network conditioned on lighting-aware feature representations

Methods, systems, and non-transitory computer readable storage media are disclosed for generating digital images with a diffusion-based generative neural network conditioned on background-extracted lighting features. The disclosed system determines, in response to a request to generate a digital image, a target background image for inserting a foreground object into the target background image. The disclosed system generates, from the target background image and utilizing a lighting conditioning neural network, a lighting feature representation indicating one or more lighting parameters of the target background image. Additionally, the disclosed system generates, utilizing a diffusion-based generative neural network conditioned on the lighting feature representation, the digital image including the foreground object inserted into the target background image based on a composite image comprising the foreground object and the target background image with a foreground mask corresponding to the foreground object.
Owner:ADOBE INC

Control method of video monitoring system

The invention discloses a control method of a video monitoring system, which relates to the technical field of video monitoring, and comprises the following steps of: acquiring continuous image frames through the video monitoring system, extracting brightness channels, edge structures, direction gradients and contrast changes of images, constructing a reflective perception vector group, and calculating an included angle change trend by combining a target motion direction vector, determining whether a reflection offset condition consistent with the target direction appears in a picture in a camera attitude control process; after it is determined that the reflection offset condition consistent with the target direction appears in the picture, the inter-frame change tensor of the target area and the reflection area is extracted, and a time sequence track consistency matrix is constructed; according to the invention, the problem of wrong adjustment of the camera caused by misjudgment of light reflection in video monitoring is solved, attitude regulation and control based on image displacement abnormal mode recognition are realized, and the tracking stability and the monitoring accuracy are improved.
Owner:ANHUI HUIDI INTELLIGENT TECHNOLOGY CO LTD

Visual special effect video generation method and system

The invention discloses a method and system for generating a visual special effect video, and relates to the technical field of video generation, and the method comprises the steps: carrying out the collection and preprocessing of multi-modal data, synchronously collecting an RGB video, a depth map and environment illumination data, and generating a registered 3D point cloud; carrying out time-space domain combined motion compensation image matting, and extracting a moving object Alpha mask; by means of edge purification of self-adaptive illumination separation, illumination pollution and edge color overflow of a green screen scene are eliminated; the multi-scale transparency estimation network outputs a transparency graph; dynamic special effect synthesis and output guarantee video fluency; and the special effect video is intelligently recommended, and a self-adaptive output strategy is realized. The image matting and edge processing precision is improved, the transparency estimation and special effect synthesis effect is enhanced, intelligent recommendation and resource optimization are realized, and the defects in dynamic scene processing in the prior art are overcome.
Owner:XIXIAN TECH CO LTD

Multi-modal fusion green port digital management and control system and method

The invention discloses a multi-modal fusion green port digital management and control system and method, and relates to the technical field of port management and control, and the method comprises the following steps: collecting the signal-to-noise ratio data of a radio channel of a video link, and generating a port area electromagnetic interference thermodynamic diagram through spatial interpolation; and dynamically adjusting a synchronous clock of each camera according to the electromagnetic interference thermodynamic diagram, mapping the frame triggering offset into a pixel compensation matrix, and performing sub-pixel-level space-time correction on the image frame to obtain a corrected video frame sequence. According to the method, target credibility judgment and path optimization under multi-modal perception are realized through construction of an interference thermodynamic diagram, space-time correction, artifact recognition and point cloud fusion, an interference model and a recognition threshold are dynamically regulated and controlled through closed-loop feedback, and an end-to-end self-adaptive port digital management and control method is constructed. The artifact identification accuracy and scheduling stability are significantly improved, and the anti-interference and green efficient operation capabilities of the port management and control system are enhanced.
Owner:TIANJIN RES INST FOR WATER TRANSPORT ENG M O T

Video data processing method and device, storage medium and electronic equipment

The invention discloses a video data processing method and device, a storage medium and electronic equipment, and relates to the technical field of smart home, and the video data processing method comprises the steps: determining a target distance boundary value according to the region type of a region where monitoring equipment is located; when it is detected that a target object exists in the target video data collected by the monitoring equipment, determining a first distance between the target object and the monitoring equipment, and determining a size relationship between the first distance and a target distance boundary value; and under the condition that the size relationship indicates that the first distance is greater than the target distance boundary value, privacy processing is performed on a to-be-privatized area in the target video data, and the to-be-privatized area is an area determined according to the target object.
Owner:QINGDAO HAIER INTELLIGENT HOME APPLIANCE TECHNOLOGY CO LTD

Multi-protocol high-speed image data receiving and aligning method and system based on FPGA (Field Programmable Gate Array)

The invention discloses a multi-protocol high-speed image data receiving and aligning method and system based on an FPGA (Field Programmable Gate Array), and relates to the technical field of electronic information and image processing. Serial data streams from an image sensor are received, an FPGA serial input primitive is dynamically called to convert serial data into parallel data; a low-voltage differential signal protocol, a mobile industry processor interface protocol and a serial low-voltage signal protocol are supported; a configurable sliding window mechanism is adopted to compare synchronous codes in real time, and byte boundaries are locked; by identifying the time difference of arrival of data of each channel, multi-channel phase alignment is realized by adopting hardware-level delay compensation, and data arriving at the channel in advance is compensated by adopting a first-level register cache; and analyzing frame header and frame tail information of the image data, reconstructing an original image matrix and generating a video time sequence signal. The problems of poor compatibility and high resource consumption in multi-protocol high-speed image data receiving are solved.
Owner:AI TUER

Establishing a video conference during a phone call

Some embodiments provide a method for initiating a video conference using a first mobile device. The method presents, during an audio call through a wireless communication network with a second device, a selectable user-interface (UI) item on the first mobile device for switching from the audio call to the video conference. The method receives a selection of the selectable UI item. The method initiates the video conference without terminating the audio call. The method terminates the audio call before allowing the first and second devices to present audio and video data exchanged through the video conference.
Owner:APPLE INC

Children story video generation method and system based on AI

The invention discloses an AI-based child story video generation method and system, and relates to the technical field of artificial intelligence and multimedia crossing, and the method comprises the steps: generating a script from an original text input by a user through constructing an AI model fusing an emotion modeling capability; constructing an image generation combination model, defining a joint loss function, calculating an edge intensity graph of the contour image by using a Sobel edge detection algorithm, calculating an optical flow field of frame change by using a block matching algorithm, and performing color image dynamic frame alignment; a fine tuning WaveNet model is used to generate audio; through constructing an image generation combination model, combining a StyleGAN3-T model and an LDM model, defining a joint loss function, and using a Sobel edge detection algorithm and a block matching algorithm to calculate an edge intensity graph and an optical flow field, dynamic frame alignment of a color image is realized, and inter-frame continuity of a generated video is improved.
Owner:KUAISHANGYUN (SHANGHAI) NETWORK TECHNOLOGY CO LTD

Dynamic split mirror generation system and method based on controllable diffusion model

The invention discloses a dynamic split mirror generation system and method based on a controllable diffusion model, and belongs to the technical field of film and television production. The implementation method comprises the following steps of: 1, setting a keyword text by a user, and inputting the keyword text into ChatGPT to generate a script; 2, converting the script into a scene category, a camera motion mode, a character role position and action and image description in a shot language by utilizing ChatGPT; 3, using a CLIP model to carry out contrast training on the image encoder and the text encoder; 4, performing an OpenPose model on the action reference image to obtain skeleton key points of the character role, and converting the skeleton key points into image character actions; 5, performing action fine-grained control on the character action of the image by using a Stable Diffusion model and a ControlNet model, and generating a split image; 6, utilizing a Pika Labs model to generate a dynamic video from the split image; compared with the prior art, the accuracy of user role action matching under the scene based on multi-text and high-complexity actions is improved.
Owner:BEIJING UNIV OF POSTS & TELECOMM +2