Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

94 results about "Adaptive video" patented technology

Multi-mode large model video content understanding reasoning acceleration method and system

The invention discloses a multi-mode large model video content understanding reasoning acceleration method and system, and mainly relates to the technical field of artificial intelligence reasoning acceleration. Comprising the following steps: inputting video data and preprocessing the video data to generate a video frame sequence; performing adaptive video Token compression on the generated video frame sequence, and outputting a compressed visual Token set; performing visual feature coding and Key-Value generation on the compressed visual Token set to obtain visual KV data; performing video KV cache partition management on the visual KV data; executing cross-modal reasoning based on the vLLM framework to generate a video content understanding result; and outputting a video content understanding result, and carrying out post-processing and structured mapping. The method has the beneficial effects that the obvious reasoning acceleration and throughput improvement can be realized on the premise of keeping the precision of the original large model.
Owner:海看网络科技(山东)股份有限公司

Adaptive video restoration method based on digital video technology

The invention discloses an adaptive video restoration method based on a digital video technology, and relates to the technical field of image communication. The method comprises the following steps: preprocessing an original video containing hard subtitles and automatically removing the hard subtitles to obtain a video to be repaired; screening a reference frame and a to-be-repaired frame by comparing the original video with the to-be-repaired video; determining a to-be-repaired area according to the pixel gray scale difference; extracting matched feature points, calculating a reference motion vector, estimating possible positions and motion differences of points in the to-be-repaired region in combination with object structure segmentation and region prediction, and generating a corrected motion vector; performing sub-pixel-level interpolation by using the vector, and constructing a corrected reference image; and finally, fusing the corrected image and the to-be-repaired area to generate a repaired video frame. The method can effectively improve the restoration quality of the subtitle shielded area.
Owner:EC INNOVATIONS (SHENYANG) INC

High performance and low complexity adaptive video image defogging

An apparatus comprising an interface and a processor. The interface may be configured to receive pixel data of an environment. The processor may be configured to process the pixel data arranged as video frames, generate a luminance distribution map of the video frames in response to a low-pass filter operation, determine a plurality of defogging intensity weights for the luminance distribution map, perform adaptive smoothing to each of the plurality of defogging intensity weights, and generate defogged video frames in response to the video frames and the plurality of defogging intensity weights with the adaptive smoothing. The plurality of defogging intensity weights may each correspond to one of a plurality of luminance intervals of the luminance distribution map. The adaptive smoothing may be configured to prevent brightness differences in the defogged video frames.
Owner:AMBARELLA INT LP

High-fidelity generation type video stream transmission system based on visual base model

The invention relates to a high-fidelity generative video stream transmission system based on a visual base model, which belongs to the field of image communication, and is characterized in that a visual enhancement-oriented generative codec is designed, and high-fidelity video reconstruction under a high compression ratio is realized through an asymmetric space-time compression strategy and time sequence consistency enhancement; a resolution scaling module is provided, the calculation complexity is remarkably reduced through a video super-resolution recovery module of adaptive resolution control and joint optimization, and real-time high-definition video processing is achieved; and constructing a network adaptive video stream transmission controller, an intelligent token discarding mechanism based on semantic importance and a mixed packet loss processing strategy to realize code rate scalable control and robust transmission under network fluctuation.
Owner:THE CHINESE UNIV OF HONG KONG (SHENZHEN) +1

Dynamic adaptive video black edge real-time detection method based on multistage scanning

The invention discloses a dynamic self-adaptive video black edge real-time detection method based on multistage scanning, and relates to the technical field of image processing, and the method comprises the following specific steps: 1, inputting a video frame, and selecting 28 detection regions in the video frame according to the requirements of four-corner region positioning, edge center region center positioning and cross region positioning; 2, sparse sampling is carried out on the 28 detection areas one by one, whether the detection areas are black areas or not is judged according to sampling results, if yes, the step 3 is executed, and if not, the detection areas are output; 3, preliminarily determining the orientation of the black edge of the video frame and the starting and ending positions of the black edge in the four-edge direction of the video frame; the method supports real-time updating of the position of the black edge, adapts to resolution switching, transverse and vertical screen change and picture cutting, realizes efficient and accurate black edge detection, and updates the detection result in real time. Meanwhile, the method supports a mainstream video format (RGB / YUV), and the false detection rate is lower.
Owner:NANJING DUOWEIXINLIAN TECH

Self-adaptive video low-delay hard decoding display method and system

The invention relates to the technical field of video decoding, in particular to a self-adaptive video low-delay hard decoding display system, which identifies a video stream format and a switching state through a self-adaptive demultiplexing module, extracts naked stream data and judges a decoding mode; the dynamic decoder allocation module dynamically allocates decoding resources according to the mode information, and supports concurrent decoding of a whole-frame single decoder or N-fragment multi-decoders; and the frame synchronization and VCXO dynamic adjustment module outputs a latest complete frame in a whole frame mode, aligns and recombines fragmentation data according to a frame number in an N fragmentation mode, and dynamically adjusts a VCXO clock frequency based on a decoding and display time difference. According to the invention, multi-format compatibility, seamless switching of video streams and ultra-low delay display are supported, and black fields, static frames and picture tearing are effectively avoided.
Owner:南京威翔科技有限公司

Adaptive video compression coding method, system and device, and mobile platform

The invention discloses a self-adaptive video compression coding method and system, and the method comprises the steps: monitoring the type of a current automatic driving scene in real time based on perception information, and obtaining a target coding parameter matched with the type of the current automatic driving scene according to the dynamically determined type of the current automatic driving scene; and performing video compression coding on image data acquired in real time under the corresponding automatic driving scene type by using the target coding parameter matched with the current automatic driving scene type. According to the method provided by the invention, the target coding parameters can be flexibly and adaptively adjusted in combination with different scenes, so that the adaptive target coding parameters can be adaptively selected under different automatic driving scene types to carry out video compression coding on the image data under the corresponding scenes, and therefore, better image quality can be maintained, and the video coding efficiency can be improved. Storage occupation can be compressed as much as possible, and comprehensive balance of image quality and storage size of different scenes is realized.
Owner:SZ ZHUOYU TECH CO LTD

Bidirectional adaptive video super-resolution method based on frame difficulty index

The invention discloses a bidirectional adaptive video super-resolution method based on frame difficulty index, and belongs to the field of computer vision. The invention provides a frame-level dynamic reconstruction method for solving the problems that simple frame calculation is redundant and difficult frame reconstruction is insufficient due to the fact that an existing model adopts a fixed calculation strategy for video frames with different difficulties. According to the method, a motion detail decoupling propagation network is constructed, motion information is efficiently transmitted by utilizing a shallow forward propagation branch, and a deep backward propagation branch focuses on recovering texture details so as to decouple a time sequence propagation task; meanwhile, a frame reconstruction difficulty evaluation network is introduced to generate a global difficulty index, so that the receptive field weight of the adaptive time sequence fusion network and the refining depth of the dynamic refining network are regulated and controlled. According to the method, through explicit modeling frame-level reconstruction difficulty, adaptive matching of the model capacity and the video frame feature complexity is realized, and the video reconstruction performance is remarkably improved under limited computing power.
Owner:GUILIN UNIV OF ELECTRONIC TECH

HTTP3-based GB / T28181 monitoring video stream transmission and decoding method and electronic equipment

The invention discloses a GB / T28181 monitoring video stream efficient transmission and decoding method based on HTTP3 (Hyper Text Transfer Protocol 3) and electronic equipment. According to the method, the problems of high delay and high resource consumption of a traditional GB / T 28181 video monitoring Web end playing scheme based on WebSocket are solved by utilizing a new generation Web API (Application Program Interface) of WebTransport and WebCodec. According to the invention, through the UDP transmission capability of WebTransport based on HTTP / 3, the direct transmission of the RTP data packet from the server to the Web client is realized, and protocol conversion does not need to be carried out; the WebCodes is adopted to realize efficient hardware acceleration decoding of the client; and meanwhile, a multiplexing real-time monitoring and self-adaptive video quality control mechanism is provided. Compared with the traditional scheme, the method has the advantages that the end-to-end delay is remarkably reduced, the resource consumption of the server is reduced, the concurrency capability is improved, the video parameters can be dynamically adjusted according to the network condition, and the method is suitable for video monitoring application in various network environments.
Owner:武汉市公安局科技信息化支队 +1

Space-time adaptive video lane line detection method and system based on differential memory and wavelet guidance

The invention discloses a space-time adaptive video lane line detection method based on differential memory and wavelet guidance, and mainly solves the problems of difficulty in collaborative modeling of space-time characteristics, easy loss of shallow geometric details and insufficient utilization of deep time sequence information in a dynamic scene in the prior art. According to the implementation scheme, the method comprises the following steps: acquiring marked video sequence data from a public video lane line detection data set, preprocessing the marked video sequence data, and dividing the preprocessed video sequence data into a training set and a test set; constructing a video lane line detection network comprising a feature extraction unit, a differential memory time sequence alignment unit, a wavelet-guided direction perception feature enhancement unit, a space-time adaptive fusion unit and a feature decoding unit; iteratively training the video lane line detection network through back propagation by using the training set; and inputting the test set into the trained video lane line detection network, and outputting a lane line. According to the method, the detection precision and robustness in challenging environments such as shielding, strong light and motion blur are remarkably improved, and the method can be used for realizing accurate extraction and stable tracking of lane lines in a dynamic driving scene.
Owner:XIDIAN UNIV

Self-adaptive video arm support adjusting structure

The utility model discloses a self-adaptive video arm support adjusting structure which comprises a base, a main arm with one end rotationally connected with the base, an auxiliary arm telescopically connected in the main arm and a photographing piece installed on the auxiliary arm, and the photographing piece comprises an installation plate fixedly installed at the first end, away from the main arm, of the auxiliary arm. A mounting groove is formed in the mounting plate, and a first camera and a second camera are movably connected into the mounting groove. The angle of the arm body can be controlled through hinged rotation of the main arm on the base, and the length of the photographing piece connected to one end of the auxiliary arm can be adjusted through telescopic movable connection of the auxiliary arm in the main arm, so that the optimal photographing height and angle in a dredging detection area are formed. And meanwhile, the shooting height can be adjusted in real time according to the change of an actual scene, and the first camera and the second camera are movably connected on the auxiliary arm, so that an optimal shooting angle is formed in a working state.
Owner:HUNAN ZHONGKE HENGQING ENVIRONMENTAL MANAGEMENT CO LTD

Constraints and unit types to simplify video random access

Disclosed herein are innovations for bitstreams having clean random access (CRA) pictures and / or other types of random access point (RAP) pictures. New type definitions and strategic constraints on types of RAP pictures can simplify mapping of units of elementary video stream data to a container format. Such innovations can help improve the ability for video coding systems to more flexibly perform adaptive video delivery, production editing, commercial insertion, and the like.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Self-adaptive video transmission control method based on bandwidth perception of unmanned aerial vehicle

The invention discloses a self-adaptive video transmission control method based on unmanned aerial vehicle bandwidth perception. The method comprises the steps that S1, an unmanned aerial vehicle video system sends a bandwidth perception data frame to a ground receiving end; s2, the ground end returns a bandwidth state data frame, wherein the bandwidth state data frame comprises a received signal strength indication (RSSI) and a frame loss rate (FLR) index; s3, the unmanned aerial vehicle video system analyzes the bandwidth state data frame and evaluates the channel quality grade; s4, the unmanned aerial vehicle video system calculates a buffer proportion factor of the annular buffer area; s5, the unmanned aerial vehicle video system calls a pre-constructed adaptive transmission control model, transmission parameters are updated according to the channel quality grade and the buffer proportion factor, and the transmission parameters comprise transmission video quality and coding parameters; s6, transmitting video data according to the updated transmission parameters of the unmanned aerial vehicle video system; and S7, the ground end receives the video data and decodes and plays the video data. According to the invention, the problem of poor adaptability in different network environments when the unmanned aerial vehicle sends the video data with the fixed channel parameters is solved, and the video transmission quality is improved.
Owner:XIAN FLIGHT SELF CONTROL INST OF AVIC

Adaptive video filter

A video coder may receive video data and apply a filter to a plurality of types of samples of the video data. The plurality of types of samples may include two or more of: an input of a fixed filter, an output of a fixed filter, an input of a signaled filter, an output of a signaled filter, an input of an adaptive loop filter, an output of an adaptive loop filter, an input of a sample adaptive offset (SAO) filter, the method comprises the following steps: receiving the input of an SAO filter, the output of an SAO filter, the input of a bilateral filter, the output of the bilateral filter, the input of a cross component SAO filter, the output of the cross component SAO filter, the input of a deblocking filter, the output of the deblocking filter, filtered reconstructed residual data, a dequantization coefficient, a filtered dequantization coefficient, a predictor or a filtered predictor.
Owner:QUALCOMM INC

Adaptive video stream code rate decision-making method and system based on neural symbol reinforcement learning buffer area perception

The invention provides a perception adaptive video stream code rate decision-making method and system based on a neural symbol reinforcement learning buffer area, and belongs to the technical field of computer network and multimedia transmission. According to the method, a neural symbol reinforcement learning agent is constructed, a comprehensive state vector is generated by collecting a playing state, video content information and user preference, and double decisions of code rate selection and buffer area adjustment are jointly output; security rules are formalized in a first-order logic mode, security exploration is achieved through a security monitor and a security layer strategy, and illegal risks such as underflow of a buffer area are avoided; the bandwidth prediction precision is improved by adopting HTTP / 3 stable connection sampling and a hybrid prediction model; and designing a multi-target reward function to balance user experience quality and buffer occupation optimization. According to the method, high-video-quality, low-lagging and low-delay transmission can be stably realized in a low-speed to 5G high-speed network, and unnecessary buffer accumulation can be remarkably reduced on the premise of ensuring smooth playing, so that end-to-end delay is reduced.
Owner:GUANGXI UNIV

Adaptive video transmission method based on fountain code, electronic device and storage medium

ActiveCN117014697BComputer hardwareFountain code
The application provides a fountain code-based adaptive video transmission method, an electronic device and a storage medium, and the method comprises the following steps: receiving historical packet loss rates of all data packets corresponding to video image sent frames which are counted and fed back by a receiving end of a video transmission system; updating coding redundancy for fountain code encoding of a current frame of a video image based on the historical packet loss rates of the video image sent frames; compressively encoding the current frame of the video image to be transmitted to generate a compressed code stream data frame; fountain code encoding the compressed code stream data frame according to the coding redundancy to generate coding data packets corresponding to the current frame of the video image; and transmitting the coding data packets corresponding to the current frame of the video image to the receiving end of the video transmission system. The application can realize real-time video transmission with low computational complexity, low performance consumption and high decoding success rate.
Owner:BEIJING UNIV OF POSTS & TELECOMM

A serverless computing based adaptive video streaming method and system

ActiveCN116962414BVideo deliveryReinforcement learning algorithm
The application relates to a kind of adaptive video streaming method and system based on serverless computing, belong to streaming media transmission technical field.System is realized by fine-grained serverless pipeline Video delivery, use stateless function to strengthen the response to video request event, use a kind of near-end strategy optimization PPO of three-end clipping based on deep reinforcement learning algorithm to solve the bit rate adaptive sequence decision problem in video playing process.In addition, dynamic video block quality factor is included in user experience QoE index, to configure QoE model, for each video block is assigned a priority weight, improves the robustness of video bit rate decision, thereby reduces video streaming delay, improves user viewing experience.
Owner:BEIJING INST OF TECH

Constraints and unit types to simplify video random access

Disclosed herein are innovations for bitstreams having clean random access (CRA) pictures and / or other types of random access point (RAP) pictures. New type definitions and strategic constraints on types of RAP pictures can simplify mapping of units of elementary video stream data to a container format. Such innovations can help improve the ability for video coding systems to more flexibly perform adaptive video delivery, production editing, commercial insertion, and the like.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Self-adaptive video watermarking method based on Web front end

The invention discloses a self-adaptive video watermarking method based on a Web front end. Establishing a standardized percentage coordinate system based on the video effective area; the method comprises the following core steps: acquiring a preview canvas size and a device pixel ratio (DPR), and executing rendering precision compensation calculation to generate a vectorization logic component; sensing transparency in real time, and triggering self-adaptive visualization enhancement assistance based on background brightness contrast when a threshold value is met; and capturing an interaction instruction, updating a proportional parameter in real time based on bidirectional mapping, and executing boundary clamping correction. Atomized attribute stripping is performed in response to the save instruction to restore the original visual attributes and generate a structured configuration information stream. According to the method, engineering pain points such as heterogeneous resolution geometric distortion, high-transparency watermark interaction difficulty and cross-end display fuzziness are solved through a visual means of DPR compensation and threshold perception, and watermark positioning precision, editing interaction efficiency and cross-platform visual consistency are remarkably improved.
Owner:HANGZHOU ARTECH

Short video intelligent editing method, system and terminal based on sentiment analysis

The invention discloses a short video intelligent editing method, system and terminal based on sentiment analysis, and relates to the technical field of artificial intelligence, computer vision and digital media processing. The method comprises a multi-modal emotion feature extraction step, a unified emotion curve generation step, an intelligent editing strategy generation step and a video synthesis and post-processing step. According to the method, intelligent and self-adaptive video editing with the emotion flow as core driving force is realized, and the emotion expressive force and professional degree of the automatically edited video are remarkably improved.
Owner:李嘉祺

Intelligent dynamic adaptive video service system

The application discloses a kind of intelligent dynamic adaptation video service systems, including video service capability registration unit, video service function requirement reporting unit, terminal parameter reporting unit, terminal capability analysis unit, terminal capability fusion unit, video capability matching unit and video service configuration issuing unit, the position of terminal parameter reporting unit connection terminal capability analysis unit, the position of terminal capability analysis unit connection terminal capability fusion unit, the position of terminal capability fusion unit connection video capability matching unit, the position of video capability matching unit connection video service capability registration unit.The intelligent dynamic adaptation video service system described in the application is low in interfacing cost, high fault tolerance, intelligent dynamic adaptation, low in average use cost, when terminal condition is good, the highest cost-effective video service is preferentially used to reduce average use cost, and the longer the use time is, the lower the total cost is.
Owner:JIANGSU GAREA HEALTH TECH

Forgery face video detection method and system based on Mobilnet-GRU

The invention provides a Mobilene-GRU-based forged face video detection method and system, and relates to the technical field of forged face video detection, and the method specifically comprises the steps: carrying out the partitioning processing of a face image sequence, obtaining the motion consistency characteristics, decomposing each frame of image into a low-frequency region and a plurality of high-frequency regions through wavelet transformation, and carrying out the recognition of the low-frequency region and the plurality of high-frequency regions; the method comprises the following steps: acquiring an energy ratio of a low-frequency region to a high-frequency region, inputting each frame of image in a face image sequence into a Mobilenet model, obtaining a spatial feature vector, setting an initial proportion of frame image replacement, inputting the spatial feature vector into a GRU model for time sequence prediction, iteratively optimizing the proportion of frame image replacement, and adopting a pre-trained discrimination model to discriminate the face image. And carrying out counterfeiting judgment on the to-be-detected video. According to the method, the initial proportion of frame image replacement is set, and the proportion is iteratively optimized, so that the spatial features can dynamically adapt to the change of the video content, and the capturing capability and the expression capability of the model on the dynamic change features of the video are remarkably enhanced.
Owner:HEFEI UNIV

A dynamic processing method and device for image adaptive video memory and a related medium thereof

This invention discloses a dynamic processing method, apparatus, and related medium for image adaptation to video memory. The method includes acquiring image information and video memory information of the hardware device and calculating correction coefficients; using the correction coefficients to correct the image to obtain scaling data; determining whether block segmentation is needed based on the scaling data, and processing the data to obtain output data; performing edge padding on the output data and then performing model inference to obtain model inference data; removing the edge padding from the model inference data to obtain de-edged model data; determining whether the original input image corresponding to the de-edged model data has undergone block segmentation, and processing the output image restoration data based on the determination. This invention dynamically adapts to operating devices with different computing power by segmenting and scaling the image, reducing image loss on the operating device and improving the display effect of the image on the operating device.
Owner:深圳牛学长科技有限公司

A method and system for detecting fake face video based on Mobilenet-GRU

The application provides a kind of based on Mobilenet-GRU's fake face video detection method and system, it is related to fake face video detection technical field, specific steps include: by being handled to face image sequence block, obtain motion consistency feature, each frame image is decomposed into low frequency and multiple high frequency regions by wavelet transform, obtain the energy ratio of low frequency and high frequency region, each frame image in face image sequence is input into Mobilenet model, obtain spatial feature vector, set the initial proportion of frame image replacement, spatial feature vector is input into GRU model and carries out time series prediction, iteration optimization the proportion of frame image replacement, using pre-trained discriminant model, the judgment of being detected video is fake.This application sets the initial proportion of frame image replacement and iteration optimizes the proportion, so that spatial feature can dynamically adapt to the change of video content, thereby significantly enhance the capture ability and expression ability of model to video dynamic change feature.
Owner:HEFEI UNIV

Self-adaptive video frame extraction method, equipment and medium

The invention discloses a self-adaptive video frame extraction method and device and a medium, and relates to the technical field of data processing. The method comprises the following steps: extracting a target frame set from a video at a fixed rate; a difference index between target frames is calculated, and the index quantifies the motion change amplitude of the target objects by calculating the mean value of the minimum matching Euclidean distances of the center positions of bounding boxes of all the target objects between the two frames; and finally, a preset difference threshold value is set, an adaptive traversal strategy is adopted, images with the target object difference between adjacent frames larger than the preset difference threshold value are screened out from the target frame set, and a final image set is formed. And the quantity of the similar image data extracted in the frame extraction process is reduced.
Owner:浪潮智慧科技有限公司 +2

Systems and methods for dynamically generating manifests that enable dynamic insertion of content during adaptive streaming of video

Techniques for serving a manifest file of an adaptive streaming video include receiving a request for the manifest file from a user device. The video is encoded at different reference bitrates and each encoded reference bitrate is divided into segments to generate video segment files. The manifest file includes an ordered list of universal resource locators (URLs) that reference a set of video segment files encoded at a particular reference bitrate. A source manifest file that indicates the set of video segment files is identified based on the request. An issued manifest file that includes a first URL and a second URL is generated based on the source manifest file. The first URL references a first domain and the second URL references a second domain that is different from the first domain. The issued manifest file is transmitted to the user device as a response to the request.
Owner:ADEIA MEDIA HOLDINGS INC

Information processing method and system based on spatialization of unmanned aerial vehicle video

The application provides an information processing method and system based on unmanned aerial vehicle video spatialization, relates to the unmanned aerial vehicle video processing technical field, and first collects camera optical parameters, camera pose information and sensor type identification and synchronously encodes into a non-display data section of an unmanned aerial vehicle video frame structure to form an encoded video with multi-dimensional space information, and then forms an encoded video stream through streaming packaging. The encoded video stream is divided into scenes for transcoding adaptation to obtain an adaptive video stream, a dynamic space mapping model is constructed by combining an image correction model after hierarchical decoding and splitting, the geographical range of a video frame is calculated to obtain a live video with a multi-reference geographical range identifier. Based on the live video, geographical data with a ground feature attribute and an identification confidence is generated and transmitted through multiple links, layered update processing is performed on the basis of the feedback modification trajectory, bidirectional dynamic synchronization and conflict resolution of the geographical data and the live video are realized, and the quality and reliability of the geographical information are improved.
Owner:JILIN PROVINCIAL PUBLIC SECURITY BUREAU

Adaptive adjustment video stream encoding optimization method, device, equipment and product

The application discloses a self-adaptive video stream coding optimization method and device, equipment and product, and relates to the technical field of multimedia communication. The method comprises the following steps: inputting image basic data of a video stream into a video feature extraction model; calculating image feature differences between video frames by the video feature extraction model to obtain target image features; determining user viewing preferences by using a pre-trained binary classifier according to the image basic data and the target image features; and performing coding optimization on a current video by a video stream coding model based on a current network state and the user viewing preferences. The pre-trained binary classifier is used to determine the user viewing preferences, the viewing preferences of different users for different video contents are determined, the viewing preferences of the users and real-time network conditions are combined, the optimization points are considered from multiple angles, the best coding configuration is selected for real-time interactive video streams, and the self-adaptability of the interactive video streams in different scenes is improved, and the user experience is improved.
Owner:PENG CHENG LAB