Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

314 results about "Video capture" patented technology

Video capture is the process of converting an analog video signal—such as that produced by a video camera, DVD player, or television tuner—to digital video and sending it to local storage or to external circuitry. The resulting digital data are referred to as a digital video stream, or more often, simply video stream. Depending on the application, a video stream may be recorded as computer files, or sent to a video display, or both.

Bluetooth communication intelligent speech translation method and system based on multi-mode enhancement

The invention relates to the technical field of artificial intelligence, and discloses a Bluetooth communication intelligent speech translation method and system based on multi-mode enhancement, and the method comprises the steps: collecting a multi-channel audio signal through a built-in multi-microphone array of a Bluetooth device, carrying out the dynamic direction self-adaptive beam forming of the multi-channel audio signal, and carrying out the self-adaptive beam forming of the multi-channel audio signal; extracting a Mel spectrogram feature of the direction enhancement signal, identifying lip regions of a plurality of candidate speakers in each frame of real-time speaking video captured by a camera, performing time sequence convolution on the lip regions to obtain a lip movement time sequence embedded vector, calculating a correlation score with the Mel spectrogram feature, separating the direction enhancement signal, and obtaining a lip movement time sequence embedded vector; and performing text transcription and conversion on the high-confidence separation voice to obtain a translation language text, and sending the synthesized target translation voice to a preset mobile terminal through the Bluetooth device to obtain a target translation result. According to the method, the real-time performance and accuracy of speech translation are improved in a multi-person scene, far-field speech, noise interference and accent difference.
Owner:SHENZHEN DIE MICRO SEMICON CO LTD

Integrated hardware-software computer vision system for autonomous control and inspection of substrate processing systems

A substrate processing system comprises an edge computing device including processor that executes instructions stored in a memory to process an image or video captured by camera(s) of at least one of a substrate and a component of the substrate processing system. The component is associated with a robot transporting the substrate between processing chambers of the substrate processing system or between the substrate processing system and a second substrate processing system. The cameras are located along a travel path of the substrate. The instructions configure the processor to transmit first data from the image to a remote server via a network and to receive second data from the remote server via the network in response to transmitting the first data to the remote server. The instructions configure the processor to operate the substrate processing system according to the second data in an automated or autonomous manner.
Owner:LAM RES CORP

Remote tele-biometrics for liveness detection and deepfake video identification

A system comprises a video capture module to acquire a video, a preprocessing module to enhance video quality and isolate regions of interest within the video, and a biometric data extraction module using remote photoplethysmography to extract heartbeat and SpO2 levels from the video. A machine learning module analyzes the extracted biometric data for liveness and deepfake detection, a verification module to compare the analyzed data against known biometric signatures, and a user interface to display analysis results and alerts.
Owner:PURECIPHER INC

Source state determination using machine-learning models

An audiovisual system uses a spatial position detection module to process audio signals and determine the location and orientation of a speaking participant. Based on this data, the system dynamically controls sensors to optimize audio and video capture. Behavioral and contextual information may also be used to train intelligence models for improved system performance. Further, sensor arrays may be utilized to identify gaze vectors of participants to select a camera sensor of the sensor array. Meeting content collected with the sensor array may be organized into meeting metrics in accordance with an analytics strategy before training an intelligence module with the meeting metrics. Meeting content may also be configured into digital tiles in accordance with a tile strategy. At least one of the digital tiles may be altered in response to a meeting condition detected by the sensor array.
Owner:QSC LLC

Scanning optical codes with improved energy efficiency

Systems and methods for scanning optical codes with improved energy efficiency are disclosed. In some embodiments, a disclosed method includes: obtaining a video captured by a camera; determining, based on a machine learning model, whether an optical code is included in any frame image of the video; presenting a viewable area on a display operatively coupled to the camera, once the video starts to include at least a portion of the optical code; and dynamically adjusting the viewable area based on a size of the optical code in each frame image of the video.
Owner:WALMART APOLLO LLC

System and method for online measurement of slag state in electric arc furnace steelmaking

The embodiment of the invention provides a system and method for online measurement of the slag state in electric arc furnace steelmaking, and relates to the field of steelmaking, the system comprises a front-end audio acquisition unit, a signal processing unit, a video splashing detection unit and a picture debugging display unit, the front-end audio acquisition unit is used for acquiring a noise signal in a furnace, and the signal processing unit is used for processing the noise signal; the signal processing unit is used for carrying out frequency division processing on the received noise signal, outputting a standard signal and obtaining a voiceprint slagging curve based on the standard signal, the video splashing detection unit comprises a video filtering camera and a video acquisition card and is used for acquiring an electric furnace mouth image picture, and the picture debugging display unit is used for displaying the electric furnace mouth image picture. The method is used for displaying the voiceprint slagging curve and the electric furnace mouth image. Intelligent measurement and control of the slag state in the steelmaking process of the electric arc furnace are achieved, manual intervention is reduced, the steelmaking efficiency and the molten steel quality are improved, the production cost is reduced, and wide application prospects are achieved.
Owner:新余钢铁股份有限公司 +1

Inside cabin sensor system

An inside cabin sensor system, having an integrated, single hardware part comprising a camera sub-system and an integrated radar sub-system is proposed. A proposed system provides related information for covering the following applications inside a vehicle: Child Presence Detection (CPD), Driver Drowsiness & Fatigue (F), Driver Distraction (DD) as a safety-relevant function, complemented by Intrusion & Proximity Alert (IPA), Seat Occupancy Detection (SOD), Face Recognition (FR), and optional applications such as Driver Emotion (ES), Passenger Classification (PC), Airbag Suppression (AS), Airbag Activation (AA), Mobile Phone Detection (MP), Gesture Detection (GD) and Vital Signs Detection (VS). A proposed system utilizes Artificial Intelligence (AI)-related data processing for data processing, using radar point cloud data, calculated by said radar sensors. The proposed system utilizes Artificial Intelligence (AI)-related data processing for data processing, using video-captured data. The proposed system utilizes Artificial Intelligence (AI) -related data processing for sensor data fusion, using radar sensor and video sensor data.
Owner:NOVELIC DOO BEOGRAD-ZVEZDARA

Scene-adaptive online learning for video post processing

A video capture device may be encode a set of original pictures to create encoded video data, decode the encoded video data to create a set of reconstructed pictures, determine a subset of parameters to update from among a plurality of parameters of a post-processing filter network, including using the set of original pictures as ground truth and the set of reconstructed pictures as input to the post-processing filter network, update the subset of parameters to generate updated parameters, and send the encoded video data and the updated parameters to a playback device.
Owner:QUALCOMM INC

Display method and display system of video streams

The present disclosure relates to a display method and a display system of video streams. The display method includes: obtaining a plurality of video streams associated with a plurality of video capture devices; determining a first plurality of video streams associated with a first escalator among the plurality of video streams; determining positions of a first plurality of video capture devices associated with the first plurality of video streams among the plurality of video capture devices over the first escalator; determining a first display order of the first plurality of video streams in a first plurality of display windows based on the determined positions of the first plurality of video capture devices over the first escalator; and displaying the first plurality of video streams in the first plurality of display windows according to the first display order.
Owner:KONE OYJ

Multi-stream peak bandwidth dispersal

A system may be configured to perform multi-stream bandwidth dispersal. In some aspects, the system may receive, via a communication network, a plurality of frame collections from a video capture device and one or more other video capture devices, and detect a congestion context based upon the plurality of frame collections. Further, the system may determine a schedule notification for the video capture device, the schedule notification providing instruction for transmitting a frame of a frame collection, and transmit, via the communication network, the schedule notification to the video capture device.
Owner:TYCO FIRE & SECURITY GMBH

Method for submitting video proof of the state of a trading card for remote certification

This invention relates to a method allowing collectors to submit a video proof of the condition of their collection cards to obtain remote certification by a competent third party of the condition of their filmed cards.The process mainly involves the following steps:E1—Recording by a user device of a video capturing the following steps performed by said user:E1.1—Presentation of the cardE1.2—Insertion into the protection deviceE1.3—Closure and sealing of the protection deviceE1.4—Presentation of the different sides of the sealed protection deviceE1.5—Presentation of the unique identifier of the protection deviceE2—Automatic transfer by said user device connected to a remote network of said video to a remote storage server accessible to the competent third party.
Owner:MICHAUD ADRIEN +1

System

PendingJP2026033086AData processing applicationsAddress bookEngineering
An object of a system according to an embodiment is to reduce the burden on a nursery teacher and effectively manage the safety and growth of a child.SOLUTION: A system includes a camera, an analysis unit, a generation unit, a distribution unit, an alert unit, and an analysis unit. The camera captures an image of the entire classroom. The analysis unit analyzes the video captured by the camera. The generation unit creates an address book or a daily childcare record based on the data analyzed by the analysis unit. The distribution unit distributes the address book or the daily childcare record created by the generation unit to the protector. The alert unit issues an alert based on the dangerous behavior detected by the analysis unit. The analysis unit analyzes the growth of each child based on the data collected by the analysis unit.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Information processing apparatus, control method of information processing apparatus, non-transitory computer readable medium, and system

An information processing apparatus acquires a video captured by an imaging apparatus, acquires a plurality of sounds picked up by a plurality of sound pickup apparatuses in sync with capturing of the video, acquires state information regarding an attention state of a viewer to the video, and generates audio to be reproduced with the video by combining the plurality of sounds, wherein, in a case where the state information indicates that the viewer is having an overall view of the video, a plurality of sounds picked up by a plurality of sound pickup apparatuses set inside an area being watched by the viewer in the video are combined with an equal ratio, and sounds picked up by sound pickup apparatuses set outside the area are combined with a ratio that is reduced as the distance from the area to the sound pickup apparatuses increases.
Owner:CANON KK

Spatial recall from videos

A technique creates entries in a spatiotemporal data structure that describe objects and activities in videos captured by a plurality of cameras. For instance, each entry in the spatiotemporal data structure includes different kinds of embeddings associated with a particular video captured by a camera. Each entry is further associated with a particular pose in a three-dimensional map and a particular time. In some implementations, the different kinds of embeddings include text embeddings, audio embeddings, and action embeddings, all produced using a neural network (such as a multi-modal language model). Another technique interrogates the spatiotemporal data structure by: receiving a query; mapping the query into query embeddings using the neural network; finding a particular entry in the spatiotemporal data structure that matches the query embeddings; and retrieving information associated with the particular entry.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

System and method for processing video frames depicting an animal for detection of health conditions exhibited by the animal

One variation of a method includes: during a video capture session for an animal, accessing a video feed captured at a mobile device of a user affiliated with the animal; characterizing quality of a frame of the video feed; and, in response to quality of the frame falling below a threshold quality, generating a prompt to modify a characteristic of video capture and serving the prompt to the mobile device; discarding frames of the video corresponding to quality issues to generate a filtered sequence of frames; assembling the filtered sequence of frames into a sequence of clips; predicting a condition exhibited by the animal based on body data extracted from the sequence of clips; populating a report, describing prediction of the condition, with a subset of clips, in the sequence of clips, correlated with diagnosis of the first condition; and transmitting the report to an animal health professional.
Owner:COMPANION PROFESSIONAL

Driver fatigue behavior detection method and system based on video pose invariance

The application discloses a driver fatigue behavior detection method and system based on video posture invariance, and relates to the technical field of computer vision. The application proposes a key frame selection model based on facial geometric information and a head and face action information fusion space-time network. First, the driver video captured by the vehicle-mounted camera is subjected to sequential processing, and image preprocessing is performed. Then, the key frame selection model based on facial geometric information is constructed based on the geometric features of the facial key points and a two-stage decision mechanism, and the key frames in the video sequence are extracted. Finally, the facial action modalities under any posture are extracted based on the facial forward processing, and the head posture attributes obtained based on the head posture estimation are combined to construct the head and face action information fusion space-time network, which is used for detecting the yawning, speaking, normal and other driver states. The application fully considers the head posture attributes, has high posture robustness, and can effectively distinguish the yawning and other fatigue behaviors from other driver states.
Owner:SHANDONG UNIV

Real-time evidence management system for code enforcement using body-worn and vehicle-mounted ai cameras

A method, apparatus, and system of automated code enforcement monitoring using geospatially tagged video capture is disclosed. In one embodiment, a data acquisition device is provided comprising at least one of a body-worn camera worn by a code enforcement officer, a vehicle-mounted camera, or a drone deployed from the vehicle. The device captures video data of real property together with geospatial coordinates and timestamps within a jurisdictional boundary. An evidence management server is communicatively coupled to the device through a network to store the video data, coordinates, and timestamps. The server identifies a parcel number associated with the captured property based on geospatial coordinates. A violation detection module compares timestamped video data of the parcel with previously captured data to determine modifications to a physical structure or landscaping, thereby identifying potential violations of jurisdictional codes.
Owner:GOVERNMENTGPT INC

Method, apparatus, computational equipment, and storage medium for video capture

Provided is a method for capturing video, the method including during a process of capturing video content, based on detecting missing-object content in which a target object is not included in a frame, using video content recorded prior to the missing-object content to obtain information regarding a location of the target object, a timestamp of the video content recorded prior to the missing-object content being earlier than a timestamp of the missing-object content, based on the location information obtained for the target object, adjusting a recording equipment's capturing direction, obtaining a post-adjustment recording direction of the recording equipment, decreasing a zoom factor of the recording equipment and obtaining a post-decrease zoom factor, and performing reidentification of the target object and continuing to capture video content with the target object included in the frame based on at least one of the post-decrease zoom factor and the post-adjustment recording direction.
Owner:ARASHI VISION INC

Color space based video ppg measurement method

The application relates to a color space-based video PPG measurement method, wherein the method comprises the following steps: acquiring a face video captured by an RGB camera; processing the face video, extracting a face region from the face video, and determining a region of interest; based on the region of interest, adopting a proxy variable of illumination intensity obtained from a B channel pixel value in the RGB to perform weighted processing on pixels in the region of interest, calculating a PPG value, and acquiring a PPG waveform. According to the embodiment of the application, the proxy variable of illumination intensity obtained from the B channel pixel value in the RGB can be adopted to perform weighted processing on the pixels in the region of interest, calculate the PPG value, and acquire the PPG waveform according to the region of interest, so that the interference of motion on the video PPG measurement accuracy is greatly reduced, the practicability of the video PPG is improved, the measurement accuracy of the video PPG under a complex motion scene is improved, and support is provided for reliable measurement of non-contact physiological signals.
Owner:TSINGHUA UNIVERSITY

Coal mine dynamic area personnel safety management and control method and system

The invention discloses a coal mine dynamic area personnel safety management and control method and system, and is applied to an underground coal mine fully mechanized coal mining face, and the method comprises the steps: shooting an operation site video in real time through a camera; on the basis of a KNN background modeling algorithm, extracting movement change areas which respectively appear in each video frame and generate movement change relative to a static background in the operation site video in real time; when it is judged that the camera moves and changes, re-identifying a new dangerous area displayed in a work site video through a YOLOv8 segmentation model according to the work site video shot when the camera moves and changes; and when the operator is identified to enter the new dangerous area, the alarm system is triggered to generate an alarm signal. The problem of false alarm caused by angle change or position movement of a camera in an existing personnel management and control method is solved, meanwhile, the complex terrain boundary recognition precision is improved, and the accuracy and reliability of personnel intrusion monitoring in a coal mine dangerous area are improved.
Owner:SHENHUA SHENDONG COAL GRP +1

Endoscope fault detection and processing method and system

The invention relates to the technical field of endoscope imaging, in particular to an endoscope fault detection and processing method and system. According to the scheme, firstly, a current image frame in a video shot by a camera of an endoscope is obtained; then, inputting the obtained current image frame into a pre-trained image discrimination neural network model for judgment, and finally outputting a classification type and confidence corresponding to the image information; and then, according to the classification type and the confidence coefficient, judging whether an image abnormal fault occurs in the endoscope, if the image abnormal fault occurs, determining a corresponding fault processing measure according to the classification type, performing corresponding processing on the endoscope according to the determined fault processing measure, and if the image abnormal fault does not occur, performing corresponding processing on the endoscope. And if so, continuing to acquire the next image frame. By adopting the scheme of the invention, real-time fault judgment can be carried out, and adaptive judgment and correction can be carried out on possible image abnormity, so that the endoscope fault can be effectively prevented from being continued or expanded.
Owner:MACROLUX MEDICAL TECH CO LTD

system

The system according to this embodiment aims to automate the creation of business records and reduce the burden on the user. [Solution] The system according to the embodiment comprises a recording unit, a capture unit, a recognition unit, a documentation unit, a summarization unit, and a record creation unit. The recording unit records video. The capture unit captures the video recorded by the recording unit at regular intervals. The recognition unit performs face recognition based on the video captured by the capture unit. The documentation unit documents the situation based on the information recognized by the recognition unit. The summarization unit summarizes the information documented by the documentation unit. The record creation unit automatically creates a business record based on the information summarized by the summarization unit.
Owner:SOFTBANK GROUP CORP

Video monitoring device, video monitoring system, video monitoring method, and storage medium storing video monitoring program

A video monitoring device includes processing circuitry to acquire position information indicating a position of a mobile object; to command image capturing directions of a plurality of movable cameras provided at predetermined positions; to evaluate a viewability level of the mobile object in a video captured by each of the plurality of movable cameras based on the position of the mobile object and the positions of the plurality of movable cameras; to select a movable camera for use for image capturing of the mobile object from the plurality of movable cameras based on the viewability level; and to display the video of the mobile object captured by the selected movable camera, wherein the viewability level is evaluated based on an angular speed of swiveling of each of the plurality of movable cameras necessary for each of the plurality of movable cameras to keep sight of the mobile object.
Owner:MITSUBISHI ELECTRIC CORP

Video editing support device, video editing support method, and recording medium

A video editing support device includes: a captured video acquirer that acquires a captured video capturing a state of a sport being played, in which the captured video includes an object as a subject; a stock video storage unit that stores a stock video based on the captured video; a condition storage unit that stores multiple attention mark assignment conditions for the object; an attention scene detector that detects an attention scene that satisfies an attention mark assignment condition in the stock video; an attention mark assignment unit that relates, to a detected attention scene, an attention mark corresponding to the type of an attention mark assignment condition associated with the detection; and a stock video display controller that displays the stock video on a certain display device and that displays, on the stock video, an attention mark related to an attention scene in the stock video, together with playback time information.
Owner:ASICS CORP

Selfie Display (R1021)

1. Name of the product in this design: Selfie Display Screen (R1021). 2. Purpose of this design: After connecting to a mobile phone, it is used to display images, photos and videos captured by the mobile phone, as well as various interactive information. 3. The key design feature of this product is its shape. 4. The image or photograph that best illustrates the design's key points: 3D view 1.
Owner:SHENZHEN XIANZHI IOT TECH CO LTD

Image processing device, display device, image processing system, control method and program for image processing device

The objective is to provide an image processing device, a display device, an image processing system, a control method for the image processing device, and a program that can complete image processing in accordance with the frame rate during distribution. [Solution] The image processing device 110 is connected to the HMD 120 which displays the left eye image and the right eye image in a communicative manner. The image processing device 110 includes an acquisition means (control unit 205) capable of acquiring a fisheye image, which is a video captured using a fisheye lens, and an image processing means (control unit 205) that performs image processing to reduce distortion on the fisheye image. If the fisheye image includes a left eye image and a right eye image, the image processing means performs image processing on one of the left eye image and the right eye image, and if the image processing on one image is performed within a predetermined time, it performs image processing on the other image.
Owner:CANON KK

Method for locating a vehicle in a geographical area

The invention relates to methods for locating moving objects in a geographical area. A method for locating a wheeled vehicle in a geographical area includes receiving information by means of video surveillance devices, performing recognition on said information, and determining the location of the vehicle. Information about the vehicle is obtained by means of video cameras, the locations of which are known. Subsequently, video capture results are assigned a feature confirming the location of the vehicle, and data containing the video capture results and the coordinate-confirming feature are transmitted to a remote server. Recognition is then performed on the received video capture results and the vehicle is identified. In the event of a positive identification, information about the location of the vehicle is transmitted to the user. The invention makes it possible to more accurately locate a wheeled vehicle in the case of interference impacting the operation of satellite navigation systems and in the case of a spoofing risk.
Owner:ZADOROZHNY ARTEM ANATOLYEVICH

Gaze stability test systems and methods thereof

A system and method for assessing gaze stability and vestibular function using a mobile device are disclosed. The system includes a mobile application configured to execute gaze-stability testing protocols, a display module for presenting visual targets, a head-movement and eye-tracking module utilizing real-time video captured via the device's camera, a speech-recognition module for processing verbal responses, and a data-processing module to analyze head-movement and visual-acuity data. The system enables remote patient assessments and includes protocols such as static visual acuity, visual processing, and mobile gaze stabilization tests to evaluate metrics like peak head velocity and visual acuity. Results can be processed in real-time and can be securely transmitted to clinicians for remote evaluation. The disclosed system provides a cost-effective, user-friendly telehealth solution for vestibular function assessment, eliminating the need for specialized equipment or clinical visits.
Owner:DZ BALANCE INNOVATIONS LLC