Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

1200 results about "Videoconferencing" patented technology

Videoconferencing is the conduct of a videoconference by a set of telecommunication technologies which allow two or more locations to communicate by simultaneous two-way video and audio transmissions. It has also been called 'visual collaboration' and is a type of groupware. Videoconferencing differs from videophone calls in that it's designed to serve a conference or multiple locations rather than individuals. It is an intermediate form of videotelephony, first used commercially in Germany during the late-1930s and later in the United States during the early 1970s as part of AT&T's development of Picturephone technology. With the introduction of relatively low cost, high capacity broadband telecommunication services in the late 1990s, coupled with powerful computing processors and video compression techniques, videoconferencing has made significant inroads in business, education, medicine and media. Like all long distance communications technologies, by reducing the need to travel, which is often carried out by aeroplane, to bring people together the technology also contributes to reductions in carbon emissions, thereby helping to reduce global warming.

Face-translator: end-to-end system for speech-translated lip-synchronized and voice preserving video generation

A neural end-to-end system is provided for the face and voice preserving translation of videos. The system is a pipeline of multiple models that produces a video of the original speaker speaking in the target language with modified lip movement to match the target speech, while preserving emphases and prosody of the original speech, and voice characteristics of the original speaker. The pipeline starts with automatic speech recognition including emphasis detection, followed by the translation model. The translated text is then synthesized by a Text-to-Speech model that recreates the original emphases in the target sentence. The resulting synthetic speech is then converted back to the original speakers' voice using a voice conversion model. Finally, to synchronize the lips of the speaker with the translated audio, a generative model generates frames of adapted lip movements which are combined with the audio to produce the final output. The disclosure further describes several use-cases and configurations that apply these techniques to video conferencing, dubbing, low-bandwidth transmission, speech enhancement and assistive technology for the hearing impaired.
Owner:WAIBEL ALEXANDER

Multi-language cross-culture communication auxiliary method and system based on large model

The invention provides a multi-language cross-culture communication assisting method and system based on a large model. The method comprises the following steps: receiving a source language audio stream during a call, calling a multi-language sound frequency harmonic modulation feature library to extract fundamental frequency harmonic intensity distribution and tone turning features, and generating a cultural acoustic fingerprint vector; based on the vector, controlling a microphone array phase difference, directionally enhancing a fundamental frequency harmonic component of a speaker and suppressing noise, and outputting a high signal-to-noise ratio spectrogram; analyzing the pronunciation rhythm and tone turning characteristics of the spectrogram, capturing the pitch jump and duration of the syllable boundary, and generating an acoustic culture label; associating the spectrogram with a target semantic library, matching harmonic distribution and a cultural context rule based on a large model, and outputting a cultural interpretation prompt containing an ambiguity resolution suggestion; and generating a calibration result according to the acoustic tag and the semantic prompt, and overlapping the dynamic floating subtitles to the face area of the speaker in the video conference picture. According to the invention, cultural tone ambiguity in multi-language communication is eliminated.
Owner:LUSTER LIGHTWAVE CO LTD

Online meeting summarization for videoconferencing

Systems and methods for generating online meeting summaries for videoconferencing are provided. For example, a computing device establishes a video conference for a plurality of participants. While the video conference is in progress, the computing device receives a first portion of a transcript of the video conference and generates a first meeting summary based on the first portion of the transcript. The computing device causes the first meeting summary to be presented in a user interface accessible by a client computing device associated with at least one of the plurality of participants. The computing device receives a second portion of the transcript of the video conference and generates a second meeting summary based on the first portion and the second portion of the transcript. The computing device causes the second meeting summary to be presented in the user interface.
Owner:ZOOM VIDEO COMM INC

Block chain-based double-stream video conference privacy tracing method

The invention provides a double-stream video conference privacy tracing method based on a block chain, and the method comprises the steps: building a tracing evidence obtaining model which comprises information uplink and tracing evidence obtaining; the information uploading comprises the following steps: in a video conference process, audio and video data and related conference information are uploaded to an interstellar file system for storage and verification after being subjected to encryption and signature processing through a threshold ring signature mechanism, and key evidence is uploaded to a block chain; the intelligent contract is used for automatically managing data storage; the traceability evidence obtaining comprises the following steps: after leaked picture information is obtained, detecting a traceability identifier in a video or a picture, and comparing the traceability identifier with information stored on a chain; a secret divulging block point is automatically tracked and positioned through an intelligent contract, and the real identity of a secret divulging party is decrypted. According to the method, tampering resistance, transparency, decentration and anonymity of chain information in the traceability evidence obtaining process are achieved, and automatic evidence obtaining and traceability operation is supported.
Owner:NANJING UNIV OF INFORMATION SCI & TECH +1

Expanding online chat communications based on chat context

Systems and methods for expanding chat communication groups based on a chat context are provided. In an example, a chat and video conference provider establishes a first chat communication group for exchanging chat messages between a plurality of client devices and generates a chat summary for a subset of the chat messages within the first chat communication group. The chat and video conference provider determines a relevant user of the first chat communication group based on the chat summary and provides a recommendation for inviting the relevant user a second chat communication group. The chat and video conference provider establishes the second chat communication group and presents the chat summary in the second chat communication group in response to the relevant user joining the second chat communication group.
Owner:ZOOM COMMUNICATIONS INC

Generating speaker video and audio in multiple languages for videoconferencing

Systems and methods for generating speaker video and audio in multiple languages for videoconferencing are provided. For example, a computing device can access a speaker speech audio signal that includes a speaker speech in a first language, a video of the speaker and a translated speech audio signal of the speaker speech in a second language. The computing device generates, based on the translated speech audio signal, a converted translated speech audio signal that includes a speech in the second language having voice characteristics in the speaker speech. The computing device further generates a lip-synched speaker video based on the video of the speaker and the converted translated speech audio signal. Lip movements in the lip-synched speaker video correspond to the converted translated speech audio signal. The converted translated speech audio signal and the lip-synched speaker video are transmitted to a video conference provider configured to host the video conference.
Owner:ZOOM VIDEO COMM INC

Generating A Unified Virtual Background Image For Multiple Video Conference Participants

A unified virtual background image is generated for multiple participants of a video conference to create an immersive conference experience based on its use within video streams of those multiple participants. Generative artificial intelligence software associated with a conferencing system obtains input associated with a video conference. The generative artificial intelligence software generates a virtual background image based on the input. The virtual background image is then for use within multiple participant video streams during the video conference
Owner:ZOOM COMMUNICATIONS INC

Identity data analysis method based on spatial frequency sensing fusion network

The invention relates to the technical field of data analysis and processing, in particular to an identity data analysis method based on a spatial frequency sensing fusion network, and the method specifically comprises the following steps: collecting identity data to be detected, and marking real identity information; preprocessing the collected identity data, and dividing the collected identity data into a training set and a test set in proportion; constructing a spatial frequency sensing multi-scale network for real-time deep identity analysis, and training the network; inputting the identity data in the test set into the trained network for forgery detection and prediction, and outputting a probability score that each identity is true or false so as to obtain an identity analysis result. By constructing a lightweight deep neural network structure which combines space and frequency feature perception and has a multi-scale fusion capability, the method can be used for efficiently detecting identity information in actual scenes such as video conferences and social media.
Owner:QILU UNIVERSITY OF TECHNOLOGY (SHANDONG ACADEMY OF SCIENCES)

Establishing a video conference during a phone call

Some embodiments provide a method for initiating a video conference using a first mobile device. The method presents, during an audio call through a wireless communication network with a second device, a selectable user-interface (UI) item on the first mobile device for switching from the audio call to the video conference. The method receives a selection of the selectable UI item. The method initiates the video conference without terminating the audio call. The method terminates the audio call before allowing the first and second devices to present audio and video data exchanged through the video conference.
Owner:APPLE INC

Network adaptive video conference transmission optimization method and system

The invention relates to the field of network communication, and particularly discloses a network self-adaptive video conference transmission optimization method and system, and the method comprises the steps: systematically analyzing a historical data sequence of a network state, and extracting a deep trend and mode of a network behavior from the historical data sequence; meanwhile, different requirements of the type of the current conference video frame on transmission parameters (especially code rates) are fully considered, and the key information is fused into a decision-making process. Through a specially designed fusion model, the historical evolution trend of the network state and the specific requirement of the current video content can be intelligently subjected to interactive perception and cooperative processing, so that a video target code rate which can adapt to future network fluctuation and can meet the current picture quality requirement is generated. Based on the dynamically optimized target code rate, the system further determines an appropriate target resolution and an appropriate target frame rate, and the video is coded and transmitted.
Owner:SHENZHEN MINRRAY IND CORP LTD

Real-time document collaborative editing and synchronizing method in video conference

The invention relates to the technical field of video conferences, and discloses a real-time document collaborative editing and synchronizing method in a video conference, which is used for solving the problem of insufficient concurrent editing conflict solving mechanism in the traditional method. The method comprises the following steps: firstly, receiving document editing operations submitted by a plurality of participants through a terminal, and establishing an operation sequence; comparing the timestamps with the position information, identifying concurrent operation groups as conflict sets, and arranging the conflict sets according to a time sequence; types and contents are extracted from the conflict set, semantic matching is carried out, matching operation is converted into a compatible sequence, and subsequent position offset is adjusted; distributing the compatible sequence to a synchronous buffer area, updating a server document state, and generating an incremental update package; a packet is pushed to the online terminal, application changes are buffered locally, and sequence playback is provided for a new participant; the monitoring terminal confirms the signal, re-deduces the unconfirmed change, and circularly checks the Hash consistency; according to the method, conflict resolution is optimized through semantic matching and offset adjustment, and efficient synchronization is achieved.
Owner:WUHAN HONGYANGUO TECH CO LTD

Accentuating communications in a video conference

Communications between participants of a video conference are accentuated based on a participant's focus. A first participant of a video conference is determined to be focused on a user interface element associated with a second participant of the video conference. Responsive to the determining that the first participant is focused on the user interface element associated with the second participant, relative audio levels of communications are changed for the second participant such that a first audio level for communications from the first participant to the second participant is higher than a second audio level for other communications received by the second participant.
Owner:ZOOM COMMUNICATIONS INC

Video conference multi-modal data alignment method and device based on causal mask, equipment and medium

The invention discloses a video conference multi-modal data alignment method and device based on a causal mask, equipment and a medium, and relates to the technical field of computers, and the method comprises the steps: carrying out the feature extraction and fusion of an original audio, an original video stream and an original document in an online video conference, time sequence division is carried out based on the obtained multi-modal fusion features to obtain a triple time sequence window; determining an initial weight value corresponding to the triple time sequence window, and performing normalization adjustment on the initial weight value by using a preset constraint condition to obtain an adjusted weight; indexing a preset time sequence offset matrix by using a speaking identifier of a speaking party, correcting an original time sequence of the triple time sequence window based on an indexing result, and determining a target attention result corresponding to the triple time sequence window by using a preset causal mask mechanism, and performing multi-level alignment fusion on the multi-modal fusion features based on the target attention result to obtain a multi-modal alignment result. The precision of the multi-mode alignment technology is improved, and future information leakage is avoided.
Owner:SHANDONG INSPUR SCI RES INST CO LTD

Dynamic Participant Engagement and Representation in a Video Conferencing Environment

Systems and methods are provided for optimizing a user interface display in a video conferencing environment. In some embodiments, the systems and methods receive conference feeds for each of a plurality of participants in a video conference. In some embodiments, the systems determine, based on historical interaction data for one or more participants of the plurality of participants, an interaction score for the one or more participants of the plurality of participants. In some embodiments, the systems generate, based on the determined interaction score, a first arrangement of representations of the conference feeds in a user interface. In some embodiments, the systems provide for presentation in the user interface the first arrangement of the conference feeds.
Owner:ADEIA GUIDES INC

Universal Identity Verification for Video Conferencing

Systems, methods, and apparatuses are described for verifying a user identity in a video conference. A computing device may receive user data and a plurality of security parameters associated with accessing a video conference based on a confidentiality level of the video conference. The computing device may generate a security code that is encoded with user data. The computing device might cause the security code to be displayed on the mobile device for a predetermined time period. The computing device may receive an indication that the first device scanned the security code by using a camera. To verify the identity of a user, the computing device may decode the security code, compare the decoded user data of the decoded security code and expected user data associated with the video conference. The computing device may determine the authenticity of a user video and allow access to the video conference.
Owner:CAPITAL ONE SERVICES LLC

Privacy preserving online video capturing and recording

Systems and methods are provided herein for only including portions of a user's environment that have been approved by a user in a video conference while excluding portions that have not been approved. This may be accomplished by a device receiving a policy identifying one or more approved objects of a scene of a video stream. The device may then generate a filtered video stream by only including portions of the scene that comprise the one or more objects that were approved by the policy in the filtered video stream. The filtered video stream may be combined with other video streams to generate a video conference that is transmitted and / or stored by one or more devices participating in the video conference.
Owner:ADEIA GUIDES INC

Performing integrity verification of content in a video conference using lighting adjustment

Systems and methods for performing integrity verification of content in a video conference using lighting adjustment are provided. An example method includes determining that an integrity verification of video content generated by a first client device of a plurality of client devices of a plurality of participants of a video conference is to be performed; causing a modified UI comprising one or more visual items, each corresponding to a video stream, to be presented on the first client device, wherein the UI was modified using a color pattern encoding; receiving, from the first client device, a video stream generated by the first client device subsequent to a presentation of the modified UI on the first client device; and verifying the integrity of the video content generated by the first client device based on the video stream generated by the first client device and the color pattern encoding.
Owner:GOOGLE LLC

Method and system for evaluating health degree of audio and video conference equipment

The invention provides a method and a system for evaluating the health degree of audio and video conference equipment, which comprehensively utilize a QFD thought and AHP and AI methods to construct an equipment fault model and a fault risk transfer model, solve the field problem of fault influence dependency analysis among an equipment principle, a real-time operation state and a historical operation state in equipment fault detection, and improve the reliability of the equipment fault detection. The evaluation method comprises the following steps: scanning a basic state in real time and detecting an environment in real time; verifying the integrity of the signal link; carrying out equipment performance standard test; the influence of the hardware loss on functionality is evaluated; performing actual use state testing, including limiting condition testing and fault injection testing; and performing health degree modeling and grading, and establishing a multi-dimensional comprehensive evaluation model. According to the method, a traditional regular maintenance mode is exceeded through multi-dimensional dynamic modeling and a predictive intervention mechanism, real-time environment sensing and hardware loss modeling are fused, a fault window can be pre-judged, the equipment operation fault rate is reduced, and preventive maintenance is achieved.
Owner:BEIJING BOSHU ZHIYUAN ARTIFICIAL INTELLIGENCE TECH CO LTD

Video processing method and related device

The invention provides a video processing method and a related device, which can be applied to a video conference scene. The method comprises the following steps: acquiring a first video stream, and performing motion estimation and motion compensation processing according to two adjacent frames of images in the first video stream to obtain an intermediate frame of image; the coding formats of the two adjacent frames of images are respectively I frames and P frames, or both frames are P frames, and the middle frame of image is coded in a B frame format. And then inserting the target number of intermediate frame images into the corresponding adjacent frames based on a frame rate required by a receiving end device to obtain a second video stream. According to the method and the device, the middle frame image is obtained through MEMC processing based on the adjacent front and back frames, the middle frame image is coded by adopting the format of the B frame, the MEMC and the B frame are calculated based on the difference value of the front and back frames, and the calculation method is similar, so that only a small number of code streams are added after frame insertion processing.
Owner:HUAWEI TECH CO LTD

Machine learning-based audio manipulation using virtual backgrounds for virtual meetings

In one embodiment, a videoconference service determines a selection of a virtual background for a videoconference from a particular participant of a plurality of participants in the videoconference. The videoconference service determines an audio context filter that is associated with a visual context of the virtual background. The videoconference service modifies an audio stream of the videoconference into a modified audio stream according to the audio context filter. The videoconference service presents, to the plurality of participants during the videoconference, the particular participant using the virtual background and the modified audio stream. In an embodiment, the videoconference service ascertains the visual context of the virtual background based on applying a machine learning model to the virtual background.
Owner:CISCO TECHNOLOGY INC

Video information processing method and system

The invention discloses a video information processing method, and relates to video transmission, and the method comprises the steps: obtaining a video stream of a video conference participant, and obtaining auxiliary information from a video conference system; dividing the auxiliary information into control type information and content type information; establishing a first transmission channel independent of a video stream for the control type information through a WebSocket or QUIC protocol; establishing a second transmission channel associated with the video stream for the content information through an SEI mechanism or an RTCP channel of the video encoder; respectively transmitting the control type information and the content type information according to the first transmission channel and the second transmission channel, and adding a serial number and a timestamp to the transmitted information; and receiving the auxiliary information at a receiving end, and performing timestamp alignment and synchronization processing on the auxiliary information and the video stream according to the serial number and the timestamp to generate a synchronous data structure. Aiming at low transmission reliability of the auxiliary information under the condition that the network condition is unstable, the application improves the transmission reliability of the auxiliary information under the condition that the network is unstable.
Owner:CHONGQING IND POLYTECHNIC COLLEGE +1

Securely recording and retrieving encrypted video conferences

One disclosed example method includes obtaining a meeting cryptographic key; transmitting, from a client device to a video conference provider, a request to initiate an encrypted video conference, the encrypted video conference including a plurality of participants; distributing the meeting cryptographic key to each participant of the plurality of participants; obtaining a public cryptographic key of a key pair, the key pair including the public cryptographic key and a private cryptographic key; encrypting the meeting cryptographic key using the public cryptographic key; transmitting, from the client device to the video conference provider, a request to record the video conference; encrypting audio and video from a microphone and image sensor of the client device using the meeting cryptographic key; transmitting the encrypted audio and video to the video conference provider; and providing the encrypted meeting cryptographic key to the video conference provider.
Owner:ZOOM COMMUNICATIONS INC

Video conference processing method and system, electronic equipment and storage medium

The embodiment of the invention provides a video conference processing method and system, and the method comprises the steps: obtaining the audio data of each conference terminal in real time in a conference process; generating structured conference summary data according to the audio data and a pre-trained artificial intelligence large model; and outputting the conference summary data. According to the embodiment of the invention, the audio data of each conference terminal is acquired in real time, and the structured conference summary data is generated by using the pre-trained artificial intelligence large model, so that the automation of the conference recording process is realized. Compared with manual recording or automatic recording only generating an original text in a traditional method, the method has the advantages that the requirement of manual participation is reduced, the conference content can be intelligently processed, and the conference summary data can be generated and output.
Owner:VISIONVERA INFORMATION TECH CO LTD

Cloud-Based Application of Visual Effects to Video

A server system receives, from a first client device, a video stream relating to a videoconferencing session and receives, from the first client device, visual effects information relating to one or more visual effects to be applied to the video stream. The server system applies, based on the received visual effects information, the one or more visual effects to the video stream to generate one or more modified video streams, and transmits the one or more modified video streams to one or more other client devices participating in the videoconferencing session.
Owner:GOOGLE LLC

Establishing private communication channels

Examples are described herein for establishing private communication channels. In various examples, camera-acquired eye tracking data if participants in a first video conference call may be monitored. Based on the monitoring, it may be determined that gazes of a subset of the participants collectively satisfy a criterion, wherein the subset includes a plurality of the participants in the first video conference call. In response to the determining that the gazes of the subset of participants collectively satisfy the criterion, a private second video conference call between the plurality of the participants is established.
Owner:HEWLETT PACKARD DEVELOPMENT COMPANY LP

Systems, methods, and apparatus for virtual meetings

In accordance with some embodiments, systems, apparatus, interfaces, methods, and articles of manufacture are provided for providing information incorporating additional data feeds, creating common arrangements, and improving performance in a virtual meeting. In some embodiments, a video conferencing system hosts a customized video conferencing and includes video conference settings data, participant display device data, supplemental data feed data, and processing instructions. The video conferencing system may identify (based on the video conference settings data) a number of data feeds for display on a display screen of each participant in a video conference and (based on the participant display device data) a screen specification for each display screen of each participant. The system may then compute an arrangement for the number of data feeds that fits within an available display area.
Owner:SCIENCE HOUSE LLC

Chat overlay in video conferences

Techniques for overlaying messages are disclosed. In an example, a method involves receiving, from a first client participating as a first virtual participant in a first virtual meeting hosted by a virtual conference provider, a first message from a first messaging channel of the first virtual meeting. The method further involves receiving, from a second client participating as a second virtual participant in a second virtual meeting, a second message from a second messaging channel of the second virtual meeting. The method involves generating an overlay that includes the first and second messages. The method further involves modifying a first video stream of the first virtual conference based on the overlay. The method further involves outputting the modified first video stream to a third participant of the first or second virtual meeting.
Owner:ZOOM COMMUNICATIONS INC

Video conference device

A video conference device, includes a display device, the display device provided with a displayer, support seat and control module, the control module including a transmitting and receiving module, touch circuit, and transmission device, the transmitting and receiving module in electric connection with the displayer, the displayer in electric connection with the transmission device, the touch circuit in in electric connection with the transmission device, the displayer provided with a lens and sound device, the displayer provided with a plurality of display screens, a support frame provided between the support seat and displayer, the support frame respectively in connection with the support seat and displayer, and the support frame provided with a press lifting device, no one will be unable to fully view the content on the display device due to a blind spot, and there will be no neck pain caused by keeping looking up at the displayer.
Owner:OXTI PTE LTD