Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

2660results about "Two-way working systems" patented technology

Method and apparatus for reducing the number of control messages transmitted by a set top terminal in an SDV system

A method is provided by which a subscriber accesses an SDV channel using a set top terminal. The method begins when the set top terminal receives a user request to tune to a first SDV channel. An active services list is also received over an access network. The active services list includes an entry for each currently available SDV program and a time-to-live (TTL) associated therewith. Tuning information is identified for the first SDV channel from its entry in the active services list. The set top terminal tunes to the first SDV channel using the identified tuning information. The channel change information associated with the user request is locally stored in set top terminal for transmission over the access network at a later time.
Owner:GENERAL INSTR CORP

Visual call information processing method and system based on 5G

The invention relates to the field of data processing, and provides a 5G-based video call information processing method and system, and the method comprises the steps: continuously obtaining a real-time video frame sequence and 5G network environment perception data in a video call scene, carrying out the multi-dimensional state mapping processing of the 5G network environment perception data, constructing a network transmission adaption model, and carrying out the real-time video frame sequence and 5G network environment perception data. Generating a video coding control instruction based on the network transmission adaptation model, performing content-aware coding conversion on the real-time video frame sequence, and outputting a coding optimization stream; in the transmission process of the coding optimization stream, link state fluctuation information is obtained through a 5G network feedback channel, transmission strategy dynamic calibration is performed on the coding optimization stream according to the link state fluctuation information, and a calibration transmission stream is obtained; and carrying out decoding time sequence alignment processing on the calibration transport stream, generating a visual call output sequence which is synchronous with the time of the original video stream unit, and pushing the visual call output sequence to a receiving end presentation device.
Owner:CHENGDU IKE IND CO LTD

Generating participant-specific information in a virtual meeting

A method includes providing, for display on a first client device of a first participant of a plurality of participants of a virtual meeting, a user interface (UI) during the virtual meeting. The UI includes multiple regions each presenting a visual item corresponding to a video stream generated by a client device of a respective participant of the virtual meeting. The method includes detecting engagement of the first participant with a first visual item corresponding to a video stream generated by a second client device of a second participant of the virtual meeting. The method further includes generating one or more information items associated with the second participant. The method further includes causing the one or more information items to be presented within the UI on the first client device of the first participant during the virtual meeting.
Owner:GOOGLE LLC

Personalized digital meeting agent

A digital agent is pre-trained to be a digital proxy for a user. Taking on the persona (e.g., personality, mannerisms, preferences, knowledge, and in some cases, a realistic visual appearance and voice of the human), the digital agent can effectively act on behalf of the human. During a virtual meeting, the pre-trained digital agent can listen to what the team has to say, ask clarifying questions, answer questions on the human's behalf, and raise points the human would want the team to consider. Since the digital agent visually resembles, sounds like, and acts like the human, the digital agent appears much like other remote participants, thereby improving the meeting experience of the other attendees and facilitating meeting productivity in the absence of a human team member.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

User interfaces for managing live communication sessions

The present disclosure generally relates to managing live communication sessions. A computer system optionally displays an option to invite the respective user to join the ongoing communication session. A computer system optionally displays one or more options to modify an appearance of an avatar representing the user of the computer system. A computer system optionally transitions a communication session from a spatial communication session to a non-spatial communication session. A computer system optionally displays information about a participant in a communication session.
Owner:APPLE INC

Generating A Unified Virtual Background Image For Multiple Video Conference Participants

A unified virtual background image is generated for multiple participants of a video conference to create an immersive conference experience based on its use within video streams of those multiple participants. Generative artificial intelligence software associated with a conferencing system obtains input associated with a video conference. The generative artificial intelligence software generates a virtual background image based on the input. The virtual background image is then for use within multiple participant video streams during the video conference
Owner:ZOOM COMMUNICATIONS INC

Intelligent conference video frame dynamic coding method based on multi-mode semantic understanding

The invention relates to the technical field of computer vision, in particular to an intelligent conference video frame dynamic coding method based on multi-modal semantic understanding, which comprises the following steps: acquiring a video stream sequence and a synchronous audio stream in a conference scene in real time; performing semantic analysis and decoupling on the video stream sequence, and extracting key frames and subsequent frames; extracting a sparse motion field from a subsequent frame, and segmenting a video frame into candidate visual areas including a face, a mouth shape and a background; extracting audio semantic features, executing cross-modal semantic correlation analysis, calculating semantic correlation between the sparse motion field distribution features and the audio semantic features, and positioning a pronunciation area highly related to the voice content; and calculating a quantization offset value of each candidate visual area according to the semantic relevancy, applying the quantization offset values in different areas, and packaging the quantization offset values into a variable-code-rate video code stream. According to the invention, the multi-mode semantic understanding model is constructed to carry out deep semantic analysis on the video frame content so as to realize the dynamic coding of the conference video frame.
Owner:SHENZHEN JIKEYUAN ELECTRONIC TECH CO LTD

Establishing a video conference during a phone call

Some embodiments provide a method for initiating a video conference using a first mobile device. The method presents, during an audio call through a wireless communication network with a second device, a selectable user-interface (UI) item on the first mobile device for switching from the audio call to the video conference. The method receives a selection of the selectable UI item. The method initiates the video conference without terminating the audio call. The method terminates the audio call before allowing the first and second devices to present audio and video data exchanged through the video conference.
Owner:APPLE INC

Intelligent network control method for low-delay video return and related equipment

The invention relates to the field of multimedia communication and network control, in particular to an intelligent network control method for low-delay video return and related equipment. The intelligent network control method comprises the following steps: acquiring network key indexes including bandwidth, delay, jitter and packet loss rate of a network link in real time, and providing real-time network environment data support for transmission strategy adjustment. According to the method, an intelligent control mechanism combining network state perception and video content feature recognition is constructed, so that the technical problem of low-delay video return in a complex network environment is effectively solved. Specifically, key indexes of a network link are collected in real time, and a lightweight CNN model is introduced to analyze the video content activeness, so that dual perception capabilities for a network environment and content features are formed, data support is provided for dynamic adjustment of coding parameters, and accurate balance between video quality and network adaptability is realized.
Owner:IFREECOMM TECH CO LTD

Apparatus and method for providing healthcare services remotely or virtually with or using an electronic healthcare record and / or a communication network

An apparatus, including a computer including a database which stores a controllable healthcare record and information contained in a master records file, and a distributed ledger and Blockchain technology system. The apparatus facilitates a video call between a user device and a provider device. The computer, after processing information for updating the controllable healthcare record, processes information for identifying a plurality of electronic records for the individual and a second healthcare provider associated with each electronic record. The computer updates each of the plurality of electronic records, and generates a record update message. The computer transmits the record update message to each of a plurality of second provider devices associated with each second healthcare provider. Information regarding the update to the controllable healthcare record and the update to each of the second records is stored in the distributed ledger and Blockchain technology system.
Owner:JOAO RAYMOND ANTHONY +1

Online debate platform and method

The present invention comprises a novel social media video debating web and mobile application. The platform will provide a space for users to debate uninterrupted by both the audience and the opponent whereby each participant is given a set time to express their thoughts on a subject matter. The online debate platform provides a controlled setting for the participants to have their debates viewed, voted on and subsequently ranked by the other users of the platform. The online debate platform is also monitored by a unique AI system that updates debate “winners,” flags offensive content, and moderates each debate on the platform in real time. The disclosed platform and following figures will provide a space for individuals to debate subjects in a uniformed structure and have real-time results from active user viewership. The online debate platform aims to provide an established place for constructive debating.
Owner:VURBIL INC

Video conference multi-modal data alignment method and device based on causal mask, equipment and medium

The invention discloses a video conference multi-modal data alignment method and device based on a causal mask, equipment and a medium, and relates to the technical field of computers, and the method comprises the steps: carrying out the feature extraction and fusion of an original audio, an original video stream and an original document in an online video conference, time sequence division is carried out based on the obtained multi-modal fusion features to obtain a triple time sequence window; determining an initial weight value corresponding to the triple time sequence window, and performing normalization adjustment on the initial weight value by using a preset constraint condition to obtain an adjusted weight; indexing a preset time sequence offset matrix by using a speaking identifier of a speaking party, correcting an original time sequence of the triple time sequence window based on an indexing result, and determining a target attention result corresponding to the triple time sequence window by using a preset causal mask mechanism, and performing multi-level alignment fusion on the multi-modal fusion features based on the target attention result to obtain a multi-modal alignment result. The precision of the multi-mode alignment technology is improved, and future information leakage is avoided.
Owner:SHANDONG INSPUR SCI RES INST CO LTD

Online speaker affiliation method and system based on voiceprint recognition

The invention provides an online speaker affiliation method and system based on voiceprint recognition, and the method comprises the steps: obtaining a continuous audio stream, extracting voiceprint features, comparing the voiceprint features with a candidate library containing confirmed and temporary identities, and outputting a transcription result with an identity label in real time; and continuously monitoring the temporary identity accumulation data, triggering an identity confirmation event when a confirmation condition is met, upgrading the temporary identity to a formal identity, and updating the identity labels of the historical voice segments in batches in response to the event. According to the invention, the previous affiliation result can be corrected by using the subsequently accumulated voice evidence without interrupting the real-time output, and the unification of low time delay and high accuracy is realized.
Owner:VISION INTELLIGENCE CO LTD

Universal Identity Verification for Video Conferencing

Systems, methods, and apparatuses are described for verifying a user identity in a video conference. A computing device may receive user data and a plurality of security parameters associated with accessing a video conference based on a confidentiality level of the video conference. The computing device may generate a security code that is encoded with user data. The computing device might cause the security code to be displayed on the mobile device for a predetermined time period. The computing device may receive an indication that the first device scanned the security code by using a camera. To verify the identity of a user, the computing device may decode the security code, compare the decoded user data of the decoded security code and expected user data associated with the video conference. The computing device may determine the authenticity of a user video and allow access to the video conference.
Owner:CAPITAL ONE SERVICES LLC

Determination of meeting content for display by an enterprise system

Techniques for a service provider network to manage a meeting between individuals located in a conference room with one or more other individuals remote from the conference room are discussed herein. An Enterprise system can implement one or more machine learned models to generate tiles that represent different individuals associated with a meeting. A same or different machine learned model can automatically arrange a series of tiles for display that, when presented collectively, promotes inclusiveness and attention for both in-room and remote participants of the meeting. The meeting management techniques can include determining which content to include on a display device based on evaluating changes in behavior of the meeting participants over time.
Owner:AMAZON TECH INC

Secure authentication of digital humans

A video stream that depicts at least the face of an individual, and information identifying a known individual is received. Predetermined validation data derived from the known individual is accessed. An analysis of a segment of the video stream based on the predetermined validation data is performed. Based on the analysis, an output signal indicative of a confidence level that the video stream is a video stream generated by the known individual is provided.
Owner:CHARTER COMM OPERATING LLC

Selection of client connection type in a virtual meeting based on stored configuration information

Systems and methods for selection of client connection type in a virtual meeting based on stored configuration information. In response to a request of a first participant of a virtual meeting platform to join a virtual meeting, a configuration data structure associated with the virtual meeting is accessed. The configuration data structure identifies roles for participants of the virtual meeting. A first role associated with the first participant within the virtual meeting is identified using the configuration data structure. One or more client connection types are assigned to the first participant based on the first role. The first participant is allowed to join the virtual meeting using one of the one or more client connection types.
Owner:GOOGLE LLC

Data transmission control method and device, equipment and medium

The embodiment of the invention discloses a data transmission control method and device, equipment and a medium, which can be applied to scenes such as intelligent traffic, auxiliary driving, cloud technology and artificial intelligence. The method comprises the following steps: in a process of transmitting service data through a first type of network link, obtaining a network quality index of the first type of network link; if the network quality of the first type of network link is detected to be lower than a preset quality threshold value based on the network quality index of the first type of network link, selecting a target transmission strategy from candidate transmission strategies based on the network quality index of the second type of network link, the candidate transmission strategies including independent transmission of the second type of network link; the second type of network link and the first type of network link transmit together; and switching the service data from the first type of network link to a target network link corresponding to the target transmission strategy for transmission. According to the technical scheme, the reliability of controlling data transmission is high, and the data transmission efficiency is improved.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Conference state control method, system and equipment for AI digital people

The embodiment of the invention provides an AI digital human-oriented conference state control method and system, and the method comprises the steps: obtaining conference environment data; performing multi-modal analysis on the conference environment data to obtain a conference state evaluation result; generating a conference control instruction according to the conference state evaluation result; and executing a conference process control operation according to the conference control instruction. According to the embodiment of the invention, the problem of process stiffness caused by passive response in the prior art is solved, and the flexibility, naturalness and intelligence level of conference control are remarkably improved, so that the conference efficiency is effectively improved, the interaction experience of participants is enhanced, and the dynamic change requirement in a complex conference scene is met.
Owner:VISIONVERA INFORMATION TECH CO LTD

AI digital human conference proxy method and device under off-line local area network and medium

The invention discloses an AI digital human conference proxy method and device under an offline local area network and a medium, and the method comprises the steps: collecting conference voice in real time, converting the conference voice into a real-time text, and pre-judging a subject set in combination with a localized industry knowledge base and a user historical conference track; if it is detected that the user or the to-be-decided item is mentioned, reply voice is generated in combination with historical corpora of the user; and if the conference enters the pre-judgment topic set, calling the pre-loaded user feature packet to generate reply voice, and controlling the digital person to generate a corresponding audio and video stream. The invention provides an AI digital human conference proxy method and device under an offline local area network and a medium, and aims to realize topic pre-judgment in combination with a local knowledge base and a historical conference track of a user and provide preparation for real-time reply; meanwhile, for different scenes, the user historical corpus and the user feature packet are called respectively to generate the reply, so that the problem that the real-time personalized reply generation of the AI digital person conference agency and the conference issue pre-judgment are difficult to collaboratively realize in an offline local area network environment can be solved.
Owner:GUANGZHOU BAOLUN ELECTRONICS CO LTD

Video segmentation method, server, storage medium, and program product

The present application provides a video segmentation method, a server, a storage medium, and a program product. In the method of the present application, video data to be segmented is segmented into multiple data segments, unimodal features of the data segments, including text features of a text modality and visual features of a visual modality, are respectively extracted by means of a video topics segmentation model, and then the text features and visual features of the data segments are fused, so that the fusion of multimodal information can be performed at the intermediate representation level, the relationship and interaction between different modalities can be better captured, and higher-quality multimodal fusion features of the data segments are obtained. Furthermore, on the basis of the multimodal fusion features of the data segments, whether the data segments are topic boundaries is predicted, so that the topic boundaries of the video data can be accurately predicted, improving the accuracy of topic boundary recognition, thereby improving the accuracy and quality of video topics segmentation results.
Owner:ALIBABA (CHINA) CO LTD

Segmentation of video feed during video conference

In one aspect, a first device includes a processor and storage. The storage includes instructions executable by the processor to facilitate a video conference and to segment a camera feed into first video and second video as part of facilitating the video conference. The first video shows a first person but not a second person, the second video shows the second person but not the first person, and the camera feed shows both the first and second people. The instructions are also executable to use the first video and the second video as part of the video conference. This might include transmitting the first and second video to a second device, and / or controlling the display of the second device to concurrently but separately present the first and second video.
Owner:LENOVO (SINGAPORE) PTE LTD

Communication system with trust levels and related method

An electronic device and a method of operating an electronic device comprising processor circuitry is disclosed, the method comprising receiving media data associated with a media stream, the media data comprising one or both of audio data representative of audio and video data representative of video; determining a trust level associated with the media stream based on the media data; and providing an output indicative of the trust level.
Owner:GN HEARING AS

Detecting generative machine learning model content

Various embodiments of the technology described herein relate to distribution-verified and authenticated content, including obtaining content and authentication data from a user device, authenticating the content based on the authentication data, and distributing the content, including an indication that the content has been verified and / or authenticated. For example, an entity depicted in the content is verified, and data depicting the entity (e.g., video and / or audio) is authenticated and distributed to various user devices.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Visual communication system based on weak current engineering and operation method thereof

The invention relates to the technical field of intelligent industry, and discloses a visual communication system based on weak current engineering and an operation method thereof, and the system comprises a multi-mode signal collection module, an intelligent data processing module, a dual-channel communication transmission module, an interactive visualization module and an autonomous control system module. The method comprises the following steps: synchronously acquiring video, audio and environmental parameters through a distributed sensor array; eliminating multi-source signal deviation by adopting a space-time alignment algorithm; feature fusion and quality evaluation are realized based on a deep neural network; dynamically selecting a wired and wireless hybrid transmission mode; an AR enhanced visual human-computer interface is constructed; and implementing a fault self-diagnosis and fault-tolerant recovery mechanism. According to the invention, the bit error rate of the communication system is reduced, the continuity and reliability of the communication process are ensured, and the man-machine interaction efficiency and the emergency response speed are improved.
Owner:SHANGHAI FANXIANG NETWORK TECH CO LTD

UI design for patient and clinician controller devices operative in a remote care architecture

A system, method and network architecture for facilitating remote care delivery involving a patient having an implantable medical device (IMD). Upon establishing a remote care session between a patient controller device and a clinician programmer, wherein the clinician and the patient are remotely located with respect to each other, telehealth consultation may be provided to the patient by the clinician via an audiovisual (AV) communications session channel. Responsive to determining that the patient requires remote therapy, one or more (re)programming instructions may be provided to the patient's IMD via a remote therapy session channel of the remote care session using an integrated graphical user interface (GUI) that may include suitable GUI controls for AV communications, remote session control as well as the remote programming of the patient's IMD.
Owner:ADVANCED NEUROMODULATION SYSTEMS INC

Artificial intelligence driven leading and templatizing of ideation session

A data processing system implements detecting an occurrence of a trigger condition during an online meeting among a plurality of client devices associated with participants of the session, the occurrence of the trigger condition indicating that an ideation session for collecting ideas from the participants should be initiated; selecting an ideation session template based on meeting information associated with the online meeting, each ideation session template comprising a natural language prompt template that includes instructions to a language model to generate an agenda for a specific type of ideation session and to conduct the ideation session according to the agenda; constructing a prompt based on the natural language prompt template of the ideation session template; and providing the prompt as an input to the language model to cause the language model to generate the agenda for the ideation session and conduct the ideation session according to the agenda.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Cross-modal conference information association retrieval method and system and medium

The invention discloses a cross-modal conference information association retrieval method and system and a medium, and relates to the technical field of artificial intelligence, and the method comprises the following steps: carrying out feature extraction on obtained multi-source heterogeneous data to obtain multi-modal features, and uniformly mapping the multi-modal features to a first feature space of a preset dimension; in the first feature space, cross-modal deep fusion processing is performed on the multi-modal features, and a joint embedding space with consistent semantics is constructed according to the cross-modal deep fusion processing; constructing a vector index database based on a multi-modal feature vector in the joint embedding space, receiving a natural language query and mapping the natural language query to the joint embedding space, executing two-stage retrieval, and then obtaining a semantic fusion score based on calculated semantic fusion scores; multiplying a time sequence reward value based on the query time deviation and a dynamic reward value based on the core word matching degree to obtain a dynamic fusion score, and performing fusion sorting on the candidate set to obtain a final sorting result and an associated retrieval result; according to the method, refined sorting of the retrieval results is realized.
Owner:UNIV OF SCI & TECH OF CHINA

Identifying a virtual meeting for absent user participation

A method for identifying a virtual meeting for absent user participation includes identifying, by a calendar application of a first user and based on one or more calendar event criteria, a calendar event that corresponds to a virtual meeting. The first user of the calendar application may be invited to attend the virtual meeting. The method includes causing a calendar application UI to be presented to the first user. The calendar application UI may include a UI element requesting that virtual meeting information be provided to the first user who is unable to attend the virtual meeting. The method includes, in response to a user activation of the UI element, providing a virtual meeting information request of the first user to a virtual meeting application.
Owner:GOOGLE LLC