Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

2348results about "Two-way working systems" patented technology

Method and apparatus for reducing the number of control messages transmitted by a set top terminal in an SDV system

A method is provided by which a subscriber accesses an SDV channel using a set top terminal. The method begins when the set top terminal receives a user request to tune to a first SDV channel. An active services list is also received over an access network. The active services list includes an entry for each currently available SDV program and a time-to-live (TTL) associated therewith. Tuning information is identified for the first SDV channel from its entry in the active services list. The set top terminal tunes to the first SDV channel using the identified tuning information. The channel change information associated with the user request is locally stored in set top terminal for transmission over the access network at a later time.
Owner:GENERAL INSTR CORP

Visual call information processing method and system based on 5G

The invention relates to the field of data processing, and provides a 5G-based video call information processing method and system, and the method comprises the steps: continuously obtaining a real-time video frame sequence and 5G network environment perception data in a video call scene, carrying out the multi-dimensional state mapping processing of the 5G network environment perception data, constructing a network transmission adaption model, and carrying out the real-time video frame sequence and 5G network environment perception data. Generating a video coding control instruction based on the network transmission adaptation model, performing content-aware coding conversion on the real-time video frame sequence, and outputting a coding optimization stream; in the transmission process of the coding optimization stream, link state fluctuation information is obtained through a 5G network feedback channel, transmission strategy dynamic calibration is performed on the coding optimization stream according to the link state fluctuation information, and a calibration transmission stream is obtained; and carrying out decoding time sequence alignment processing on the calibration transport stream, generating a visual call output sequence which is synchronous with the time of the original video stream unit, and pushing the visual call output sequence to a receiving end presentation device.
Owner:CHENGDU IKE IND CO LTD

Personalized digital meeting agent

A digital agent is pre-trained to be a digital proxy for a user. Taking on the persona (e.g., personality, mannerisms, preferences, knowledge, and in some cases, a realistic visual appearance and voice of the human), the digital agent can effectively act on behalf of the human. During a virtual meeting, the pre-trained digital agent can listen to what the team has to say, ask clarifying questions, answer questions on the human's behalf, and raise points the human would want the team to consider. Since the digital agent visually resembles, sounds like, and acts like the human, the digital agent appears much like other remote participants, thereby improving the meeting experience of the other attendees and facilitating meeting productivity in the absence of a human team member.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Intelligent conference video frame dynamic coding method based on multi-mode semantic understanding

The invention relates to the technical field of computer vision, in particular to an intelligent conference video frame dynamic coding method based on multi-modal semantic understanding, which comprises the following steps: acquiring a video stream sequence and a synchronous audio stream in a conference scene in real time; performing semantic analysis and decoupling on the video stream sequence, and extracting key frames and subsequent frames; extracting a sparse motion field from a subsequent frame, and segmenting a video frame into candidate visual areas including a face, a mouth shape and a background; extracting audio semantic features, executing cross-modal semantic correlation analysis, calculating semantic correlation between the sparse motion field distribution features and the audio semantic features, and positioning a pronunciation area highly related to the voice content; and calculating a quantization offset value of each candidate visual area according to the semantic relevancy, applying the quantization offset values in different areas, and packaging the quantization offset values into a variable-code-rate video code stream. According to the invention, the multi-mode semantic understanding model is constructed to carry out deep semantic analysis on the video frame content so as to realize the dynamic coding of the conference video frame.
Owner:SHENZHEN JIKEYUAN ELECTRONIC TECH CO LTD

Apparatus and method for providing healthcare services remotely or virtually with or using an electronic healthcare record and / or a communication network

An apparatus, including a computer including a database which stores a controllable healthcare record and information contained in a master records file, and a distributed ledger and Blockchain technology system. The apparatus facilitates a video call between a user device and a provider device. The computer, after processing information for updating the controllable healthcare record, processes information for identifying a plurality of electronic records for the individual and a second healthcare provider associated with each electronic record. The computer updates each of the plurality of electronic records, and generates a record update message. The computer transmits the record update message to each of a plurality of second provider devices associated with each second healthcare provider. Information regarding the update to the controllable healthcare record and the update to each of the second records is stored in the distributed ledger and Blockchain technology system.
Owner:JOAO RAYMOND ANTHONY +1

Online debate platform and method

The present invention comprises a novel social media video debating web and mobile application. The platform will provide a space for users to debate uninterrupted by both the audience and the opponent whereby each participant is given a set time to express their thoughts on a subject matter. The online debate platform provides a controlled setting for the participants to have their debates viewed, voted on and subsequently ranked by the other users of the platform. The online debate platform is also monitored by a unique AI system that updates debate “winners,” flags offensive content, and moderates each debate on the platform in real time. The disclosed platform and following figures will provide a space for individuals to debate subjects in a uniformed structure and have real-time results from active user viewership. The online debate platform aims to provide an established place for constructive debating.
Owner:VURBIL INC

Online speaker affiliation method and system based on voiceprint recognition

The invention provides an online speaker affiliation method and system based on voiceprint recognition, and the method comprises the steps: obtaining a continuous audio stream, extracting voiceprint features, comparing the voiceprint features with a candidate library containing confirmed and temporary identities, and outputting a transcription result with an identity label in real time; and continuously monitoring the temporary identity accumulation data, triggering an identity confirmation event when a confirmation condition is met, upgrading the temporary identity to a formal identity, and updating the identity labels of the historical voice segments in batches in response to the event. According to the invention, the previous affiliation result can be corrected by using the subsequently accumulated voice evidence without interrupting the real-time output, and the unification of low time delay and high accuracy is realized.
Owner:VISION INTELLIGENCE CO LTD

Universal Identity Verification for Video Conferencing

Systems, methods, and apparatuses are described for verifying a user identity in a video conference. A computing device may receive user data and a plurality of security parameters associated with accessing a video conference based on a confidentiality level of the video conference. The computing device may generate a security code that is encoded with user data. The computing device might cause the security code to be displayed on the mobile device for a predetermined time period. The computing device may receive an indication that the first device scanned the security code by using a camera. To verify the identity of a user, the computing device may decode the security code, compare the decoded user data of the decoded security code and expected user data associated with the video conference. The computing device may determine the authenticity of a user video and allow access to the video conference.
Owner:CAPITAL ONE SERVICES LLC

Secure authentication of digital humans

A video stream that depicts at least the face of an individual, and information identifying a known individual is received. Predetermined validation data derived from the known individual is accessed. An analysis of a segment of the video stream based on the predetermined validation data is performed. Based on the analysis, an output signal indicative of a confidence level that the video stream is a video stream generated by the known individual is provided.
Owner:CHARTER COMM OPERATING LLC

Selection of client connection type in a virtual meeting based on stored configuration information

Systems and methods for selection of client connection type in a virtual meeting based on stored configuration information. In response to a request of a first participant of a virtual meeting platform to join a virtual meeting, a configuration data structure associated with the virtual meeting is accessed. The configuration data structure identifies roles for participants of the virtual meeting. A first role associated with the first participant within the virtual meeting is identified using the configuration data structure. One or more client connection types are assigned to the first participant based on the first role. The first participant is allowed to join the virtual meeting using one of the one or more client connection types.
Owner:GOOGLE LLC

AI digital human conference proxy method and device under off-line local area network and medium

The invention discloses an AI digital human conference proxy method and device under an offline local area network and a medium, and the method comprises the steps: collecting conference voice in real time, converting the conference voice into a real-time text, and pre-judging a subject set in combination with a localized industry knowledge base and a user historical conference track; if it is detected that the user or the to-be-decided item is mentioned, reply voice is generated in combination with historical corpora of the user; and if the conference enters the pre-judgment topic set, calling the pre-loaded user feature packet to generate reply voice, and controlling the digital person to generate a corresponding audio and video stream. The invention provides an AI digital human conference proxy method and device under an offline local area network and a medium, and aims to realize topic pre-judgment in combination with a local knowledge base and a historical conference track of a user and provide preparation for real-time reply; meanwhile, for different scenes, the user historical corpus and the user feature packet are called respectively to generate the reply, so that the problem that the real-time personalized reply generation of the AI digital person conference agency and the conference issue pre-judgment are difficult to collaboratively realize in an offline local area network environment can be solved.
Owner:GUANGZHOU BAOLUN ELECTRONICS CO LTD

Video segmentation method, server, storage medium, and program product

The present application provides a video segmentation method, a server, a storage medium, and a program product. In the method of the present application, video data to be segmented is segmented into multiple data segments, unimodal features of the data segments, including text features of a text modality and visual features of a visual modality, are respectively extracted by means of a video topics segmentation model, and then the text features and visual features of the data segments are fused, so that the fusion of multimodal information can be performed at the intermediate representation level, the relationship and interaction between different modalities can be better captured, and higher-quality multimodal fusion features of the data segments are obtained. Furthermore, on the basis of the multimodal fusion features of the data segments, whether the data segments are topic boundaries is predicted, so that the topic boundaries of the video data can be accurately predicted, improving the accuracy of topic boundary recognition, thereby improving the accuracy and quality of video topics segmentation results.
Owner:ALIBABA (CHINA) CO LTD

Segmentation of video feed during video conference

In one aspect, a first device includes a processor and storage. The storage includes instructions executable by the processor to facilitate a video conference and to segment a camera feed into first video and second video as part of facilitating the video conference. The first video shows a first person but not a second person, the second video shows the second person but not the first person, and the camera feed shows both the first and second people. The instructions are also executable to use the first video and the second video as part of the video conference. This might include transmitting the first and second video to a second device, and / or controlling the display of the second device to concurrently but separately present the first and second video.
Owner:LENOVO (SINGAPORE) PTE LTD

Detecting generative machine learning model content

Various embodiments of the technology described herein relate to distribution-verified and authenticated content, including obtaining content and authentication data from a user device, authenticating the content based on the authentication data, and distributing the content, including an indication that the content has been verified and / or authenticated. For example, an entity depicted in the content is verified, and data depicting the entity (e.g., video and / or audio) is authenticated and distributed to various user devices.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Visual communication system based on weak current engineering and operation method thereof

The invention relates to the technical field of intelligent industry, and discloses a visual communication system based on weak current engineering and an operation method thereof, and the system comprises a multi-mode signal collection module, an intelligent data processing module, a dual-channel communication transmission module, an interactive visualization module and an autonomous control system module. The method comprises the following steps: synchronously acquiring video, audio and environmental parameters through a distributed sensor array; eliminating multi-source signal deviation by adopting a space-time alignment algorithm; feature fusion and quality evaluation are realized based on a deep neural network; dynamically selecting a wired and wireless hybrid transmission mode; an AR enhanced visual human-computer interface is constructed; and implementing a fault self-diagnosis and fault-tolerant recovery mechanism. According to the invention, the bit error rate of the communication system is reduced, the continuity and reliability of the communication process are ensured, and the man-machine interaction efficiency and the emergency response speed are improved.
Owner:SHANGHAI FANXIANG NETWORK TECH CO LTD

UI design for patient and clinician controller devices operative in a remote care architecture

A system, method and network architecture for facilitating remote care delivery involving a patient having an implantable medical device (IMD). Upon establishing a remote care session between a patient controller device and a clinician programmer, wherein the clinician and the patient are remotely located with respect to each other, telehealth consultation may be provided to the patient by the clinician via an audiovisual (AV) communications session channel. Responsive to determining that the patient requires remote therapy, one or more (re)programming instructions may be provided to the patient's IMD via a remote therapy session channel of the remote care session using an integrated graphical user interface (GUI) that may include suitable GUI controls for AV communications, remote session control as well as the remote programming of the patient's IMD.
Owner:ADVANCED NEUROMODULATION SYSTEMS INC

Artificial intelligence driven leading and templatizing of ideation session

A data processing system implements detecting an occurrence of a trigger condition during an online meeting among a plurality of client devices associated with participants of the session, the occurrence of the trigger condition indicating that an ideation session for collecting ideas from the participants should be initiated; selecting an ideation session template based on meeting information associated with the online meeting, each ideation session template comprising a natural language prompt template that includes instructions to a language model to generate an agenda for a specific type of ideation session and to conduct the ideation session according to the agenda; constructing a prompt based on the natural language prompt template of the ideation session template; and providing the prompt as an input to the language model to cause the language model to generate the agenda for the ideation session and conduct the ideation session according to the agenda.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Cross-modal conference information association retrieval method and system and medium

The invention discloses a cross-modal conference information association retrieval method and system and a medium, and relates to the technical field of artificial intelligence, and the method comprises the following steps: carrying out feature extraction on obtained multi-source heterogeneous data to obtain multi-modal features, and uniformly mapping the multi-modal features to a first feature space of a preset dimension; in the first feature space, cross-modal deep fusion processing is performed on the multi-modal features, and a joint embedding space with consistent semantics is constructed according to the cross-modal deep fusion processing; constructing a vector index database based on a multi-modal feature vector in the joint embedding space, receiving a natural language query and mapping the natural language query to the joint embedding space, executing two-stage retrieval, and then obtaining a semantic fusion score based on calculated semantic fusion scores; multiplying a time sequence reward value based on the query time deviation and a dynamic reward value based on the core word matching degree to obtain a dynamic fusion score, and performing fusion sorting on the candidate set to obtain a final sorting result and an associated retrieval result; according to the method, refined sorting of the retrieval results is realized.
Owner:UNIV OF SCI & TECH OF CHINA

Identifying a virtual meeting for absent user participation

A method for identifying a virtual meeting for absent user participation includes identifying, by a calendar application of a first user and based on one or more calendar event criteria, a calendar event that corresponds to a virtual meeting. The first user of the calendar application may be invited to attend the virtual meeting. The method includes causing a calendar application UI to be presented to the first user. The calendar application UI may include a UI element requesting that virtual meeting information be provided to the first user who is unable to attend the virtual meeting. The method includes, in response to a user activation of the UI element, providing a virtual meeting information request of the first user to a virtual meeting application.
Owner:GOOGLE LLC

Personal monitoring apparatus and methods

An apparatus, including a personal monitoring device and a drone. The personal monitoring device monitors a movement of the individual and detects a deviation from an expected position or location, or detects a deviation from an expected travel route, by the individual, generates a message containing information regarding when the individual deviated from the expected position or location or deviated from the expected travel route, and transmits the message or information contained in the message to the drone. An operation of the drone is activated, the drone travels to the position or location of the personal monitoring device, records video information or video and audio information, and transmits the video information or the video and audio information to a user communication device or to a central processing computer.
Owner:JOAO RAYMOND ANTHONY

Shared augmented reality experience in video chat

Methods and systems are disclosed for performing operations for providing a shared augmented reality experience in a video chat. A video chat can be established between a plurality of client devices. During the video chat, videos of users associated with the client devices can be displayed. During the video chat, a request from a first client device to activate a first AR experience can be received, and in response, and body parts of users depicted in the videos are modified to include one or more AR elements associated with the first AR experience.
Owner:SNAP INC

An audio and video signal processing system and a video conference terminal device using the same

The application provides an audio and video signal processing system and a video conference terminal device using the same, and relates to the technical field of Internet.The system comprises an audio and video processing module, a signal preprocessing module, a signal delay measurement module, a signal synchronization processing module and a master control module, the audio and video processing module is used for separating audio signals and video signals in input audio and video data signals, the signal preprocessing module is used for preprocessing the audio signals, the signal delay measurement module is used for measuring relative delay amounts of the audio signals and the video signals, the signal synchronization processing module is used for performing delay compensation according to the relative delay amounts, realizing the synchronization of audio and video display, and the video conference terminal device further comprises a microphone, a camera, a display screen and a sound box.The system and the video conference terminal device can test the relative delay time of audio and video signals, perform delay compensation, ensure the synchronization of output audio and video pictures, and improve the video conference quality and user experience.
Owner:BEIJING ZOBO ELECTRONIC TECH CO LTD

Video conference multi-mode real-time abstract generation method

The invention relates to the technical field of video conference data processing, and discloses a video conference multi-mode real-time abstract generation method, which comprises the following steps of: synchronously acquiring an audio stream, a video stream and a text chat record of a conference, and converting the audio stream, the video stream and the text chat record into a time-aligned text, a key frame sequence and effective chat content through preprocessing; then text semantic features, visual scene features and interactive intention features are extracted, cross-modal correlation analysis is carried out through a multi-modal fusion model, and a fusion feature set is generated; and based on the identified core issue, the key conclusion and the action item, performing structured organization according to the time sequence and the importance degree to form a real-time abstract and performing dynamic updating. According to the method, multi-dimensional information is integrated, the one-sidedness problem of a traditional single-mode abstract is solved, the integrity, accuracy and timeliness of the abstract are improved, participants are assisted in mastering key points of a conference in real time, and the conference efficiency and decision-making quality are improved.
Owner:SHENZHEN ZHONG XUN WANG LIAN SCI & TECH CO LTD

Communication method based on router bridging and router

The invention discloses a communication method based on router bridging and a router, and relates to the technical field of wireless communication. The method and the device are used for solving the problems of link stability and switching efficiency of video conference services in a multi-interference environment. A master control node scans an interference source in an environment, detects the link delay and load of bridging equipment, calculates a resource allocation priority based on interference intensity distribution and node performance, screens low-delay nodes to form a core backhaul candidate set, and marks a continuous high-interference channel as a forbidden frequency band; when the video conference flow exceeds a threshold value, dynamically allocating a DFS channel by utilizing reinforcement learning and combining with a forbidden frequency band, and establishing a pre-synchronization security key data channel; on the basis of a terminal motion state and a service label, a moving track is predicted through return link quality and a motion vector, and target node pre-binding and service data flow mirror image caching of the high-speed mobile terminal are realized; in the switching stage, a target node injects cache data and verifies continuity, rolls back a source link when abnormity occurs, and dynamically adjusts a threshold value and a track parameter.
Owner:HANGZHOU FENGHENG ELECTROMECHANICAL

Methods for implementing artificial intelligence capabilities in software applications

The invention relates to methods and systems for integrating generative artificial intelligence (AI) capabilities into Software as a Service (SaaS) platforms. It comprises maintaining AI agents with varying credentials, enabling their interaction with alphanumeric data in table structures, and implementing a hierarchical access control scheme. The system displays table structures, provides interfaces for user inputs, and allows AI agents to be added as platform users. The generative AI agents can analyze data, identify actions, and perform tasks autonomously. The invention also includes methods for proactive information gathering, interactive analysis of AI outputs, and management of AI resources as limited assets. This approach enhances SaaS functionality by enabling AI-driven task completion, data analysis, and decision-making while maintaining data security and user-specific access controls.
Owner:MONDAY COM LTD

In-vehicle conferencing with user tracking and speaker prioritization

Systems and methods for implementing smart in-vehicle conferencing with user tracking and speaker prioritization. A controller system of a vehicle may determine, based at least in part on one or more cameras, a plurality of candidate best image sources, determine, from the plurality of candidate best image sources, a best image source for a speaker in the vehicle, determine, based at least in part on the best image source, a video feed, determine, based at least in part on one or more microphones, a plurality of candidate best audio sources, determine, from the plurality of candidate best audio sources, a best audio source for the speaker, determine, based at least in part on the best audio source, an audio feed, and provide, to a conference call line, the video feed and the audio feed.
Owner:FORD GLOBAL TECH LLC

Smart visual sign language translator, interpreter and generator

The present invention relates to Smart visual sign language translator, interpreter, and generator. The invention is a process of two-way real time communication between person with hearing and / or speech disabilities (PwD) and normal person, wherein sign language of PwD (100) is captured as video, converted into image frames, interpreted using DCN trained model (104) and context based Natural language generation (105); and then converted into text (106) and speech (107) for a normal person to understand; similarly, communication of normal person is captured using mic (108), and is converted to text using speech to text converter (109) wherein speech recognition is done using Deep Stack Network. Module of Text and context analysis using Natural Language Understanding (110) analyses text received from (109). Interpreted text from (110) is used for predicting text to nearest sign image. These sign images are used for creating meaningful sequence of sign language gestures in (112).
Owner:DATACRUX INSIGHTS PTE LTD

Method and system for conducting a virtual meeting between at least one first person and one second person

Method for conducting a virtual meeting between at least one first person (4) and a second person (6), wherein the virtual meeting is displayed by means of a first display device (2) assigned to the first person (4) in the form of a first virtual environment (8) and by means of a second display device (3) assigned to the second person (6) in the form of a second virtual environment (11), wherein - the first person (4) is in reality located in a moving vehicle (5), whose vehicle movement is detected and the representation of the first virtual environment (8) is selected according to the detected vehicle movement; - the second person (6) is in reality in a reference system (7) that differs from the vehicle (5), whose movement is detected and the representation of the second virtual environment (11) is chosen according to the detected movement of the reference system (7).
Owner:AUDI AG