Multi-Camera Video Calling for Keeping Multiple Subjects in Frame

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Video calls are disrupted by network issues and camera placement, leading to subjects being outside the field-of-view or capturing unwanted distractions, which affect the call stability and continuity.

Innovation Solution

A system that utilizes multiple cameras on a device to capture and merge frames of multiple subjects, allowing for seamless inclusion and exclusion of subjects based on user input, and adjusts camera feeds to maintain all participants in the field-of-view during a video call.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If a single camera is used for video calling, then the device complexity is low, but multiple subjects cannot be captured simultaneously leading to call continuity issues

Engineering Contradiction:
Improvecall continuityVSAvoidcamera system
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The video feed is segmented into multiple separate camera feeds, each captured by a different image capture device. The system displays a preview showing multiple visually separated video frames from different cameras, allowing the user to select specific subjects from different camera views. This segmentation enables multiple subjects to be captured simultaneously while maintaining call continuity.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system merges multiple camera feeds into a unified video call stream. After the user selects desired subjects from the preview of separated camera feeds, the system generates a single frame that combines the selected subjects from different camera feeds, ensuring all participants remain visible while managing the complexity through automated composition.

Inventive Principle:
Principle #5Merging (Combining)

2Reliability

If camera placement is fixed to capture one subject, then the device complexity is low, but other subjects may be outside the field-of-view causing call disruptions

Engineering Contradiction:
Improvecall stabilityVSAvoidcamera system
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

Multiple image capture devices are employed, each capable of capturing different subjects. The system provides a preview showing multiple camera feeds that can be selected based on which subjects need to be included. This multi-functionality ensures call stability by allowing flexible subject selection without requiring complex automated tracking systems.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Reliability

If the camera field-of-view is widened to include multiple subjects, then all participants can be visible, but the image quality and focus decrease

Engineering Contradiction:
Improvecall continuityVSAvoidimage quality
Core Design Contradiction:
ReliabilityVSManufacturing precision

Solution Approach 1:

Instead of using a single wide-angle shot that compromises image quality, the system segments the video feed into multiple high-quality camera feeds. Each camera maintains optimal focus and image quality for its specific subject, and the system allows selection of desired subjects from these segmented feeds, preserving image quality while ensuring call continuity.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS20260052220A1Video calling experience for multiple subjects on a device
Publication Date: 2026.02.19 QUALCOMM INC
  • US20260052220A1 patent drawing
  • US20260052220A1 patent drawing
  • US20260052220A1 patent drawing

AI summary

Systems, methods, and computer-readable media are provided for video calling. An example method can include establishing a video call between a first device and a second device; displaying a preview of a first camera feed and a second camera feed, the first camera feed including a first video frame captured by a first image capture device of the first device and a second video frame captured by a second image capture device of the first device, the first video frame and the second video frame being visually separated within the preview; receiving a selection of a set of subjects depicted in the preview; and generating, based on the first camera feed and the second camera feed, a single frame depicting the set of subjects.