Drone Multi-View Media Capture With Location-Based Image Fusion
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Traditional digital media formats, such as 2D images and videos, are inadequate for capturing and reproducing memories with high fidelity and do not support comprehensive search and indexing mechanisms, limiting user interaction and immersive experiences.
Innovation Solution
A drone-based system and method for capturing multi-view interactive digital media representations (MIDMRs) that fuse images with location information to create content and context models, applying enhancement algorithms to generate immersive, interactive 3D-like experiences without rendering actual 3D models.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If traditional 2D digital media formats are used, then the system complexity is low, but the user interaction and immersive experience are limited
Solution Approach 1:
The patent transitions from traditional 2D digital media to multi-view interactive digital media representations (MIDMRs) that incorporate depth information and multiple viewing angles. This dimensional enhancement allows users to interact with and explore visual data from different perspectives, significantly improving user interaction and immersive experience while maintaining manageable system complexity through efficient capture and processing methods.
2Loss of information
If multi-view interactive digital media representations are generated, then the immersive experience is improved, but the memory footprint increases
Solution Approach 1:
The patent segments the visual data capture process into distinct phases: capturing a set of images from multiple viewpoints, selectively choosing a subset of these images, and generating MIDMRs only from the selected images. This segmentation allows the system to achieve immersive multi-view experiences while controlling memory usage by processing and storing only the essential image subsets rather than all captured images.
Solution Approach 2:
The patent employs parameter changes by selectively choosing images based on specific criteria (such as viewpoint quality, overlap, and relevance) to generate MIDMRs. This selective approach optimizes the balance between immersive experience quality and memory footprint by transforming the complete image set into a optimized subset that maintains essential viewing characteristics while reducing storage requirements.
3Difficulty of detecting and measuring
If comprehensive search and indexing mechanisms are implemented, then the visual data query capability is improved, but the device complexity increases
Solution Approach 1:
The patent implements preliminary action by organizing and indexing visual data during the capture and processing phase, rather than attempting to search and organize data after capture. The system establishes search-capable structures when generating MIDMRs from selected images, enabling efficient querying without requiring complex post-processing search mechanisms.
Data Source
AI summary
Various embodiments of the present disclosure relate generally to systems and methods for drone-based systems and methods for capturing a multi-media representation of an entity. In some embodiments, the multi-media representation is digital, or multi-view, or interactive, and/or the combinations thereof. According to particular embodiments, a drone having a camera is controlled or operated to obtain a plurality of images having location information. The plurality of images, including at least a portion of overlapping subject matter, are fused to form multi-view interactive digital media representations.


