Spoken Mobile Engine for Multimedia Stream Analysis
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current mobile messaging services, such as SMS, are limited in their ability to handle rich media and require cumbersome keypad input for text-based queries, making it difficult for users to efficiently search for services, people, or products on-the-go.
Innovation Solution
A system that captures user speech, converts it into speech symbols, and transmits these over a wireless channel to an engine for analysis, allowing for voice-based searches and multimedia messaging, including location-specific queries, and providing customized search results.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If keypad input is used for text-based queries, then text messaging capability is maintained, but user input efficiency deteriorates and operation complexity increases
Solution Approach 1:
The patent replaces the mechanical keypad input system with a voice-based input system. The mobile device captures user speech through a microphone, converts it to speech symbols, and transmits these symbols for processing. This substitution eliminates the need for manual keypad pressing, significantly improving input efficiency and ease of operation while maintaining text messaging capability.
2Adaptability or versatility
If SMS is used for messaging, then basic text communication is achieved, but rich media handling capability is limited
Solution Approach 1:
The patent extends the basic SMS messaging function to support multiple media types including text, images, audio, and video. The system processes and transmits various media formats through the speech symbol conversion mechanism, enabling a single messaging system to handle diverse content types and fulfill multiple communication needs.
3Productivity
If voice-based search is implemented, then search efficiency is improved, but speech recognition accuracy must be maintained under varying conditions
Solution Approach 1:
The patent incorporates feedback mechanisms where the system analyzes usage patterns from multiple users and refines speech recognition based on this collective data. The engine continuously improves its accuracy by learning from real-world usage, adapting to different accents, backgrounds, and speaking styles, thereby maintaining high recognition accuracy under varying conditions.
4Measurement precision
If location-specific queries are supported, then service relevance is improved, but position determination complexity increases
Solution Approach 1:
The patent introduces an intermediary position determination system that uses multiple location sources including triangulation, WiFi positioning, GPS, and assisted GPS. This intermediary layer handles the complexity of position determination, providing accurate location data to the search engine without requiring the mobile device itself to become overly complex.
Data Source
AI summary
Systems and methods are disclosed to operate a mobile device by capturing user input, transmitting the user input over a wireless channel to an engine, analyzing at the engine a music clip or video in a multimedia stream, and sending an analysis wirelessly to the mobile device.


