Multi-Modal Browser Form Synchronization
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing browsers do not allow users to fill out forms using both verbal and tactile interactions simultaneously, limiting accessibility for individuals who may find visual navigation difficult with audio presentations.
Innovation Solution
A multi-modal browser that uses Embedded Browser Markup Language (EBML) to provide a synchronized verbal/visual presentation, enabling users to fill out forms using verbal or tactile interaction, or a combination of both, by navigating and submitting forms through verbal or tactile commands.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If a conversational model is used for audio presentation of forms, then verbal interaction is improved, but tactile interaction capability deteriorates
Solution Approach 1:
The patent merges verbal and visual modalities into a unified form presentation system. The form is displayed visually with fields and input areas while simultaneously being presented through audio narration. Users can interact with the visual form elements (click, type, tab) while the audio provides context and guidance, combining the benefits of both conversational audio and traditional visual forms without forcing a choice between modalities.
2Adaptability or versatility
If standard visual presentation is used, then tactile interaction is maintained, but accessibility for visual navigation difficulties worsens
Solution Approach 1:
The audio presentation acts as an intermediary that bridges the gap between visual form display and user interaction. The audio narrates the current field, provides instructions, and confirms user input, making the visual form accessible to users with visual navigation difficulties while preserving the ability to interact with the visual interface through keyboard, mouse, or other tactile input methods.
3Adaptability or versatility
If multi-modal interaction is enabled, then accessibility is improved, but system complexity increases
Solution Approach 1:
The patent segments the form presentation into distinct visual and audio components that operate independently but coordinate through synchronization. The visual form contains the structural elements (fields, labels, input areas) while the audio component handles narration and interaction guidance. This segmentation allows each component to be developed and optimized separately while maintaining a unified user experience through synchronized presentation.
Data Source
AI summary
A method of synchronizing an audio and visual presentation in a multi-modal browser. A form is transmitted over a network having at least one field requiring user supplied information to a multi-modal browser. Blank fields within the form are filled in by user who provides either verbal or tactile interaction, or a combination of verbal and tactile interaction. The browser moves to the next field requiring user provided input. Finally, the form exits after the user has supplied input for all required fields. The method also provides a synchronized verbal and visual presentation by said browser by having the headings for the fields to be filled out and typing in what the user says.


