Bi-directional Text Conversion Using Punctuation Idioms
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current text processing methods adulterate original text with visible markup languages like HTML and XML, leading to issues of veracity, intellectual ownership, complexity, longevity, subjectivity, clarity, and generality, making it difficult to maintain and read text while benefiting from computer processing.
Innovation Solution
A method for bi-directional conversion between language content and documents using a readable text and a text grammar, combining additional information with language content through punctuation idioms, allowing for declarative and rigorous processing without additional actions, and storing instructions on a computer-readable memory device.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Extent of automation
If visible markup languages like HTML and XML are used to process text, then computer processing capabilities are improved, but text readability and veracity deteriorate
Solution Approach 1:
The patent introduces an intermediary conversion layer between the original text and the markup language representation. Instead of directly adulterating the text with visible markup, the system converts the text into an internal representation format that can be processed by computers, then converts this representation to the desired output format. This intermediary layer preserves the original text's veracity while enabling automated processing.
Solution Approach 2:
The patent creates a copy of the text in a standardized internal representation format that separates the original text content from the processing markup. This copy can be manipulated, processed, and converted to various output formats without modifying the original text, thus preserving veracity while enabling extensive computer processing capabilities.
2Loss of information
If markup languages are used to add information to text, then information richness is improved, but text complexity and device complexity worsen
Solution Approach 1:
The patent segments the text processing into distinct components: the original text, an internal representation layer, and the markup language layer. This segmentation allows information to be added and processed separately from the original text, reducing the complexity of the text itself while maintaining information richness through the structured internal representation.
Solution Approach 2:
The patent changes the representation parameters of the text by converting it into a standardized internal format with specific structural properties. This parameter transformation enables rich information storage and processing without increasing the complexity of the original text, as the enriched representation is created through systematic parameter expansion rather than direct text modification.
3Adaptability or versatility
If multiple computer grammars are applied to text, then processing capability is improved, but text clarity and ease of operation worsen
Solution Approach 1:
The patent implements a universal internal representation format that can serve multiple functions and accommodate various processing requirements. This single standardized representation can be converted to different output formats and processed by different systems without requiring multiple separate grammars to be applied to the original text, thus maintaining clarity while providing versatile processing capability.
Data Source
AI summary
Embodiments are directed at processing language content by a method of bi-directional conversion between language content with additional information to and from documents, using a readable text and a text grammar. A method combines additional information with the language content using punctuation idioms. The combined language content and additional information remains readable by one ordinarily skilled in the art of reading and also remains allowable according to a text grammar; that is embodiments are rigorous and may be declarative. The document is compliant with a format drawn from a set which comprises SGML, XML, TEI, HTML, DOC, DOCX, ODX, PDF and XPS. The document is publishable in a medium drawn from a set which comprises a book, a magazine, a journal, a newspaper, an article and a web page. A computer-readable memory device and a computing device are also claimed.


