Website Building System Spider Simulation for Third-Party App Indexing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Website building systems that incorporate third-party applications face challenges in making dynamic content searchable by search engines, as existing systems fail to provide deep links to specific configurations of pages and do not allow spiders to index content from third-party applications effectively, leading to incomplete or improper indexing.
Innovation Solution
A method and system for a website building system that detects the presence of search engine spiders, parses and extracts encoded text from non-text components, and creates a search engine-friendly page by substituting HTML iframe tags, ensuring that content from third-party applications is indexed by search engines.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If third-party applications are integrated into website building systems, then functionality and content diversity are improved, but search engine indexing capability deteriorates due to inability to index dynamic content and encoded text
Solution Approach 1:
The patent introduces an intermediary component (search engine spider simulation module) that mediates between the website building system and search engines. This module captures rendered page images, extracts text from visual elements, and generates structured data that search engines can index, thereby bridging the gap between dynamic third-party content and static search engine requirements
Solution Approach 2:
The system creates a copy of the rendered webpage (as an image or structured representation) that can be processed independently. By capturing the visual output and extracting text from this copy, the system makes dynamic content from third-party applications searchable without modifying the original application code
2Ease of manufacture
If HTML iframe tags are used to embed third-party applications, then integration ease is improved, but search engine visibility deteriorates due to inability to crawl content within iframes
Solution Approach 1:
The patent introduces an intermediary processing layer that sits between the iframe content and the search engine. This layer captures the rendered content from within iframes, extracts text from visual elements, and presents it in a search-engine-friendly format, thereby making iframe content visible without changing the iframe implementation itself
Solution Approach 2:
The system transitions from relying solely on traditional HTML text dimensions to also capturing visual/textual information from rendered images. By extracting text from the visual representation of iframe content, the system adds another dimension of content accessibility for search engines
3Productivity
If dynamic content is used from third-party applications, then content freshness is improved, but search engine optimization deteriorates due to lack of proper deep linking support
Solution Approach 1:
The system performs preliminary actions by pre-generating and storing structured representations of dynamic content before search engines request it. By capturing and indexing content in advance, the system ensures that fresh dynamic content is immediately available for search engines without requiring real-time processing during crawling
Data Source
AI summary
A method and a system for a website building system (WBS) includes enabling a user to create a website page with the WBS, enabling the user to add at least one instance of a third party application (TPA) to the website page, assigning a specific instance name to the instance of the TPA and providing a permalink to specific mini-pages in the instance of the TPA including the specific instance name.


