Systems and methods for enabling user voice interaction with a host computing device
By recognizing and processing voice interactions through content management computing devices, the problem of users being unable to interact directly in online content systems has been solved, enabling delayed interaction and resource optimization.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- GOOGLE LLC
- Filing Date
- 2016-03-18
- Publication Date
- 2026-04-21
AI Technical Summary
Existing online content systems cannot effectively delay interaction when users cannot interact directly, resulting in wasted resources and a decline in user experience.
The system retrieves online content items through a content management computing device, identifies voice interactions associated with content metadata, instructs the user's computing device to collect voice response data, identifies user requests based on the data, and delays the interaction until it is convenient for the user.
It enables delayed interaction of online content, reduces resource waste, improves user experience, and optimizes the use of computing resources.
Smart Images

Figure CN113987377B_ABST
Abstract
Description
[0001] Case Analysis
[0002] This application is a divisional application of Chinese invention patent application 201680016683.7, filed on March 18, 2016. Technical Field
[0003] This specification relates to voice interaction with computing devices, and more specifically, to methods and systems for creating and managing voice-interactive content configured to respond to voice interactions from a user. Background Technology
[0004] At least some online content (i.e., content presented to customers through online publication or online applications) is configured to be interactive, receiving actual actions such as mouse clicks or keyboard input from users (i.e., the individuals presented with the online content). These user interactions can trigger further interactions between the user and the online content provider. For example, in the case of some text- or graphic online content (such as advertisements), users can directly respond to offers, request further information, or arrange subsequent interactions with the online content provider.
[0005] However, in many situations, users receive online content while busy with other tasks. In these cases, users are less able to interact with or otherwise engage with the online content. For example, a user driving or jogging while receiving online content cannot directly respond to it because their hands are not able to interact with it. Furthermore, users cannot record or otherwise remember details of the online content for later follow-up. Therefore, methods and systems for delivering interactive online content in such scenarios are desirable. Summary of the Invention
[0006] On one hand, a content management computing device for managing voice-interactive online content is provided. The content management computing device includes a memory for storing data and a processor in communication with the memory. The processor is programmed to: retrieve online content items including content metadata; identify at least one voice interaction associated with the content metadata; dispatch the online content items to a user computing device, wherein dispatching the online content items further includes instructing the user computing device to collect voice response data in response to the at least one voice interaction; receive the voice response data from the user computing device; identify a user request based on the voice response data; and transmit a response to the user account based on the user request.
[0007] On the other hand, a computer-implemented method for managing voice-interactive online content is provided. The method is implemented by a content management computing device in communication with a memory. The method includes retrieving an online content item including content metadata; identifying at least one voice interaction associated with the content metadata; distributing the online content item to a user computing device, wherein distributing the online content item further includes instructing the user computing device to collect voice response data in response to the at least one voice interaction; receiving the voice response data from the user computing device; identifying a user request based on the voice response data; and transmitting a response to a user account based on the user request.
[0008] On the other hand, a computer-readable storage device is provided for managing voice-interactive online content, having processor-executable instructions embodied thereon. When executed by a computing device, the processor-executable instructions cause the computing device to retrieve an online content item including content metadata; identify at least one voice interaction associated with the content metadata; dispatch the online content item to a user computing device, wherein dispatching the online content item further includes instructing the user computing device to collect voice response data in response to at least one voice interaction; receiving the voice response data from the user computing device; identifying a user request based on the voice response data; and transmitting a response to a user account based on the user request.
[0009] In another aspect, a computer-implemented method for providing voice-interactive online content on a user computing device is provided. The method is implemented by the user computing device, which communicates with a memory. The method includes receiving an online content item from a content management computing device, wherein the online content item includes content metadata; identifying at least one voice interaction associated with the content metadata; distributing the online content item via a user output interface; collecting voice response data in response to the at least one voice interaction from a user input interface; and transmitting the voice response data to the content management computing device.
[0010] On the other hand, a system for managing voice-interactive online content is provided. The system includes means for retrieving online content items including content metadata. The system also includes means for identifying at least one voice interaction associated with the content metadata. The system further includes means for distributing the online content item to a user computing device, wherein distributing the online content item further includes instructing the user computing device to collect voice response data in response to at least one voice interaction. The system also includes means for receiving voice response data from the user computing device. The system further includes means for identifying a user request based on the voice response data. The system also includes means for transmitting a response to a user account based on the user request.
[0011] On the other hand, the aforementioned system is provided, wherein the system further includes means for transmitting a request for additional voice response data to a user account when it is determined, based on a user request, that further input is required.
[0012] On the other hand, the aforementioned system is provided, wherein the system further includes means for processing voice response data into a text dataset using a voice processing algorithm; and means for identifying user requests from the text dataset by applying at least one of a regular expression algorithm and a context-independent grammar algorithm.
[0013] On the other hand, the aforementioned system is provided, wherein the system further includes means for determining that a user request represents a request for a quote; means for retrieving a set of user profile information associated with a user computing device, including at least a contact dataset; and means for generating a response using the user profile information set.
[0014] On the other hand, the aforementioned system is provided, wherein the system further includes means for determining that a user request represents a request for purchase; means for identifying a purchase dataset defining the request for purchase from the user request; means for retrieving a set of user payment information associated with a user computing device; and means for transmitting the purchase dataset and the set of user payment information to an online content provider.
[0015] On the other hand, the aforementioned system is provided, wherein the system further includes means for transmitting a security request to a user account to verify that the request for purchase is authorized; means for receiving a security response from a user computing device; and means for verifying that the request for purchase is authorized.
[0016] On the other hand, the aforementioned system is provided, wherein the system further includes means for determining that a user request represents a request for a scheduled event; means for identifying a set of calendar options associated with a user computing device; and means for transmitting the request for the scheduled event including the set of calendar options.
[0017] On the other hand, the aforementioned system is provided, wherein the system further includes means for determining that a user request represents a request for more information; means for identifying a second online content item associated with an online content item, wherein the second online content item includes more information than the online content item; and means for distributing the second online content item to a user account.
[0018] On the other hand, a system for distributing voice-interactive online content is provided. The system includes means for receiving online content items from a content management computing device, wherein the online content items include content metadata. The system also includes means for identifying at least one voice interaction associated with the content metadata. The system further includes means for distributing the online content items via a user output interface. The system additionally includes means for collecting voice response data in response to at least one voice interaction from a user input interface. The system also includes means for transmitting the voice response data to the content management computing device.
[0019] The features, functions, and advantages described herein may be implemented independently in the various embodiments of this disclosure or may be combined in other embodiments, further details of which will be understood with reference to the following description and accompanying drawings. Attached Figure Description
[0020] Figure 1 This is a diagram illustrating an exemplary online content environment.
[0021] Figure 2 Is it like this? Figure 1 A block diagram of a computing device for managing, providing, displaying, and analyzing voice-interactive online content in an online content environment.
[0022] Figure 3 Is Figure 1 In the online content environment shown, using Figure 2 An exemplary flowchart of a computing device for managing and providing voice-interactive online content.
[0023] Figure 4 Is using Figure 1 An exemplary method for managing and providing voice-interactive online content in an online content environment.
[0024] Figure 5 Is using Figure 1 The online content environment, towards Figure 2 An exemplary method for displaying and providing voice-interactive online content on a user computing device.
[0025] Figure 6 It is possible Figure 1 A diagram showing the components of one or more exemplary computing devices used in the environment illustrated.
[0026] While certain features of various embodiments are shown in some figures but not in others, this is for convenience only. Any feature of any figure may be referenced and / or combined with any feature of any other figure to claim protection. Detailed Implementation
[0027] The following detailed description of the embodiments is with reference to the accompanying drawings. The same reference numerals in different drawings may denote the same or similar elements. Furthermore, the following detailed description does not limit the scope of the claims.
[0028] The systems and methods described herein overcome the challenges of delivering interactive online content by distributing content configured to receive user-interactive voice content. More specifically, in exemplary embodiments, the systems and methods are implemented by a content management computing device configured to: (i) retrieve online content items including content metadata; (ii) identify at least one voice interaction associated with the content metadata; (iii) distribute the online content items to the user computing device, wherein distributing the online content items further includes instructing the user computing device to collect voice response data in response to at least one voice interaction; (iv) receive the voice response data from the user computing device; (v) identify a user request based on the voice response data; and (vi) transmit a response to the user account based on the user request.
[0029] As described and suggested above, the system and method thereby achieve several technical effects. First, the system and method allow users to delay joining online content. As mentioned above, many existing online contents require immediate user interaction even when they are not readily available. The system and method described herein allow users to postpone this interaction until it is more convenient for them to interact with the system. For example, a user can receive an online content message and use the method and system to request and receive subsequent messages (e.g., email messages) to re-engage with the online content provider at a later time. By delaying interaction with the online content provider until it is more convenient for the user to interact, the system and method described herein avoids sending the same online content to the user repeatedly. That is, if the user requests to receive subsequent messages of the initial online content message, the initial online content message does not need to be sent to the user repeatedly. Repeatedly sending the same online content to a user consumes the processing power of the sending and receiving computing devices. Furthermore, repeatedly sending the same online content to a user requires transmission bandwidth, thus consuming computing resources. Therefore, using the methods and systems described herein to avoid repeatedly sending the same online content reduces the use of computing resources, thereby providing more efficient data processing.
[0030] Second, the systems and methods described herein provide mechanisms and architectures that can be used to define, create, manage, and distribute interactive content, and additionally receive, process, and analyze user voice response data. Therefore, the systems and methods solve the technical problem of accessing user interaction data in response to interactive online content in situations where known interaction data is inaccessible. In the embodiments and technical implementations described herein, the technical problem of data access in computer network-specific scenarios (furthermore, content distribution scenarios) is solved. Third, the systems and methods improve the technical field of content distribution. By utilizing the described architecture and system, content management computing devices access interactive content that is unavailable to content servers, publishers, and other parties. By using the content metadata in online content items, the content management computing device identifies at least one voice interaction associated with the content metadata and distributes the online content item to the user computing device, wherein distributing the online content item further includes instructing the user computing device to collect voice response data in response to at least one voice interaction. Therefore, the content metadata facilitates the reception of this additionally inaccessible information. Fourth, the systems and methods described herein provide novel solutions with unique value in computer network scenarios, and more specifically, in content distribution scenarios.
[0031] On one hand, the content management computing device implements the method described. The content management computing device is configured to retrieve, manage, and distribute online content items such as online advertisements. Online content items can be in any suitable format, including text, graphics, audio, video, or any combination thereof. In an exemplary embodiment, the online content item includes at least some audio content. In some embodiments, the online content item may not include audio content but may still respond to voice interaction.
[0032] Online content items include content metadata. Content metadata includes a description of voice interactions that can be associated with the online content item. For example, voice interactions can include voice commands responded to by the online content item. In one example, an online content item can be configured to respond to user voice commands such as “Tell me more,” “Send me a message,” and “Buy now.” Further, as described in more detail below, this content metadata can allow for a more detailed description of the voice interactions associated with the online content. This content metadata can be analyzed, parsed, and processed by multiple systems to determine which voice interactions are associated with the online content item. In one embodiment, a content management computing device is configured to parse the content metadata and determine which voice interactions are associated with the online content item. In other embodiments, a client system, such as a user computing device, can parse the content metadata and determine which voice interactions are associated with the online content item.
[0033] The content management computing device distributes online content items (such as advertisements) to a user computing device (“user computing device”). More specifically, in the context of online publication, the online content items are provided to the user computing device. In an exemplary embodiment, the online publication is audio content, and online content items are distributed within the audio content. In one example, the online publication is a music stream, and online content items are distributed between songs in the music stream.
[0034] As described herein, during the distribution of online content items, the content management computing device also instructs the user computing device to monitor user feedback associated with voice interactions. In other words, the content management computing device instructs the user computing device to collect voice response data in response to at least one voice interaction defined by content metadata. Thus, in the example above, the content management computing device could instruct the user computing device to monitor commands spoken by the user, including “tell me more,” “send me a message,” and “buy now.” These examples are described in detail below.
[0035] In an exemplary embodiment, the user computing device receives the voice response data and transmits it to the content management computing device. Therefore, the content management computing device receives the voice response data from the user computing device. In at least some examples, the user computing device may transmit the voice response data to the content management computing device in real time. In other examples, the user computing device may transmit the voice response data at periodic intervals or when appropriate data connectivity is available.
[0036] The content management computing device further processes the voice response data to identify text information. In an exemplary embodiment, the content management computing device may use any suitable audio processing algorithm to identify the text information. Further, the content management computing device uses a language processing algorithm to process the text information to identify a user request. A user request or user intent indicates an action the user wishes to perform. The content management computing device also transmits the user request to the online content provider associated with the online content item.
[0037] As used herein, a processor can include any programmable system, including systems using microcontrollers, reduced instruction set circuitry (RISC), application-specific integrated circuits (ASICs), logic circuits, and any other circuitry or processor capable of performing the functions described herein. The examples above are merely illustrative and are not intended to limit the definition and / or meaning of the term "processor" in any way.
[0038] This document describes computer systems, such as content management computing devices, user computing devices, and related computer systems. As described herein, all such computing devices and computer systems include processors and memory. However, any processor referred to herein as a computer device may also refer to one or more processors, wherein the processor may be in one computing device or multiple computing devices operating in parallel. Furthermore, any memory referred to herein as a computer device may refer to one or more memories, wherein the memory may be in one computing device or multiple computing devices operating in parallel.
[0039] As used herein, the term "database" can refer to a data ontology, a relational database management system (RDBMS), or both. As used herein, a database can include any collection of data, including hierarchical databases, relational databases, flat file databases, object-relational databases, object-oriented databases, and any other structured collection of records or data stored in a computer system. The examples above are merely illustrative and are not intended to limit the definition and / or meaning of the term "database" in any way. Examples of RDBMS include, but are not limited to, those including... Database, MySQL DB2, SQL Server And PostgreSQL. However, any database that implements the systems and methods described in this article can be used. (Oracle is a registered trademark of Oracle Corporation, Redwood Shores, California; IBM is a registered trademark of International Business Machines Corporation, Armonk, New York; Microsoft is a registered trademark of Microsoft Corporation, Redmond, Washington; and Sybase is a registered trademark of Sybase, Dublin, California).
[0040] As described above and herein, in some embodiments, the content management computing device may store user computing device identifiers, user identifiers, user-associated geographic identifiers, and user-associated transaction and shopping data, excluding sensitive personal information also referred to as personally identifiable information (PII), in order to ensure the privacy of individuals associated with the stored data. Personally identifiable information may include any information capable of identifying an individual. For privacy and security reasons, personally identifiable information may not be granted, and only auxiliary identifiers may be used. For example, data received by the content management computing device may identify user “JohnSmith” as user “ZYX123” without any method of determining the actual name of user “ZYX123”. In some examples where privacy and security are otherwise ensured (e.g., via encryption and secure storage), or with individual consent, personally identifiable information may be received and used by the content management computing device. In these examples, personally identifiable information is required to report on online user groups. Where the system described herein collects or utilizes personal information about individuals, including online users and merchants, that individual is given the opportunity to control whether such information is collected or to control whether and / or how such information is used. Furthermore, some data may be processed in one or more ways before being stored or used, resulting in the removal of personally identifiable information. For example, an individual's identity may be processed to the point that their personal identification information cannot be determined, or the geographic location of an individual whose location data is obtained may be generalized (such as city, postal code, or state level), making it impossible to determine the individual's specific location.
[0041] In one embodiment, a computer program is provided, and the program is embodied on a computer-readable medium. In an exemplary embodiment, the system is executed on a single computer system without requiring a connection to a server computer. In a further embodiment, the system can be... The system runs in an environment (Windows is a registered trademark of Microsoft Corporation, Redmond, Washington). In yet another embodiment, the system runs in a mainframe environment and Runs in a server environment (UNIX is a registered trademark of X / Open Company Limited, located in Reading, Berkshire, United Kingdom). The application is flexible and designed to run in a variety of different environments without compromising any of its core functionality. In some embodiments, the system includes multiple components distributed across multiple computing devices. One or more components may be in the form of computer-executable instructions embodied in a computer-readable medium.
[0042] As used herein, an element or step described in the singular and followed by the word "a" should be understood to not exclude multiple elements or steps unless such exclusions are explicitly stated. Furthermore, references to "exemplary embodiments" or "an embodiment" in this disclosure are not intended to be construed as excluding the existence of additional embodiments that also include the described features.
[0043] As used herein, the terms “software” and “firmware” are used interchangeably and include any computer program stored in memory and executed by a processor, including RAM memory, ROM memory, EPROM memory, EEPROM memory, and non-volatile RAM (NVRAM). The memory types described above are merely examples and are not intended to limit the types of memory that can be used to store computer programs.
[0044] As used herein, the term "online content" can refer to any form of communication that identifies and / or promotes (or otherwise notifies) one or more products, services, ideas, messages, people, organizations, or other items. "Online content" refers to various types of information presented via the web, software applications, and / or other means, including articles, discussion topics, reports, analyses, financial statements, music, videos, graphics, search results, webpage listings, information feeds (e.g., RSS feeds), television broadcasts, radio broadcasts, print publications, or any other form of information that can be presented to a user using a computing device. In one embodiment, "online content" can refer to advertising ("advertising").
[0045] Advertising is not limited to commercial promotions or other communications. Advertising can be public service announcements or any other type of notification, such as notices published in print or electronic news or broadcast. Advertising can be referred to as sponsored content.
[0046] Advertising can be delivered through a variety of media and in a variety of formats. In some examples, advertising can be delivered through interactive media such as the Internet, and advertising can include graphic ads (e.g., banner ads), text ads, image ads, audio ads, video ads, ads combining any one or more of these components, or any form of electronic delivery advertising. Advertising can include embedded information such as embedded media, links, metadata, and / or machine-executable instructions. Advertising can also be delivered via RSS (True Simple Aggregator) feeds, radio channels, television channels, print media, and other media.
[0047] The term "advertisement" can refer to a single "ad creative" or an "ad group." An ad creative is any entity that represents an ad flash. An ad flash is any form of display of an ad so that a user can see / receive it. In some examples, an ad flash can occur when an ad is displayed on a display device accessed by the user or when an ad is played on the user's accessed device. An ad group, for example, refers to an entity that represents a set of creatives sharing common characteristics—such as having the same ad selection and recommendation criteria. Ad groups can be used to create ad campaigns.
[0048] As used herein, “content metadata” refers to “data about data” that describes voice interactions associated with content such as online content. Specifically, this content metadata can be descriptive metadata that describes individual instances of voice interactions associated with specific online content.
[0049] As used herein, "voice interaction" and related terms can refer to any interaction associated with online content. In an exemplary embodiment, content metadata describes voice interactions associated with specific online content. The system enables a user computing device to monitor and capture voice responses received by the user in conjunction with the display of online content. The user computing device captures these voice responses as "voice response data."
[0050] The systems and processes are not limited to the specific embodiments described herein. Furthermore, each component of each system and each process can be implemented independently of and separately from other components and processes described herein. Each component and process can also be used in conjunction with other packages and processes.
[0051] As described above, the content metadata used by the system defines at least one voice interaction associated with online content. These voice interactions are identified and used to capture voice response data at the user's computing device. In an exemplary embodiment, examples of descriptive content metadata are given in the following examples (Table 1):
[0052]
[0053] Table 1
[0054] Table 1 includes four illustrative examples of voice interactions associated with different online content items. As described below and herein, in other examples, multiple voice interactions may be associated with a given online content item. However, for simplicity, Table 1 identifies only one voice interaction for each online content item. Furthermore, the types of voice interactions shown in Table 1 are illustrative and not restrictive. Therefore, other voice interactions (including those described below) may be associated with other online content items.
[0055] As shown in Table 1, the online content item “ABC123” could be an advertisement for services such as a gym or fitness center. When “ABC123” is displayed to a user on their computing device, it can offer promotions with special offers (e.g., discounted gym memberships) and prompts to call the gym now to request a membership. As described herein, the user may not be able to call the gym when the advertisement is displayed. Therefore, the content management computing device enables the user's computing device to allow the user to interact with the online content item “ABC123” using the voice interaction “SEND_PH_NUMBER” (as identified in the “Interaction Label”). When executed, SEND_PH_NUMBER allows the user to request the gym’s phone number to be provided to the user. In one example, the user can respond to the voice prompt at the end of the display of “ABC123”. For example, the display of “ABC123” could end with the user's computing device (via visual or audio output) providing the message “If you want our phone number, please say yes”.
[0056] The content management computing device causes the user computing device to listen to responses and collect voice response data from the user within a configurable time period. In other words, the voice interaction begins after "ABC123" is displayed. In an exemplary embodiment, the content management computing device causes the user computing device to listen for five seconds. In other embodiments, the listening period can be configured in the content metadata. Alternatively, the listening period can be controlled using settings of the content provider, the content management computing device, and the user computing device.
[0057] Content metadata may also include “interaction parameters” that further define the voice interaction. Specifically, interaction parameters define parameters that are monitored when the content management computing device parses and analyzes the voice response data. In an exemplary embodiment, SEND_PH_NUMBER includes the interaction parameter for a phone number. Similarly, the content metadata enables the user computing device to listen for a contact number (e.g., a mobile phone number) that can be sent to it by the gym's phone number. Upon receiving a “yes” voice response from the user computing device (indicating that the user wants the gym's phone number), the content management computing device sends the gym's phone number to a messaging account associated with the user computing device. If the user provides a phone number with text messaging capabilities (in response to the phone number's interaction parameter), the message can be sent via text. Alternatively, the messaging account associated with the user computing device may be an email address associated with the user computing device detected by the content management computing device based on previous or current interactions with the user computing device. Therefore, if no phone number is provided, the gym's phone number can be sent via email. In alternative embodiments, the messaging account may be any suitable message format, including SMS, text messages, and instant messages. In further embodiments, the system's response may be sent via an application, including a web-based application.
[0058] Content metadata may also include “interactive responses.” Interactive responses reflect what might happen after the user computing device attempts to collect voice response data. In a specified example, an interactive response includes “providing opening hours via audio.” In this example, upon receiving a “yes” voice response, the user computing device could provide an audio message with the gym’s opening hours. In other examples, other forms of follow-up might occur in the interactive response. In one example, the interactive response might include a second prompt message that provides a new audio message and allows the user to listen to another voice interaction. For example, an alternative form of interactive response like “ABC123” might lead the user computing device to provide the user with the message “What type of membership are you interested in?” and allow the user to listen to a second voice interaction. Because user bandwidth may vary (e.g., when the user computing device migrates between data networks), in some examples, the audio associated with the interactive response may be pre-downloaded to avoid delayed delivery of the interactive response or to avoid using unwanted data networks (e.g., cellular roaming networks).
[0059] Content metadata may further include email format types, which allow online content providers to specify the type of subsequent email sent in response to user voice response data. In some examples, messaging accounts may have preferred certain email formats (or message formats), and, where appropriate, content management computing devices may match these email format types to messaging accounts accordingly.
[0060] In the second example, the online content item “DEF456” is associated with the voice interaction “SEND_PROD_OFFER”. DEF456 includes content describing several product offers promoted by a specific merchant. As suggested, SEND_PROD_OFFER is a voice interaction that enables the user's computing device to monitor the user's request for details of the product offers. For example, “DEF456” could include audio content describing a clothing sale and ending with the statement, “If you would like to know more about this sale, please say ‘Send me the details’ and specify the product you would like to hear about!” After parsing and analyzing DEF456 and identifying SEND_PROD_OFFER, the content management computing device enables the user's computing device to dispatch DEF456 and listen to the user's response “Send me the details” during a specified listening period. In SEND_PROD_OFFER, the interaction parameters include the product name, product attributes, and product quantity. Therefore, the content management computing device enables the user's computing device to listen to these parameters in the voice response data. Upon completion, SEND_PROD_OFFER also provides an audio offer summary describing the available offers. In this example, SEND_PROD_OFFER can be sent as a plain text email or an HTML email.
[0061] In the third example, the online content item “GHI789” is associated with the voice interaction “RESERVE_LOCATION”. “GHI789” includes content describing the services offered by industries such as restaurants. As suggested, RESERVE_LOCATION is a voice interaction that enables the user computing device to listen to a user's reservation request in an advertising business. For example, “GHI789” could include audio content describing a restaurant's special discounts and end with the statement “Reserve now!” RESERVE_LOCATION includes several interaction parameters, including date, time, location, and number of people. In one example, the content management computing device enables the user computing device to listen to these interaction parameters and, if possible, create a reservation for the restaurant. In the second example, the user computing device can communicate with the user's calendar. In these examples, the user computing device can identify and access the user's calendar and identify available space on the user's calendar that can be offered to the restaurant. In at least some examples, RESERVE_LOCATION may also include follow-up requests for information not present in the calendar, such as the number of people.
[0062] In the fourth example, the online content item “JKL012” is associated with the voice interaction “PRODUCT_PURCHASE”. “JKL012” includes content describing the services offered for a product that may be purchased. Although JKL012 is similar to DEF456, PRODUCT_PURCHASE allows the user to specifically request the purchase of a product or service. Compared to SEND_PROD_OFFER, PRODUCT_PURCHASE also collects interaction parameters for the payment method, allowing the user's computing device to provide payment data. In the first example, the payment method is provided based on the user's voice interaction. In the second example, the payment method is provided through the user's computing device or software associated with the user's computing device—including, for example, an e-wallet or a web-based wallet. In these examples, the content management computing device may also require the user's computing device to receive secure input (e.g., a password or PIN) to verify the user's authorization to access the payment method specified in the payment method.
[0063] In an optional example, as described above, the content management computing device facilitates deferred interaction between the user and the online content provider. In the first example (e.g., the example of the online content item "JKL102" associated with the voice interaction "PRODUCT_PURCHASE"), the content management computing device (or an associated device including the online content provider's computing device) can send order details to the user. These order details can be sent to the user via email or any other suitable medium. The order details can be sent to the user's computing device or another computing device accessible to the user. Upon receipt, the order details are configured to allow the user (via the accessed computing device) to review and approve, cancel, or modify the order by interacting with the order details.
[0064] In the second example (e.g., the example of the online content item "GHI789" associated with the voice interaction "RESERVE_LOCATION"), the content management computing device (or related devices including the online content provider's computing device) can send appointment details to the user. These appointment details can be sent to the user via email or any other suitable medium. The appointment details can be sent to the user's computing device or another computing device accessible to the user. Upon receipt, the appointment details are configured to allow the user (via the accessed computing device) to review and approve, cancel, or modify the appointment through interaction with the order details.
[0065] As described herein, optional voice interaction and combinations of voice interaction can be provided. Further additional interaction parameters can be collected for the aforementioned voice interaction types or any alternative types.
[0066] As described above and herein, the content management computing device is configured to transmit responses to an account associated with the user (“user account”) based on a user request. As noted above, these responses may include messages with contact information of an online content provider, confirmation of appointment details, confirmation of order details, quote details, or any other follow-up messages created based on voice interaction. The user account can be any account associated with a user identified based on user characteristics such as a user profile. In one example, the user account is an online application account. In a second example, the user account is an email account. In a third example, the user account is a messaging account used for any suitable messaging protocol. As described herein, the user account can be accessed via the user computing device or other computing devices, including auxiliary user computing devices, as described below.
[0067] The foregoing and Table 1 describe several variations of voice interaction types that can be used in content metadata associated with online content. In addition to the descriptive content metadata described above, structural metadata and associated syntax are defined to ensure that online content can use a consistent data format and structure when communicating with content management computing devices and / or user computing devices.
[0068] Structured metadata can be provided to online content publishers and providers (e.g., advertisers) by content management computing devices. This structured metadata may also include acceptable metadata syntax. In some examples, structured metadata is defined and provided, including standardization tools, which include, but are not limited to, controlled vocabularies, taxonomies, thesauri, data dictionaries, and metadata registries. Structured metadata can be provided using any suitable format, including plain text, Rich Data Format (RDF), Hypertext Markup Language (HTML), and Extensible Markup Language (XML).
[0069] In an exemplary embodiment, the structural metadata defines the confirmed set of voice interaction types, the parameters associated with each voice interaction type, the interaction response associated with each voice interaction type, and the email format associated with each voice interaction type. Furthermore, the structural metadata defines the layout, format, and syntax of the content metadata.
[0070] As described above, multiple parties can receive structured metadata used to create content metadata. In at least one example, a content provider (e.g., an advertiser) can create content metadata and embed it into online content items. In other examples, content publishers, content management computing devices, and other parties can create content metadata and embed it into online content items. In one example, a content provider can send a request to a content management computing device. This request can be used to modify a specific online content item created by the content provider to include specific voice response data. In these examples, the content management computing device can edit online content to include voice interaction metadata.
[0071] As described above, multiple systems can analyze online content to determine the presence of content metadata. In an exemplary embodiment, a content management computing device can scan online content items to identify content metadata. Because the content metadata is constructed in a manner specified or published by the content management computing device, the content management computing device can identify this content metadata. Specifically, the content metadata records at least one version of the content metadata format and definition in a memory or accessible memory that can be used when scanning online content items.
[0072] After identifying the presence of content metadata in online content items, the content management computing device analyzes the content metadata to identify voice interactions associated with the online content items. Furthermore, the content management computing device can identify interaction parameters, email formats, and interaction responses associated with the online content items. These identified voice interactions and other attributes are used when the content management computing device dispatches online content items to a user computing device. Specifically, as described, the content management computing device dispatches online content items to the user computing device and sends instructions to the user computing device to listen for or monitor voice response data for a period of time after the online content items are dispatched. Additionally, the content management computing device sends instructions to the user computing device to send the collected voice response data back to the content management computing device. Furthermore, the content management computing device can send additional instructions to dispatch interaction responses based on the collected voice response data.
[0073] In at least some examples, the content management computing device may also use the user computing device to identify and analyze content metadata. In these examples, the user computing device at least partially identifies, parses, and analyzes the content metadata and determines how to dispatch voice response data associated with the content metadata. Therefore, the user computing device at least partially dispatches voice interactions with online content items. In these examples, the content management computing device may provide the user computing device with programming (e.g., scripts, plugins, or applications) that can be used to identify and dispatch voice interactions.
[0074] After collecting voice response data, the user computing device transmits this data to the content management computing device. The content management computing device processes the voice response data to identify user requests. In other words, the content management computing device processes the voice response data based on the voice interaction (as shown in Table 1, for example) and identifies the meaning of the voice response data. In one example, the content management computing device uses a voice processing algorithm to process the voice response data into a text dataset and further identifies user requests from this text dataset by applying at least one of a regular expression algorithm and a context-independent grammar algorithm.
[0075] In some examples, the content management computing device can access user profile information associated with a user's computing device. This user profile information may include, for example, user calendar information, user contact information, and user payment information. In at least one example, the content management computing device determines that a user request identified based on voice response data represents a request for a quote. For example, the content management computing device may determine that the voice response data responds to SEND_PROD_OFFER (as shown in Table 1 above). In these examples, the content management computing device may also retrieve a set of user profile information associated with the user's computing device, including at least a contact dataset, and use this user profile information set to generate a response.
[0076] In another example, the content management computing device can specifically access user payment information associated with the user's computing device (e.g., data associated with payment methods, as shown above). For example, the user's computing device can determine that a user request represents a purchase request because the voice response data responds to the voice interaction PRODUCT_PRUCHASE (as shown in Table 1 above). The content management computing device can also identify a purchase dataset from the user request that defines the purchase request. In other words, the content management computing device can identify the purchase item requested in the voice response data (e.g., the product and quantity being attempted to purchase). The content management computing device can also retrieve a set of user payment information associated with the user device and send this purchase dataset and user payment information set to the online content provider. Therefore, the content management computing device can allow online content providers to sell goods based on the collected voice response data.
[0077] In some examples, the use of payment data may have security limitations. In at least one example, the content management computing device is configured to send a security request to the user device to verify that the request for purchase is authorized. For example, the content management computing device may send the authentication request based on a password, biometric data, a PIN code, or any other suitable security protocol. In this exemplary embodiment, the content management computing device sends a general request to the user computing device to verify the user, without requesting actual security data. In this example, the content management computing device receives a secure response from the user device indicating whether the user is authorized to purchase goods or services (but not indicating the user's private information). The content management computing device verifies that the request for purchase is authorized.
[0078] In some examples, the user computing device can also be used to analyze collected voice response data. For instance, by using a client-server architecture, the content management computing device can provide the user computing device with software or other tools that can process the voice response data on the user computing device. Therefore, the user computing device can analyze the voice response data and send the parsed and analyzed data to the content management computing device in a non-audio format, such as a text file. In these examples, less data consumption can be achieved because the voice data file is not transferred from the user computing device to the content management computing device.
[0079] In some examples, the content management computing device may determine that the voice response data is incomplete. For example, the collected information may not fully respond to the voice interaction. In these examples, the content management computing device may determine that this voice response data is incomplete and, if it determines that further input is needed based on the user request, further transmit a request for additional voice response data to the user device. In some embodiments, the user computing device may also be configured to analyze the voice response data and determine whether a request for additional voice response data is necessary.
[0080] In some examples, the content management computing device can also determine that a user request represents a request for a scheduled event. For example, the user computing device can determine that a user request represents a request for a scheduled event because the voice response data is in response to the voice interaction RESERVE_LOCATION (as shown in Table 1 above). In these examples, the content management computing device can determine that a user request represents a request for a scheduled event, thereby identifying a set of calendar options associated with the user device and transmitting a request for a scheduled event that includes the set of calendar options. Identification of the calendar options set can be performed by retrieving user profile information that includes the user's calendar.
[0081] In a further example, the content management computing device is configured to determine that a user request represents a request for more information. For example, a user might provide a response to a voice interaction that is a question requesting more information. In these examples, the content management computing device can identify a second online content item associated with an online content item, wherein the second online content item includes more information than the online content item, and dispatch the second online content item to the user device.
[0082] Based on a user request, the content management computing device can identify at least one response. For example, based on a user request associated with SEND_PH_NUMBER, the content management computing device can determine that at least one response includes sending a phone number associated with an online content item to the user's computing device. In the case of a user request associated with SEND_PROD_OFFER, the content management computing device can determine that at least one response includes sending a product quote for a user-identified product in the voice response data. In the case of a user request associated with RESERVE_LOCATION, the content management computing device can determine that at least one response includes sending a reservation request to an online content provider (e.g., a merchant) and sending a confirmation to the user's computing device when it is determined that the merchant can fulfill the reservation. In the case of a user request associated with PRODUCT_PURCHASE, the content management computing device can determine that at least one response includes sending a purchase request to an online content provider (e.g., a merchant) and also sending a confirmation to the user's computing device after processing the purchase. Alternatively, in some examples, the content management computing device can determine that at least one response includes sending a purchase request to an online content provider (e.g., a merchant) and also sending a confirmation to a second user computing device (different from the first user's computing device) after processing the purchase. Sending at least one response to a second user computing device essentially allows users to interact with online content providers deferredly and using multiple computing devices. As mentioned above, in many examples, users may prefer to use different devices and interact with online content providers (and online content items) at different times. Similarly, in all the examples described herein, the content management computing device can be configured to communicate with these second user computing devices. Because different computing devices have different display and interaction characteristics (e.g., varying screen sizes and input interfaces), users may prefer to redirect interactions from the user computing device to these second user computing devices.
[0083] In at least some examples, the content management computing device is configured to receive requests to redirect communications, including responses, to these second user computing devices. For example, the content management computing device may identify the second user computing devices based on user profile information or based on voice response data. Thus, in some examples, the content management computing device requests information based on a user profile used to identify a set of contact information, including information to identify the second user computing device or methods of contacting these second user computing devices (including, for example, email addresses, account names, and other identifiers). In other examples, voice interactions (such as those described above) may be configured to prompt the user to identify the second user computing device in the voice interaction. Therefore, the content management computing device is configured to transmit responses based on that set of contact information to these second user computing devices. Thus, the content management computing device allows the user (via a second user computing device or any other computing device) to interact with responses with a delay compared to the time of initial delivery of online content items.
[0084] As described herein, the user computing device is also configured to perform several steps to display voice interaction content. Specifically, the user computing device is configured to at least: (i) receive online content items from a content management computing device, wherein the online content items include content metadata; (ii) identify at least one voice interaction associated with the content metadata; (iii) dispatch online content via a user output interface; (iv) collect voice response data in response to at least one voice interaction from a user input interface; and (v) transmit the voice response data to the content management computing device.
[0085] In some embodiments, the user computing device is further configured to receive a second online content item, determine, based on the collected voice response data, that the second online content item should be distributed, and distribute the second online content item via a user output interface.
[0086] The methods and systems described herein can be implemented using computer programming or engineering techniques, including computer software, firmware, hardware, or any combination or subset thereof, wherein the technical effects can be achieved by performing one of the following steps: (a) retrieving an online content item containing content metadata; (b) identifying at least one voice interaction associated with the content metadata; (c) distributing the online content item to a user computing device, wherein distributing the online content item further includes instructing the user computing device to collect voice response data in response to at least one voice interaction; (d) receiving the voice response data from the user computing device; (e) identifying a user request based on the voice response data; (f) transmitting a response to the device user account based on the user request; (g) transmitting a request for additional voice response data to the user account when it is determined based on the user request that further input is required; (h) processing the voice response data into a text dataset using a voice processing algorithm; (i) identifying the user request from the text dataset by applying at least one of a regular expression algorithm and a context-independent grammar algorithm; (j) determining that the user request represents a quote. The user requests the following: (k) retrieves a user profile information set associated with the user's computing device, which includes at least a contact dataset; (i) uses the user profile information set to generate a response; (m) determines that the user request represents a request for a purchase; (n) identifies a purchase dataset from the user requests that define the request for a purchase; (o) retrieves a user payment information set associated with the user's computing device; (p) transmits the purchase dataset and the user payment information set to the online content provider; (q) transmits a security request to the user account to verify that the request for the purchase is authorized; (r) receives a security response from the user's computing device; (s) verifies that the request for the purchase is authorized; (t) determines that the user request represents a request for a scheduled event; (u) identifies a calendar options set associated with the user's computing device; (v) transmits a request for a scheduled event including the calendar options set; (w) determines that the user request represents a request for more information; (x) identifies a second online content item associated with an online content item, wherein the second online content item includes more information than the information in the online content item; and (y) dispatches the second online content item to the user account.
[0087] Figure 1 This is a diagram illustrating an exemplary online content environment 100. The online content environment 100 can be integrated into a context where online advertising is distributed to users, including those using mobile computing devices. (Reference) Figure 1 The exemplary environment 100 may include one or more online content providers 102 (e.g., advertisers), one or more publishers 104, an online content management system (OCMS) 106, and one or more user access devices 108 that can be coupled to network 110. The user access devices may be used by users 150, 152, and 154. Figure 1Each of elements 102, 104, 106, 108, and 110 is implemented or associated with a hardware component, software component, or firmware component, or any combination thereof. Elements 102, 104, 106, 108, and 110 may be implemented or associated with, for example, a general-purpose server, software processing and engine, and / or various embedded systems. Elements 102, 104, 106, and 110 may, for example, be used as an advertising distribution network. Although referenced to advertising distribution, environment 100 may be adapted to distribute other forms of content, including other forms of sponsored content. OCMS 106 may also be referred to as content management system 106.
[0088] Online content provider 102 may include any entity associated with online content such as advertisements (“ads”). An advertisement or “ad” refers to any form of communication that identifies and promotes (or otherwise conveys) one or more products, services, ideas, messages, people, organizations, or other items. Advertising is not limited to commercial promotions or other communications. An advertisement may be a public service announcement or any other type of notification, such as a notice published in print or electronic news or broadcast. An advertisement may be referred to as sponsored content.
[0089] Advertising can be delivered through a variety of media and in a variety of formats. In some examples, advertising can be delivered through interactive media such as the Internet, and advertising can include graphic ads (e.g., banner ads), text ads, image ads, audio ads, video ads, content combining one or more of these components, or content delivered electronically in any form. Advertising can include embedded information, such as embedded media, links, metadata, and / or machine-executable instructions. Advertising can also be delivered via RSS (Real Easy Aggregator) feeds, radio channels, television channels, print media, and other media.
[0090] The term "advertisement" can refer to both a single "ad creative" and an "ad group." An ad creative is any entity that represents an ad flash. An ad flash is any form of display of an ad that allows a user to see / receive the ad. In some examples, an ad flash may occur when an ad is displayed on a display device accessed by a user. An ad group, for example, refers to an entity that represents a group of ad creatives sharing common characteristics—such as having the same ad selection and recommendation criteria. Ad groups can be used to create ad campaigns.
[0091] Online content provider 102 may provide (or be associated with) products and / or services related to the advertisement. Online content provider 102 may include, for example, retailers, wholesalers, warehouses, manufacturers, distributors, healthcare providers, educational institutions, financial institutions, technology providers, electricity providers, infrastructure providers, or any other providers or distributors of products or services, or be associated with them.
[0092] Online content provider 102 may directly or indirectly generate and / or maintain advertisements that may be related to products or services offered by or otherwise associated with advertisers. Online content provider 102 may include or maintain one or more data processing systems 112 coupled to network 110, such as servers or embedded systems. Online content provider 102 may include or maintain one or more processes running on one or more data processing systems.
[0093] Publisher 104 may include any entity that generates, maintains, provides, presents, and / or otherwise processes content within environment 100. "Publisher" specifically includes the author of the content, where the author may be an individual, or, in some cases, an owner who employs an individual responsible for creating the online content, for works created for hire. The term "content" refers to various types of web-based, software application-based, and / or otherwise presented information, including articles, discussion threads, reports, analyses, financial statements, music, videos, graphics, search results, web page listings, information feeds (e.g., RSS feeds), television broadcasts, radio broadcasts, print publications, or any other form of information presented to a user using a computing device such as one of the user access devices 108.
[0094] In some implementations, publisher 104 may include content providers with Internet presentation, such as online publishing and news providers (e.g., online newspapers, online magazines, television websites, etc.), online service providers (e.g., financial service providers, health service providers, etc.), and so on. Publisher 104 may include software application providers, television broadcasters, radio broadcasters, satellite broadcasters, and other content providers. One or more of publishers 104 may represent a content network associated with CMS 106.
[0095] Publisher 104 may receive requests from user access device 108 (or other elements in environment 100) and provide or present content to the requesting device. Publisher may provide or present content via various media or in various forms, including web-based and / or non-web-based media and forms. Publisher 104 may generate and / or maintain such content and / or retrieve content from other network resources.
[0096] In addition to the content, publisher 104 can be configured to integrate or combine the retrieved content with other sets of content related to or associated with the retrieved content, such as advertisements, to display to users 150, 152, and 154. As further described below, these related advertisements can be provided from OCMS 106, and these related advertisements can be combined with the content to be displayed to users 150, 152, and 154. In some examples, publisher 104 can retrieve content to be displayed on a specific user access device 108, and then forward that content along with code that causes one or more advertisements from OCMS 106 to be displayed to users 150, 152, and 154. As used herein, user access device 108 may be referred to as client computing device 108. In other examples, publisher 104 can retrieve content (e.g., from OCMS 106 or online content provider 102), retrieve one or more related advertisements, and then integrate the advertisements and articles to form a content page to be displayed to users 150, 152, or 154.
[0097] As described above, one or more of the publishers 104 can represent a content network. In such an implementation, the online content provider 102 is able to present advertisements to users through this content network.
[0098] Publisher 104 may include or maintain one or more data processing systems 114 coupled to network 110, such as servers or embedded systems. These data processing systems 114 may include or maintain one or more processes running on the data processing system. In some examples, publisher 104 may include one or more content libraries 124 for storing content and other information.
[0099] OCMS 106 manages advertising and provides various services to online content providers 102, publishers 104, and user access devices 108. OCMS 106 can store advertisements in an advertisement library 126 and facilitates the distribution or selective delivery and recommendation of advertisements to user access devices 108 through environment 100. In some configurations, OCMS 106 may include or access functionality associated with managing online content and / or online advertising, particularly functionality associated with distributing online content and / or online advertising to mobile computing devices.
[0100] OCMS 106 may include one or more data processing systems 116, such as servers or embedded systems, coupled to network 110. OCMS 106 may also include one or more processes, such as server processes. In some examples, OCMS 106 may include an ad delivery system 120 and one or more back-end processing systems 118. The ad delivery system 120 may include one or more data processing systems 116 and may perform functions associated with delivering ads to publishers or user access devices 108. The back-end processing system 118 may include one or more data processing systems 116 and may perform functions associated with identifying relevant ads to be delivered, processing various rules, performing filtering processes, generating reports, maintaining account and usage information, and other back-end system processing. OCMS 106 can use the back-end processing system 118 and the ad delivery system 120 to selectively recommend and deliver relevant ads from online content providers 102 to publishers 104 and then to user access devices 108.
[0101] OCMS 106 may include or access one or more crawling, indexing, and searching modules (not shown). These modules can browse accessible resources (e.g., the World Wide Web, publisher content, data feeds, etc.) to identify, index, and store information. Modules can browse information and create copies of the browsed information for subsequent processing. Modules can also verify links, validate codes, produce results on information, and / or perform other maintenance or other tasks.
[0102] The search module can search for information from various resources such as the World Wide Web, publisher content, intranets, newsgroups, databases, and / or directories. The search module can employ one or more known search or other processes to search data. In some implementations, the search module can index the crawled content and / or content received from data feeds to build one or more search indexes. Search indexes can be used to facilitate the rapid retrieval of information relevant to the search query.
[0103] OCMS 106 may include one or more interfaces or front-end modules for providing various features to advertisers, publishers, and user access devices. For example, OCMS 106 may provide one or more Publisher Front-End Interfaces (PFEs) to allow publishers to interact with OCMS 106. OCMS 106 may also provide one or more Advertiser Front-End Interfaces (AFEs) to allow advertisers to interact with OCMS 106. In some examples, this front-end interface may be configured as a web application to provide users with web access to features available in OCMS 106.
[0104] OCMS 106 provides online content providers 102 with a variety of advertising management features. As described herein, OCMS 106 advertising features allow users to create user accounts, set account preferences, create ads, select keywords for ads, create or launch campaigns for multiple products or services, view reports associated with their accounts, analyze costs and ROI, selectively identify consumers in different regions, selectively recommend and deliver ads to specific publishers, analyze financial information, analyze ad performance, evaluate ad traffic, access keyword tools, and add graphics and animations to ads, among other things.
[0105] OCMS 106 allows online content provider 102 to create advertisements and input keywords or other ad place descriptors for which those advertisements will appear. In some examples, OCMS 106 can serve advertisements to users accessing devices or publishers when keywords associated with those advertisements are included in user requests or requested content. OCMS 106 also allows online content provider 102 to set bids for advertisements. A bid can represent the maximum amount an advertiser is willing to pay for each ad flash, user click, or other interaction with the ad. A click can include an action taken by the user who selected the ad. Other actions include generating haptic or gyroscope feedback for the click. Online content provider 102 can also select a currency and a monthly budget.
[0106] OCMS 106 also allows online content providers 102 to view information about ad flashes that can be maintained by OCMS 106. OCMS 106 can be configured to determine and maintain the number of ad flashes relative to a specific website or keyword. OCMS 106 can also determine and maintain the number of ad clicks and the click-to-flash ratio.
[0107] OCMS 106 can also allow online content providers 102 to select and / or create conversion types for advertisements. A "conversion" can occur when a user completes a transaction related to a given advertisement. A "conversion" can be defined as occurring when a user (e.g., through haptic or gyroscope feedback) directly or implicitly clicks on an advertisement on a webpage called an advertiser and completes a purchase before leaving the webpage. In another example, a conversion can be defined as showing an advertisement to a user within a predetermined time (e.g., 7 days) and making a corresponding purchase on the advertiser's webpage. OCMS 106 can store conversion data and other information in a conversion database 136.
[0108] OCMS 106 allows online content providers 102 to input descriptive information associated with advertisements. This information can be used to assist publishers 104 in determining which advertisements to publish. Online content providers 102 can also input costs / values associated with the selected conversion type, such as a five-dollar payment to the publisher for each product or service purchased.
[0109] OCMS 106 can provide various features to publisher 104. When a user accesses content from publisher 104, OCMS 106 can deliver advertisements (associated with online content provider 102) to the user's access device 108. OCMS 106 can be configured to deliver content relevant to the publisher's site, site content, and publisher's audience.
[0110] In some examples, OCMS 106 can crawl content provided by publisher 104 and deliver advertisements relevant to the publisher's site, site content, and publisher's audience based on the crawled content. OCMS 106 can also selectively recommend and / or serve advertisements based on user information and behavior, such as specific search queries performed on a search engine website as described herein, or advertisements specified for subsequent comments. OCMS 106 can store user-related information in a general database 146. In some examples, OCMS 106 can add a search service to the publisher's site and deliver advertisements configured to provide appropriate and relevant content relative to search results generated by requests from visitors to the publisher's site. Combinations of these and other methods can be used to deliver relevant advertisements.
[0111] OCMS 106 allows publisher 104 to search for and select specific products and services, as well as associated advertisements to be displayed along with the content provided by publisher 104. For example, publisher 104 can search for advertisements in the advertisement library 126 and select certain advertisements to be displayed along with its content.
[0112] OCMS 106 can be configured to selectively recommend and serve advertisements created by online content provider 102 to user access device 108, either directly or through publisher 104. When a user requests search results or loads content from publisher 104, CMS 106 can selectively recommend and serve advertisements to a specific publisher 104 (as described in further detail herein) or to the user access device 108 that is making the request.
[0113] In some implementations, OCMS 106 can manage and process financial transactions between elements within environment 100. For example, OCMS 106 can pay accounts associated with publisher 104 and deduct from accounts of online content provider 102. These and other transactions can be based on conversion data, flash information, and / or click-through rates received and maintained by OCMS 106.
[0114] For example, a "computing device" for user access device 108 can include any device capable of receiving information from network 110. User access device 108 can include general-purpose computing components and / or embedded systems optimized using specific components for performing specific tasks. Examples of user access devices include personal computers (e.g., desktop computers), mobile computing devices, cellular phones, smartphones, head-mounted computing devices, media players / recorders, music players, game consoles, media centers, media players, tablets, personal digital assistants (PDAs), television systems, audio systems, radio systems, removable storage devices, navigation systems, set-top boxes, and other electronic devices. User access device 108 can also include various other elements, such as processes running on various machines.
[0115] Network 110 may include any elements or systems that facilitate communication among and between network nodes, such as elements 108, 112, 114, and 116. Network 110 may include one or more telecommunications networks, such as computer networks, telephone or other communication networks, the Internet, etc. Network 110 may include shared, public, or private data networks covering a wide area (e.g., WAN) or a local area (e.g., LAN). In some embodiments, network 110 may facilitate data exchange using the Internet Protocol (IP) in a packet-switched manner. Network 110 may facilitate wired and / or wireless connectivity and communication.
[0116] For illustrative purposes only, see reference. Figure 1 The discrete elements illustrated herein depict certain aspects of this disclosure. The number, identification, and arrangement of elements in environment 100 are not limited to those shown. For example, environment 100 may include any number of geographically dispersed online content providers 102, publishers 104, and / or user access devices 108, which may be discrete, integrated modules, or distributed systems. Similarly, environment 100 is not limited to a single OCMS 106, but may include any number of integrated or distributed AMS systems or elements.
[0117] Furthermore, additional and / or different elements, not shown, may be included in or coupled to. Figure 1Some of the illustrated components may be missing. In some examples, the functionality provided by the illustrated components may be performed by fewer components than those illustrated, or even by a single component. The illustrated components may be implemented as individual processes running on discrete machines, or as a single process running on a single machine.
[0118] Figure 2 As shown in online content environment 100 ( Figure 1 The diagram shows a block diagram of a computing device 200 for managing, providing, displaying, and analyzing voice-interactive online content. The computing device 200 is intended to represent various forms of digital computers, such as laptops, desktops, workstations, personal digital assistants, servers, blade servers, mainframes, and other suitable computers. The computing device 200 is also intended to represent various forms of mobile devices, such as personal digital assistants, cellular phones, smartphones, and other similar computing devices. The components shown herein, their connections and relationships, and their functions are merely illustrative and are not intended to limit the implementation of the topics described and / or claimed in this document. Therefore, the computing device 200 can represent a user computing device, a content management computing device, an online content provider computing device, and an online content publishing computing device (in...). Figure 2 (Not shown in the image). As described, all user computing devices, content management computing devices, online content provider computing devices, and online content publishing computing devices can be in use. Figure 2 In the network communication capabilities of the aforementioned system.
[0119] In an example embodiment, computing device 200 may be user access device 108 or data processing device 112, 114 or 116. Figure 1 Any of the components shown in the diagram. The computing device 200 may include a bus 202, a processor 204, a main memory 206, a read-only memory (ROM) 208, a storage device 210, an input device 212, an output device 214, and a communication interface 216. The bus 202 may include a pathway that allows communication between components of the computing device 200.
[0120] Processor 204 may include any type of conventional processor, microprocessor, or processing logic that interprets and executes instructions. Processor 204 is capable of processing instructions for execution within computing device 200, including instructions stored in memory 206 or on storage device 210 for displaying graphical information for a GUI on external input / output devices, such as output device 214 coupled to a high-speed interface. In other implementations, multiple memories and various types of memory may be combined, multiple processors and / or multiple buses may be used, where appropriate. Furthermore, multiple computing devices 200 may be connected, each providing a portion of the necessary operation (e.g., server groups, blade server groups, or multiprocessor systems).
[0121] Main memory 206 may include random access memory (RAM) or another dynamic storage device that stores information and instructions executed by processor 204. ROM 208 may include a common ROM device or another static storage device that stores static information and instructions used by processor 204. Main memory 206 stores information within computing device 200. In one implementation, main memory 206 is a volatile storage unit. In another implementation, main memory 206 is a non-volatile storage unit. Main memory 206 may also be another form of computer-readable medium, such as magnetic or optical disk.
[0122] Storage device 210 may include magnetic and / or optical recording media and their corresponding drives. Storage device 210 is capable of providing mass storage for computing device 200. In one implementation, storage device 210 may be or contain computer-readable media such as floppy disk devices, hard disk devices, optical disk devices or tape devices, flash memory or other similar solid-state storage devices, or device arrays, including devices in storage area networks or other configurations. A computer program product can be tangibly embodied in an information carrier. The computer program product may also contain instructions that, when executed, perform one or more methods, such as those described above. The information carrier is a computer or machine-readable medium, such as main memory 206, ROM 208, storage device 210, or memory on processor 204.
[0123] The high-speed controller manages bandwidth-intensive operations for computing device 200, while the low-speed controller manages less bandwidth-intensive operations. This functional allocation is for illustrative purposes only. In one implementation, the high-speed controller is coupled to main memory 206, display 214 (e.g., via a graphics processor or accelerator), and a high-speed expansion port that can accommodate various expansion cards (not shown). In another implementation, the low-speed controller is coupled to storage device 210 and a low-speed expansion port. The low-speed expansion port may include various communication ports (e.g., USB, Bluetooth, Ethernet, Wireless Ethernet) and may be coupled to one or more input / output devices, such as keyboards, pointing devices, scanners, or networking devices such as switches or routers, for example, via network adapters.
[0124] Input device 212 may include common mechanisms that allow computing device 200 to receive commands, instructions, or other inputs, including visual, audio, touch, button presses, stylus clicks, etc., from users 150, 152, or 154. Additionally, the input device may receive location information. Therefore, input device 212 may include, for example, a camera, microphone, one or more buttons, a touchscreen, and / or a GPS receiver. Output device 214 may include common mechanisms for outputting information to the user, including a display (including a touchscreen) and / or speakers. Communication interface 216 may include any transceiver-like mechanism that enables computing device 200 to communicate with other devices and / or systems. For example, communication interface 216 may include a mechanism for communicating via, for example, network 110 (… Figure 1 (As shown in the image) is an organization that communicates with another device or system via a network.
[0125] As described herein, computing device 200 facilitates the presentation to a user of content from one or more publishers, as well as one or more sets of sponsored content, such as advertisements. Computing device 200 can perform these and other operations in response to processor 204 executing software instructions contained in a computer-readable medium, such as memory 206. A computer-readable medium can be defined as a physical or logical storage device and / or a carrier wave. Software instructions can be read into memory 206 from another computer-readable medium, such as data storage device 210, or from another device via communication interface 216. The software instructions contained in memory 206 can cause processor 204 to perform the processes described herein. Alternatively, hardwired circuitry can be used in place of or in combination with the software instructions to implement processes consistent with the subject matter herein. Thus, implementations consistent with the principles of the subject matter disclosed herein are not limited to any particular combination of hardware circuitry and software.
[0126] The computing device 200 can be implemented in many different forms, as shown in the figure. For example, it can be implemented as a standard server, or more often as a group of such servers. It can also be implemented as part of a rack server system. Furthermore, it can be implemented as a personal computer, such as a laptop computer. Each of such devices can contain one or more computing devices 200, and the entire system can consist of multiple computing devices 200 communicating with each other.
[0127] Processor 204 can execute instructions within computing device 200, including instructions stored in main memory 206. The processor can be implemented as a chip comprising individual and multiple analog and digital processors. The processor can provide, for example, coordination of other components of device 200, such as control of the user interface, applications running by device 200, and wireless communication of device 200.
[0128] In addition to components such as receivers and transceivers, computing device 200 includes a processor 204, main memory 206, ROM 208, input device 212, output device such as display 214, and communication interface 216. Device 200 may also provide storage device 210, such as microdrives or other devices, to provide additional storage. Each of the components is interconnected using various buses, and several components may be mounted on a common motherboard or otherwise suitably mounted.
[0129] The computing device 200 can communicate wirelessly via communication interface 216, which may include digital signal processing circuitry if necessary. Communication interface 216 can provide communication under various modes or protocols, such as GSM voice calls, SMS, EMS, MMS messages, CDMA, TDMA, PDC, WCDMA, CDMA 2000, or GPRS. Such communication can occur, for example, via a radio frequency transceiver. Furthermore, short-range communication can occur using transceivers such as Bluetooth, WiFi, or others (not shown). Additionally, a GPS (Global Positioning System) receiver module can provide the device 200 with additional navigation and location-related wireless data, which can be used by applications running on the device 200, where appropriate.
[0130] Figure 3 It uses an online content environment 100 ( Figure 1 Example data flow diagram 300 shows computing devices 112, 116, 303, and 114 (as shown) used to manage and provide voice-interactive online content. Figure 2 As shown, the structures of computing devices 112, 116, 303 and 104 are similar to the structure of computing device 200.
[0131] As described above, the content management computing device 116 defines structure metadata 310 that can be used to create content metadata 325. The content management computing device 116 provides the structure metadata 310 to multiple systems, including the online content provider computing device 112. The online content provider computing device 112 uses the structure metadata 310 to create content metadata 325 and distributes online content items 320 that include the content metadata 325. More specifically, as described above, the content metadata 325 associates the online content items 320 with at least a voice interaction type.
[0132] Content management computing device 116 identifies at least one voice interaction associated with content metadata 325 and transmits online content item 32, including content metadata 325, to user computing device 303. Content management computing device 116 also dispatches online content item 320 by instructing user computing device 303 to collect voice response data 350 in response to the identified at least one voice interaction.
[0133] In some examples, in conjunction with online content item 320, online publisher computing device 114 also distributes publication 330 to user computing device 303.
[0134] User computing device 303 displays and / or provides online content items 320 to user 301, and also dispatches at least one voice interaction to user 301. User 301 provides user input 340, which is processed into voice response data 350.
[0135] Content management computing device 116 receives voice response data 350 and identifies user requests based on the voice response data 350. Content management computing device 116 also generates content response 360 based on the voice response data 350 and transmits it to an appropriate party including at least one of the following: user computing device 303, online content provider computing device 112, online publisher computing device 114, and other systems (not shown).
[0136] Figure 4 It uses an online content environment 100 ( Figure 1 As shown), an example method 400 for managing and providing online content with voice interaction. In an example embodiment, method 400 is performed by a content management computing device 116 (as shown). Figure 3 (as shown) is executed. In an alternative embodiment, as described above, some steps of method 400 may also be performed using a user computing device 303 (in... Figure 3 Other systems (shown in the diagram).
[0137] Content management computing device 116 retrieves online content items including content metadata, such as online content item 320 including online content metadata 325.
[0138] Content management computing device 116 also identifies 420 at least one voice interaction associated with content metadata 325 and communicates it to the user computing device (e.g., Figure 3 The user computing device 303 shown provides 430 online content items 320. Distributing online content items 320 further includes instructing the user computing device 303 to collect voice response data 350 in response to at least one voice interaction. Figure 3 (as shown in the image).
[0139] The content management computing device 116 also receives 440 voice response data 350 from the user computing device 303, and identifies 450 user requests based on the voice response data 350. The content management computing device 116 also transmits 460 responses to the user computing device 303 based on the user requests.
[0140] Figure 5 It uses an online content environment 100 ( Figure 1 As shown), to user computing device 303 ( Figure 3 An exemplary method for displaying and providing voice-interactive online content (shown). User computing device 303 is configured to receive content from content management computing device 116 (shown). Figure 3 (As shown) Received 510 online content items 320 (in Figure 3 As shown in the figure), online content item 320 includes content metadata 325. Figure 3 (As shown). User computing device 303 is also configured to recognize at least one voice interaction 520 associated with content metadata 325. User computing device 303 is also configured to dispatch 530 online content items 320 via a user output interface. User computing device 303 is further configured to collect 540 voice response data 350 in response to at least one voice interaction from a user input interface. Figure 3 (As shown in the diagram). User computing device 303 is also configured to transmit voice response data 350 to content management computing device 116.
[0141] Figure 6 Figure 600 shows a component of one or more exemplary computing devices used for managing and providing voice-interactive online content.
[0142] For example, one or more of the computing devices 200 can form an advertising management system (AMS) 106 and a client computing device 108 (both in Figure 1 (shown in the image), content management computing device 116 and user computing device 303 (both in...) Figure 3 (as shown in the image). Figure 6 Further examples are shown in databases 126 and 146 (in...). Figure 1The configuration is shown in the figure. Databases 126 and 146 are coupled to several separate components within the content management computing device 120, the content provider data processing system 112, and the client computing device 108, which perform specific tasks.
[0143] Content management computing device 120 includes a retrieval component 602 for retrieving online content items including content metadata. Content management computing device 120 includes a first recognition component 604 for identifying at least one voice interaction associated with the content metadata. Content management computing device 120 includes a distribution component 605 for distributing online content items to a user computing device, wherein distributing online content items further includes instructing the user computing device to collect voice response data in response to at least one voice interaction. Content management computing device 120 includes a receiving component 606 for receiving voice response data from the user computing device. Content management computing device 120 includes a second recognition component 607 for identifying a user request based on the voice response data. Content management computing device 120 includes a transmission component 608 for transmitting a response to a user account based on a user request.
[0144] In an exemplary embodiment, databases 126 and 146 are divided into multiple parts, including but not limited to a content metadata description unit 610, a metadata structure unit 612, and a voice interaction processing unit 614. These parts within databases 126 and 146 are interconnected to update and retrieve information as needed.
[0145] These computer programs (also referred to as programs, software, software applications, or code) include machine instructions for a programmable processor and may be implemented in high-level procedural and / or object-oriented programming languages, and / or in assembly / machine language. As used herein, the terms “machine-readable medium” and “computer-readable medium” mean any computer program product, apparatus, and / or device (e.g., disk, optical disk, memory, programmable logic device (PLD)) used to provide machine instructions and / or data to a programmable processor, including machine-readable media that receive machine instructions as machine-readable signals. However, the terms “machine-readable medium” and “computer-readable medium” do not include transient signals. The term “machine-readable signal” refers to any signal used to provide machine instructions and / or data to a programmable processor.
[0146] Furthermore, the logical flow illustrated in the figures does not require a specific order or sequence to achieve the desired result. Additionally, other steps may be provided, or steps may be removed from the described flow, and other components may be added to or removed from the system. Therefore, other embodiments are within the scope of the following claims.
[0147] It should be recognized that the embodiments described in the detailed description above are merely examples or possible embodiments, and many other combinations, additions or alternatives may be included.
[0148] Furthermore, specific naming of components, capitalization of words, attributes, data structures, and any other programming or structural aspects are not mandatory or important, and mechanisms for implementing the topics or features described herein may have different names, formats, or protocols. Moreover, the system may be implemented via a combination of hardware and software (as described) or entirely within hardware components. Furthermore, the specific division of functionality among the various system components described herein is for illustrative purposes only and is not mandatory; functionality performed by a single system component may alternatively be performed by multiple components, and functionality performed by multiple components may alternatively be performed by a single component.
[0149] Some of the elements described above are characterized in terms of the algorithms and symbolic representations of information operations. Those skilled in the art of data processing can use these algorithmic descriptions and representations to most effectively communicate the substance of their work to others skilled in the art. Although these operations are described functionally or logically, they are understood to be implemented by computer programs. Furthermore, it has been shown that it is sometimes convenient, without loss of generality, to refer to the arrangement of these operations as modules or their functional names.
[0150] Unless otherwise stated, as is apparent from the discussion above, it can be understood that throughout this specification, the use of terms such as “processing” or “computing” or “operating” or “determining” or “displaying” or “providing” refers to the actions and processes of a computer system or similar electronic computing device that manipulate and transform data represented as physical (electronic) quantities in computer system memory or registers or other such information storage, transmission or display devices.
[0151] Based on the foregoing description, the above embodiments can be implemented using computer programming or engineering techniques including computer software, firmware, hardware, or any combination or subset thereof. Any resulting program having computer-readable and / or computer-executable instructions can be embodied or provided in one or more computer-readable media, thereby creating a computer program product, i.e., an article of manufacture. The computer-readable medium can be, for example, a fixed (hard) drive, a card cartridge, an optical disc, a magnetic tape, a semiconductor memory such as read-only memory (ROM) or flash memory, or any transmission / reception medium such as the Internet or other communication networks or links. An article of manufacture containing computer code can be made and / or used by executing instructions directly from a medium, by copying the code from one medium to another, or by transmitting the code over a network.
[0152] Although this disclosure has been described with reference to various specific embodiments, it should be recognized that modifications can be made to this disclosure within the spirit and scope of the claims.
Claims
1. A system for managing online content, comprising: a memory device storing data, including machine executable instructions; and one or more processors in communication with the memory device, wherein the processors are configured to execute the executable instructions, the executable instructions causing the one or more processors to perform operations comprising: retrieving an online content item including content metadata specifying one or more interaction tags enabled for the online content item; causing a client device to present the online content item at a client device and to wait for an audio response from a user after presenting the online content item; detecting a first audio response submitted through the client device; determining that the audio response matches a first interaction tag of the one or more interaction tags; and in response to determining that the audio response matches the first interaction tag, deferring an interaction with the user to a later time, the interaction including communicating additional information specified for the matched first interaction tag to the user based on account information of the user, wherein the additional information is different from the online content item presented at the client device.
2. The system of claim 1, wherein, the instructions causing the one or more processors to perform operations comprising requesting an additional audio response from the user prior to deferring the interaction, wherein deferring the interaction is performed in response to determining that the audio response matches the first interaction tag and the additional audio response received from the user provides information needed to communicate the additional information to the user.
3. The system of claim 1, wherein: the instructions cause the one or more processors to perform operations comprising: processing the first audio response into a text data set using a speech processing algorithm; and identifying a voice interaction tag from the text data set by applying at least one of a regular expression algorithm or a context free grammar algorithm.
4. The system of claim 1, wherein, the instructions cause the one or more processors to perform operations comprising: determining that the audio response represents a request for an offer; and retrieving contact data for the user from a profile of the user, wherein communicating the audio response includes communicating the audio response using the contact information of the user.
5. The system of claim 1, wherein, the instructions cause the one or more processors to perform operations comprising: determining that the audio response represents a request for a purchase; retrieving a user payment information set for the user; and initiating an order based on the payment information and the request for a purchase.
6. The system of claim 5, wherein, communicating the additional information to the user includes communicating order details to the user, wherein the order details enable the user to review, approve, cancel, or modify the order.
7. The system of claim 1, wherein, the instructions cause the one or more processors to perform operations comprising: determining that the audio response represents a request for scheduling an event; and identifying a calendar option set for the user, wherein: communicating the additional information includes communicating information about the scheduled event based on the calendar option set.
8. A method for managing online content, comprising: retrieving, by one or more processors, an online content item including content metadata specifying one or more interaction tags enabled for the online content item; causing, by the one or more processors, a client device to present the online content item at the client device and wait for an audio response from a user after presenting the online content item; detecting, by the one or more processors, a first audio response submitted through the client device; determining, by the one or more processors, that the audio response matches a first interaction tag of the one or more interaction tags; and responsive to determining that the audio response matches the first interaction tag, deferring, by the one or more processors, an interaction with the user to a later time, the interaction including communicating additional information specified for the matched first interaction tag to the user based on account information of the user, wherein the additional information is different from the online content item presented at the client device.
9. The method of claim 8, further comprising: requesting an additional audio response from the user prior to deferring the interaction, wherein deferring the interaction is performed responsive to determining that the audio response matches the first interaction tag and the additional audio response received from the user provides information needed to communicate the additional information to the user.
10. The method of claim 8, wherein, The method further comprising: processing, using a speech processing algorithm, the first audio response into a text data set; and identifying, from the text data set, a voice interaction tag by applying at least one of a regular expression algorithm or a context free grammar algorithm.
11. The method of claim 8, further comprising: determining that the audio response represents a request for an offer; and retrieving contact data for the user from a profile of the user, wherein communicating the audio response includes using the contact information of the user to communicate the audio response.
12. The method of claim 8, further comprising: determining that the audio response represents a request for a purchase; retrieving a user payment information set for the user; and initiating an order based on the payment information and the request for a purchase.
13. The method of claim 12, wherein, communicating the additional information to the user includes communicating order details to the user, wherein the order details enable the user to review, approve, cancel, or modify the order.
14. The method of claim 8, further comprising: determining that the audio response represents a request for scheduling an event; and identifying a calendar option set for the user, wherein: communicating the additional information includes communicating information about the scheduled event based on the calendar option set.
15. A non-transitory computer-readable storage device storing processor-executable instructions that, when executed by one or more processors, cause the one or more processors to perform operations comprising: retrieving an online content item including content metadata specifying one or more interaction tags enabled for the online content item; causing a client device to present the online content item at the client device and wait for an audio response from a user after presenting the online content item; detecting a first audio response submitted through the client device; determining that the audio response matches a first interaction tag of the one or more interaction tags; and in response to determining that the audio response matches the first interaction tag, deferring an interaction with the user to a later time, the interaction including communicating additional information designated for the matched first interaction tag to the user based on account information of the user, wherein the additional information is different from the online content item presented at the client device.
16. The non-transitory computer-readable storage device of claim 15, wherein, the instructions cause the one or more processors to perform operations including requesting an additional audio response from the user prior to deferring the interaction, wherein deferring the interaction is performed in response to determining that the audio response matches the first interaction tag and the additional audio response received from the user provides information needed to communicate the additional information to the user.
17. The non-transitory computer-readable storage device of claim 15, wherein: the instructions cause the one or more processors to perform operations including: processing, using a speech processing algorithm, the first audio response into a text data set; and identifying, from the text data set, a voice interaction tag by applying at least one of a regular expression algorithm or a context free grammar algorithm.
18. The non-transitory computer-readable storage device of claim 15, wherein, the instructions cause the one or more processors to perform operations including: determining that the audio response represents a request for a quote; and retrieving contact data for the user from a profile of the user, wherein communicating the audio response includes communicating the audio response using the contact information of the user.
19. The non-transitory computer-readable storage device of claim 15, wherein, the instructions cause the one or more processors to perform operations including: determining that the audio response represents a request for a purchase; retrieving a user payment information set for the user; and initiating an order based on the payment information and the request for a purchase.
20. The non-transitory computer-readable storage device of claim 19, wherein, communicating the additional information to the account of the user includes communicating order details to the user, wherein the order details enable the user to review, approve, cancel, or modify the order. the instructions cause the one or more processors to perform operations including: determining that the audio response represents a request for a purchase; retrieving a user payment information set for the user; and initiating an order based on the payment information and the request for a purchase. communicating the additional information to the account of the user includes communicating order details to the user, wherein the order details enable the user to review, approve, cancel, or modify the order.
Citation Information
Patent Citations
Managing internet advertising and promotional content
CN102216947A
Method for enabling the voice interaction with a web page
US20040141597A1