Content output device, content output method, program, and storage medium
The content output device addresses the lack of correlation between audio and display content by dynamically generating and adjusting content detail and frequency, improving user retention and memory of facility information.
Patent Information
- Application Number
- JP2024502714
- Authority / Receiving Office
- JP · JP
- Patent Type
- Patents
- Current Assignee / Owner
- Filing Date
- 2022-02-28
- Publication Date
- 2025-11-21
- Estimated Expiration
- 2042-02-28
AI Technical Summary
Existing systems fail to effectively correlate audio and display content for facility information, leading to poor user retention and memory of provided information.
A content output device that generates audio and display content based on feature information, adjusting content detail and frequency to enhance user memory, including a mechanism to change audio and display content based on output frequency and content changes.
Enhances user memory and retention of facility information by providing correlated and progressively detailed audio and display content.
Smart Images

Figure 0007774706000001 
Figure 0007774706000002 
Figure 0007774706000003
Abstract
Description
[Technical Field]
[0001] The present invention relates to a technology that can be used in outputting content. [Background technology]
[0002] BACKGROUND ART There is known a technique for acquiring the current location of an information terminal carried by a user and providing the user with information about facilities located in the vicinity of the current location (see, for example, Patent Document 1). [Prior art documents] [Patent documents]
[0003] [Patent Document 1] Japanese Patent Application Laid-Open No. 2018-18299 Summary of the Invention [Problem to be solved by the invention]
[0004] In some cases, facility information is provided to users using information terminals through both audio content and display content. In this case, if there is no correlation between the audio content and display content for the same facility, the provided facility information will not leave an impression on the user. Furthermore, even if content for the same facility is provided multiple times, users often do not remember the content that was provided in the past.
[0005] The present invention has been made to solve the above-mentioned problems, and its main object is to provide information about facilities and other features in a way that is likely to leave an impression or remain in the user's memory. [Means for solving the problem]
[0006] The claimed invention is a content output device including: an acquisition unit that acquires feature information related to a feature based on a vehicle position; an audio content generation unit that generates audio content related to the feature based on the feature information; and a display content generation unit that generates display content indicating details of the audio content based on the feature information and predetermined words included in the audio content. The audio content generation unit changes the audio content to be generated in accordance with the number of times the audio content is output, and the display content generation unit changes the display content to be generated in accordance with changes in the content of the audio content. .
[0007] The claimed invention also includes: 1. A computer-implemented content output method, comprising: Acquire feature information related to a feature based on the vehicle position, generate audio content related to the feature based on the feature information, and generate display content indicating the details of the audio content based on the feature information and predetermined words included in the audio content. The generated audio content is changed according to the number of times the audio content is output, and the generated display content is changed according to changes in the content of the audio content. .
[0008] The claimed invention is also a program that causes a computer to execute a process of acquiring feature information related to a feature based on a vehicle position, generating audio content related to the feature based on the feature information, and generating display content that shows the details of the audio content based on the feature information and predetermined words included in the audio content. The generated audio content is changed according to the number of times the audio content is output, and the generated display content is changed according to changes in the content of the audio content. . [Brief explanation of the drawings]
[0009] [Figure 1] 1 is a diagram illustrating an example of the configuration of an audio output system according to an embodiment. [Figure 2] 1 is a block diagram showing a schematic configuration of an audio output device. [Figure 3] FIG. 2 is a block diagram showing a schematic configuration of a server device. [Figure 4] FIG. 2 is a block diagram showing the functional configuration of a server device for providing facility information. [Figure 5] An example of facility information is shown below. [Figure 6]10 shows an example of display content. [Figure 7] 10 is an example of facility information classified by level of detail. [Figure 8] 10 is a flowchart of an information providing process. DETAILED DESCRIPTION OF THE INVENTION
[0010] In one preferred embodiment of the present invention, the content output device includes an acquisition unit that acquires feature information related to a feature based on the vehicle position, an audio content generation unit that generates audio content related to the feature based on the feature information, and a display content generation unit that generates display content indicating the content of the audio content based on the feature information and predetermined words included in the audio content.
[0011] The content output device acquires feature information related to features based on the vehicle position, and generates audio content related to the features based on the feature information. The content output device also generates display content that indicates the details of the audio content based on the feature information and predetermined words included in the audio content. This makes it possible to generate and output display content related to the audio content.
[0012] In one aspect of the content output device, the audio content generation unit changes the audio content to be generated in accordance with the number of times the audio content is output, and the display content generation unit changes the display content to be generated in accordance with changes in the content of the audio content. In this aspect, different audio content and display content can be output depending on the number of times the audio content is output.
[0013] In another aspect of the content output device, the audio content generation unit increases the level of detail of the audio content as the number of times the audio content is output increases. In this aspect, as the number of times the audio content is output increases, more detailed information about the feature is provided as the audio content and the display content.
[0014] In another aspect of the content output device, the display content includes the first and most recent output dates of the audio content and the display content, and by looking at the output dates of the display content, the user can recall past output of audio content.
[0015] In another aspect of the content output device, the display content generation unit generates the display content using the same words as the audio content. In a preferred example, the display content generation unit generates the display content by arranging the same words as the audio content. This makes it possible to generate display content related to the audio content.
[0016] In another preferred embodiment of the present invention, a content output method includes acquiring feature information related to a feature based on a vehicle position, generating audio content related to the feature based on the feature information, and generating display content indicating details of the audio content based on the feature information and predetermined words included in the audio content, thereby generating and outputting display content related to the audio content.
[0017] In another preferred embodiment of the present invention, a program causes a computer to execute processes of acquiring feature information related to a feature based on the vehicle's position, generating audio content related to the feature based on the feature information, and generating display content indicating the details of the audio content based on the feature information and predetermined words included in the audio content. By executing this program on a computer, the above-mentioned content output device can be realized. This program can be stored in a storage medium and used. [Example]
[0018] Preferred embodiments of the present invention will now be described with reference to the drawings. <System configuration> [Overall configuration] 1 is a diagram illustrating an example of the configuration of an audio output system according to an embodiment. The audio output system 1 according to the embodiment includes an audio output device 100 and a server device 200. The audio output device 100 is mounted on a vehicle Ve. The server device 200 communicates with a plurality of audio output devices 100 mounted on a plurality of vehicles Ve.
[0019] The audio output device 100 basically performs route guidance processing, information provision processing, and the like for a user who is a passenger in a vehicle Ve. For example, when a destination or the like is input by the user, the audio output device 100 transmits an upload signal S1 including location information of the vehicle Ve and information related to the specified destination to the server device 200. The server device 200 calculates a route to the destination by referring to map data, and transmits a control signal S2 indicating the route to the destination to the audio output device 100. The audio output device 100 provides route guidance to the user by audio output based on the received control signal S2.
[0020] Furthermore, the audio output device 100 provides various types of information to the user through dialogue with the user. For example, when the user makes an information request, the audio output device 100 supplies the server device 200 with an upload signal S1 including information indicating the content or type of the information request and information regarding the running state of the vehicle Ve. The server device 200 acquires and generates the information requested by the user and transmits it to the audio output device 100 as a control signal S2. The audio output device 100 provides the received information to the user by audio output.
[0021] [Audio output device] The audio output device 100 travels with the vehicle Ve and provides route guidance primarily through audio so that the vehicle Ve travels along the guidance route. Note that "route guidance primarily through audio" refers to route guidance that allows the user to understand information necessary for driving the vehicle Ve along the guidance route at least through audio alone, and does not exclude the audio output device 100 supplementarily displaying a map or the like around the current location. In this embodiment, the audio output device 100 outputs various driving-related information by audio, such as at least points on the route where guidance is required (also referred to as "guidance points"). Herein, guidance points include, for example, intersections where the vehicle Ve must turn right or left, and other important passing points for the vehicle Ve to travel along the guidance route. The audio output device 100 provides audio guidance regarding guidance points, such as the distance from the vehicle Ve to the next guidance point and the direction of travel at that guidance point. Hereinafter, audio guidance regarding the guidance route will also be referred to as "route audio guidance."
[0022] The audio output device 100 is attached, for example, to the top of the windshield or on the dashboard of the vehicle Ve. The audio output device 100 may also be incorporated into the vehicle Ve.
[0023] 2 is a block diagram showing a schematic configuration of the audio output device 100. The audio output device 100 mainly includes a communication unit 111, a storage unit 112, an input unit 113, a control unit 114, a sensor group 115, a display unit 116, a microphone 117, a speaker 118, an exterior camera 119, and an interior camera 120. The elements within the audio output device 100 are connected to each other via a bus line 110.
[0024] The communication unit 111 performs data communication with the server device 200 under the control of the control unit 114. The communication unit 111 may receive, for example, map data for updating a map database (hereinafter, the database will be referred to as "DB") 4 described later from the server device 200.
[0025] The storage unit 112 is configured with various types of memory such as a RAM (Random Access Memory), a ROM (Read Only Memory), and a non-volatile memory (including a hard disk drive, a flash memory, etc.). The storage unit 112 stores programs for the audio output device 100 to execute predetermined processes. The above-mentioned programs may include an application program for providing route voice guidance, an application program for playing music, an application program for outputting content other than music (such as television), etc. The storage unit 112 is also used as a working memory for the control unit 114. The programs executed by the audio output device 100 may be stored in a storage medium other than the storage unit 112.
[0026] The storage unit 112 also stores a map DB4 and a facility information DB5. The map DB4 stores various data necessary for route guidance. The map DB4 stores, for example, road data that represents a road network using a combination of nodes and links. The facility information DB5 stores information about shops and the like that is provided to the user. The map DB4 and the facility information DB5 may be updated based on map information that the communication unit 111 receives from a map management server under the control of the control unit 114.
[0027] The input unit 113 is a button, a touch panel, a remote controller, or the like that is operated by the user. The display unit 116 is a display or the like that displays information under the control of the control unit 114. The microphone 117 collects sounds inside the vehicle Ve, particularly sounds uttered by the driver. The speaker 118 outputs audio for route guidance to the driver or the like.
[0028] The sensor group 115 includes an external sensor 121 and an internal sensor 122. The external sensor 121 is one or more sensors for recognizing the surrounding environment of the vehicle Ve, such as a lidar, radar, ultrasonic sensor, infrared sensor, or sonar. The internal sensor 122 is a sensor for measuring the position of the vehicle Ve, such as a Global Navigation Satellite System (GNSS) receiver, a gyro sensor, an Inertial Measurement Unit (IMU), a vehicle speed sensor, or a combination thereof. Note that the sensor group 115 may include any sensor that allows the control unit 114 to directly or indirectly (i.e., by performing estimation processing) derive the position of the vehicle Ve from the output of the sensor group 115.
[0029] Exterior camera 119 is a camera that captures the outside of vehicle Ve. Exterior camera 119 may be only a front camera that captures the view in front of the vehicle, or may include a rear camera that captures the view behind the vehicle in addition to the front camera, or may be an omnidirectional camera that can capture the entire periphery of vehicle Ve. On the other hand, interior camera 120 is a camera that captures the interior of vehicle Ve, and is installed in a position that can capture at least the area around the driver's seat.
[0030] The control unit 114 includes a CPU (Central Processing Unit), a GPU (Graphics Processing Unit), etc., and controls the entire audio output device 100. For example, the control unit 114 estimates the position (including the direction of travel) of the vehicle Ve based on the output of one or more sensors in the sensor group 115. When a destination is specified by the input unit 113 or the microphone 117, the control unit 114 generates route information indicating a guidance route to the destination, and provides route voice guidance based on the route information, the estimated position information of the vehicle Ve, and the map DB 4. In this case, the control unit 114 outputs the guidance voice from the speaker 118. The control unit 114 also provides the user with facility information regarding facilities located near the current location of the vehicle Ve. The control unit 114 also controls the display unit 116 to display information about the music being played, video content, a map of the area around the current location, etc.
[0031] The processing performed by the control unit 114 is not limited to being realized by software programs, but may be realized by any combination of hardware, firmware, and software. The processing performed by the control unit 114 may also be realized by a user-programmable integrated circuit, such as an FPGA (field-programmable gate array) or a microcomputer. In this case, the program executed by the control unit 114 in this embodiment may be realized by using this integrated circuit. In this way, the control unit 114 may be realized by hardware other than a processor.
[0032] The configuration of the audio output device 100 shown in FIG. 2 is an example, and various modifications may be made to the configuration shown in FIG. 2. For example, instead of storing the map DB4 and the facility information DB5 in the memory unit 112, the control unit 114 may receive information necessary for route guidance from the server device 200 via the communication unit 111. In another example, instead of including the speaker 118, the audio output device 100 may be connected electrically or by a known communication means to an audio output unit configured separately from the audio output device 100, and audio may be output from the audio output unit. In this case, the audio output unit may be a speaker provided in the vehicle Ve. In yet another example, the audio output device 100 may not include the display unit 116. In this case, the audio output device 100 may be electrically connected, via wired or wireless connections, to a display unit provided in the vehicle Ve, or a user's smartphone, and thereby execute a predetermined display. Similarly, instead of including the sensor group 115, the audio output device 100 may acquire information output by sensors provided in the vehicle Ve from the vehicle Ve based on a communication protocol such as CAN (Controller Area Network).
[0033] [Server device] The server device 200 generates route information indicating a guide route along which the vehicle Ve should travel, based on an upload signal S1 including a destination and the like received from the audio output device 100. The server device 200 then generates a control signal S2 related to information output in response to the user's information request, based on the user's information request and the traveling state of the vehicle Ve indicated in the upload signal S1 transmitted thereafter by the audio output device 100. The server device 200 then transmits the generated control signal S2 to the audio output device 100.
[0034] Furthermore, the server device 200 generates content for providing information to the user of the vehicle Ve and for dialogue with the user, and transmits the content to the audio output device 100. The provision of information to the user mainly includes push-type information provision initiated by the server device 200 side when triggered by the vehicle Ve entering a predetermined driving situation. Furthermore, the dialogue with the user is basically pull-type dialogue initiated by a question or inquiry from the user. However, the dialogue with the user may also be initiated by push-type information provision.
[0035] 3 is a diagram showing an example of a schematic configuration of the server device 200. The server device 200 mainly includes a communication unit 211, a storage unit 212, and a control unit 214. The elements within the server device 200 are connected to each other via a bus line 210.
[0036] The communication unit 211 performs data communication with external devices such as the audio output device 100 under the control of the control unit 214. The storage unit 212 is configured with various types of memory such as RAM, ROM, and non-volatile memory (including a hard disk drive, flash memory, etc.). The storage unit 212 stores programs for the server device 200 to execute predetermined processes. The storage unit 212 also includes a map DB4 and a facility information DB5.
[0037] The control unit 214 includes a CPU, a GPU, and the like, and controls the entire server device 200. The control unit 214 also executes programs stored in the storage unit 212 to operate together with the audio output device 100 and perform processes such as route guidance and information provision for the user. For example, the control unit 214 generates route information indicating a guidance route or a control signal S2 related to information output in response to an information request from the user, based on an upload signal S1 received from the audio output device 100 via the communication unit 211. The control unit 214 then transmits the generated control signal S2 to the audio output device 100 via the communication unit 211.
[0038] <Push-type information provision> Next, push-type information provision will be described. Push-type information provision refers to the audio output device 100 audibly outputting information related to a driving situation of the vehicle Ve to the user when the vehicle Ve is in a predetermined driving situation. Specifically, the audio output device 100 acquires driving situation information indicating the driving situation of the vehicle Ve based on the output of the sensor group 115 as described above, and transmits the information to the server device 200. The server device 200 stores table data for performing push-type information provision in the storage unit 212. The server device 200 refers to the table data, and when the driving situation information received from the audio output device 100 installed in the vehicle Ve matches a trigger condition defined in the table data, the server device 200 acquires output information using text data corresponding to the trigger condition and transmits the information to the audio output device 100. The audio output device 100 audibly outputs the output information received from the server device 200. In this way, information corresponding to the driving situation of the vehicle Ve is audibly output to the user.
[0039] The driving situation information may include at least one piece of information that can be acquired based on the functions of each unit of the audio output device 100, such as the position of the vehicle Ve, the direction of the vehicle, traffic information around the position of the vehicle Ve (including speed limits and congestion information), the current time, the destination, etc. The driving situation information may also include any of the voice (excluding the user's speech) acquired by the microphone 117, the image captured by the external camera 119, and the image captured by the internal camera 120. The driving situation information may also include information received from the server device 200 via the communication unit 111.
[0040] <Provision of facility information> Next, the provision of facility information will be described as an example of the above-mentioned push-type information provision. The provision of facility information refers to recommending to the user facilities that are located on or near the route that the vehicle Ve is scheduled to travel. Specifically, in this embodiment, the server device 200 provides the user with information about facilities that are located near the current location of the vehicle Ve by push-type information provision.
[0041] [Function Configuration] 4 shows the functional configuration of the server device 200 for providing facility information. As shown in the figure, the server device 200 functionally comprises an audio content generation unit 231, a TTS (Text To Speech) engine 232, and a display content generation unit 233. The audio content generation unit 231 and the display content generation unit 233 acquire facility information from the facility information DB 5.
[0042] Facility information is also called POI (Point Of Interest) data. Fig. 5 shows an example of facility information. In the example of Fig. 5, the facility information stores location (latitude, longitude), area name, name, major category, minor category, characteristics, rating, etc. in association with a facility ID assigned to each facility.
[0043] "Location (latitude, longitude)" is the location information of the facility. "Area name" is information indicating the geographical area to which the facility belongs, such as the name of a region. "Name" is the name of the facility; for example, if the facility is a company, the company name is used, and if the facility is a store, the store name is used. "Major category" is the large category, or major classification, that indicates the facility, and "minor category" is the small category, or minor classification, that indicates the facility. "Characteristics" are the characteristics of the facility. "Rating" is the rating of multiple users for the facility, and is expressed, for example, as a number from "0" to "5."
[0044] The audio content generation unit 231 references the facility information DB5 and acquires facility information about facilities within a predetermined range from the current location of the vehicle Ve. The audio content generation unit 231 then generates audio content by selecting and combining portions of the information included in the facility information. For example, if the current location of the vehicle Ve is near Kawagoe, the audio content generation unit 231 acquires data such as the name, subcategory, and characteristics of the facility with facility ID "060" shown in FIG. 5 and combines them to generate audio content such as "We have 'BIG PIZZA,' an Italian restaurant popular for its large pizzas." The audio content generation unit 231 may also generate audio content by inserting the "name," "characteristics," and other information from the facility information into multiple pre-prepared template sentences to create sentences. The audio content generation unit 231 then outputs the text data of the generated audio content to the TTS engine 232 and the display content generation unit 233.
[0045] The TTS engine 232 converts the text data of the input audio content into an audio file and outputs it. This audio file is sent to the audio output device 100.
[0046] The display content generation unit 233, like the audio content generation unit 231, refers to the facility information DB 5 and acquires facility information about facilities within a predetermined range from the current position of the vehicle Ve. Then, the display content generation unit 233 generates display content using the facility information acquired from the facility information DB 5 and the text data of the audio content input from the audio content generation unit 231.
[0047] FIG. 6 shows a display example of display content. The example in FIG. 6 is an example in which a smartphone is used as the display unit 116. As shown in the figure, a display example 70 includes map data 71 and display content 72. The map data 71 shows a map of the area around the current position of the vehicle Ve. In the display example 70 in FIG. 6, a plurality of display contents 72 (72a, 72b, ...) are displayed. Each display content 72 relates to a different facility.
[0048] The display content generation unit 233 generates display content that summarizes the details of the audio content. For example, the display content generation unit 233 references the text data of the audio content and generates display content using the same words as the audio content. In the display example 70 of FIG. 6, display content 72a is display content related to the Italian restaurant "BIG PIZZA." For example, as described above, it is assumed that the audio content generation unit 231 generates audio content about this facility saying, "We have an Italian restaurant called "BIG PIZZA" that is popular for its large pizzas." In this case, the display content generation unit 233 generates display content such as "An Italian restaurant called "BIG PIZZA" that is popular for its large pizzas" using the same words as the audio content, as shown in the figure.
[0049] The display content generation unit 23 may generate the display content by combining some of the words included in the text data of the audio content with words acquired from the facility information DB 5 that are not included in the audio content. In this way, the display content generation unit 233 generates display data for the display content. The display data for the display content is image data for the display content or data required to reproduce the display content in the audio output device 100. The display data for the display content is transmitted to the audio output device 100.
[0050] The audio output device 100 plays back the audio file received from the server device 200, thereby outputting the audio content from the speaker 118. The audio output device 100 also displays the display content on the display unit 116 using the display data received from the server device 200.
[0051] [Content Features] (Relationship between audio content and display content) Next, the relationship between the audio content and the display content will be explained. The audio content generation unit 231 extracts a plurality of words from the facility information exemplified in FIG. 5, creates a sentence using those words, and generates the audio content. In contrast, the display content generation unit 233 generates the display content by arranging at least a portion of the plurality of words included in the audio content. In this way, the display content generation unit 233 generates the display content for a certain facility using the same words as the audio content, so that the audio content and the display content have similar expressions and are more likely to be memorable to the user.
[0052] (Change the output content) Next, we will explain how to change the content of the audio content and the display content. For convenience of explanation, the audio content and the display content will be collectively referred to as "content", and the audio content generation unit 231 and the display content generation unit 233 will be collectively referred to as "content generation unit".
[0053] When displaying facility information about a certain facility, the content generation unit outputs different facility information each time. For example, when a vehicle Ve passes near a certain facility X for the first time (first time), the content generation unit generates content using four words extracted from the facility information. If the vehicle Ve then passes near the same facility X again (second time), the content generation unit generates content using multiple words, at least one of which is different from the four words used to generate the content the first time. This means that different content about facility X is output each time the vehicle Ve passes near the same facility X. Therefore, by repeatedly passing near facility X, the user can obtain various information about facility X.
[0054] Here, we will explain one method for changing the content to be output for each facility. In one method, the content generation unit changes the words extracted from the facility information for generating content. For example, in the case of the facility information shown in FIG. 5, the content generation unit uses the same words each time for basic information such as "area name" and "name," but for other information, the content generation unit varies the content by, for example, selecting one of "major category" and "minor category," or randomly using multiple words included in "characteristics."
[0055] In another method, the content generation unit may gradually change the content to more detailed information depending on the number of times the vehicle Ve passes near a facility. In other words, the content generation unit changes the content from general information to detailed information depending on the number of times content about a certain facility is output. For example, the first time the vehicle passes near a certain facility X, the content generation unit generates content that provides an overview of the facility X. Then, as the number of times the vehicle passes near the same facility X increases, the content generation unit includes more detailed information about the facility X in the content. This allows the user to learn more about a certain facility X by passing near the facility X multiple times.
[0056] Fig. 7 shows an example of facility information prepared when more detailed information is output depending on the number of times content is output. The example in Fig. 7 is facility information about a certain facility X, which is prepared by classifying it into multiple detail levels (six levels in this example) as shown. Note that in Fig. 7, the smaller the numerical value of the detail level, the lower the degree of detail, i.e., the more general the information, and the larger the numerical value of the detail level, the higher the degree of detail, i.e., the more detailed the information.
[0057] Specifically, level 1 is the least detailed level, and level 1 information corresponds to a category such as an overview of a facility. An example of level 1 information is "There is a restaurant up ahead."
[0058] Level 2 is the second lowest level of detail, and level 2 information corresponds to categories such as facility location, distance, and directions. Examples of level 2 information include "It's about 5 minutes from here," and "Turn left at the next intersection and it's about 500 meters away."
[0059] Level 3 is the third least detailed level, and level 3 information corresponds to the category of information registration recommendations. Examples of level 3 information include "BB has bookmarked restaurant AA. How about you?" and "You can bookmark restaurant AA on your smartphone."
[0060] Level 4 is the third most detailed level, and level 4 information corresponds to the category of detailed facility information. Examples of level 4 information include "There is parking for 10 cars," "The lunch set is 1,200 yen," and "The cake set is 500 yen."
[0061] Level 5 is the second most detailed level, and level 5 information corresponds to the category of information for visiting the facility. Examples of level 5 information include "You can add restaurant AA to your current route" and "You can make a reservation for restaurant AA now."
[0062] Level 6 is the most detailed level, and level 6 information corresponds to the category of follow-up information after a visit to a facility. Examples of level 6 information include "Visit three more times and get a free cup of coffee," "Was it a good experience?", and "Rate the facility on a 5-point scale."
[0063] In this way, as the number of times content is output for a certain facility increases, the information contained in the content is changed to more detailed information, making it possible to provide the user with more detailed information about facilities that are repeatedly passed by.
[0064] (Display of content release date) Next, the content output date included in the display content will be described. As shown in FIG. 6, the display content 72 includes the date on which the display content was output. In the example of FIG. 6, the vehicle Ve first passed near the Italian restaurant "BIG Pizza" and the display content 72a related to the facility "BIG Pizza" was first output, i.e., the initial output date, is February 16, 2021. Furthermore, the vehicle Ve most recently passed near the same facility and the display content 72a was output, i.e., the previous output date, is November 15, 2021. Therefore, the display content 72a includes the initial output date of the content, "2021 / 2 / 16," and the previous output date, "2021 / 11 / 15."
[0065] The previous output date of the display content is updated every time the display content related to the facility is output. The server device 200 stores, for each facility, history information (hereinafter also referred to as "content output history") relating to the date or date and time when the display content for that facility was output, in association with the user ID or vehicle ID, and updates the content output history every time the display content related to that facility is output. When generating display content related to a certain facility, the display content generation unit 233 refers to the content output history and includes the first and previous output dates of the display content related to that facility in the display content. Note that if the display content has only been output once in the past, the output date of that one output may be displayed as either or both of the "first" and "previous" output dates.
[0066] This allows the user to see the date when they passed near the facility in the past and recall the audio content that was output in the past. Note that in the example of Fig. 6, the date when the content was output in the past is displayed as part of the displayed content, but it is also possible to display not only the "date" but also the "time." In other words, it is also possible to display the date and time when the content was output.
[0067] [Information provision processing] Fig. 8 is a flowchart of the information provision process. This process is realized by the server device 200 shown in Fig. 3 controlling the audio output device 100. This process is realized by the control unit 214 executing a program prepared in advance and operating as each element shown in Fig. 4. This process is also repeatedly executed at predetermined time intervals.
[0068] First, the control unit 114 of the audio output device 100 acquires driving situation information related to the current driving situation of the vehicle Ve and transmits it to the server device 200. The server device 200 acquires the driving situation information from the audio output device 100 and acquires facility information around the current position of the vehicle Ve by referring to the facility information in the facility information DB5 (step S11). Specifically, the server device 200 acquires facility information about facilities that exist within a predetermined range from the current position of the vehicle Ve. Note that the "predetermined range" is determined in advance as, for example, a range of a radius of Xm from the current position.
[0069] Next, the audio content generation unit 231 of the server device 200 generates audio content using the acquired facility information (step S12). Specifically, the audio content generation unit 231 extracts multiple words included in the facility information as described above and generates text data of the audio content by organizing them into sentences. Then, the TTS engine 232 converts the text data of the audio content into an audio file.
[0070] Next, the display content generation unit 233 generates display content based on the facility information and the audio content generated by the audio content generation unit 231 (step S13). Specifically, the display content generation unit 233 generates display data for the display content by arranging words included in the audio content. Note that the display content generation unit 233 refers to the content output history described above, and includes, for each facility, the past output dates of display content related to that facility, specifically the first and previous output dates, in the display content.
[0071] Next, the server device 200 transmits the audio file of the audio content and the display data of the display content to the audio output device 100 (step S14). When the audio output device 100 receives the audio file of the audio content, it plays the audio file and outputs the audio content. When the audio output device 100 receives the display data of the display content, it displays the display content on the display unit 116. In this way, the audio content and the display content are output to the user.
[0072] Next, the server device 200 stores information about the audio content and display content transmitted to the audio output device 100 in step S14 in the storage unit 212 as a content output history (step S15). Specifically, the server device 200 stores the details of the audio content and display content, the date or time when they were transmitted to the audio output device 100, and the like as a content output history. Then, the information providing process ends.
[0073] [Variations] (Variation 1) In the above embodiment, the configuration for generating content shown in Fig. 4 is provided in the server device 200, and the audio content and display content are generated on the server device 200 side. Alternatively, the configuration for generating content may be provided in the audio output device 100, and the audio content and display content may be generated and output on the audio output device 100 side.
[0074] (Variation 2) In the above embodiment, facility information such as stores is provided, but instead, information about various objects such as rivers and bridges that should be provided while the vehicle Ve is traveling may be provided in the same manner as facility information. In this specification, objects for which information is provided, including facilities such as stores, rivers, bridges, etc., are collectively referred to as "land features."
[0075] (Variation 3) In the above-described embodiments, the program can be stored using various types of non-transitory computer-readable media and supplied to a control unit, such as a computer. The non-transitory computer-readable media includes various types of tangible storage media. Examples of non-transitory computer-readable media include magnetic storage media (e.g., flexible disks, magnetic tapes, hard disk drives), magneto-optical storage media (e.g., magneto-optical disks), CD-ROMs (Read Only Memory), CD-Rs, CD-R / Ws, and semiconductor memories (e.g., mask ROMs, PROMs (Programmable ROMs), EPROMs (Erasable PROMs), flash ROMs, and RAMs (Random Access Memory)).
[0076] Although the present invention has been described above with reference to the embodiments, the present invention is not limited to the above embodiments. Various modifications within the scope of the present invention that would be understood by those skilled in the art can be made to the configuration and details of the present invention. In other words, the present invention naturally includes various modifications and alterations that would be possible for those skilled in the art based on the entire disclosure, including the claims, and the technical ideas. Furthermore, the disclosures of the above-cited patent documents and other documents are incorporated herein by reference. [Explanation of symbols]
[0077] 100 Audio output device 200 Server device 111, 211 Communications Department 112, 212 Storage section 113 Input section 114, 214 Control unit 115 Sensor Group 116 Display section 117 Mike 118 Speaker 119 Exterior Camera 120 In-car camera
Claims
1. an acquisition unit that acquires feature information related to features based on the vehicle position; an audio content generation unit that generates audio content related to the feature based on the feature information; a display content generation unit that generates display content indicating details of the audio content based on the feature information and predetermined words included in the audio content; Equipped with the audio content generation unit changes the audio content to be generated in accordance with the number of times the audio content is output; The display content generating unit changes the display content to be generated in response to a change in the content of the audio content.
2. The content output device according to claim 1 , wherein the audio content generation unit increases the level of detail of the audio content as the number of times the audio content is output increases.
3. 3. The content output device according to claim 1, wherein the display content includes the first output date and the latest output date of the audio content and the display content.
4. The content output device according to claim 1 , wherein the display content generation unit generates the display content using the same words as the audio content.
5. The content output device according to claim 4 , wherein the display content generation unit generates the display content by arranging the same words as those in the audio content.
6. A content output method executed by a computer, comprising: Acquire feature information related to the feature based on the vehicle position; generating audio content related to the feature based on the feature information; generating display content that indicates the details of the audio content based on the feature information and predetermined words included in the audio content; The generated audio content is changed according to the number of times the audio content is output. A content output method in which the generated display content is changed in response to a change in the content of the audio content.
7. Acquire feature information related to the feature based on the vehicle position; generating audio content related to the feature based on the feature information; causing a computer to execute a process of generating display content indicating the details of the audio content based on the feature information and predetermined words included in the audio content; The generated audio content is changed according to the number of times the audio content is output. A program in which the generated display content is changed in response to changes in the content of the audio content.
8. A storage medium storing the program according to claim 7.
Citation Information
Patent Citations
On-vehicle equipment, data creation device, data creation program, and vehicle-mounted agent system
JP2004163232A
Information guidance system
JP2016184242A
Server and information terminal
JP2018018299A
Voice guidance device, voice guidance server, and voice guidance method
JP2020204574A