Voice SNS management device and program
The voice SNS management system facilitates two-way communication by suggesting location-based audio data playback and managing user reactions, enhancing the pedestrian experience by integrating voice data discovery and interaction.
Patent Information
- Authority / Receiving Office
- JP · JP
- Patent Type
- Applications
- Current Assignee / Owner
- ARROWS CO LTD
- Filing Date
- 2025-05-19
- Publication Date
- 2026-05-11
AI Technical Summary
Conventional social networking services (SNS) primarily offer one-way communication of voice memos based on vehicle location, limiting the user's experience and requiring them to sift through vast amounts of information, detracting from the joy of movement, especially for pedestrians.
A voice SNS management system that allows pedestrians to discover location-associated audio data by walking, enabling two-way communication through voice data playback suggestions, response acquisition, and management, using a device with components like a control unit, storage unit, and communication unit to manage and analyze voice data and user reactions.
Enables a wandering-type SNS where pedestrians can discover and interact with location-associated audio data, balancing the joy of movement with two-way communication, allowing for comments and tags, and filtering information based on user preferences.
Smart Images

Figure 2026076100000001_ABST
Abstract
Description
Technical Field
[0006] , ,
[0001] The present invention relates to a management device and a program for a voice SNS.
Background Art
[0002] Based on the current position obtained using a global positioning system (GPS) or the like, route guidance information, advertisement information, emergency information, sightseeing information, traffic information, weather information, news information, and other information are distributed. The news information distributed in this way may include information related to articles posted on a social networking service (SNS) in addition to news information. These information are distributed not only in a visual mode exemplified by display of a map, an image, a video, and text, but also in voice.
[0003] Regarding a technique for distributing voice data associated with the current position in a conventional service and a conventional technique, etc., Patent Document 1 discloses a system in which a control unit of a vehicle navigation device stores voice memo data received from a distribution server in a storage unit, and when the vehicle is located at a position or a playback range associated with the voice memo data, the voice memo data is played by a speaker unit.
[0004] The technique described in Patent Document 1 can cause a voice memo to be played according to the current position of the vehicle.
Prior Art Documents
Patent Documents
[0005]
Patent Document 1
Summary of the Invention
Problems to be Solved by the Invention
[0006] The technology described in Patent Document 1 can enable the unilateral distribution of voice memos from a distribution server based on the vehicle's current location. Incidentally, in social networking services (SNS) and the like, instead of unilaterally distributing information, two-way communication across national borders is carried out through reactions such as comments or ratings on posted articles.
[0007] However, the technology described in Patent Document 1 is limited to enabling one-way communication by delivering voice memos that are played back according to the vehicle's current location. Therefore, there is room for further improvement in the technology described in Patent Document 1 in terms of enabling two-way communication.
[0008] However, unlike the technology described in Patent Document 1, conventional SNSs enable two-way communication that is not bound by geographical requirements, resulting in a vast amount of communication information being delivered to the user's device. Therefore, users may be required to interrupt their travel and spend a long time selecting and filtering the information delivered to their device.
[0009] Thus, conventional social networking services (SNS) can deprive users of the joy of movement in exchange for two-way communication that is not bound by geographical constraints. The joy of movement is especially important for pedestrians, whose travel routes are not limited to roadways. Examples of such joys include the joy of walking, strolling, and exploring. Therefore, there is a need for technological means to realize an SNS that balances the joy of movement for pedestrians with two-way communication.
[0010] The objective of this invention is to realize an audio social networking service (SNS), which could also be called a wandering-type SNS, in which pedestrians discover location-associated audio data by walking around themselves and communicate with the posters of such audio data and other users. [Means for solving the problem]
[0011] As a result of diligent research to solve the above problems, the inventors of the present invention have found that the above objectives can be achieved by proposing the playback of audio data associated with a location according to the location of the terminal, acquiring the response to this audio data from the terminal, and managing it in association with the audio data. Thus, the inventors of the present invention have completed the present invention.
[0012] One aspect of the present invention provides a voice SNS management device comprising: a voice data management unit that manages voice data in association with location; a positioning information acquisition unit that acquires first positioning information indicating the location of a first terminal carried by a pedestrian from the first terminal; a suggestion display command unit that instructs the first terminal to display a suggestion for playback of voice data associated with the location when the location indicated by the first positioning information corresponds to the location; a response acquisition unit that acquires data indicating the pedestrian's response to the voice data from the first terminal; and a response management unit that manages the data indicating the response in association with the voice data.
[0013] The management device of this type can determine whether the location of the first terminal carried by the pedestrian corresponds to the aforementioned location, based on the audio data associated with the location managed by the audio data management unit and the first positioning information acquired by the positioning information acquisition unit. The management device of this type then instructs the first terminal, via the suggestion display command unit, to display a suggestion to play the audio data associated with the location if the location of the first terminal corresponds to the location associated with the audio data. As a result, the management device of this type enables the pedestrian carrying the first terminal to discover the audio data associated with the location by walking around themselves. This allows the pedestrian to experience the joy of movement that cannot be obtained through ordinary strolls, such as discovering audio data by walking around themselves.
[0014] The management device of this type then acquires data from the first terminal indicating the pedestrian's reaction to the audio data using a reaction acquisition unit. Furthermore, the management device of this type manages the data indicating the reaction in association with the audio data using a reaction management unit. Through the above series of processes, the management device of this type enables users who have discovered audio data by walking around to communicate with the poster of the audio data and other users.
[0015] Based on the above, this embodiment realizes an audio SNS, which could be called a wandering SNS, in which pedestrians discover location-associated audio data by walking around themselves and communicate with the posters of such audio data and other users.
[0016] Furthermore, the present invention can take various forms, as exemplified by embodiments that realize communication via tags and comments, embodiments that propose voice data associated with a route, embodiments that acquire voice data associated with positioning information from a second terminal carried by a pedestrian, and embodiments of a program that causes a first terminal carried by a pedestrian to cooperate with the management device described above. Each of these embodiments contributes to the realization of a voice SNS, which could also be called a wandering-type SNS, through the effects brought about by the addition of unique configurations. [Effects of the Invention]
[0017] Based on the above, the present invention makes it possible to realize an audio SNS, which could also be called a wandering SNS, in which pedestrians discover location-associated audio data by walking around themselves and communicate with the poster of such audio data and other users. [Brief explanation of the drawing]
[0018] [Figure 1] Figure 1 is a block diagram showing an example of the hardware and software configuration of system S in this embodiment. [Figure 2] Figure 2 shows an example of the voice database 131. [Figure 3]FIG. 3 is an example of the reaction database 132. [Figure 4] FIG. 4 is a main flowchart showing an example of a preferred flow of the management process executed by the management device 1 of the present embodiment. [Figure 5] FIG. 5 is a figure following the previous figure. [Figure 6] FIG. 6 is a figure following the previous figure. [Figure 7] FIG. 7 is a figure following the previous figure. [Figure 8] FIG. 8 is a flowchart showing an example of a preferred flow of the voice SNS usage process executed by the terminal T of the present embodiment. [Figure 9] FIG. 9 is a figure following the previous figure. [Figure 10] FIG. 10 is a figure following the previous figure. [Figure 11] FIG. 11 is an example of the playback proposal screen D displayed on the terminal T.
MODE FOR CARRYING OUT THE INVENTION
[0019] First of all, although the following disclosure, charts, and / or claims, etc. are described as being alone or in combination with one or more other aspects, the subject matter of the immediate disclosure is not intended to be limited in that way. That is, the immediate disclosure, charts, and claims are intended to include the various aspects described herein, either alone or in one or more combinations with each other. For example, even when the immediate disclosure describes and illustrates the first embodiment, the second embodiment, and the third embodiment in such a way that the first embodiment is described and illustrated particularly in relation to the second embodiment, or the second embodiment is described and illustrated only in relation to the third embodiment, the immediate disclosure and illustration are not limited in that way, and only the first embodiment, only the second embodiment, only the third embodiment, or one or more combinations of the first, second, and / or third embodiments, for example, the first embodiment and the second embodiment, the first embodiment and the third embodiment, the second embodiment and the third embodiment, or the first, second, and third embodiments may be included.
[0020] In this text, the phrase "or" is used to mean a "non-exclusive" arrangement unless explicitly specified otherwise. For example, when we say "item x is A or B," it means either (1) item x is either A or B, or (2) item x is both A and B. In other words, the word "or" is not used to define an "exclusive" arrangement.
[0021] Furthermore, when the phrases "contain at least one" or "contain at least one of the following" are used in the text, they mean that the system or element contains one or more of the elements listed after the phrase. For example, if there are three types of elements, from element 1 to element 3, the phrases "contain at least one" or "contain at least one of the following" are interpreted as any of the following structural arrangements: a device containing element 1, a device containing element 2, a device containing element 3, a device containing element 1 and element 2, a device containing element 1 and element 3, a device containing element 2 and element 3, or a device containing element 1, element 2, and element 3.
[0022] The same interpretation is intended when the phrase "used in at least one of the following" is used in the text. Furthermore, "and / or" as used in the text is used as a linguistic conjunction to indicate that one or more of the listed elements or conditions are included or occur. For example, a device containing the first element, the second element, and / or the third element is interpreted as any of the following structural arrangements: a device containing the first element, a device containing the second element, a device containing the third element, a device containing the first and second elements, a device containing the first and third elements, a device containing the second and third elements, or a device containing the first, second, and third elements.
[0023] Furthermore, the use of the phrase "and / or" in this text signifies a "non-exclusive" arrangement, as stipulated in the Japanese Industrial Standard (JIS) "Format and Preparation Method of Standards Documents JIS Z 8301".
[0024] The following describes in detail an example of an embodiment of the present invention with reference to the drawings.
[0025] <System S> Figure 1 is a block diagram showing an example of the hardware and software configuration of system S in this embodiment. The voice SNS management system (system S) according to this embodiment is configured to include a voice SNS management device 1. The management device 1 is configured to communicate with terminal T via network N.
[0026] [Management device 1] The management device 1 comprises various hardware components such as a control unit 11, a storage unit 13, and a communication unit 14. The management device 1 manages a voice social networking service (SNS), which could also be called a wandering SNS, in which pedestrians discover location-associated voice data by walking around themselves and communicate with the poster of the voice data and other users. The type of management device 1 is not particularly limited and may be, for example, a server device, a cloud server, etc.
[0027] [Control Unit 11] The control unit 11 includes a Central Processing Unit (CPU), Random Access Memory (RAM), and Read Only Memory (ROM), among other things.
[0028] The control unit 11 cooperates with at least one of the storage unit 13 and the communication unit 14 as needed. The control unit 11 then implements the software components of the program of this embodiment executed by the management device 1, such as the recording data acquisition unit 111, the positioning information acquisition unit 112, the voice data management unit 113, the suggestion display command unit 114, the voice data transmission unit 115, the response acquisition unit 116, the response management unit 117, and the voice data analysis unit 118.
[0029] The details of the management processes implemented by the aforementioned software components will be explained later using Figures 4 to 7.
[0030] [Storage section 13] The storage unit 13 is a device on which data and / or files are stored, and has a storage unit that stores data non-temporarily using a hard disk, semiconductor memory, recording medium, and memory card, etc. The storage unit 13 stores programs executed by a microcomputer, a voice database 131, a response database 132, a voice analysis database 133, etc.
[0031] (Audio Database 131) The audio database 131 stores audio data associated with locations where playback is suggested. The audio data associated with locations is managed by the audio data management unit 113.
[0032] To facilitate playback on terminal T, it is preferable that the audio data is associated with various attributes related to audio data playback, such as playback time, bitrate, playback frequency, number of audio channels, compression format, and audio data format. To facilitate data manipulation, it is preferable that the audio data is associated with information that identifies the audio data (e.g., an audio ID).
[0033] In the voice database 131, the data format of the location associated with the voice data is not particularly limited, as long as it is a format that can indicate a geographically extended range. For example, the data format may be configured to include one or more formats such as a format that specifies the range by parameters including a pair of center point and radius in the coordinate system related to the positioning system, a format that specifies the range by a figure on the coordinate system related to the positioning system (e.g., a polygon represented by multiple vertices), a format that specifies the range by a geographic mesh corresponding to the coordinate system related to the positioning system, a format that specifies the range by an address, or a format that specifies the range by a place name (e.g., block name, street name, facility name, river name, etc.).
[0034] The content of the audio data is not particularly limited. Preferably, the audio data includes information about nearby shops, etc. (e.g., advertisements, user reviews, etc.) so that pedestrians can obtain information about nearby shops, etc. by walking around on their own. Preferably, the audio data includes information about places (e.g., historical introductions, current situation introductions, usage guides, information about famous people associated with the place, trivia, etc.) so that pedestrians can obtain information about nearby places by walking around on their own. Preferably, the audio data includes emergency information about places (e.g., weather warnings, earthquake information, evacuation information, landslide information, flood information, crime information, etc.) so that pedestrians can obtain emergency information about the neighborhood by walking around on their own. Preferably, the audio data includes music appropriate for the place so that pedestrians can enjoy music appropriate for the place by walking around on their own.
[0035] Figure 2 shows an example of the voice database 131. In this example, the voice database 131 stores the first voice data identified by the voice ID "S0001", the second voice data identified by "S0002", the third voice data identified by "S0003", and so on.
[0036] The first audio data is an advertisement for a restaurant associated with the location "△△-△, △△-cho, Omiya-ku, Saitama City," and says, "Restaurant △△ is having a new sweets festival starting today!..." The second audio data is an introduction to a historical site associated with the location "35.940703785,139.707281881,r50m" (within a circle with a radius of 50m centered at 35 degrees 94.0703785 minutes north latitude and 139 degrees 70.7281881 minutes east longitude), and says, "During the Jomon period, people who made clay figurines lived around here..." The third audio data is an emergency information audio associated with the location "Shibakawa River, from Shinbashi to the confluence with the Koyama River," and says, "Due to the torrential rain associated with the approach of Typhoon △, the Shibakawa River..."
[0037] By storing this audio data in the database, the user of terminal T1 can discover information about an interesting new menu being announced near restaurant △△, learn about the history of a historical site near a historical site, and discover urgent information about the Shibakawa River near the river. The user can also add comments and tags to this audio data.
[0038] (Reaction Database 132) The response database 132 stores data (response data) that indicates pedestrians' reactions to audio data. This response data includes, for example, tag data indicating the category of the audio data, and comment data on the audio data. By storing data indicating pedestrian reactions, including tag and comment data, in the response database 132, the management device 1 can provide users with a platform for communication via tags or comments.
[0039] To allow for filtering through evaluation, the response data preferably includes evaluations of the audio data. The evaluation may be, for example, the number of times a user has given a positive rating. To keep the influence of a single user's response within an appropriate range, it is preferable that the management device 1 manages so that one user can give a positive rating to one audio data a predetermined number of times.
[0040] The response data is stored in association with the aforementioned audio data. To facilitate data manipulation, it is preferable that the response data is associated with information that identifies the response data (e.g., a response ID).
[0041] Figure 3 shows an example of a reaction database 132. In this example, the reaction database 132 stores the first reaction data identified by reaction ID "R0001", the second reaction data identified by "R0002", the third reaction data identified by "R0003", the fourth reaction data identified by "R0004", and so on.
[0042] The first reaction data shows the reaction to the first audio data identified by the audio ID "S0001," and indicates that the tag "Gourmet" was attached as a reaction to that audio data. The second reaction data also shows the reaction to the first audio data, and indicates that the comment "The new dish I ate after my usual horse mackerel set meal was..." was attached as a reaction to that audio data. The third reaction data shows the reaction to the second reaction data, and indicates that the comment "Was that combination a good choice...?" was attached as a reaction to the comment related to that reaction data. The fourth reaction data shows the reaction to the third audio data identified by the audio ID "S0003," and indicates that the rating "Positive rating: 100 times" was attached as a reaction to that audio data.
[0043] By storing this reaction data in the database of this example, users of the first terminal T1 can filter the audio data by the tag "Gourmet" or the range of the number of positive ratings. Furthermore, the user can check the number of comments attached to the audio data or reaction data and read each comment.
[0044] (Voice Analysis Database 133) The voice analysis database 133 stores data (voice analysis data) output by the voice data analysis unit 118. The voice analysis data is, for example, voice data labeled as inappropriate or the voice recognition results of voice data. An "inappropriate post" here refers to, for example, a post that hinders the communication that the management device 1 aims to achieve. Examples of such posts include discriminatory remarks, remarks that incite crime, remarks that violate compliance, insulting remarks, remarks that are contrary to public order and morals, remarks that are contrary to the facts, remarks that constitute a violation of privacy, and voice content that violates licenses. The voice analysis data is used as training data for machine learning related to the voice data analysis unit 118.
[0045] For machine learning that enables reason-based filtering, the above-mentioned labels preferably include data indicating the reason why they are inappropriate. To enable the use of training data provided by external devices, the speech analysis data may include not only the data output by the speech data analysis unit 118 but also similar data created by external devices. To enable the use of training data that reflects human insights, the speech analysis data may also include similar data created through processes that include manual work.
[0046] [Communications Section 14] The communication unit 14 is not particularly limited as long as it enables communication by connecting the management device 1 to the network N. Examples of the communication unit 14 include a network card compatible with the Ethernet standard and a communication device compatible with wireless LAN.
[0047] [Network N] The type of network N is not particularly limited as long as it enables communication with the management device 1, etc. Examples of network N include the internet, a mobile phone network, a wireless LAN, etc.
[0048] [Terminal T] Terminal T is carried and used by pedestrians. Preferably, terminal T is a portable device that can be carried and used by pedestrians (e.g., a smartphone, tablet, laptop, etc.). In the following, terminal T may be referred to as the first terminal T1 when playing audio related to a voice SNS, and as the second terminal T2 when recording said audio.
[0049] Terminal T contains a program that causes terminal T to perform various processes related to posting and using voice data on the voice-based social networking service.
[0050] The various processes related to posting audio data to an audio SNS include, for example, a recording detection step, a recording step, a recording positioning step, and a registration execution step.
[0051] The various processes related to the use of submitted audio data include, for example, a positioning information transmission step, a data reception step, a suggestion display step, a playback step, an input acquisition step, and a response transmission step.
[0052] The voice SNS usage processing performed by the program and the details of each of the steps described above will be explained later using Figures 8 to 10.
[0053] In order to enable two-way communication on a voice SNS, it is preferable that terminal T has the functions of both the first terminal T1 and the second terminal T2.
[0054] [Main flowchart of management process] Figure 4 is a main flowchart showing an example of a preferred flow of management processing performed by the management device 1 of this embodiment. Figure 5 is a continuation from the previous figure. Figure 6 is a continuation from the previous figure. Figure 7 is a continuation from the previous figure. The following is an example of a preferred flow of management processing performed by the management device 1 of this embodiment, using Figures 4 to 7.
[0055] [Step S1: Determine if you were asked to register voice data] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the recording data acquisition unit 111. The control unit 11 then performs a process to determine whether the recording data acquisition unit 111 has requested the registration of voice data (registration determination step). If the control unit 11 determines that it has been requested, it moves the process to step S2; otherwise, it moves the process to step S5.
[0056] The recording data acquisition unit 111 achieves the above-mentioned determination by, for example, determining that a request for registration of voice data has been received from the second terminal T2, through a procedure such as that.
[0057] [Step S2: Obtain recording data] The control unit 11 uses the recording data acquisition unit 111 to acquire recording data recorded by the second terminal T2 carried by the pedestrian (recording data acquisition step). The control unit 11 then moves the process to step S3.
[0058] [Step S3: Obtain second positioning information] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the positioning information acquisition unit 112. The control unit 11 then uses the positioning information acquisition unit 112 to acquire second positioning information from the second terminal T2, which indicates the location of the second terminal T2 when the above-mentioned recording data was recorded (second positioning information acquisition step). The control unit 11 then moves the process to step S4.
[0059] [Step S4: Start managing audio data] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the voice data management unit 113. The control unit 11 then performs the process of starting the management of the above-mentioned recording data (voice data) in association with the location including the location indicated by the above-mentioned second positioning information (voice data management start step). The control unit 11 then moves the process to step S5.
[0060] This allows the management device 1 to pin voice data to location information. In the voice data management start step, the voice data management unit 113 starts managing the recorded data by, for example, associating the location including the location indicated by the second positioning information with the recorded data and storing it in the voice database 131.
[0061] (Provision of incentives) Preferably, in the voice SNS implemented by the management device 1, when voice data related to an advertiser's account is registered by a user who follows the advertiser's account, the management device 1 executes a process to provide an incentive to the user. This allows the management device 1 to encourage general users, who are different from advertisers who operate accounts with the intention of advertising, to spread word-of-mouth about the advertiser. The content of the incentive is not particularly limited and may include, for example, the provision of discount coupons or the gift of presents.
[0062] (Audio data analysis step) The management process preferably includes an audio data analysis step that uses machine learning or the like to identify inappropriate posts that hinder the communication targeted by the management device 1, using recorded audio data as input. Here, machine learning or the like refers to processing using a machine learning model that identifies the content of audio data, or processing that corresponds to a similar classifier. This allows the management device 1 to avoid starting the management of audio data that constitutes an inappropriate post in the audio data management start step, based on the results of the identification. The management device 1 can then realize a healthy audio SNS that does not contain inappropriate posts. Note that "inappropriate posts" here may be the same as those listed in the audio analysis database 133.
[0063] The voice data analysis step is implemented, for example, by the control unit 11 working in cooperation with the storage unit 13 and the communication unit 14 to execute the voice data analysis unit 118, and then by the voice data analysis unit 118 inputting the above-mentioned recording data (voice data) into a machine learning model, etc., and outputting data (voice analysis data) that includes the determination result of whether or not the post is inappropriate. The voice analysis data may be the same as that described in the voice analysis database 133. If the voice analysis data includes the speech recognition result of the voice data, the above-mentioned machine learning model, etc., includes a speech recognition model that converts the voice data into text data. For subsequent machine learning, it is preferable that the voice data analysis unit 118 stores the voice analysis data in the voice analysis database 133.
[0064] (Machine learning related to the voice data analysis unit 118) The audio data analysis unit 118 preferably performs machine learning using the audio analysis data stored in the audio analysis database 133 as training data. This machine learning improves the accuracy of the machine learning model in the audio data analysis step in determining whether a post is inappropriate by using audio analysis data, in which labels indicating whether a post is inappropriate are associated with the content of the audio data, as training data.
[0065] The management process executes a series of operations that propose the playback of audio data associated with the pedestrian's location. As a result, the terminal T used by the pedestrian can automatically detect and play the scattered audio data associated with the location by moving while executing the program of this embodiment, which will be described later. Steps S5 to S9 are an example of this process.
[0066] [Step S5: Determine whether to acquire the first positioning information] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the positioning information acquisition unit 112. The control unit 11 then performs a process to determine whether to acquire first positioning information indicating the location of the first terminal T1 carried by the pedestrian using the positioning information acquisition unit 112 (first positioning information acquisition determination step). If the control unit 11 determines that it will acquire the information, it moves the process to step S6; otherwise, it moves the process to step S10.
[0067] In the first positioning information acquisition determination step, the positioning information acquisition unit 112 determines whether to acquire the first positioning information if a certain amount of time has elapsed since the previous acquisition, or if it has received data related to the first positioning information from the first terminal T1, etc., by following the above determination procedure.
[0068] [Step S6: Obtain the first positioning information] The control unit 11 uses the positioning information acquisition unit 112 to perform a process to acquire first positioning information from the first terminal T1, which indicates the location of the first terminal T1 carried by the pedestrian (first positioning information acquisition step). The control unit 11 then moves the process to step S7.
[0069] [Step S7: Determine if the location in question is under management.] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the suggestion display command unit 114. The control unit 11 then performs a process to determine whether the location indicated by the first positioning information corresponds to a location managed by the voice data management unit 113 (location determination step). That is, the suggestion display command unit 114 determines whether the location indicated by the first positioning information corresponds to a location associated with any voice data managed by the voice data management unit 113. If the control unit 11 determines that it is managed, it moves the process to step S8; otherwise, it moves the process to step S10.
[0070] [Step S8: Obtain audio data associated with the location] The control unit 11, acting on the suggestion display command unit 114, executes a process to acquire audio data associated with the location indicated by the first positioning information (first audio data acquisition step). The control unit 11 then moves the process to step S9. If there are multiple corresponding audio data, the suggestion display command unit 114 acquires some or all of them.
[0071] In order to filter audio data in advance using tags, if tags have been specified by the user of the first terminal T1, it is preferable for the suggestion display command unit 114 to acquire only audio data whose location indicated by the first positioning information is associated with the corresponding location and whose tags match the specified tags.
[0072] In order to narrow down the audio data in advance based on the evaluation, if the user of the first terminal T1 has specified the range of evaluation, it is preferable for the proposed display command unit 114 to acquire only the audio data for which the location indicated by the first positioning information is associated with the relevant location and which has been assigned an evaluation that falls within the specified range of evaluation.
[0073] In order to narrow down the audio data in advance based on the number of comments, if the user of the first terminal T1 has specified a range of the number of comments, it is preferable for the suggestion display command unit 114 to acquire only the audio data in which the location indicated by the first positioning information is associated with the corresponding location and which has a number of comments that falls within the specified range of the number of comments.
[0074] In order to filter audio data in advance based on a time range, if a time range has been specified by the user of the first terminal T1, it is preferable for the suggestion display command unit 114 to acquire only audio data for posting times that are associated with the location indicated by the first positioning information and that fall within the specified time range.
[0075] [Step S9: Send a command to display the playback suggestion screen] The control unit 11 executes a process to send a command via the suggestion display command unit 114 to display a playback suggestion screen D that suggests playback of audio data associated with the location (first suggestion display command step). The control unit 11 then moves the process to step S10.
[0076] In order to ensure that pedestrians notice the playback suggestion screen D is displayed even when they are not looking at the first terminal T1, it is preferable that the first suggestion display command step instructs the first terminal T1 to notify the user of the first terminal T1 about the playback suggestion using non-visual means such as sound or vibration.
[0077] Preferably, the playback suggestion screen D is a screen that lists multiple articles A, each corresponding to a playback suggestion for an audio data, so that playback suggestions for multiple audio data can be viewed in a list. In this case, preferably, article A includes a display element M indicating the poster, an appendix N indicating various attributes such as posting time, playback time, and tags, a button F to send a positive rating as a reaction, a button P to command playback, and a reaction R such as comments and the number of positive ratings. The display element M is, for example, the poster's name and the poster's icon. Preferably, the playback suggestion screen D includes an upload button U for recording audio data and sending it to the management device 1, so that audio data can be posted from the same screen.
[0078] In the steps that command the display of a playback proposal, such as the first proposal display command step, it is preferable that the proposal display command unit 114 commands to display a display that proposes playback of audio data associated with a tag specified by a pedestrian, and to display a display that proposes playback in a manner that allows for the display of comments when comments are associated with the audio data. An example of such a manner is a playback proposal screen D that displays only the article A associated with the audio data associated with the specified tag from among a plurality of article A, and includes a button for the response R to command the display of comments.
[0079] (Display of advertising information) To enhance the appeal of products related to audio data, it is preferable that the playback suggestion screen D is configured to display advertisements associated with the audio data. This display may be, for example, a direct display of an advertisement for a product or service associated with the audio data, a display element that transitions to an advertisement page, or a transition to an advertisement page along with the playback of the audio data. In particular, it is preferable that the display includes a transition to an advertisement page indicated by a uniform resource locator (URL), etc., along with the playback of the audio data. This allows the management device 1 to enhance the appeal of products through URL transitions.
[0080] Although not mandatory, the management process preferably includes a series of processes that instruct the system to display a message suggesting playback of audio data associated with a route if the route of the first terminal T1 substantially includes a route associated with audio data. Steps S10 to S14 are an example of such processes.
[0081] [Step S10: Determine whether to obtain the travel path] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the positioning information acquisition unit 112. The control unit 11 then performs a process to determine whether to acquire the movement path of the first terminal T1 carried by the pedestrian using the positioning information acquisition unit 112 (movement path acquisition determination step). If the control unit 11 determines that it will acquire the path, it moves the process to step S11; otherwise, it moves the process to step S15. In the movement path acquisition determination step, the procedure for determination by the positioning information acquisition unit 112 may be the same as in the first positioning information acquisition determination step.
[0082] [Step S11: Obtain the travel path] The control unit 11 uses the positioning information acquisition unit 112 to perform a process to acquire the movement path of the first terminal T1 carried by the pedestrian from the first terminal T1 (movement path acquisition step). The control unit 11 then moves the process to step S12.
[0083] [Step S12: Determine if the route, including the travel path, is managed.] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the suggestion display command unit 114. The control unit 11 then performs a process to determine whether the route including the aforementioned travel route is managed by the voice data management unit 113 (location determination step). That is, the suggestion display command unit 114 determines whether the aforementioned travel route substantially includes a route associated with any of the voice data managed by the voice data management unit 113. "Substantially includes" here means, for example, that the difference between the travel route and the associated route is within the positioning error. If the control unit 11 determines that it is managed, it moves the process to step S13; otherwise, it moves the process to step S15.
[0084] [Step S13: Obtain audio data associated with the route] The control unit 11 executes a process to acquire audio data associated with the above-mentioned route, as instructed by the suggestion display command unit 114 (second audio data acquisition step). The control unit 11 then moves the process to step S14. If there are multiple audio data files, the suggestion display command unit 114 acquires some or all of them.
[0085] [Step S14: Send a command to display the playback suggestion screen] The control unit 11 executes a process to send a command to the suggestion display command unit 114 to display a playback suggestion screen D that suggests playback of audio data associated with the route (second suggestion display command step). The control unit 11 moves the process to step S15. The second suggestion display command step may be the same as the first suggestion display command step, except that the object to which the suggested audio data is associated is included.
[0086] The second suggestion display command step may be combined with the first suggestion display command step. That is, the management device 1 may be instructed to display a playback suggestion screen D that suggests playback of audio data after it has identified and acquired both audio data associated with a location that satisfies the predetermined conditions described above and audio data associated with a route that satisfies the specific conditions described above.
[0087] [Step S15: Determine if playback of audio data was requested] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the voice data transmission unit 115. The control unit 11 then performs a process to determine whether the voice data transmission unit 115 has requested playback of the voice data related to the playback suggestion screen D described above (voice data transmission determination step). If the control unit 11 determines that playback has been requested, it moves the process to step S16; otherwise, it moves the process to step S17.
[0088] [Step S16: Send audio data] The control unit 11 executes the process of transmitting the above-mentioned voice data to the first terminal T1 using the voice data transmission unit 115 (voice data transmission step). The control unit 11 then moves the process to step S17.
[0089] The control process includes a series of processes related to the acquisition and control of the reaction. Steps S17 to S19 are an example of such a process.
[0090] [Step S17: Determine whether to obtain the reaction] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the response acquisition unit 116. The control unit 11 then performs a process to determine whether to acquire a response to the voice data using the response acquisition unit 116 (response acquisition determination step). If the control unit 11 determines that a response should be acquired, it moves the process to step S18; otherwise, it returns the process to step S1 and repeats the process from step S1 to step S19.
[0091] In the response acquisition determination step, the response acquisition unit 116 determines whether a response has been acquired when it receives data indicating a response to the audio data from the first terminal T1, by following a procedure or the like. This data may include, for example, data indicating a positive evaluation, data indicating a comment, data indicating a tag, etc.
[0092] [Step S18: Obtain the reaction] The control unit 11 executes a process to acquire a response to the audio data using the response acquisition unit 116 (response acquisition step). The control unit 11 then moves the process to step S19.
[0093] [Step S19: Start managing the reaction] The control unit 11 works in cooperation with the storage unit 13 and the communication unit 14 to execute the reaction management unit 117. The control unit 11 then executes the process to start managing the above-mentioned reaction using the reaction management unit 117 (reaction management start step). The control unit 11 returns to step S1 and repeats the process from step S1 to step S19.
[0094] In the reaction management initiation step, the reaction management unit 117 starts the above-mentioned management by, for example, associating audio data with reactions and storing the associated data in the reaction database 132.
[0095] [Effects of management processing] In the management process described above, the voice data management unit 113 manages voice data in association with location (steps S1 to S4), the positioning information acquisition unit 112 acquires first positioning information from the first terminal T1 indicating the location of the first terminal T1 carried by the pedestrian (steps S5 to S6), the suggestion display command unit 114 commands the first terminal T1 to display a suggestion to play voice data associated with the location if the location indicated by the first positioning information corresponds to the location (steps S7 to S9), the response acquisition unit 116 acquires data indicating the pedestrian's response to the voice data from the first terminal T1 (steps S17 to S18), and the response management unit 117 manages data indicating the response in association with the voice data (step S19), and a series of processes are executed.
[0096] As a result, the management device 1 can determine whether the location of the first terminal T1 carried by the pedestrian corresponds to the above-mentioned location based on the voice data associated with the location managed by the voice data management unit 113 and the first positioning information acquired by the positioning information acquisition unit 112 (step S7).
[0097] Then, the management device 1, using the suggestion display command unit 114, instructs the first terminal T1 to display a suggestion to play the audio data associated with a location if the location of the first terminal T1 matches the location associated with the audio data (step S9). In this way, the management device 1 enables pedestrians carrying the first terminal T1 to discover the audio data associated with a location by walking around themselves.
[0098] Furthermore, pedestrians can experience the joy of movement that cannot be obtained through ordinary strolls, by discovering audio data by walking around themselves. When a pedestrian commands the playback of the audio data they have discovered, the management device 1 provides the audio data to the first terminal T1 (steps S15 to S16). This allows the pedestrian to enjoy the audio data they have discovered.
[0099] Furthermore, the management device 1 acquires data from the first terminal T1 indicating the pedestrian's reaction to the audio data using the reaction acquisition unit 116 (steps S17 to S18). Then, the management device in this configuration manages the data indicating the reaction in association with the audio data using the reaction management unit 117 (step S19). Through the above series of processes, the management device 1 enables a user who plays the audio data they discovered by walking around to communicate with the poster of the audio data and other users.
[0100] Based on the above, the management process described above provides an audio SNS, which could be called a wandering-type SNS, in which pedestrians discover location-associated audio data by walking around themselves and communicate with the posters of such audio data and other users.
[0101] Furthermore, the management process can be implemented in a manner that enables communication via tags and comments, such as by providing a screen that allows for filtering by tags (steps S7 and S9) and displaying comments (step S9).
[0102] In order to encourage interaction using voice data corresponding to the route a pedestrian takes, in addition to the places they visit, the management process may propose voice data associated with the route in addition to voice data associated with the location (steps S10 to S14).
[0103] In order to provide an audio SNS consisting of viewing and posting, the management process may take the form of acquiring audio data associated with positioning information from a second terminal T2 carried by a pedestrian for the audio data provided on the audio SNS (steps S1 to S4).
[0104] Based on the above, the management process contributes to the realization of a voice-based social networking service (SNS), which could also be called a browsing-type SNS.
[0105] [Flowchart for processing voice SNS usage] Figure 8 is a flowchart illustrating an example of a preferred flow of voice SNS usage processing performed on terminal T in this embodiment. Figure 9 is a continuation of the previous figure. Figure 10 is a continuation of the previous figure. The following is an example of a preferred flow of voice SNS usage processing performed on terminal T in this embodiment, using Figures 8 to 10.
[0106] Note that the following explanation will omit any mention of the hardware components of terminal T (CPU, RAM, communication components, input unit, display unit, GPS device, microphone, speaker, etc.).
[0107] [Step S101: Determine whether to provide first positioning information, etc.] Terminal T performs a process to determine whether to provide the first positioning information, etc. (positioning information transmission determination step). If Terminal T determines to provide the information, it moves the process to step S102; otherwise, it moves the process to step S105. Terminal T achieves the above determination by, for example, determining to provide the first positioning information, etc. when a certain amount of time has elapsed since the previous provision.
[0108] [Step S102: Determine current location] Terminal T performs a process to determine its current location (positioning step). Terminal T then moves the process to step S103.
[0109] [Step S103: Provide first positioning information] Terminal T performs the process of providing the first positioning information to the management device 1 based on the current location described above (positioning information transmission step). Terminal T then moves the process to step S104.
[0110] [Step S104: Provide a route] Terminal T performs a process to provide the management device 1 with its travel path based on the time-series data of its current location described above (travel path transmission step). Terminal T then moves the process to step S105. This step is not mandatory, but it contributes to the playback of audio data corresponding to Terminal T's travel path.
[0111] [Step S105: Determine if the display of the playback suggestion screen has been instructed] Terminal T performs a process to determine whether it has been instructed to display the playback suggestion screen D (data reception step). If Terminal T determines that it has been instructed, it moves the process to step S106; otherwise, it moves the process to step S107. In this step, Terminal T receives data from the management device 1 that instructs it to display a screen suggesting the playback of audio data associated with the location of Terminal T, and determines that it has been instructed to display the above screen.
[0112] [Step S106: Display the playback suggestion screen] Terminal T executes the process of displaying the playback suggestion screen D based on the data described above (suggestion display step). Terminal T then moves the process to step S107. If the pedestrian has specified display conditions such as tags, range of likes, range of comments, etc., Terminal T displays only the suggestions that meet the conditions from among the suggestions included in the playback suggestion screen D.
[0113] [Step S107: Determine if playback of audio data has been instructed] Terminal T performs a process to determine whether it has been commanded by a pedestrian using Terminal T to play audio data (playback command determination step). If Terminal T determines that it has been commanded, it moves to step S108; otherwise, it moves to step S110. The form of the command is not particularly limited. The command may be a command transmitted via input to a touch panel or a voice command.
[0114] [Step S108: Acquire audio data] Terminal T performs the process of acquiring voice data related to the above-mentioned command from the management device 1 (voice data acquisition step). Terminal T then moves the process to step S109.
[0115] [Step S109: Play audio data] Terminal T performs the process of playing the aforementioned audio data (audio data acquisition step). That is, when terminal T is commanded to play, it plays the audio data related to the command. Terminal T then moves the process to step S110.
[0116] [Step S110: Determine if a tag has been entered] Terminal T performs a process to determine whether a tag has been entered for the audio data by the pedestrian using Terminal T (tag input acquisition step). If Terminal T determines that a tag has been entered, it moves the process to step S111; otherwise, it moves the process to step S112. In the tag input acquisition step, which is one aspect of the input acquisition step, Terminal T acquires input indicating the pedestrian's response to the audio data (input related to the tag) when such input is made and determines that a tag has been entered.
[0117] [Step S111: Send the tag] Terminal T performs the process of sending the aforementioned tag to the management device 1 (tag response transmission step). Terminal T then moves the process to step S112. In the tag response transmission step, which is one aspect of the response transmission step, Terminal T sends data indicating the response of the tag type to the management device 1, in association with the audio data that is the target of the response.
[0118] [Step S112: Determine if a comment has been entered] Terminal T performs a process to determine whether a comment has been entered into the audio data by the pedestrian using Terminal T (comment input acquisition step). If Terminal T determines that a comment has been entered, it moves the process to step S113; otherwise, it moves the process to step S114. In the comment input acquisition step, which is one aspect of the input acquisition step, Terminal T acquires input indicating the pedestrian's reaction to the audio data (input related to a comment) when such input is made and determines that a comment has been entered.
[0119] [Step S113: Submit a comment] Terminal T performs the process of sending the above-mentioned comment to the management device 1 (comment response transmission step). Terminal T then moves the process to step S114. In the comment response transmission step, which is one aspect of the response transmission step, Terminal T sends data indicating the type of response of the comment, associated with the audio data that is the subject of the response, to the management device 1.
[0120] [Step S114: Determine if recording was ordered] Terminal T performs a process to determine whether it has been instructed by the pedestrian using Terminal T to record audio data (recording determination step). If Terminal T determines that it has been instructed to record, it moves the process to step S115; otherwise, it returns the process to step S101 and repeats the process from step S101 to step S117. For example, Terminal T determines that it has been instructed to record when the upload button U included in the playback suggestion screen D is operated.
[0121] [Step S115: Record audio data] Terminal T performs the process of recording audio data using a microphone or other device installed in Terminal T (recording step). The recorded audio data is also called the audio recording data. Terminal T then moves the process to step S116.
[0122] [Step S116: Determine current location] Terminal T performs a process to determine its current location using a GPS device or the like installed in Terminal T (recording positioning step). Terminal T then moves the process to step S117.
[0123] [Step S117: Perform registration] Terminal T transmits data linking the aforementioned recording data with data indicating its current location to the management device 1 and performs the process of registering the recording data with the voice SNS (registration execution step). Terminal T returns to step S101 and repeats the processes from step S101 to step S117.
[0124] [Effects of processing voice SNS usage] Terminal T performs the above-described voice SNS usage processing using a program stored in a non-temporary memory area, thereby providing the management device 1 with data indicating its current location or travel route (steps S101 to S104), and displays a suggestion to play voice data corresponding to its current location or travel route (steps S105 to S106). Then, in response to a playback command from the pedestrian, Terminal T plays the voice data corresponding to one of the suggestions (steps S107 to S109).
[0125] This enables terminal T to discover location-associated audio data as the pedestrian carrying it walks around. In other words, by moving while executing the program of this embodiment, terminal T can automatically detect and play audio data scattered in association with locations. This allows pedestrians to experience the joy of movement that cannot be obtained through ordinary walks, such as discovering audio data by walking around. Furthermore, even when pedestrians are engaged in other activities such as hiking or glamping, they can connect with other users or obtain information through audio data.
[0126] Furthermore, when a response (tag, comment, etc.) to the audio data is entered by terminal T, it transmits the response to the management device 1 (steps S110 to S113). This enables terminal T to enable users who discover audio data by walking around to communicate with the poster of the audio data and other users.
[0127] In the voice SNS usage process, terminal T executes a series of processes in which a pedestrian posts voice data to the voice SNS (steps S114 to S117). This enables terminal T to achieve two-way communication on the voice SNS.
[0128] Based on the above, the aforementioned voice SNS usage processing contributes to realizing a voice SNS that could be called a wandering-type SNS, in which pedestrians discover location-associated voice data by walking around themselves and communicate with the posters of that voice data and other users.
[0129] <Usage example> The following is an example of using System S of this embodiment.
[0130] [Registering audio data] The first pedestrian runs the program of this embodiment on a terminal T such as a smartphone. The first pedestrian then goes for a walk with the terminal T in hand. The first pedestrian wants to share with others the wild birds they saw in a park they stopped at during their walk, and registers audio data about the wild birds in the management device 1. The management device 1 begins managing data that associates the audio data with locations around the park.
[0131] [Discovery of audio data] The second pedestrian runs the program of this embodiment on a terminal T such as a smartphone. The second pedestrian then goes for a walk with the terminal T in their possession. When the second pedestrian stops by the aforementioned park, they notice that a playback suggestion screen D is displayed on the terminal T, which suggests playing multiple audio data, including the aforementioned audio data.
[0132] Figure 11 shows an example of a playback suggestion screen D displayed on terminal T. This example displays six articles A and an upload button U. Each article A includes a display element M indicating the poster, an appendix N indicating various attributes such as posting time, playback time, and tags, a button F to send a positive rating as a reaction, a button P to command playback, and reactions R such as comments and the number of positive ratings.
[0133] The six articles A in this example consist of, from top to bottom: the first article, posted by the first poster "Aji" using a fish icon, containing 23 minutes of audio data categorized under the tag "Gourmet"; the second article, posted by the second poster "City Tourism Division" using an icon with the character "City" containing 10 minutes of audio data categorized under the tag "Tourism"; the third article, posted by the first poster, containing 5 minutes of audio data categorized under the tag "Monologue"; the fourth article, posted by the third poster "Cabinet Office" using an icon with the character "Inside" containing 30 minutes of audio data categorized under the tag "Emergency Information"; the fifth article, posted by the fourth poster "Prefectural Tourism Division" using an icon with the character "Prefecture" containing 3 minutes of audio data categorized under the tag "Directions"; and the sixth article, posted by the fifth poster "Ajisuki" using an icon with the character "Ajisuki" containing 30 seconds of audio data categorized under the tag "Bargain Information".
[0134] Articles 1 through 6 each display tags, comments, and the number of likes. The audio data registered by the first pedestrian (first poster) is displayed as article 3. The second pedestrian plays the audio data by operating button P, which is attached to article 3 and commands playback. The second pedestrian then finds the wild bird that the first pedestrian spoke of and gives a like to this audio data by operating button F, which sends a like. Management device 1 changes the number of likes from 2 to 3. The second pedestrian also adds the comment "Thank you. I found it too!" to this audio data. Management device 1 adds this comment to the comment list.
[0135] As a result of the above-described process, the first pedestrian and the second pedestrian can interact through the park as a place, even though they did not visit the park at the same time. Therefore, the system S of this embodiment realizes a voice SNS in which a pedestrian (the second pedestrian in this example) communicates with the poster of the audio data (the first pedestrian in this example) and other users through audio data associated with a place (the park in this example) (audio data introducing wild birds that were in the park in this example).
[0136] In this interaction, the audio data posted by the first pedestrian on the audio SNS is associated with the park location. Therefore, the second pedestrian can discover this audio data as part of the enjoyment of walking, without having to stop and search for it among the vast number of posts on the SNS. Thus, the aforementioned audio SNS can be called a wandering-type SNS where pedestrians discover location-associated audio data by walking around and communicate with the poster of that audio data and other users.
[0137] [Use in advertising] The administrator of System S listens to the advertiser's challenges related to advertising. The advertiser proposes an advertising plan that utilizes the voice SNS implemented using System S of this embodiment. Such plans include, for example, a plan that links locations around the advertiser's store with voice data related to the store and provides it as voice data, a plan that links routes leading to the store with voice data related to the store and provides it as voice data, and a plan that provides incentives to those who register such voice data. The advertiser considers the advertising plan and starts advertising in collaboration with the voice SNS according to this embodiment.
[0138] Within the scope of the concept of this invention, those skilled in the art can conceive of various modifications and alterations. Therefore, such modifications and alterations are understood to fall within the scope of this invention. For example, any addition, deletion, or design change of components, or addition, omission, or modification of processes, made by a person skilled in the art to the above-described embodiments, is also included within the scope of this invention, as long as it retains the essence of this invention. [Explanation of symbols]
[0139] S System 1 Management device 11 Control Unit 111 Recording data acquisition unit 112 Positioning Information Acquisition Unit 113 Voice Data Management Department 114 Proposal Display Command Department 115 Voice data transmission unit 116 Reaction acquisition section 117 Reaction Control Department 118 Voice Data Analysis Department 13 Storage section 131 Audio Databases 132 Reaction Database 133 Voice Analysis Database 14 Communications Department D Playback suggestion screen N Network T terminal T1 Terminal 1 T2 Second Term
Claims
1. The voice data management unit manages voice data in association with location, A positioning information acquisition unit that acquires first positioning information from a first terminal that indicates the location of a first terminal carried by a pedestrian, A suggestion display command unit that, when the location indicated by the first positioning information corresponds to the location, instructs the first terminal to display a suggestion to play audio data associated with that location, A response acquisition unit that acquires data from the first terminal indicating the pedestrian's response to the aforementioned audio data, A reaction management unit manages data indicating the reaction in association with the aforementioned audio data, A management device for a voice-based social networking service (SNS).
2. The response acquisition unit is configured to acquire data indicating a response that includes a tag indicating the category of the audio data, and is further configured to acquire data indicating a response that includes a comment on the audio data. The aforementioned proposal display command unit, If a tag has been specified by the aforementioned pedestrian, the system will be instructed to display a message suggesting the playback of audio data associated with that specified tag. If the comment is associated with the audio data, the system is instructed to display a display that suggests playback in a manner that allows the display of the comment to be instructed. The control device according to claim 1.
3. The aforementioned voice data management unit manages voice data in association with the route, The positioning information acquisition unit acquires data indicating the movement path of the first terminal, The proposed display command unit instructs the first terminal to display a suggestion to play audio data associated with the route when the travel route substantially includes the route. The control device according to claim 1.
4. The system further includes an audio data acquisition unit that acquires audio data recorded by a second terminal carried by a pedestrian, When the recording data is acquired, the positioning information acquisition unit acquires second positioning information from the second terminal indicating the location of the second terminal at the time the recording data was recorded. The aforementioned voice data management unit manages the recorded data in association with the location, including the location indicated by the second positioning information. The control device according to claim 1.
5. A first terminal that communicates with a management device according to any one of claims 1 to 4 and is configured to be carried by a pedestrian, A positioning information transmission step of transmitting positioning information indicating the location of the first terminal to the management device, A data reception step in which data is received from the management device that commands a display to propose the playback of audio data associated with the location mentioned above, A proposal display step that displays a suggestion for playback of the audio data based on the aforementioned data, When the aforementioned playback command is given, a playback step is performed to play the audio data, An input acquisition step to acquire input indicating the pedestrian's response to the aforementioned audio data, A reaction transmission step of transmitting data indicating the reaction to the control device, A program that executes something.