Message Processing Method and Message Processing Device

The message processing method addresses the risk of harming the mood of occupants in other vehicles by correcting message information to remove provocative expressions before presentation, thereby enhancing the interaction environment.

JP7697257B2Active Publication Date: 2025-06-24NISSAN MOTOR CO LTD
View PDF 6 Cites 0 Cited by

Patent Information

Application Number
JP2021074905
Authority / Receiving Office
JP · JP
Patent Type
Patents
Current Assignee / Owner
Filing Date
2021-04-27
Publication Date
2025-06-24
Estimated Expiration
2041-04-27

AI Technical Summary

Technical Problem

Existing message processing systems risk harming the mood of occupants in other vehicles by presenting raw voice, video, and emotion data from one vehicle to another.

Method used

A message processing method that receives message information from a first occupant, determines if it contains predefined provocative expressions, corrects the information to remove such expressions, and outputs the corrected message for presentation to a second occupant.

Benefits of technology

This approach effectively prevents the mood of occupants in other vehicles from being negatively affected by the presented messages, promoting a more positive interaction environment.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 0007697257000001
    Figure 0007697257000001
  • Figure 0007697257000002
    Figure 0007697257000002
  • Figure 0007697257000003
    Figure 0007697257000003
Patent Text Reader

Abstract

To avoid offending occupants of one vehicle by messages, when presenting verbal or nonverbal messages detected from occupants of the other vehicle to the occupants of the one vehicle.SOLUTION: A message processing method receives message information containing a verbal or non-verbal message detected from a first occupant, an occupant of a first vehicle (S10); determines whether the received message information contains a predefined provocative expression (S11 to S16, S20, S31); the received message information is corrected to message information that does not contain provocative expressions, if it is determined that the received message information contains provocative expressions (S21-S25, S32-S33); and the corrected message information is output as presentation information to be presented to a second vehicle occupant, an occupant of a second vehicle (S28, S35).SELECTED DRAWING: Figure 5
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present invention relates to a message processing method and a message processing apparatus.

Background Art

[0002] Patent Document 1 proposes a technique of receiving information indicating the voice, video, and emotion of the driver of a vehicle that is the transmission source of inter-vehicle communication at a vehicle that is the reception destination and displaying it on a display device.

Prior Art Document

Patent Document

[0003]

Patent Document 1

Summary of the Invention

Problems to be Solved by the Invention

[0004] However, if information indicating the voice, video, and emotion of the driver of the vehicle that is the transmission source is presented to the driver of the other party as it is, there is a risk of harming the mood of the driver of the other party due to this information. An object of the present invention is to avoid harming the mood of the occupants of another vehicle by these messages when presenting a language message or a non-language message detected from an occupant of a certain vehicle to the occupants of another vehicle.

Means for Solving the Problems

[0005] According to one aspect of the present invention, a message processing method causes a computer to execute: a process of receiving message information including a language message or a non-verbal message detected from a first occupant who is an occupant of a first vehicle; a process of determining whether there is a predefined provocative expression in the received message information; a process of correcting the received message information to message information not including the provocative expression when it is determined that there is a provocative expression in the received message information; and a process of outputting the corrected message information as message information to be presented to a second occupant who is an occupant of a second vehicle.

Advantages of the Invention

[0006] According to the present invention, when presenting a language message or a non-verbal message detected from an occupant of a certain vehicle to an occupant of another vehicle, it is possible to avoid damaging the mood of the occupant of the other vehicle with these messages.

Brief Description of the Drawings

[0007]

Figure 1

Figure 2

Figure 3

Figure 4

Figure 5

Figure 6

Embodiments for Carrying Out the Invention

[0008] Hereinafter, embodiments of the present invention will be described with reference to the drawings. Note that each drawing is schematic and may differ from the actual one. Further, the embodiments of the present invention shown below exemplify apparatuses and methods for embodying the technical idea of the present invention, and the technical idea of the present invention does not specify the structure, arrangement, etc. of the components as follows. The technical idea of the present invention can be variously modified within the technical scope defined by the claims described in the claims.

[0009] (Configuration) Referring to FIG. 1, the vehicle - to - vehicle communication system 100 of the embodiment includes an in - vehicle device 10 mounted on vehicle 1, an in - vehicle device 20 mounted on vehicle 2, and a server device 3. The in - vehicle device 10 includes a positioning device 11, an in - vehicle camera 12, a vehicle sensor 13, a user interface 14, a communication device 15, and a controller 16. In the drawings, the interface is denoted as "I / F". The configuration of the in - vehicle device 20 of vehicle 2 may be the same as that of the in - vehicle device 10. The positioning device 11 measures the current position of vehicle 1. The positioning device 11 may include, for example, a global navigation satellite system (GNSS) receiver. The GNSS receiver is, for example, a global positioning system (GPS) receiver or the like, which receives radio waves from a plurality of navigation satellites and measures the current position of vehicle 1. The measurement result of the current position of vehicle 1 is output to the controller 16 as vehicle state information. The in - vehicle camera 12 includes an out - vehicle camera that photographs the surrounding environment of vehicle 1 and an in - vehicle camera that photographs the occupants (e.g., the driver) of vehicle 1. The in - vehicle camera 12 outputs the captured image generated by the in - vehicle camera to the controller 16. Also, the captured image of the out - vehicle camera is output to the controller 16 as vehicle state information.

[0010] The vehicle sensor 13 detects various information obtained from the vehicle 1. The vehicle sensor 13 includes, for example, a vehicle speed sensor that detects the vehicle speed of the vehicle 1, a wheel speed sensor that detects the rotational speed of each tire provided on the vehicle 1, a three-axis acceleration sensor (including deceleration) that detects the acceleration of the vehicle 1 in three axial directions (G sensor), a steering angle sensor that detects the steering angle (including the turning angle), a gyro sensor that detects the angular velocity generated in the vehicle 1, and a yaw rate sensor that detects the yaw rate. The vehicle sensor 13 outputs, as vehicle state information, information on the vehicle speed, wheel speed, acceleration, steering angle, and angular velocity of the vehicle 1, as well as information on the operating states of the vehicle 1's siren (horn) and headlamps, to the controller 16.

[0011] The user interface 14 is a human-machine interface device that exchanges information between the controller 16 and the vehicle occupants. The user interface 14 includes a display device visible to the vehicle 1's occupants (e.g., the display screen of a navigation system), a speaker and a buzzer for outputting alarm sounds, notification sounds, and voice information. The user interface 14 also includes an operator for receiving operation inputs from the vehicle occupants to the controller 16. The operator may be a mechanical interface device such as a button, switch, lever, dial, keyboard, etc., or may be a button, switch, lever, dial, keyboard, etc. displayed on a touch panel. The user interface 14 also includes a voice input device (e.g., a microphone) for acquiring the voice (sound, words) uttered by the vehicle occupants.

[0012] The communication device 15 provides a communication function between the in-vehicle device 10 and an external device. The communication method by the communication device 15 may be, for example, wireless communication via a public mobile phone network, vehicle-to-vehicle communication, road-to-vehicle communication, or satellite communication. In this embodiment, the in-vehicle device 10 communicates with the server device 3 via the communication network NW by the communication function of the communication device 15. The controller 16 acquires a voice signal uttered by the vehicle 1's occupants from the voice input device of the user interface 14. The controller 16 also acquires a captured image of the vehicle 1's occupants captured by the in-vehicle camera of the in-vehicle camera 12. The controller 16 transmits information including the voice signal and captured image of the occupant of the vehicle 1 to the server device 3 via the communication device 15. Further, the controller 16 transmits the vehicle state information of the vehicle 1 to the server device 3 via the communication device 15.

[0013] Here, the voice signal of the occupant of the vehicle 1 includes the words spoken by the occupant as a language message. Further, the voice signal includes the tone (e.g., mood, intonation) of the voice uttered by the occupant as a non-verbal message. On the other hand, the captured image of the occupant of the vehicle 1 includes the expression and gesture of the occupant as a non-verbal message. Therefore, the information including the voice signal and captured image of the occupant of the vehicle 1 is referred to as "message information". The controller 16 is a computer including a processor 17 and peripheral components such as a storage device 18. The processor 17 may be, for example, a CPU (Central Processing Unit) or an MPU (Micro-Processing Unit). The storage device 18 may include a semiconductor storage device, a magnetic storage device, an optical storage device, or the like. The functions of the controller 16 described in this specification are realized by the processor 17 executing a computer program stored in the storage device 18.

[0014] Similarly, the in-vehicle device 20 of the vehicle 2 transmits message information including the voice signal and captured image of the occupant of the vehicle 2 and the vehicle state information of the vehicle 2 to the server device 3. When the server device 3 receives the message information and the vehicle state information from the in-vehicle device 10, it converts the received message information into information (hereinafter "presentation information") to be presented to the occupants (e.g., drivers) of other vehicles (e.g., vehicle 2) around the vehicle 1, and transmits it to other vehicles around the vehicle 1. The presentation information may include a voice signal generated based on the words included in the voice signal of the occupant of the vehicle 1. Further, the presentation information may include a control signal for the avatar of the occupant of the vehicle 1. The server device 3 generates a control signal for the avatar of the occupant of the vehicle 1 based on the tone of the voice signal of the occupant of the vehicle 1, the expression and gesture represented by the captured image of the occupant of the vehicle 1.

[0015] Similarly, the in-vehicle device 20 also receives message information and vehicle state information, converts them into presentation information to be presented to the passengers of other vehicles (e.g., vehicle 1) around vehicle 2, and transmits the presentation information to other vehicles around vehicle 2. The server device 3 is a computer including a processor 30 and peripheral components such as a storage device 31 and a communication interface 32. The processor 30 may be, for example, a CPU or an MPU. The storage device 31 may include a semiconductor storage device, a magnetic storage device, an optical storage device, etc. The server device 3 communicates with the in-vehicle devices 10 and 20 via a communication network NW by the communication function of the communication interface 32. The functions of the server device 3 described in this specification are realized by the processor 30 executing a computer program stored in the storage device 31. Details of the functions realized by the server device 3 will be described later.

[0016] When the controllers 16 of the in-vehicle devices 10 and 20 of vehicle 1 and vehicle 2 receive the presentation information from the server device 3, they output the received presentation information from the user interface 14 and present it to the passengers. For example, the voice information included in the presentation information may be output from the speaker of the user interface 14. Also, for example, an avatar may be displayed on the display device of the user interface 14 based on the control signal of the avatar included in the presentation information. For example, when the message information transmitted from vehicle 1 is converted to generate presentation information and transmitted to vehicle 2, voice information may be output from the speaker of the user interface 14 of vehicle 2, and the avatar of the passenger of vehicle 1 may be displayed on the display device. When the message information transmitted from vehicle 2 is converted to generate presentation information and transmitted to vehicle 1, voice information may be output from the speaker of the user interface 14 of vehicle 1, and the avatar of the passenger of vehicle 2 may be displayed on the display device. Thereby, a verbal message or a non-verbal message detected from the passenger of a certain vehicle (e.g., vehicle 1) can be presented to the passengers of other vehicles (e.g., vehicle 2).

[0017] Thus, when presenting a message detected from a passenger of a certain vehicle to passengers of other vehicles, there is a risk that the presented message may harm the mood of the passengers of the other vehicles. For example, when a passenger of Vehicle 1 performs an inappropriate driving operation that violates driving manners maliciously (e.g., forceful cut-in, sudden braking, sudden acceleration, sudden steering), it is conceivable that the passenger shows a provocative attitude towards other surrounding vehicles. In such a case, if a message (e.g., words, tone of voice, expression, gesture) detected from the passenger of Vehicle 1 is presented to the passengers of other vehicles as it is, there is a risk of harming the mood of the passengers of the other vehicles. On the other hand, when the passenger of Vehicle 1 has no malice and is reflecting on the inappropriate driving operation, it is conceivable that the passenger shows a feeling of apology towards other surrounding vehicles. In this case, it is preferable to generate presentation information corresponding to the message detected from the passenger of Vehicle 1 and present it to the passengers of other vehicles.

[0018] Therefore, the server device 3 executes a process of determining whether there is a predefined provocative expression in the message information received from a certain vehicle, a process of correcting the received message information to message information that does not include the provocative expression when it is determined that there is a provocative expression in the received message information, and a process of outputting the corrected message information as presentation information to be presented to the passengers of other vehicles. The predefined provocative expressions are stored in a storage device accessible from the server device 3. Thereby, when presenting a message detected from a passenger of a certain vehicle to passengers of other vehicles, it is possible to avoid harming the mood of the passengers of the other vehicles due to these messages.

[0019] Hereinafter, the functional configuration of the server device 3 will be described. FIG. 2 is a block diagram of an example of the functional configuration of the server device 3. The server device 3 functions as an information reception unit 40, a driving operation determination unit 41, a provocation determination unit 42, and an information correction unit 43. The information reception unit 40 receives message information and vehicle state information transmitted from Vehicle 1 and Vehicle 2 that are in motion. The driving operation determination unit 41 selects, as a vehicle to be determined for provocation, a vehicle that has performed an inappropriate driving operation that violates driving manners (for example, forced cut-in, sudden braking, sudden acceleration, sudden steering). The vehicle to be determined for provocation is an example of the "first vehicle" described in the claims.

[0020] FIG. 3 is a flowchart of an example of the processing by the driving operation determination unit 41. When the information reception unit 40 receives vehicle state information transmitted from a traveling vehicle, in step S1, the driving operation determination unit 41 acquires the vehicle state information from the information reception unit 40. In step S2, the driving operation determination unit 41 determines, based on the acquired vehicle state information, whether any vehicle has performed an inappropriate driving operation that violates driving manners (for example, aggressive driving, sudden braking, sudden acceleration, sudden steering wheel operation, abnormal operation of a siren, or abnormal passing operation). If an inappropriate driving operation has been performed (step S2: Y), the process proceeds to step S3. If an inappropriate driving operation has not been performed (step S2: N), the processing of the driving operation determination unit 41 ends.

[0021] For example, the driving operation determination unit 41 may detect signals such as sudden braking, sudden acceleration, and sudden steering from the vehicle state information, or detect sudden approach or cut-in between vehicles from the captured image of an external camera, and determine whether an inappropriate driving operation (an operation that violates driving manners), such as aggressive driving, has been performed. Also, for example, the driving operation determination unit 41 may detect the operating state of the siren from the vehicle state information, detect an abnormality based on the operating frequency, and detect an abnormal operation of the siren as an inappropriate driving operation.

[0022] Also, for example, the driving operation determination unit 41 may detect the operating state of the headlight from the vehicle state information, detect an abnormality when the flashing frequency is high and the duration is long, and detect an abnormal passing operation as an inappropriate driving operation. In step S3, the driving operation determination unit 41 selects, as a vehicle to be determined for provocation, the vehicle that has performed the inappropriate driving operation. Thereafter, the processing of the driving operation determination unit 41 ends. Hereinafter, an example in the case where the vehicle to be determined for provocation is Vehicle 1 will be described in this specification.

[0023] Refer to FIG. 2. The provocation determination unit 42 determines whether the message information detected from the occupant of Vehicle 1, which is the vehicle to be determined for provocation, includes a provocative expression. For example, when the message information includes a captured image of the face of the occupant of Vehicle 1, the provocation determination unit 42 analyzes the face image to estimate the expression of the occupant. For example, the positions and shapes of parts such as the occupant's eyes, eyebrows, nose, cheeks, and mouth may be recognized, and expressions such as joy, surprise, fear, sadness, anger, and disgust may be estimated based on these recognition results. When the expression of the occupant of Vehicle 1 is a predefined and stored provocative expression (for example, an expression of anger or disgust), the provocation determination unit 42 determines that a provocative expression is included.

[0024] Also, for example, when the message information includes a captured image representing the movement of the occupant of Vehicle 1, the provocation determination unit 42 may estimate the intention of the occupant's gesture. When the gesture of the occupant of Vehicle 1 is a predefined and stored provocative gesture (for example, when it indicates irritation such as hitting the steering wheel), the provocation determination unit 42 determines that a provocative expression is included. Also, for example, when the message information includes an audio signal uttered by the occupant of Vehicle 1, the provocation determination unit 42 may estimate whether the tone of the occupant is a provocative tone. For example, when the amplitude of the waveform of the audio signal uttered by the occupant is greater than a predefined and stored determination threshold (for example, the normal amplitude of the occupant of Vehicle 1), it may be determined that the tone of the occupant is a provocative tone. It is preferable that the determination threshold is individually calibrated according to the occupant of Vehicle 1 and the audio input device. When the tone of the occupant is a provocative tone, the provocation determination unit 42 determines that a provocative expression is included.

[0025] For example, when the message information includes a voice signal uttered by a passenger in Vehicle 1, the provocation determination unit 42 may convert the voice signal into text information by performing voice recognition. When the converted text information includes provocative words defined and stored in advance, such as swear words, imperative expressions (e.g., "Get out of the way"), or critical expressions (e.g., "You're slow"), the provocation determination unit 42 determines that a provocative expression is included. FIG. 4 is a flowchart of an example of the processing by the provocation determination unit 42. When the information receiving unit 40 receives the message information transmitted from the vehicle to be determined for provocation (Vehicle 1 in this example), in step S10, the provocation determination unit 42 acquires the message information from the information receiving unit 40.

[0026] In step S11, the provocation determination unit 42 determines whether the expression of the passenger in Vehicle 1 included in the message information is provocative. If it is a provocative expression (step S11: Y), the process proceeds to step S15. If it is not a provocative expression (step S11: N), the process proceeds to step S12. In step S12, the provocation determination unit 42 determines whether the gesture of the passenger in Vehicle 1 included in the message information is provocative. If it is a provocative gesture (step S12: Y), the process proceeds to step S15. If it is not a provocative gesture (step S12: N), the process proceeds to step S13.

[0027] In step S13, the provocation determination unit 42 determines whether the tone of voice of the passenger in Vehicle 1 included in the message information is provocative. If it is a provocative tone of voice (step S13: Y), the process proceeds to step S15. If it is not a provocative tone of voice (step S13: N), the process proceeds to step S14. In step S14, the provocation determination unit 42 determines whether the words spoken by the passenger in Vehicle 1 included in the message information are provocative. If it is a provocative word (step S14: Y), the process proceeds to step S15. If it is not a provocative word (step S14: N), the process proceeds to step S16.

[0028] In step S15, the provocation determination unit 42 determines that the message information transmitted from the vehicle 1 includes a provocative expression. After that, the process of the provocation determination unit 42 ends. In step S16, the provocation determination unit 42 determines that the message information transmitted from the vehicle 1 does not include a provocative expression. After that, the process of the provocation determination unit 42 ends.

[0029] Referring to FIG. 2, when it is determined that the message information received from the provocation determination target vehicle (vehicle 1 in this example) includes a provocative expression, the information correction unit 43 corrects the received message information into message information that does not include a provocative expression, and generates presentation information to be presented to the passengers of other vehicles around the provocation determination target vehicle. The information correction unit 43 transmits the presentation information generated from the corrected message information to the in-vehicle device (for example, the in-vehicle device 20 of vehicle 2) of other vehicles around the provocation determination target vehicle. The vehicle to which the presentation information is output is an example of the "second vehicle" described in the claims. Here, an example of presenting the presentation information to the passengers of vehicle 2 will be described.

[0030] As described above, the presentation information may include an audio signal generated based on the words included in the audio signal of the passenger of vehicle 1. The presentation information may also include a control signal for the avatar of the passenger of vehicle 1. The in-vehicle device 20 of vehicle 2 may output the audio signal included in the received presentation information from the speaker of the in-vehicle device 20 user interface 14. Further, based on the control signal of the avatar included in the received presentation information, the in-vehicle device 20 user interface 14 display device may display the avatar of the passenger of vehicle 1.

[0031] FIG. 5 is a flowchart of the first example of the process by the information correction unit 43. In this example, presentation information including an audio signal and a control signal for the avatar is generated. In step S20, the information correction unit 43 determines whether the provocation determination unit 42 has determined that the message information from the vehicle 1 includes a provocative expression. If it is determined that the message information includes a provocative expression (S20: Y), the process proceeds to step S21. If the message information does not include a provocative expression (S20: N), the process proceeds to step S26. In step S21, the information correction unit 43 converts the voice signal included in the message information from the vehicle 1 into text information by performing voice recognition on the voice signal. If the provocation determination unit 42 has already converted the voice signal into text information, the information correction unit 43 may obtain the text information from the provocation determination unit 42.

[0032] In step S22, the information correction unit 43 corrects the text information including the provocative expression into text information not including the provocative expression. For example, swear words included in the text information may be deleted. Also, for example, when the text information includes an imperative expression or a critical expression, these expressions may be replaced with corresponding proposals. For this reason, the information correction unit 43 may define and store in advance text information of various proposals corresponding to imperative expressions or critical expressions. For example, a proposal "Go to the adjacent lane" may be associated and stored in advance corresponding to the imperative expression "Get out of the way". Also, a proposal "Observe the minimum driving speed" may be associated and stored in advance corresponding to the critical expression "Slow". The information correction unit 43 may obtain the driving environment around the vehicle 1 (lane information, regulated speed information, traffic signal information, etc.) from the map database based on the current position information included in the vehicle state information of the vehicle 1, and adopt a proposal that conforms to the driving environment around the vehicle 1. Also, for example, when the text information is in plain language, the information correction unit 43 may convert it into honorific language (for example, polite language).

[0033] In step S23, the information correction unit 43 performs voice synthesis processing on the text information corrected in step S22 to generate a synthesized voice having the voice feature amount of the avatar. In this way, by generating a synthesized voice from the text information, even if the message information from the vehicle 1 includes a voice signal with a provocative tone, it can be corrected to a voice signal that does not include a provocative tone (that is, a provocative expression). In step S24, the information correction unit 43 generates a gentle expression (for example, a smiling expression, a kind expression, a happy expression, or a joyous expression) as the expression of the avatar of the occupant of the vehicle 1. For example, the information correction unit 43 defines and stores in advance the positions and shapes of parts such as the eyes, eyebrows, nose, cheeks, and mouth of the avatar as parameters for determining the expression of the avatar, and sets these parameters as control signals for controlling the expression of the avatar of the occupant of the vehicle 1.

[0034] In step S25, the information correction unit 43 changes the expression parameters (that is, the positions and shapes of parts such as the eyes, eyebrows, nose, cheeks, and mouth) generated in step S24 so that the voice signal generated in step S23 appears as if the avatar is speaking, and synchronizes the expression of the avatar with the voice signal. The information correction unit 43 generates the parameters synchronized with the voice signal as control signals for the expression of the avatar. Thereafter, the process proceeds to step S28.

[0035] On the other hand, when the message information from the vehicle 1 does not include a provocative expression in the determination in step S20 (step S20: N), in step S26, the information correction unit 43 converts the voice signal included in the message information from the vehicle 1 into a synthesized voice having the voice feature amount of the avatar. In step S27, the information correction unit 43 warps the expression of the occupant of the vehicle 1 included in the message information from the vehicle 1 with the face of the avatar. For example, the positions and shapes of parts such as the eyes, eyebrows, nose, cheeks, and mouth are recognized from the face image of the occupant of the vehicle 1, and control signals for the expression of the avatar are generated so that the positions and shapes of parts such as the eyes, eyebrows, nose, cheeks, and mouth of the avatar become the recognized positions and shapes. Thereafter, the process proceeds to step S28. In step S28, the information correction unit 43 outputs the presentation information including the synthesized speech generated in step S23 or S26 and the control signal for the expression of the avatar generated in step S25 or S27 to the in-vehicle device 20 of the vehicle 2. After that, the processing of the information correction unit 43 ends.

[0036] Figure 6 is a flowchart of a second example of the processing by the information correction unit 43. In this example, presentation information that does not include an audio signal and an avatar control signal is generated. In step S30, the information correction unit 43 converts the audio signal included in the message information from vehicle 1 into text information by performing speech recognition. The information correction unit 43 may acquire the converted text information from the provocation determination unit 42. In step S31, the information correction unit 43 determines whether or not the provocation determination unit 42 has determined that the message information from vehicle 1 includes a provocative expression. If a provocative expression is included (S31: Y), the process proceeds to step S32. If a provocative expression is not included (S31: N), the process proceeds to step S34.

[0037] In step S32, the information correction unit 43 corrects the text information including the provocative expression into text information that does not include the provocative expression by the same processing as step S22 in FIG. 5. In step S33, the information correction unit 43 performs speech synthesis processing on the text information corrected in step S32 to generate a synthesized speech having a polite (or gentle) voice feature. After that, the process proceeds to step S35.

[0038] On the other hand, when it is determined in step S31 that the message information from vehicle 1 does not include a provocative expression (step S31: N), in step S34, the information correction unit 43 estimates the emotion of the occupant of vehicle 1 (for example, calm, relaxed, apologetic) from the occupant and gestures included in the message information. The information correction unit 43 performs speech synthesis processing on the text information converted in step S30 to generate a synthesized speech having the estimated emotional voice feature. After that, the process proceeds to step S35. In step S35, presentation information including the synthesized speech generated in step S33 or step S34 is output to the in-vehicle device 20 of the vehicle 2. After that, the processing of the information correction unit 43 ends. Note that when the message information from the vehicle 1 includes provocative expressions, instead of steps S32 and S33, background music that calms the mood of the occupant of the vehicle 2 may be acquired, and in step S35, presentation information including the background music may be output to the in-vehicle device 20 of the vehicle 2.

[0039] In the above description, the case where the server device 3 functions as the information reception unit 40, the driving operation determination unit 41, the provocation determination unit 42, and the information correction unit 43 has been described. However, the vehicle-to-vehicle communication system 100 of the present embodiment is not limited to this. For example, the in-vehicle devices 20 mounted on the vehicles 1 and 2 may function as the information reception unit 40, the driving operation determination unit 41, the provocation determination unit 42, and the information correction unit 43.

[0040] (Effect of the embodiment) (1) The computer executes a process of receiving message information including a language message or a non-language message detected from a first occupant who is an occupant of the first vehicle, a process of determining whether there is a predefined provocative expression in the received message information, a process of correcting the received message information to message information that does not include the provocative expression when it is determined that there is a provocative expression in the received message information, and a process of outputting the corrected message information as presentation information to be presented to a second occupant who is an occupant of the second vehicle. Thereby, when presenting the language message or the non-language message detected from the occupant of the first vehicle to the occupant of the second vehicle, it is possible to avoid harming the mood of the occupant of the second vehicle due to these messages.

[0041] (2) The message information may include voice information, image information, or text information. Thereby, since the message information is defined, the voice information, the image information, or the text information can be used for the determination of the provocative expression. (3) As provocative expressions of verbal messages, candidates for words that can cause anger or disgust may be predefined and stored in a storage device accessible by a computer. As provocative expressions in non-verbal messages, facial expressions, tones of voice, or gestures that can cause anger or disgust may be predefined and stored in the storage device. Since provocative expressions are defined in this way, it is possible to determine whether expressions from the occupants of the first vehicle, such as gestures, words, expressions, and tones of voice, are expressions that provoke the occupants of the second vehicle.

[0042] (4) The provocative expression of a verbal message may be an abusive word, an imperative expression, or a critical expression. Thereby, it is possible to determine a provocative expression based on voice information. (5) The computer may correct the message information to a message information that does not include a provocative expression by deleting the abusive words included in the verbal message. Thereby, it is possible to convey information by deleting negative feelings from the occupants of the second vehicle. (6) The computer may correct the message information to a message information that does not include a provocative expression by replacing the imperative expression or critical expression included in the verbal message with a proposal predefined corresponding to the imperative expression or critical expression. Thereby, it is possible to convey information in a more polite and calming manner to the occupants of the second vehicle.

[0043] (7) The computer may output a control signal of an avatar that represents the corrected message information as presentation information to be presented to the second occupant. Thereby, an avatar that conveys the corrected information to the occupants of the second vehicle can be output. (8) When the received message information includes a face image of the first occupant, the computer may estimate whether the expression of the first occupant based on the face image represents disgust or anger, thereby determining whether there is a provocative expression in the expression of the first occupant included in the received message information. Thereby, a provocative expression can be determined based on the face image of the first occupant.

[0044] (9) The computer may execute a process of outputting, as presentation information to be presented to the second occupant, a control signal of an avatar that expresses the facial expression of the first occupant included in the received message information. When the computer estimates that the facial expression of the first occupant represents disgust or anger, the computer may correct the received message information into message information not including provocative expressions by not reflecting the facial expression representing disgust or anger in the facial expression of the avatar. Thereby, it is possible to delete negative feelings from the occupants of the second vehicle and convey information. (10) The computer receives vehicle state information regarding the state of the traveling vehicle, determines whether the vehicle has performed a driving operation that violates driving manners based on the vehicle state information, and may select the vehicle as the first vehicle when the vehicle has performed a driving operation that violates driving manners. Thereby, it is possible to determine which vehicle is performing inappropriate driving and may perform a provocative expression.

Explanation of Signs

[0045] 100… Vehicle - to - vehicle communication system, 1, 2… Vehicles, 3… Server device, NW… Network, 10, 20… In - vehicle devices, 11… Positioning device, 12… In - vehicle camera, 13… Vehicle sensor, 14… User interface, 15… Communication device, 16… Controller, 17… Processor, 18… Storage device, 30… Processor, 31… Storage device, 32… Communication interface, 40… Information receiving unit, 41… Driving operation determination unit, 42… Provocation determination unit, 43… Information correction unit

Claims

1. A process of receiving vehicle state information regarding the state of a moving vehicle, determining whether a driving operation that violates driving manners has been performed based on the vehicle state information, and selecting as a first vehicle when a driving operation that violates driving manners has been performed; a process of receiving message information including a language message or a non-verbal message detected from a first occupant who is an occupant of the first vehicle; a process of determining whether there is a predefined provocative expression in the received message information; when it is determined that the received message information has the provocative expression, a process of correcting the received message information into message information that does not include the provocative expression; a process of outputting the corrected message information as presentation information to be presented to a second occupant who is an occupant of a second vehicle; A message processing method characterized by causing a computer to execute the above.

2. The message processing method according to claim 1, wherein the message information is voice information, image information, or text information.

3. As the provocative expression of the language message, candidates for words that can cause anger or disgust are predefined and stored in a storage device accessible by the computer, and as the provocative expression in the non-verbal message, facial expressions, tones, or gestures of the face that can cause anger or disgust are predefined and stored in the storage device. The message processing method according to claim 1 or 2, characterized by the above.

4. The message processing method according to claim 3, wherein the provocative expression of the language message is an abusive word, an imperative expression, or a critical expression.

5. The message processing method according to claim 4, characterized by correcting the message information not including the provocative expression by deleting the abusive words included in the language message.

6. The message processing method according to claim 4, characterized by correcting the message information not including the provocative expression by replacing the imperative expression or the critical expression included in the language message with a predefined proposal corresponding to the imperative expression or the critical expression.

7. The message processing method according to any one of claims 1 to 6, characterized by outputting a control signal of an avatar representing the corrected message information as the presentation information.

8. When the received message information includes the face image of the first occupant, it is determined whether there is a provocative expression in the expression of the first occupant included in the received message information by estimating whether the expression of the first occupant represented by the face image is disgust or anger. The message processing method according to claim 3, characterized by the above.

9. A process of receiving vehicle state information regarding the state of a traveling vehicle, Determining whether the vehicle has performed a driving operation that violates driving manners based on the vehicle state information, and selecting it as the first vehicle when the vehicle has performed a driving operation that violates driving manners, A process of receiving message information including a language message or a non-verbal message detected from a first occupant who is an occupant of the first vehicle, A process of determining whether there is a predefined provocative expression in the received message information, When it is determined that the received message information has the provocative expression, a process of correcting the received message information to message information that does not include the provocative expression, A process of outputting the corrected message information as presentation information to be presented to a second occupant who is an occupant of a second vehicle, A message processing apparatus characterized by comprising a computer that executes the above.

Citation Information

Patent Citations

  • Image processing method

    JP2002077592A

  • Information processing apparatus, and voice correction program

    JP2010183444A

  • Communication vehicle display device

    JP2010218568A

  • Voice control system, voice controller, voice control method, and voice control program

    JP2013046088A

  • Drive evaluation device and drive evaluation program

    JP2017211703A