Image display method, program, and remote dialogue system
The system addresses stress and discomfort in video conferencing by enabling participants to select alternative images for others, enhancing communication comfort and effectiveness.
Patent Information
- Authority / Receiving Office
- JP · JP
- Patent Type
- Applications
- Current Assignee / Owner
- Filing Date
- 2024-09-30
- Publication Date
- 2026-04-09
AI Technical Summary
Existing video conferencing systems cause stress and hinder smooth communication by displaying participant images that may not be preferred by recipients, leading to varied perceptions and discomfort.
An image display method and system that allows participants to select alternative images, such as avatars or registered images, to represent other participants during calls, reducing stress by offering customizable and personalized image representations.
The system effectively reduces stress and enhances communication by allowing participants to choose images that are more comfortable for them, thereby improving the communication experience.
Smart Images

Figure 2026061235000001_ABST
Abstract
Description
Technical Field
[0001] The present invention relates to an image display method, a program, and a remote interaction system.
Background Art
[0002] Conventionally, a video conferencing system that displays the face images of participants participating in a video conference is known. For example, in Patent Document 1, face images of individual participants are detected from a video conference image in which a plurality of participants participating in a video conference from one base are reflected, and the face images of each of the detected participants are displayed in a list on a terminal device at another base. A video conferencing system is disclosed. In this video conferencing system, when a participant moves out of the video conference image, the face image of the participant in the above list display is replaced with an image such as an illustration or avatar preferred by the participant.
Prior Art Documents
Patent Documents
[0003]
Patent Document 1
Summary of the Invention
Problems to be Solved by the Invention
[0004] In the video conferencing system of Patent Document 1, at another base that receives an image of one base, only the face image of the participant transmitted from one base, the illustration or avatar preferred by the participant, etc. are simply displayed.
[0005] On the other hand, people's perceptions of images vary widely. Even with images of the same person's face or avatar, some people may feel a sense of familiarity, while others may feel intimidated or resistant. Therefore, simply sending images as they are, or sending images that suit the sender's preferences, can cause problems such as participants feeling stressed by the transmitted images and being unable to communicate smoothly with the sender. For this reason, in systems that use images for communication, such as video conferencing systems, there is a need to realize image display that can reduce the stress that images cause to recipients. [Means for solving the problem]
[0006] The image display method of the present disclosure includes receiving an input via a first terminal device used by a first participant to select a first image to be used as an image representing a second participant different from the first participant, from among candidate images including at least one image, and displaying the first image on a first display device used by the first participant when a call is made between the first terminal device and a second terminal device used by the second participant.
[0007] The program of this disclosure causes a computer to perform the following actions: receive input via a first terminal device used by a first participant to select a first image to be used as an image representing a second participant different from the first participant, from among candidate images including at least one image; and display the first image on a first display device used by the first participant when a call is made between the first terminal device and a second terminal device used by the second participant.
[0008] The remote dialogue system of this disclosure includes an input device used by a first participant, which receives input via a first terminal device used by the first participant to select a first image to be used as an image representing a second participant different from the first participant, from among candidate images including at least one image; and a first display device used by the first participant, which displays the first image when a call is made between the first terminal device and a second terminal device used by the second participant. [Brief explanation of the drawing]
[0009] [Figure 1] A diagram showing the configuration of the remote dialogue system according to the embodiment. [Figure 2] An explanatory diagram showing an overview of how the remote dialogue system works. [Figure 3] A block diagram showing the configuration of the first terminal device. [Figure 4] A block diagram showing the configuration of the second terminal device. [Figure 5] A block diagram showing the configuration of the third terminal device. [Figure 6] A block diagram showing the server configuration. [Figure 7] A flowchart illustrating the processing steps for displaying an image. [Figure 8] A diagram showing an example of the configuration of the call screen displayed on the first terminal device. [Figure 9] A diagram showing an example of the mode selection screen displayed on the first terminal device. [Figure 10] A diagram showing an example of the image selection screen displayed on the first terminal device. [Figure 11] This figure shows an example of an image notification screen displayed on the second terminal device. [Figure 12] A diagram showing an example of a call screen displayed on the first terminal device. [Figure 13] A diagram showing an example of a call screen displayed on the second terminal device. [Modes for carrying out the invention]
[0010] [1. Configuration of the Remote Dialogue System] This embodiment will be described below with reference to the drawings. Figure 1 shows the configuration of the remote dialogue system according to this embodiment. The remote conversation system 1 is a system in which a plurality of participants conduct remote conversations using terminal devices. Here, "remote conversation" refers to so-called remote communication, in which a plurality of people who are not physically in the same place use tools such as terminal devices connected communicably to communicate through a call using voice and images.
[0011] The remote conversation system 1 includes a first terminal device 2A used by a first participant U1, a second terminal device 2B used by a second participant U2, a third terminal device 2C used by a third participant U3, and a server 3. The first terminal device 2A, the second terminal device 2B, the third terminal device 2C, and the server 3 are communicably connected to each other by a communication network 4. The first terminal device 2A is installed at a base S1, the second terminal device 2B is installed at a base S2, and the third terminal device 2C is installed at a base S3.
[0012] Hereinafter, when the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C are not distinguished, they are referred to as the terminal device 2. Also, when the first participant U1, the second participant U2, and the third participant U3 are not distinguished, they are referred to as the participant U. Also, when the bases S1, S2, and S3 are not distinguished, they are referred to as the base S.
[0013] With the remote conversation system 1, each participant U uses their respective terminal device 2 to share voice and images with other participants U and participate in a remote conversation. On each terminal device 2, an image indicating the participant U who uses the terminal device 2 as the communication partner is displayed. The image indicating the participant U can be an image such as a logo, icon, avatar, or animation that can identify the participant U in addition to the face image of the participant U.
[0014] The first participant U1 who uses the first terminal device 2A can be a plurality of people. The first participant U1 includes, for example, those with high sensitivity to aspects such as a person's face, expression, or tone of voice, and is a participant who wants to reduce the stress received from the image displayed on the first terminal device 2A.
[0015] As an example, the base S1 where the first terminal device 2A is installed is the living room of a house where there is a medical care child C who requires medical care. Also, the first participant U1 is the medical care child C and the person N who takes care of the medical care child C. For example, the medical care child C is a participant with high sensitivity to human faces and displays, and the person N who takes care of the child wishes to reduce the stress that the medical care child C receives from the images displayed on the first terminal device 2A.
[0016] On the other hand, the second participant U2 who uses the second terminal device 2B and the third participant U3 who uses the third terminal device 2C are, for example, nursing students who learn about medical care. The base S2 where the second terminal device 2B is installed and the base S3 where the third terminal device 2C is installed are, respectively, conference rooms in the nursing school attended by the second participant U2 and the third participant U3 who are nursing students.
[0017] The terminal device 2 is a computer having a communication function. Specifically, the terminal device 2 is a desktop PC (Personal Computer), a tablet PC, a smartphone, or the like. The server 3 may be a single computer, may be composed of multiple computers, or may be a cloud server.
[0018] The communication network 4 may be a LAN (Local Area Network) or a WAN (Wide Area Network). It may also be a global network including a dedicated line, a public switched telephone network, the Internet, or the like.
[0019] The first terminal device 2A is composed of a first processing device 21a and a plurality of devices connected to the first processing device 21a by wire or wirelessly. As devices connected to the first processing device 21a, the first terminal device 2A includes a display device 22a, a keyboard 23a, a mouse 24a, a camera 25a, a microphone 26a, a speaker 27a, and a communication device 28a. At least one of these devices may be integrally incorporated into the housing that houses the first processing device 21a.
[0020] The keyboard 23a and mouse 24a are input devices used by the first participant U1 for input operations.
[0021] The display device 22a is, for example, a liquid crystal display, an organic EL (Electro-Luminescence) display, or a plasma display. The display device 22a is the main display primarily used by caregiver N among the first participant U1, and is used, for example, when caregiver N performs input operations using a keyboard 23a or mouse 24a. The display device 22a may also be equipped with a touch sensor (not shown) as an input device, which is installed on top of its display panel.
[0022] In addition to the display device 22a used as the main display, one or more display devices used as sub-displays may also be connected to the first processing unit 21a of the first terminal device 2A. In this embodiment, the display device 221a, which is a sub-display, is connected to the first processing unit 21a.
[0023] Display device 221a may be used to display an image within the field of view of medically fragile child C. Display device 221a is, for example, a projector that displays an image on the ceiling within the field of view of medically fragile child C who is lying supine in bed. The first processing device 21a may display a copy of the image to be displayed on display device 22a using display device 221a, or it may display a screen on display device 221a that is a copy of only a part of the image to be displayed on display device 22a.
[0024] Camera 25a captures video of base S1, including the first participant U1 using the first terminal device 2A. The images of base S1 captured by camera 25a may include video of the first participant U1, who is a child requiring medical care C, or the caregiver N, as well as a background image of the first participant U1. The video from base S1 is transmitted to the second terminal device 2B and the third terminal device 2C, and can be displayed on the second terminal device 2B and the third terminal device 2C as an image representing the first participant U1. Instead of the video from base S1 captured by camera 25a, the first processing device 21a may transmit a logo or still image predetermined by the first participant U1 to the second terminal device 2B and the third terminal device 2C, in accordance with prior art.
[0025] The second terminal device 2B consists of a second processing unit 21b and a plurality of devices connected to the second processing unit 21b by wire or wireless means. The devices connected to the second processing unit 21b include a display device 22b, a keyboard 23b, a mouse 24b, a camera 25b, a microphone 26b, a speaker 27b, and a communication device 28b. At least one of these devices may be integrated into the housing that houses the second processing unit 21b.
[0026] The display device 22b is, for example, a liquid crystal display, an organic EL display, or a plasma display. Keyboard 23b and mouse 24b are input devices used by the second participant U2 for input operations. Camera 25b captures images of base S2 where the second terminal device 2B is installed. The images of base S2 captured by camera 25b may include images of the second participant U2 and background images.
[0027] The third terminal device 2C consists of a third processing unit 21c and a plurality of devices connected to the third processing unit 21c by wire or wireless means. The devices connected to the third processing unit 21c include a display device 22c, a keyboard 23c, a mouse 24c, a camera 25c, a microphone 26c, a speaker 27c, and a communication device 28c. At least one of these devices may be integrated into the housing that houses the third processing unit 21c.
[0028] The display device 22c is, for example, a liquid crystal display, an organic EL display, or a plasma display. The keyboard 23c and mouse 24c are input devices used by the third participant U3 for input operations. Camera 25c captures images of site S3 where the third terminal device 2C is installed. The images of site S3 captured by camera 25c may include images of the third participant U3 and background images.
[0029] Figure 2 is an explanatory diagram illustrating the overview of the operation of the remote dialogue system 1. In the remote dialogue system 1, participants U view and hear each other's images and sounds using their respective terminal devices 2. The terminal devices 2 send and receive image and sound data, as well as information about inputs and other operations performed on each terminal device 2, via the server 3.
[0030] Specifically, the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C each transmit terminal data D11, D12, and D13 to the server 3. The server 3 then transmits server data D21, D22, and D23 to the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C, respectively.
[0031] In the following, terminal data D11, D12, and D13 will be referred to as terminal data D1 if they are not distinguished. Similarly, server data D21, D22, and D23 will be referred to as server data D2 if they are not distinguished.
[0032] Terminal data D11 includes image data, audio data, and operation data. The image data included in terminal data D11 may be video data captured by camera 25a or data of a predetermined image. The audio data includes audio data collected by microphone 26a. The operation data may include data representing input operations or selection operations performed by keyboard 23a or mouse 24a. In this embodiment, in particular, the image data may include notifications of the first and second images, described later, showing the second participant U2 and the third participant U3, which were selected in the first terminal device 2A.
[0033] Terminal data D12 includes image data, audio data, and operation data from the second terminal device 2B. Image data includes video data captured by camera 25b. Audio data includes audio data collected by microphone 26b. Operation data may include data representing input operations and selection operations performed by keyboard 23b or mouse 24b.
[0034] Terminal data D13 includes image data, audio data, and operation data from the third terminal device 2C. Image data includes video data captured by camera 25c. Audio data includes audio data collected by microphone 26c. Operation data may include data representing input and selection operations performed by keyboard 23c or mouse 24c.
[0035] Server 3 generates display data, including the video or predetermined image, based on the image data contained in the terminal data D1 received from each terminal device 2. Server 3 also generates integrated audio data based on the audio data contained in the terminal data D1 received from each terminal device 2. The integrated audio data includes the audio collected by microphones 26a and 26b, respectively. Server 3 also generates setting data, which integrates the information contained in the operation data, based on the operation data contained in the terminal data D1 received from each terminal device 2. Server 3 generates server data D2, which includes the generated display data, integrated audio data, and setting data.
[0036] Furthermore, in order to prevent feedback at site S, server 3 may use different integrated audio data for each of the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C. For example, server 3 sends server data D21, which does not include the audio picked up by microphone 26a, to the first terminal device 2A. In this case, speaker 27a does not output the audio picked up by microphone 26a. As a result, the first participant U1 does not hear their own voice from speaker 27a, thus avoiding any discomfort for the first participant U1. Server 3 can also perform similar control for server data D22 and D23.
[0037] The first terminal device 2A receives server data D21 and outputs sound from speaker 27a based on the integrated audio data contained in server data D21. The first terminal device 2A also displays images on display devices 22a and 221a based on the display data contained in server data D21. Similarly, the second terminal device 2B outputs sound from speaker 27b and displays on display device 22b based on server data D22. The third terminal device 2C outputs sound from speaker 27c and displays on display device 22c based on server data D23.
[0038] As a result, all participants U using the remote dialogue system 1 can view images and audio based on video or predetermined images and audio transmitted by the terminal devices 2 of other participants U using their own terminal devices 2, and communicate with each other.
[0039] In this embodiment, the server data D22 transmitted to the second terminal device 2B may include notification of a first image, described later, which is used as an image representing the second participant U2 selected in the first terminal device 2A. Furthermore, the server data D23 transmitted to the third terminal device 2C may include notification of a second image, described later, which is used as an image representing the third participant U3 selected in the first terminal device 2A.
[0040] Figure 8 shows an example of the configuration of the call screen displayed on the display devices 22a and 221a of the first terminal device 2A when making a call with the second terminal device 2B and the third terminal device 2C. In the illustrated example, a participant display area IAa1 is displayed on the left side of the middle of the screen to display an image of the second participant U2, who is the other party to the call. Also, a participant display area IAa2 is displayed to the right of the participant display area IAa1 to display an image of the third participant U3, who is the other party to the call.
[0041] Furthermore, a participant display area IAb is displayed in the lower right corner of the screen to show an image of the first participant U1 using the first terminal device 2A. Hereinafter, the image of the second participant U2 displayed in participant display area IAa1 will be referred to as the "second participant display image," and the image of the third participant U3 displayed in participant display area IAa2 will be referred to as the "third participant display image." The image of the first participant U1 displayed in participant display area IAb will also be referred to as the "first participant display image."
[0042] In Figure 8, a microphone button 70 for turning the microphone 26a on and off, and an end button 71 for instructing the first terminal device 2A to end the current call are also displayed at the top of the screen.
[0043] In the remote dialogue system 1 of this embodiment, input is received via the first terminal device 2A used by the first participant U1 to select a first image to be used as an image representing a second participant U2, which is different from the first participant U1, from among candidate images that include at least one image. Furthermore, via the first terminal device 2A used by the first participant U1, input is received to select a second image from among candidate images, each containing at least one image, to be used as an image representing a third participant U3 who is different from the first participant U1 and the second participant U2.
[0044] In the remote dialogue system 1, when a call is made between the first terminal device 2A and the second terminal device 2B used by the second participant U2, the first image selected above is displayed on the first display device used by the first participant U1. Furthermore, when a call is made between the first terminal device 2A and the third terminal device 2C used by the third participant U3, the second image selected above is displayed on the first display device used by the first participant U1. Display devices 22a and 221a of the first terminal device 2A are examples of the first display device.
[0045] Specifically, the first selected image is displayed as the second participant display image in the participant display area IAa1 of the display devices 22a and 221a when a call is made between the first terminal device 2A and the second terminal device 2B. Similarly, the second selected image is displayed as the third participant display image in the participant display area IAa2 of the display devices 22a and 221a when a call is made between the first terminal device 2A and the third terminal device 2C.
[0046] In this embodiment, when a call is made between the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C, the first image and the second image are displayed on the display devices 22a and 221a. Here, the first image and the second image are displayed in the participant display areas IAa1 and IAa2 of the display devices 22a and 221a as display images for the second participant U2 and the third participant U3, respectively, replacing the actual video footage of the second participant U2 and the third participant U3. Furthermore, in this embodiment, the candidate images from which the first and second images are selected include at least one of the avatar image and the image registered by the first participant U1.
[0047] This allows the remote dialogue system 1 to reduce the stress on the first participant U1 caused by the images displayed on the display devices 22a and 221a of the first terminal device 2A.
[0048] [2. Configuration of the first terminal device] Figure 3 is a block diagram showing the functional configuration of the first terminal device 2A that constitutes the remote dialogue system 1. The first terminal device 2A has a first processing unit 21a, to which a display device 22a, keyboard 23a, mouse 24a, camera 25a, microphone 26a, speaker 27a, and communication device 28a are connected. A display device 221a is also connected to the first processing unit 21a. In Figure 3, the display device 221a is depicted as an extension device connected to the first terminal device 2A, but it may also be considered as part of the first terminal device 2A.
[0049] The keyboard 23a and mouse 24a are input devices used by the first participant U1 for input operations. In addition to the keyboard 23a and mouse 24a, the first terminal device 2A may also be equipped with a touch sensor (not shown) as an input device, which is mounted on top of the display panel of the display device 22a. The communication device 28a is a transceiver that communicates with the server 3 via the communication network 4 using wired or wireless communication.
[0050] The first processing unit 21a is a computer comprising a first processor 40 and a first memory 45. The first memory 45 is composed of a magnetic memory device, a semiconductor memory element such as flash ROM (Read Only Memory), or other types of non-volatile memory devices. The first memory 45 may also include RAM (Random Access Memory) which constitutes the work area of the first processor 40.
[0051] The first memory 45 stores the first program 46 and candidate image information 47. The candidate image information 47 will be described later.
[0052] The first processor 40 is composed of, for example, a CPU (Central Processing Unit), an MPU (Micro Processing Unit), or other integrated circuits. The first processor 40 includes, as functional elements or functional units, a communication control unit 41, a selection and acquisition unit 42, a candidate management unit 43, and a video processing unit 44.
[0053] These functional elements of the first processor 40 are realized, for example, by the first processor 40 executing a first program 46 stored in the first memory 45. Here, the first program 46 corresponds to the program in this disclosure. The first program 46 can be stored in any computer-readable storage medium. Alternatively, all or part of the above functional elements of the first processor 40 can be configured by hardware, each including one or more electronic circuit components.
[0054] The communication control unit 41 communicates with the second terminal device 2B and the third terminal device 2C via the server 3, and enables voice and video calls between the first terminal device 2A and the second terminal device 2B and the third terminal device 2C.
[0055] Specifically, the communication control unit 41 converts the sound collected by the microphone 26a into digital audio data and generates audio data based on the digital audio data. The communication control unit 41 receives input from the first participant U1 using the input device, which is the keyboard 23a or mouse 24a, and generates operation data based on the received input. The communication control unit 41 receives video data of the first participant U1 captured by the camera 25a. The communication control unit 41 generates image data including video data of the first participant U1 or still image data such as a photograph showing the first participant U1.
[0056] The communication control unit 41 transmits the terminal data D11, which includes the voice data, operation data, and image data, to the server 3 via the communication device 28a.
[0057] The communication control unit 41 also receives the server data D21 transmitted by the server 3 via the communication device 28a. The communication control unit 41 inputs the display data and audio data contained in the server data D21 to the video processing unit 44.
[0058] The selection and acquisition unit 42 receives input to select a first image from among candidate images, each containing at least one image, to be used as an image representing a second participant U2 who is different from the first participant U1. In this embodiment, the candidate images include at least one of an avatar image and an image registered by the first participant U1.
[0059] A keyboard 23a or mouse 24a is an example of an input device used by the first participant U1 to receive input for selecting the first image.
[0060] When the selection and acquisition unit 42 receives input to select a first image, it notifies the second terminal device 2B of the selected first image.
[0061] The selection and acquisition unit 42 also receives input to select a second image, different from the first image, from among the candidate images, to be used as an image representing a third participant U3, who is a different caller from the first participant U1 and the second participant U2.
[0062] When the selection and acquisition unit 42 receives input to select a second image, it notifies the third terminal device 2C of the selected second image.
[0063] In detail, the selection and acquisition unit 42 accepts the selection of a display mode that defines the participant display image before receiving the selection input for the first and second images. As described above, the participant display image refers to an image that shows participant U, who is the person on the other end of the call, and is displayed on the display devices 22a and 221a.
[0064] The display mode is selected from, for example, the following mode options: <Mode Options> Display mode [0]: This mode uses the camera feed of participant U as the participant display image. Display mode [1]: This mode uses a predetermined still image, such as a photograph or logo, as the participant display image, instead of using the participant U's camera feed. Display mode [2]: This mode uses avatar images as the participant display images. Display mode [3]: This mode uses a predetermined image registered by the first participant U1 to the first terminal device 2A as the participant display image. Hereinafter, the predetermined image will be referred to as the registered image.
[0065] Display mode [0] may be selected, for example, when the first participant U1, child C requiring medical care, is comfortable seeing strangers, or when participant U, the person on the other end of the call, is already an acquaintance and child C requiring medical care does not feel anxious.
[0066] Display mode [1] may be selected when video communication is not possible due to insufficient bandwidth on the communication line, or when the first participant U1, child C requiring medical care, feels uneasy about watching videos that suggest human faces or the presence of people.
[0067] Display modes [2] and [3] are options considered effective in preventing stress on the first participant U1, child C who requires medical care. When display mode [2] or display mode [3] is selected, the selection acquisition unit 42 receives input to select the first or second image described above.
[0068] When display mode [2] is selected, the selection and acquisition unit 42 receives input to select a first image and a second image to be used as images representing a second participant U2 and a third participant U3 that are different from the first participant U1, from among candidate images that include at least one avatar image. When display mode [3] is selected, the selection and acquisition unit 42 receives input to select a first image and a second image to be used as images representing a second participant U2 and a third participant U3 that are different from the first participant U1, from among candidate images that include at least one registered image.
[0069] The avatar image used in display mode [2] may be, for example, a character that is a stylized representation of a person or animal, such as those that appear in comics or animations, or an anthropomorphic character such as a talking box or robot.
[0070] Display mode [3] may be selected, for example, when it is considered that there is no avatar image among the candidate images of the avatar image used in display mode [2] that is suitable for the first participant U1, child C requiring medical care. The registered image to be used is pre-registered in the first terminal device 2A by the caregiver N, who is the first participant U1. The registered image may be a landscape photograph, a photograph of a character or family member at an amusement park that child C requiring medical care has visited, or a copy of an illustration from a book.
[0071] The display mode is selected, for example, by the caregiver N, who is the first participant U1. The selection acquisition unit 42 displays the above-mentioned list of mode options on the display device 22a. The caregiver N can input the selection of one display mode using the keyboard 23a or mouse 24a.
[0072] Figure 9 shows an example of a mode selection screen displayed by the selection acquisition unit 42 on the display device 22a of the first terminal device 2A. In the example shown, a list 72 of mode options, including the four display modes described above, is displayed below the message, "Which image of the other party would you like to see on this side?". The caregiver N, who is the first participant U1, uses the keyboard 23a or mouse 24a to select the desired display mode by moving the cursor 73 to the desired display mode in the list 72 and pressing the OK button 74. The selection acquisition unit 42 receives the input for selecting the display mode.
[0073] Figure 10 shows an example of an image selection screen that the selection acquisition unit 42 displays on the display device 22a of the first terminal device 2A when display mode [2] is selected. In the illustrated example, candidate images 75 containing four candidate avatar images are displayed below the message "[Sato-san]'s avatar." In the above message, "[Sato-san]" is, for example, the name that identifies the second participant U2. The caregiver N, who is the first participant U1, uses, for example, the keyboard 23a or mouse 24a to move the cursor 76 to the desired avatar image among the candidate images 75 and presses the OK button 77. As a result, the caregiver N can select the desired avatar image as the first image representing the second participant U2. The selection acquisition unit 42 receives input to select the first image.
[0074] After receiving input to select the first image, the selection and acquisition unit 42 displays an image selection screen on the display device 22a to receive input to select a second image to be used as the image representing the third participant U3. The image selection screen for receiving input to select the second image is configured similarly to the screen shown in Figure 10, and the name identifying the third participant U3 may be displayed instead of [Mr. Sato] which represents the second participant U2. The caregiver N, who is the first participant U1, can, in the same manner as above, use the keyboard 23a or mouse 24a to input input to select the second image to be used as the image representing the third participant U3. The selection and acquisition unit 42 receives input to select the second image.
[0075] The image selection screen that the selection acquisition unit 42 displays on the display device 22a of the first terminal device 2A when display mode [3] is selected can also be configured in the same way as in Figure 10. When display mode [3] is selected, the image selection screen displays candidate images that include at least one registered image registered by the first participant U1 with the first terminal device 2A, instead of the candidate avatar image 75.
[0076] Candidate avatar images and registered images may be stored in the first memory 45 as candidate image information 47.
[0077] The candidate management unit 43 manages the candidate image information 47 stored in the first memory 45. The candidate image information 47 stores one or more avatar images and one or more registered images. In addition, the candidate image information 47 includes audio data associated with each of the stored avatar images.
[0078] The candidate management unit 43 may, in accordance with instructions from the first participant U1 via the keyboard 23a or mouse 24a, add an avatar image specified by the first participant U1 and the associated audio data to the avatar image to the candidate image information 47. Alternatively, the candidate management unit 43 may, in accordance with instructions from the first participant U1 via the keyboard 23a or mouse 24a, delete an avatar image specified by the first participant U1 and the associated audio data from the candidate image information 47. Furthermore, the candidate management unit 43 may, in accordance with instructions from the first participant U1 via the keyboard 23a or mouse 24a, add a still image specified by the first participant U1 as a registered image to the candidate image information 47, or delete it from the registered images of the candidate image information 47.
[0079] The video processing unit 44 displays images on the display devices 22a and 221a and outputs sound from the speaker 27a based on the display data and sound data of the server data D21 received from the server 3.
[0080] In this embodiment, when a call is made between the first terminal device 2A and the second terminal device 2B, the video processing unit 44 displays the selected first image received by the selection acquisition unit 42 on the first display device. The first display device is a display device used by the first participant U1. Display devices 22a and 221a are examples of the first display device.
[0081] The video processing unit 44 also displays the selected second image received by the selection acquisition unit 42 on the first display device used by the first participant U1 when a call is made between the first terminal device 2A and the third terminal device 2C.
[0082] For example, when a call is made between the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C, the video processing unit 44 displays the selected first image and second image on the first display device. Here, the first image and second image are displayed on the first display device as images representing the second participant U2 and the third participant U3.
[0083] If the selected first image is an avatar image, the video processing unit 44 moves the first image in accordance with prior art, based on the video of the second participant U2 obtained from the server data D21, so that the avatar included in the first image moves in the same way as the second participant U2. Similarly, if the selected second image is an avatar image, the video processing unit 44 moves the second image in accordance with prior art, based on the video of the third participant U3 obtained from the server data D21, so that the avatar included in the second image moves in the same way as the third participant U3.
[0084] Furthermore, if the selected first image is an avatar image, the video processing unit 44 converts the voice of the second participant U2 into voice associated with the avatar image and outputs the converted voice from the speaker 27a used by the first participant U1. Specifically, the video processing unit 44 converts the voice of the second participant U2 based on the voice data associated with the avatar image stored in the candidate image information 47 and outputs the converted voice from the speaker 27a. The voice associated with the avatar image may be the voice of a character if the avatar image is a character from an anime program, for example. The voice associated with the avatar image may be a synthesized voice or the voice of a voice actor that matches the impression and atmosphere of the avatar image.
[0085] Similarly, if the selected second image is an avatar image, the video processing unit 44 converts the voice of the third participant U3 into voice associated with the avatar image, and outputs the converted voice from the speaker 27a used by the first participant U1.
[0086] This also reduces the stress that the audio during the call causes to the first participant U1. Alternatively, the video processing unit 44 may output the audio of the second participant U2 from speaker 27a without converting it to audio associated with the avatar image.
[0087] The video processing unit 44 also transmits the first image of the second participant U2 displayed on the display devices 22a and 221a, and the audio of the second participant U2 output to the speaker 27a, to the server 3 via terminal data D11. The server 3 includes the first image and audio of the second participant U2 in server data D22 and D23, and transmits the server data D22 and D23 to the second terminal device 2B and the third terminal device 2C, respectively.
[0088] Similarly, the video processing unit 44 transmits the second image of the third participant U3 displayed on the display devices 22a and 221a, and the audio of the third participant U3 output to the speaker 27a, to the server 3 via terminal data D11. The server 3 includes the second image and audio of the third participant U3 in server data D22 and D23, and transmits the server data D22 and D23 to the second terminal device 2B and the third terminal device 2C, respectively.
[0089] As a result, the first and second images displayed on the display devices 22a and 221a of the first terminal device 2A, along with their corresponding audio, can also be output to the second terminal device 2B and the third terminal device 2C.
[0090] Figure 12 shows an example of a call screen displayed by the video processing unit 44 on the display devices 22a and 221a of the first terminal device 2A when a call is made between the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C. The call screen shown in Figure 12 is configured similarly to the example of the call screen configuration shown in Figure 8. In the call screen shown in Figure 12, the first image 80, which is the avatar image selected for the second participant U2, is displayed in the participant display area IAa1, which is located on the left side of the middle of the screen and is for displaying an image of the second participant U2. In addition, the second image 81, which is the avatar image selected for the third participant U3, is displayed in the participant display area IAa2, which is located to the right of the participant display area IAa1 and is for displaying an image of the third participant U3, who is the communication partner. Furthermore, in the lower right of the screen, the image 82 representing medically fragile child C, who is one of the first participants U1, is displayed in the participant display area IAb, which is for displaying an image of the first participant U1 using the first terminal device 2A. Image 82 showing child C requiring medical care may be a video or a predetermined still image of child C requiring medical care captured by camera 25a.
[0091] In the example call screen shown in Figure 12, the names identifying the second participant U2 and the third participant U3 are displayed at the bottom of the participant display areas IAa1 and IAa2, respectively, but these names do not have to be displayed. Alternatively, instead of these names, the names of the characters in the avatar images displayed as the first image 80 and the second image 81 in the participant display areas IAa1 and IAa2 may be displayed at the bottom of the participant display areas IAa1 and IAa2.
[0092] [3. Configuration of the second terminal device] Figure 4 is a block diagram showing the functional configuration of the second terminal device 2B. The second terminal device 2B has a second processing unit 21b, to which a display device 22b, a keyboard 23b, a mouse 24b, a camera 25b, a microphone 26b, a speaker 27b, and a communication device 28b are connected.
[0093] The keyboard 23b and mouse 24b are input devices used by the second participant U2 for input operations. In addition to the keyboard 23b and mouse 24b, the second terminal device 2B may also be equipped with a touch sensor (not shown) as an input device, which is mounted on top of the display panel of the display device 22b.
[0094] The communication device 28b is a transceiver that communicates with the server 3 via the communication network 4 using wired or wireless communication.
[0095] The second processing unit 21b is a computer comprising a second processor 50 and a second memory 52. The second memory 52 is composed of a magnetic memory device, a semiconductor memory element such as flash ROM, or other types of non-volatile memory devices. The second memory 52 may also include RAM that constitutes the work area of the second processor 50. The second memory 52 stores the second program 53.
[0096] The second processor 50 is composed of, for example, a CPU, an MPU, or other integrated circuit. The second processor 50 also includes a communication control unit 51 as a functional element or functional unit.
[0097] The functional elements provided by the second processor 50 are realized, for example, by the second processor 50 executing a second program 53 stored in the second memory 52. The second program 53 can be stored in any computer-readable storage medium. Alternatively, the functional elements provided by the second processor 50 can be comprised of hardware including one or more electronic circuit components.
[0098] The communication control unit 51 communicates with the first terminal device 2A and the third terminal device 2C via the server 3 to realize remote dialogue using voice and images between the first terminal device 2A and the third terminal device 2C.
[0099] Specifically, the communication control unit 51 converts the sound collected by the microphone 26b into digital audio data and generates audio data based on the digital audio data. The communication control unit 51 receives input from the input device, which is the keyboard 23b or mouse 24b, and generates operation data based on the received input. The communication control unit 51 generates image data of the video of the second participant U2 acquired by the camera 25b. The communication control unit 51 transmits the terminal data D12, which includes the voice data, operation data, and image data, to the server 3 via the communication device 28b.
[0100] The communication control unit 51 also receives server data D22 transmitted by server 3 via the communication device 28b. The communication control unit 51 controls the display device 22b based on the display data contained in the server data D22 to display a call screen on the display device 22b. The call screen displayed on the display device 22b may include an image representing the first participant U1, as well as a first image representing the second participant U2 and a second image representing the third participant U3, which are the same as those displayed on the first terminal device 2A. The communication control unit 51 also outputs audio from the speaker 27b based on the audio data contained in the server data D22. The audio output from the speaker 27b includes the audio of the first participant U1 and the audio corresponding to the second image of the third participant U3, which is the same as that output on the first terminal device 2A.
[0101] In this embodiment, when the communication control unit 51 receives notification from the first terminal device 2A of a first image selected for a second participant U2 using the second terminal device 2B, it outputs an image notification screen to the display device 22b that notifies the second participant U2 of the first image.
[0102] Figure 11 shows an example of an image notification screen displayed by the communication control unit 51 on the display device 22b of the second terminal device 2B. In the illustrated example, below the message "This image will be displayed on the other party's screen as your image" at the top of the screen, the first image 78, which was selected by the first terminal device 2A as an image representing the second participant U2 using the second terminal device 2B, is displayed. After confirming the first image 78, the second participant U2 can press the OK button 79 displayed at the bottom of this image notification screen to transition the screen of the display device 22b to the call screen shown in Figure 13.
[0103] Figure 13 shows an example of a call screen displayed by the communication control unit 51 on the display device 22b of the second terminal device 2B. In the illustrated example, an image 82 representing the first participant U1 is displayed on the left side of the middle of the screen, and a second image 84 representing the third participant U3 is displayed to its right. The second image 84 is the same as the second image 81 displayed on the display devices 22a and 221a of the first terminal device 2A shown in Figure 12. In the lower right of the call screen shown in Figure 13, a first image 85 representing the second participant U2 is displayed. This first image 85 is the same as the first image 80 displayed on the display devices 22a and 221a of the first terminal device 2A shown in Figure 12. The call screen shown in Figure 13 also displays a microphone button 86 for turning the microphone 26b on / off and an end button 87 for instructing the second terminal device 2B to end the current call at the top of the screen.
[0104] [4. Configuration of the third terminal device] Figure 5 is a block diagram showing the functional configuration of the third terminal device 2C. The third terminal device 2C has a third processing unit 21c, to which a display device 22c, a keyboard 23c, a mouse 24c, a camera 25c, a microphone 26c, a speaker 27c, and a communication device 28c are connected.
[0105] The keyboard 23c and mouse 24c are input devices used by the third participant U3 for input operations. In addition to the keyboard 23c and mouse 24c, the third terminal device 2C may also be equipped with a touch sensor (not shown) as an input device, which is mounted on top of the display panel of the display device 22c.
[0106] The communication device 28c is a transceiver that communicates with the server 3 via the communication network 4 using wired or wireless communication.
[0107] The third processing unit 21c is a computer comprising a third processor 60 and a third memory 62. The third memory 62 is composed of a magnetic memory device, a semiconductor memory element such as flash ROM, or other types of non-volatile memory devices. The third memory 62 may also include RAM that constitutes the work area of the third processor 60. The third memory 62 stores the third program 63.
[0108] The third processor 60 is composed of, for example, a CPU, an MPU, or other integrated circuit. The third processor 60 also includes a communication control unit 61 as a functional element or functional unit.
[0109] The functional elements provided by the third processor 60 are realized, for example, by the third processor 60 executing a third program 63 stored in the third memory 62. The third program 63 can be stored in any computer-readable storage medium. Alternatively, the functional elements provided by the third processor 60 may be comprised of hardware including one or more electronic circuit components.
[0110] The communication control unit 61 communicates with the first terminal device 2A and the second terminal device 2B via the server 3 to realize remote dialogue using voice and images between the first terminal device 2A and the second terminal device 2B.
[0111] Specifically, the communication control unit 61 converts the sound collected by the microphone 26c into digital audio data and generates audio data based on the digital audio data. The communication control unit 61 receives input from the input device, which is the keyboard 23c or mouse 24c, and generates operation data based on the received input. The communication control unit 61 generates image data of the video of the third participant U3 acquired by the camera 25c. The communication control unit 61 transmits the terminal data D13, which includes the voice data, operation data, and image data, to the server 3 via the communication device 28c.
[0112] The communication control unit 61 also receives server data D23 transmitted by server 3 via the communication device 28c. Based on the display data contained in the server data D23, the communication control unit 61 controls the display device 22c to display a call screen on the display device 22c. The call screen displayed on the display device 22c may include an image representing the first participant U1, as well as a first image representing the second participant U2 and a second image representing the third participant U3, which are the same as those displayed on the first terminal device 2A. The communication control unit 61 also outputs audio from the speaker 27c based on the audio data contained in the server data D23. The audio output from the speaker 27c includes the audio of the first participant U1 and the audio corresponding to the first image of the second participant U2, which is the same as that output on the first terminal device 2A.
[0113] In this embodiment, when the communication control unit 61 receives notification from the first terminal device 2A of a selected second image for a third participant U3 using the third terminal device 2C, it outputs an image notification screen to the display device 22c that presents the second image to the third participant U3.
[0114] The image notification screen and call screen displayed by the communication control unit 61 on the display device 22c are configured in the same way as the image notification screen and call screen in the second terminal device 2B illustrated in Figures 11 and 13, respectively.
[0115] [5. Server Configuration] Figure 5 is a block diagram showing the functional configuration of server 3. Server 3 includes a fourth processing unit 30 and a communication device 35. The communication device 35 is a transceiver that communicates with the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C via the communication network 4 using wired or wireless communication.
[0116] The fourth processing unit 30 is a computer comprising a fourth processor 31 and a fourth memory 33. The fourth memory 33 is composed of a magnetic memory device, a semiconductor memory element such as flash ROM, or other types of non-volatile memory devices. The fourth memory 33 may also include RAM which constitutes the work area of the fourth processor 31. The fourth memory 33 stores the fourth program 34.
[0117] The fourth processor 31 is composed of, for example, a CPU, an MPU, or other integrated circuit. The fourth processor 31 also includes a communication control unit 32 as a functional element or functional unit.
[0118] The functional elements provided by the fourth processor 31 are realized, for example, by the fourth processor 31 executing the fourth program 34 stored in the fourth memory 33. The fourth program 34 can be stored in any computer-readable storage medium. Alternatively, the functional elements provided by the fourth processor 31 can be comprised of hardware including one or more electronic circuit components.
[0119] The communication control unit 32 communicates with the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C via the communication device 35, and enables voice and video calls between the first terminal device 2A, the second terminal device 2B, and the third terminal device 2C.
[0120] Specifically, the communication control unit 32 receives terminal data D1 transmitted by terminal device 2 via the communication device 35. Based on the image data contained in the terminal data D1 received from each terminal device 2, the communication control unit 32 generates display data that includes all the video or images contained in each image data.
[0121] Furthermore, Server 3 generates integrated voice data based on the voice data contained in the terminal data D1 received from each terminal device 2. In addition, Communication Control Unit 32 generates configuration data by integrating the information contained in the operation data contained in each operation data, based on the operation data contained in the terminal data D1 received from each terminal device 2.
[0122] The communication control unit 32 generates server data D2, which includes the generated display data, integrated voice data, and setting data. The communication control unit 32 transmits the generated server data D2 to each terminal device 2 via the communication device 35. As described above, the integrated voice data included in each server data D2 transmitted to each terminal device 2 does not need to include the voice acquired at the terminal device 2 to which it is transmitted. In addition, the purpose of the call acquired at the second terminal device 2B only needs to be included in the server data D21 transmitted to the first terminal device 2A, and does not need to be included in the server data D22 transmitted to the second terminal device 2B.
[0123] [6. Operation of the Remote Dialogue System] Next, we will explain the operation of the remote dialogue system 1. Figure 6 is a flowchart showing the processing steps for the image display method executed by the remote dialogue system 1. The image display method shown in Figure 6 is executed collaboratively by the first processing unit 21a of the first terminal device 2A, which is a computer in the remote dialogue system 1, and the fourth processing unit 30 of the server 3.
[0124] The process shown in Figure 6 starts when at least the first terminal device 2A and the server 3 are powered on. When processing begins, the first thing the communication control unit 41 of the first terminal device 2A does is determine whether or not communication with the other terminal device 2 has been established (S100). If communication is not established (S100, NO), the communication control unit 41 returns to step S100 and repeats the process, waiting until communication with the other terminal device 2 is established.
[0125] On the other hand, when communication with another terminal device 2 is established (S100, YES), the selection acquisition unit 42 of the first terminal device 2A displays a selection of display modes on the display device 22a using a mode selection screen as shown in Figure 9 (S102). For example, the caregiver N, who is the first participant U1, selects and inputs one of the displayed display modes using a keyboard or mouse. The selection acquisition unit 42 accepts the input of the above selection of display mode (S104).
[0126] The selection and acquisition unit 42 determines whether the selected display mode is a display mode that uses an avatar image [2] or a display mode that uses a registered image [3] (S106). If the selected display mode is a mode that uses an avatar or a registered image (S106, YES), the selection and acquisition unit 42 displays candidate images including an avatar or a registered image on the display device 22a according to the selected display mode (S108). For example, the selection and acquisition unit 42 may display candidate images including an avatar or a registered image on the display device 22a using an image selection screen as shown in Figure 10.
[0127] For example, caregiver N, who is the first participant U1, selects and inputs an avatar image or registered image to be used as the participant display image representing participant U of each of the other terminal devices 2 that will be communicating with that terminal device 2. This input can be performed using a keyboard 23a or a mouse 24a. The selection acquisition unit 42 receives the input of the selected avatar image or registered image (S110).
[0128] Next, the selection and acquisition unit 42 notifies each of the terminal devices 2, which are the communication partners, of the participant display image of the participant U using that terminal device 2 (S112). The participant display image is, for example, the camera image of each participant U when the display mode selected in step S104 is display mode [0]. Also, for example, when the display mode selected in step S104 is display mode [1], the participant display image is a predetermined still image for each participant U. Also, for example, when the display mode selected in step S104 is display mode [2] or [3], the participant display image is the first image showing the second participant U2 and the second image showing the third participant U3, which have been received by the selection and acquisition unit 42. The selection and acquisition unit 42 notifies the terminal device 2 used by each participant U of the participant display image of that participant. Each terminal device 2 displays an image notification screen as shown in Figure 11.
[0129] Next, the video processing unit 44 of the first terminal device 2A displays the participant display image notified above on the display devices 22a and 221a, outputs audio to the speaker 27a, and starts a call with each of the terminal devices 2 that are the communication partners (S114).
[0130] When the video processing unit 44 displays an avatar image as a participant display image on the display devices 22a and 221a, it moves the gestures and facial expressions of the avatar image in accordance with prior art based on the video from the terminal device 2 corresponding to that participant display image. For example, the video processing unit 44 can move the gestures and facial expressions of the avatar image selected for participant U in accordance with the movements and voice of participant U using the terminal device 2. Also, when the video processing unit 44 displays an avatar image as a participant display image on the display devices 22a and 221a, it outputs the audio associated with that avatar image to the speaker 27a. Specifically, the video processing unit 44 converts the audio from the terminal device 2 associated with each avatar image into audio associated with each avatar image in accordance with prior art, and outputs the converted audio to the speaker 27a.
[0131] Next, the video processing unit 44 determines whether the call has ended (S116). For example, if the end button 71 is pressed on the call screen shown in Figure 12, it can be determined that the call has ended. If the call has not ended (S116, NO), the video processing unit 44 continues the call started in step S114 and returns to step S116 to wait for the call to end. On the other hand, if the call has ended (S116, YES), the video processing unit 44 ends the call and terminates this process. Specifically, the video processing unit 44 ends the display of participant images on the display devices 22a and 221a and instructs the communication control unit 41 to disconnect the communication connection with the other terminal device 2. The communication control unit 41 disconnects the communication connection with the other terminal device 2 in accordance with the above instruction.
[0132] [6. Effects of the Embodiment] As described above, the remote dialogue system 1 performs an image display method. The image display method includes receiving an input via the first terminal device 2A used by the first participant U1 to select a first image to be used as an image representing a second participant U2, which is different from the first participant U1, from among candidate images including at least one image. The image display method also includes displaying the first image on display devices 22a and 221a, which are first display devices used by the first participant U1, when a call is made between the first terminal device 2A and the second terminal device 2B used by the second participant U2.
[0133] As a result, the image representing the second participant U2 displayed on the first terminal device 2A used by the first participant U1 is the first image selected by the first participant U1 via the first terminal device 2A. Therefore, the stress on the first participant U1 from the images displayed on the first display devices 22a and 221a can be reduced.
[0134] Furthermore, the above candidate images include at least one of the avatar image and the image registered by the first participant U1.
[0135] According to this, the first image representing the second participant U2 may use an avatar image or an image registered by the first participant U1 instead of the actual video of the second participant U2. Therefore, the stress on the first participant U1 caused by the images displayed on the first display devices, displays 22a and 221a, can be effectively reduced.
[0136] In the image display method, if the first image is an avatar image, the audio obtained by converting the voice of the second participant U2 to an audio associated with the avatar image is output from the speaker 27a used by the first participant U1.
[0137] As a result, the speaker 27a used by the first participant U1 outputs audio associated with the avatar image, which is the first image, instead of the actual voice of the second participant U2. This reduces the stress that the audio during the call puts on the first participant U1.
[0138] The image display method further includes notifying the second terminal device 2B of the first image when an input for selecting the first image is received.
[0139] As a result, when a second participant U2 receives notification of the first image, for example, if an avatar image is selected as the first image, it can send audio and video of gestures appropriate to that avatar image to the first terminal device 2A. Therefore, the stress that the images displayed on the display devices 22a and 221a of the first terminal device 2A and the audio output from the speaker 27a cause to the first participant U1 can be reduced more effectively.
[0140] The image display method also includes receiving input via the first terminal device 2A used by the first participant U1 to select a second image different from the first image from among candidate images to be used as an image representing a third participant U3, who is different from the first participant U1 and the second participant U2. The image display method also includes displaying the second image on the first display devices 22a and 221a, which are first display devices used by the first participant U1, when a call is made between the first terminal device 2A and the third terminal device 2C used by the third participant U3.
[0141] As a result, the first participant U1 can select the first image and the second image on the first terminal device 2A as images representing the second participant U2 and the third participant U3. Therefore, in simultaneous calls with the second terminal device 2B and the third terminal device 2C, the stress on the first participant U1 caused by the images displayed on the first terminal device 2A can be reduced.
[0142] Furthermore, the first program 46 is a program executed by the first terminal device 2A, which is a computer. The first program 46 causes the first terminal device 2A to receive input to select a first image from among candidate images, each containing at least one image, to be used as an image representing a second participant U2 who is different from the first participant U1. The first program 46 also causes the first terminal device 2A to display the first image on the first display device used by the first participant U1 when a call is made between the first terminal device 2A and the second terminal device 2B used by the second participant U2. The first display device may be display devices 22a and 221a.
[0143] As a result, the image representing the second participant U2 displayed on the first terminal device 2A used by the first participant U1 is the first image selected by the first participant U1 via the first terminal device 2A. Therefore, the stress on the first participant U1 from the images displayed on the first display devices 22a and 221a can be reduced.
[0144] The remote dialogue system 1 includes an input device for receiving input via a first terminal device 2A used by the first participant U1 to select a first image to be used as an image representing a second participant U2, which is different from the first participant U1, from among candidate images including at least one image. The input device is used by the first participant U1. The input device is, for example, a keyboard 23a or a mouse 24a. The remote dialogue system 1 also includes a first display device used by the first participant U1 to display the first image when making a call between the first terminal device 2A and the second terminal device 2B used by the second participant U2. The first display device is, for example, display devices 22a and 221a.
[0145] As a result, the image representing the second participant U2 displayed on the first terminal device 2A used by the first participant U1 is the first image selected by the first participant U1 via the first terminal device 2A. Therefore, the stress on the first participant U1 from the images displayed on the first display devices 22a and 221a can be reduced.
[0146] [7. Other Embodiments] The above embodiments illustrate a specific example of the application of the present invention, and the present invention is not limited thereto. In the following, components similar to those described in the embodiments are denoted by the same reference numerals, and their detailed descriptions are omitted.
[0147] In the remote dialogue system 1 of the embodiment described above, the first terminal device 2A communicates with two terminal devices 2, the second terminal device 2B and the third terminal device 2C, for the purpose of making a call. The number of terminal devices 2 with which the first terminal device 2A communicates for the purpose of making a call is not limited to two; it may be one or three or more. In this case, the selection and acquisition unit 42 of the first terminal device 2A receives input to select an image from among candidate images to be used as an image representing each participant U, for each of the three or more terminal devices 2 that are the communication partners of the first terminal device 2A. Furthermore, when the first terminal device 2A makes a call with the other terminal devices 2, the video processing unit 44 displays the selected image for each participant U on the display devices 22a and 221a used by the first participant U1.
[0148] In the embodiment described above, in the image selection screen shown as an example in Figure 10, the first image and the second image are selected in association with the second participant U2 and the third participant U3. In another embodiment, the image selection screen may simply select two images from the candidate images without associating them with the second participant U2 and the third participant U3. The selection and acquisition unit 42 can automatically associate the two selected images with the second participant U2 and the third participant U3, respectively, designating the image associated with the second participant U2 as the first image and the image associated with the third participant U3 as the second image.
[0149] In the embodiment described above, when the display mode [2] using avatar images is selected in the first terminal device 2A, only candidate avatar images are displayed on the avatar image selection screen as shown in Figure 10. In another embodiment, if the number of avatar images included in the candidate image information 47 stored in the first memory 45 is less than a predetermined number, the avatar image selection screen may include candidate registered images included in the candidate image information 47. When a registered image is selected on the avatar image selection screen, the selection acquisition unit 42 may accept input to select a registered image to be used as an image representing the second participant U2 and the third participant U3, assuming that the display mode [3] has been re-selected.
[0150] In the embodiment described above, the candidate management unit 43, the video processing unit 44, and the candidate image information 47 are provided by the first terminal device 2A. In another embodiment, the candidate management unit 43, the video processing unit 44, and the candidate image information 47 may be provided by the server 3. The video processing unit 44 can generate images representing the second participant U2 and the third participant U3 according to the selection of the first and second images received by the selection acquisition unit 42 of the first terminal device 2A, and transmit them to each terminal device 2.
[0151] Furthermore, the operational step units shown in Figure 7 are divided according to the main processing content in order to facilitate understanding of the operation of the remote dialogue system 1, and this disclosure is not limited by the way the processing units are divided or the names of those units. Depending on the processing content, the system may be divided into many step units. Alternatively, the system may be divided so that one step unit includes many processes. Also, the order of the steps may be rearranged as appropriate, as long as it does not hinder the intent of this disclosure.
[0152] [8. Summary of this disclosure] A summary of this disclosure is provided below.
[0153] (Note 1) An image display method comprising: receiving input via a first terminal device used by a first participant to select a first image to be used as an image representing a second participant different from the first participant, from among candidate images including at least one image; and displaying the first image on a first display device used by the first participant when a call is made between the first terminal device and a second terminal device used by the second participant.
[0154] As a result, the image representing the second participant displayed on the first display device used by the first participant is the first image selected by the first participant via the first display device, thus reducing the stress that the image displayed on the first display device causes to the first participant.
[0155] (Note 2) The image display method described in Appendix 1, wherein the candidate image includes at least one of an avatar image and an image registered by the first participant.
[0156] As a result, the first image representing the second participant can use an avatar image or an image registered by the first participant instead of the second participant's actual video footage, thereby effectively reducing the stress that the image displayed on the first display device causes to the first participant.
[0157] (Note 3) The image display method according to Appendix 2, further comprising, when the first image is the avatar image, outputting audio obtained by converting the voice of the second participant into audio associated with the avatar image from the speaker used by the first participant.
[0158] As a result, the speaker used by the first participant outputs audio associated with the avatar image (the first image) instead of the second participant's actual voice, thus reducing the stress that the audio during the call causes to the first participant.
[0159] (Note 4) The image display method according to any one of the appendices 1 to 3, further comprising notifying the second terminal device of the first image when an input for selecting the first image is received.
[0160] As a result, the second participant, upon receiving notification of the first image, can, for example, if an avatar image is selected as the first image, send audio and video of gestures appropriate to that avatar image to the first terminal device. Therefore, the stress caused to the first participant by the images displayed on the first display device and the audio output from the speaker can be reduced even more effectively.
[0161] (Note 5) The image display method according to any one of the appendices 1 to 4, further comprising: receiving input via the first terminal device to select a second image different from the first image to be used as an image representing a third participant different from the first participant and the second participant from among the candidate images; and displaying the second image on a first display device used by the first participant when a call is made between the first terminal device and the third terminal device used by the third participant.
[0162] This allows the first participant to select the first and second images on the first terminal device as images representing the second and third participants. Therefore, the stress on the first participant caused by the images displayed on the first display device during simultaneous calls with the second and third terminal devices can be reduced.
[0163] (Note 6) A program that causes a computer to perform the following actions: receive input via a first terminal device used by a first participant to select a first image to be used as an image representing a second participant different from the first participant, from among candidate images including at least one image; and display the first image on a first display device used by the first participant when a call is made between the first terminal device and a second terminal device used by the second participant.
[0164] As a result, the image representing the second participant displayed on the first display device used by the first participant is the first image selected by the first participant via the first display device, thus reducing the stress that the image displayed on the first display device causes to the first participant.
[0165] (Note 7) A remote dialogue system comprising: an input device used by a first participant, which receives input via a first terminal device used by the first participant to select a first image to be used as an image representing a second participant different from the first participant, from among candidate images including at least one image; and a first display device used by the first participant, which displays the first image when making a call between the first terminal device and a second terminal device used by the second participant.
[0166] As a result, the image representing the second participant displayed on the first display device used by the first participant is the first image selected by the first participant via the first display device, thus reducing the stress that the image displayed on the first display device causes to the first participant. [Explanation of Symbols]
[0167] 1...Remote dialogue system, 2...Terminal device, 2A...First terminal device, 2B...Second terminal device, 2C...Third terminal device, 3...Server, 4...Communication network, 21a...First processing unit, 21b...Second processing unit, 21c...Third processing unit, 22a, 221a, 22b, 22c...Display device, 23a, 23b, 23c...Keyboard, 24a, 24b, 24c...Mouse, 25a, 25b, 25c...Camera, 26a, 26b, 26c...Microphone, 27a, 27b, 27c...Speaker, 28a, 28b, 28c, 35...Communication device, 30...Fourth processing unit, 31...Fourth processor, 32, 41, 51, 61...Communication control unit, 33...Fourth memory, 34...Fourth program, 40...First processor, 42...Selection and acquisition unit, 43...Candidate management unit, 44...Video processing unit ,45...First memory, 46...First program, 47...Candidate image information, 50...Second processor, 52...Second memory, 53...Second program, 60...Third processor, 62...Third memory, 63...Third program, 70, 86...Microphone button, 71, 87...Exit button, 72...List, 73, 76...Cursor, 74, 77, 79...OK button, 75...Candidate image, 78, 80, 85...First image, 81, 84...Second image, 82, 83...Image, C...Child requiring medical care, D1, D11, D12, D13...Terminal data, D2, D21, D22, D23...Server data, IAa1, IAa2, IAb...Participant display area, N...Caregiver, S, S1, S2, S3...Location, U1...First participant, U2...Second participant, U3...Third participant.
Claims
1. The system receives input via a first terminal device used by the first participant to select a first image from among candidate images, each containing at least one image, to be used as an image representing a second participant who is different from the first participant. When a call is made between the first terminal device and the second terminal device used by the second participant, the first image is displayed on the first display device used by the first participant, including, Image display method.
2. The candidate image includes at least one of the avatar image and the image registered by the first participant. The image display method according to claim 1.
3. When the first image is the avatar image, the audio obtained by converting the voice of the second participant into audio associated with the avatar image is output from the speaker used by the first participant. Further including, The image display method according to claim 2.
4. When an input is received to select the first image, the first image is notified to the second terminal device. Further including, The image display method according to claim 1.
5. The first terminal device receives input to select a second image, different from the first image, from among the candidate images, to be used as an image representing a third participant, who is different from the first and second participants. When a call is made between the first terminal device and the third terminal device used by the third participant, the second image is displayed on the first display device used by the first participant, Further including, The image display method according to any one of claims 1 to 4.
6. The system receives input via a first terminal device used by the first participant to select a first image from among candidate images, each containing at least one image, to be used as an image representing a second participant who is different from the first participant. When a call is made between the first terminal device and the second terminal device used by the second participant, the first image is displayed on the first display device used by the first participant, Make the computer run it. program.
7. An input device used by the first participant, which receives input via a first terminal device used by the first participant to select a first image from among candidate images including at least one image to be used as an image representing a second participant different from the first participant, When making a call between the first terminal device and the second terminal device used by the second participant, the first display device used by the first participant displays the first image, including, Remote dialogue system.
Citation Information
Patent Citations
TV conference system, TV conference method, and program
WO2018061173A1