Terminal, processing method, communication system, program, and recording medium
The processing device enhances virtual conversations by generating control data to make avatars more likely to speak, addressing the challenge of users being unable to speak due to others, thereby facilitating easy statement-making.
Patent Information
- Application Number
- JP2022091613
- Authority / Receiving Office
- JP · JP
- Patent Type
- Patents
- Current Assignee / Owner
- Filing Date
- 2022-06-06
- Publication Date
- 2025-11-11
- Estimated Expiration
- 2041-11-02
AI Technical Summary
Existing technologies do not provide an environment where users participating in a conversation in a virtual space can easily speak up when others are speaking.
A processing device that acquires motion data of users, generates control data to make avatars more likely to speak, and transmits this data to other terminals to control avatar behavior in the virtual space, allowing users to easily make statements.
Creates an environment where users can easily make statements in virtual conversations by controlling avatar behavior to attract attention and facilitate speaking.
Smart Images

Figure 0007766922000001 
Figure 0007766922000002 
Figure 0007766922000003
Abstract
Description
[Technical Field]
[0001] The present invention relates to a terminal, a processing method, a communication system, a program, and a recording medium. [Background technology]
[0002] There is a method in which, when a predetermined trigger is detected in a virtual space in which an avatar is placed, the placement of the avatar related to the trigger is changed (see Patent Document 1). Specifically, when an avatar starts a conversation, the content distribution server changes the placement of the avatar in the virtual space. [Prior art documents] [Patent documents]
[0003] [Patent Document 1] Patent No. 6684952 Summary of the Invention [Problem to be solved by the invention]
[0004] When multiple users are conversing over a network, there are cases where users are unable to speak due to the comments of other users. However, Patent Document 1 only changes the position of the avatar that started the conversation when the conversation begins. Patent Document 1 does not provide an environment where users participating in a conversation in a virtual space can easily speak.
[0005] The present invention has been made in consideration of the above circumstances, and an object of the present invention is to provide a technology that can realize an environment in which users participating in a conversation in a virtual space can easily speak up. [Means for solving the problem]
[0006] A processing device according to one embodiment of the present invention includes an acquisition unit that acquires motion data of users participating in a conversation in a virtual space, a generation unit that generates control data for an avatar corresponding to the user from the user's motion data, and, when a predetermined motion indicating an attempt to speak is detected in the user's motion data, generates control data indicating a motion that makes it easy for the avatar to speak, a transmission unit that transmits the control data, and a playback unit that causes the avatar to act in the virtual space in accordance with the control data.
[0007] A processing method of one aspect of the present invention includes the steps of: a computer acquiring motion data of a user participating in a conversation in a virtual space; the computer generating control data for an avatar corresponding to the user from the user's motion data; and, when detecting a predetermined motion indicating an attempt to speak in the user's motion data, generating control data indicating a motion that makes the avatar more likely to speak; the computer transmitting the control data; and the computer causing the avatar to act in the virtual space in accordance with the control data.
[0008] A communication system according to one embodiment of the present invention comprises a first terminal used by a first user who is to participate in a conversation in a virtual space, and a second terminal connected to the first terminal and used by a second user who is to participate in the conversation, the first terminal comprising: an acquisition unit that acquires motion data of the first user; a generation unit that generates control data for a first avatar corresponding to the first user from the motion data of the first user, and that, when a predetermined motion indicating an attempt to speak is detected in the motion data of the first user, generates control data indicating a motion that makes it easy for the first avatar to speak; a transmission unit that transmits the control data to the second terminal; a reception unit that receives other avatar control data from the second terminal; and a playback unit that, in the virtual space, causes the first avatar to act in accordance with the control data and causes a second avatar corresponding to the second user to act in accordance with the other avatar control data.
[0009] A communication system according to one aspect of the present invention includes a first terminal used by a first user who is to participate in a conversation in a virtual space, and a second terminal connected to the first terminal and used by a second user who is to participate in the conversation, the first terminal including an acquisition unit that acquires motion data of the first user, a generation unit that generates control data for a first avatar corresponding to the first user from the motion data of the first user, and that generates instruction data indicating that the first avatar is about to speak when a predetermined motion indicating an attempt to speak is detected in the motion data of the first user, and a transmission unit that transmits the instruction data to the second terminal. The second terminal includes an acquisition unit that acquires motion data of the second user, a generation unit that generates control data of a second avatar corresponding to the second user from the motion data of the second user and, upon receiving the instruction data, generates control data of the second avatar that indicates a motion focusing on the first avatar, and a transmission unit that transmits the control data of the second avatar to the first terminal, and the first terminal further includes a playback unit that causes the first avatar to act in the virtual space according to the control data of the first avatar and causes the second avatar to act in the virtual space according to the control data of the second avatar.
[0010] A program according to one aspect of the present invention causes a computer to function as an acquisition unit that acquires motion data of users participating in a conversation in a virtual space, a generation unit that generates control data for an avatar corresponding to the user from the user's motion data, and, when a predetermined motion indicating an attempt to speak is detected in the user's motion data, generates control data indicating a motion that makes it easy for the avatar to speak, a transmission unit that transmits the control data, and a playback unit that causes the avatar to move in the virtual space in accordance with the control data.
[0011] A recording medium according to one embodiment of the present invention stores a program that causes a computer to function as an acquisition unit that acquires motion data of users participating in a conversation in a virtual space, a generation unit that generates control data for an avatar corresponding to the user from the user's motion data and, when a predetermined motion indicating an attempt to speak is detected in the user's motion data, generates control data indicating a motion that makes it easy for the avatar to speak, a transmission unit that transmits the control data, and a playback unit that causes the avatar to move in the virtual space in accordance with the control data. [Effects of the Invention]
[0012] According to the present invention, it is possible to provide a technology that can realize an environment in which users participating in a conversation in a virtual space can easily make a statement. [Brief explanation of the drawings]
[0013] [Figure 1] FIG. 1 is a diagram illustrating the system configuration of a communication system according to an embodiment of the present invention. [Figure 2] FIG. 2 is a diagram illustrating functional blocks of a terminal according to an embodiment of the present invention. [Figure 3] FIG. 3 is a flowchart showing an example of processing performed by the terminal according to the embodiment of the present invention. [Figure 4] FIG. 4 is a sequence diagram illustrating an example of processing in the communication system according to the embodiment of the present invention. [Figure 5] FIG. 5 is a diagram illustrating functional blocks of a terminal according to a modified example. [Figure 6] FIG. 6 is a flowchart showing an example of processing performed by a terminal according to a modified example. [Figure 7] FIG. 7 is a sequence diagram illustrating an example of processing in a communication system according to a modified example. [Figure 8] FIG. 8 is a diagram illustrating the hardware configuration of a computer used in a terminal. DETAILED DESCRIPTION OF THE INVENTION
[0014] Hereinafter, an embodiment of the present invention will be described with reference to the drawings. In the description of the drawings, the same parts are designated by the same reference numerals and the description thereof will be omitted.
[0015] (Communication Systems) A communication system 5 according to an embodiment of the present invention will be described with reference to Fig. 1. The communication system 5 includes a terminal 1, a remote terminal 2, and a server 3. The terminal 1, the remote terminal 2, and the server 3 can communicate with each other via a network 4. The number of terminals 1 and remote terminals 2 in the communication system 5 may be one or more.
[0016] Terminal 1 and counterpart terminal 2 have the same configuration and are general computers such as personal computers and smartphones. Each avatar converses in a virtual space by operating terminal 1 and counterpart terminal 2. Each user of terminal 1 and counterpart terminal 2 controls an avatar that is active in the virtual space. In an embodiment of the present invention, each avatar converses with other avatars, interacts with the bath-seat corresponding to each avatar, and holds web conferences.
[0017] Terminal 1 and the other terminal 2 transmit and receive control data that specifies the movements of the avatars, utterance data (not shown) that indicates what the avatars are to say, and the like. The utterance data may be text data or audio data, and is input by the user of terminal 1. Terminal 1 and the other terminal 2 play back the utterance data, control data, and the like acquired from their own terminal and other terminals. Terminal 1 and the other terminal 2 generate and display video data of the virtual space in which the avatars corresponding to terminal 1 and the other terminal 2 are active.
[0018] The server 3 relays the necessary utterance data, control data, etc. between the terminal 1 and the other terminal 2 to realize a conversation in the virtual space between each avatar. The server 3 transmits each piece of data transmitted from the terminal 1 to the other terminal 2, and transmits each piece of data transmitted from the other terminal 2 to the terminal 1. Note that in the embodiment of the present invention, the server 3 relays communication between the terminal 1 and the other terminal 2, but this is not limiting. The terminal 1 and the other terminal 2 may also communicate without going through the server, using P2P (Peer to Peer) or the like.
[0019] The embodiment of the present invention realizes an environment in which users who participate in conversations through avatars active in a virtual space can easily speak. In a virtual space, an avatar may attempt to talk to another avatar, but may be unable to start speaking because the other avatar has already started speaking. In such a case, the terminal 1 according to the embodiment of the present invention controls the behavior of the avatar attempting to speak so that the avatar can easily speak.
[0020] (Terminal) A terminal 1 according to an embodiment of the present invention will be described with reference to Fig. 2. The counterpart terminal 2 also has a similar configuration to the terminal 1 shown in Fig. 2.
[0021] Terminal 1 includes input device 11, output device 12, motion data 21, control data 22, other avatar control data 23, acquisition unit 31, generation unit 32, transmission unit 33, reception unit 34, and playback unit 35. The motion data 21, control data 22, and other avatar control data 23 are stored in a storage device such as memory 902 or storage 903. The functions of acquisition unit 31, generation unit 32, transmission unit 33, reception unit 34, and playback unit 35 are implemented in CPU 901.
[0022] The input device 11 is a device for inputting user instructions to the terminal 1, and may be a keyboard, a mouse, a motion capture device worn by the user, etc. The output device 12 is a display device, and may be a head-mounted display.
[0023] The movement data 21 is data that specifies the movement of the avatar input by the user of the terminal 1. Generally, the movement of an avatar active in a virtual space is determined by input from the user operating the avatar. The movement data 21 is generated, for example, from the user's own movements obtained by motion capture using a sensor attached to the user. The movement data 21 may also be generated from the movement of the avatar specified by the user using a keyboard, mouse, or the like.
[0024] The control data 22 is data that specifies the movements of an avatar corresponding to the user of terminal 1 in the virtual space. The control data 22 is generated by the generation unit 32. The control data 22 is played back on terminal 1 and the partner terminal 2. In the embodiment of the present invention, the control data 22 may be generated according to the movements indicated by the movement data 21, or may be generated by converting the movements into movements different from those indicated by the movement data 21.
[0025] The other avatar control data 23 is data that specifies the movement of an avatar in the virtual space corresponding to the user of the other party's terminal 2. The other avatar control data 23 is generated in the other party's terminal 2 and acquired from the other party's terminal 2. The other avatar control data 23 is played back on the terminal 1 and the other party's terminal 2.
[0026] The acquisition unit 31 acquires motion data 21 of users participating in conversations in a virtual space. The acquisition unit 31 acquires data on the motion of an avatar designated by the user by motion capture, a keyboard, etc. The method by which the acquisition unit 31 acquires the motion data 21 is not limited to this.
[0027] The generation unit 32 generates control data 22 for an avatar corresponding to the user of the terminal 1. The generation unit 32 generates the control data 22 for the avatar corresponding to the user from the user's motion data 21 under normal circumstances. Note that the control data 22 may be generated by referring to the motion data 21, and the avatar may not necessarily exhibit the same movements as the motion data 21. For example, the generation unit 32 may generate the control data 22 under normal circumstances using data on a predetermined portion of the motion data 21, such as the user's facial expression or line of sight. In this case, the generation unit 32 may generate the control data 22 by converting parameters of the motion in the motion data 21, for example by deforming the motion. The generation unit 32 generates the control data 22 by replacing portions of the motion data 21 that should not be reflected, such as gestures, with data for those portions determined by predetermined rules. The control data 22 may be generated by partially relying on the motion data 21, and the control data 22 may not be generated solely from the motion data 21. The generation unit 32 may also generate the control data 22 by referring to data other than the action data 21. For example, the generation unit 32 may estimate the user's emotion from utterance data input by the user, and generate the control data 22 to correspond to the emotion. The user's emotion may be analyzed from the content of the utterance, or, if the utterance data is voice data, may be analyzed from the frequency characteristics of the voice.
[0028] Here, when the generation unit 32 detects a predetermined movement of the user intending to speak in the movement data 21, it generates control data 22 indicating a movement that makes it easy for the avatar to speak. The transmission unit 33 transmits the control data 22 to the other terminal 2. When the generation unit 32 detects a predetermined movement of the user intending to speak in the movement data 21, it converts the data generated with reference to the movement data 21 or the data of that portion determined by a predetermined rule, for a portion related to a movement that makes it easy for the avatar to speak, into data of a movement that makes it easy for the avatar to speak, thereby generating the control data 22. Furthermore, for portions other than the portion related to a movement that makes it easy for the avatar to speak, the generation unit 32 generates the control data 22 with reference to the movement data 21, as in normal times.
[0029] Normally, the generation unit 32 generates control data 22 so that the avatar can move in the virtual space according to the avatar's movements input by the user. When it is detected in the user's movement data 21 that the user is trying to speak but has not, the generation unit 32 generates control data 22 so that the avatar can easily speak in the virtual space, regardless of the movement data 21.
[0030] If the motion identified by the motion data 21 is a predetermined motion, the generation unit 32 detects that the user is not speaking but intends to speak. The predetermined motion is, for example, half-opening the mouth for a certain period of time, slightly raising one hand, half-opening the mouth, a change in facial expression, nodding, a motion indicating agreement, or the like, which indicates that the user is not speaking but intends to speak. The predetermined motion may further include not detecting utterance data in the motion data 21 within a predetermined time.
[0031] The generation unit 32 may detect, from a detection model generated in advance, that a user is trying to speak but is unable to speak. The generation unit 32 generates a model that can detect, from past motion data, that a user is trying to speak but is unable to speak. The generation unit 32 refers to the generated model and determines whether newly input motion data 21 indicates that the user is trying to speak but is unable to speak.
[0032] When detecting that a user is about to speak, the generation unit 32 generates control data 22 so that the avatar will perform a motion that makes it easy for the avatar to speak. When detecting a predetermined motion that makes it easy for the avatar to speak in the motion data 21, the generation unit 32 generates control data 22 indicating a motion that makes it easy for the avatar to speak for a portion related to the motion that makes it easy for the avatar to speak. A motion that makes it easy for the avatar to speak is, for example, a motion that makes it easy for other avatars to pay attention to the avatar. Specific examples of motions that make it easy for the avatar to speak include raising a hand, opening a mouth, changing a facial expression, leaning forward, moving closer to the current speaker, moving closer to other participants, etc. The generation unit 32 converts the portion related to such motions into a motion that makes it easy for the avatar to speak, and generates the control data 22. Note that the generation unit 32 generates the control data 22 for other portions by referring to the motion data 21 or according to predetermined rules.
[0033] When an avatar attracts attention from other avatars, it becomes easier for the user of the avatar to input a command to make the avatar speak. Also, even if another avatar is speaking, when the avatar attracts attention, the other avatar is prompted to stop speaking, making it easier for the user of terminal 1 to input a command to make the avatar speak.
[0034] When a motion of an attempt to speak is detected in the motion data 21, the generation unit 32 converts the motion into a motion that makes the avatar corresponding to that user more likely to attract attention, and controls the motion of the avatar in the virtual space. This makes it easier for the user's avatar to attract the attention of other users' avatars and speak in the virtual space.
[0035] The generation unit 32 converts a portion of the movement data 21 to generate the control data 22. For example, the generation unit 32 may convert a portion of the movement data 21 that is related to a movement that the avatar is likely to attract attention, and may convert the other portion in accordance with the movement data 21 to generate the control data 22. Alternatively, the generation unit 32 may generate the control data 22 without referring to the movement data 21.
[0036] If the avatar does not speak after transmitting control data 22 indicating actions that make it easy for the avatar to speak to the other terminal 2, the generation unit 32 generates control data 22 from the action data 21. Specifically, if the avatar does not speak, this occurs when the user does not input any utterance data within a predetermined time after transmitting the control data 22 generated by the generation unit 32 to the other terminal 2. If the avatar does not speak despite performing actions that make it easy for the avatar to speak, the generation unit 32 generates control data 22 so as to cause the avatar to behave in accordance with the action data 21 as usual.
[0037] The control data 22 generated by the generation unit 32 is transmitted to the counterpart terminal 2 by the transmission unit 33. The user of the counterpart terminal 2 detects that the user of terminal 1 is about to speak and can encourage the avatar of terminal 1 to speak by paying attention to the avatar of terminal 1 or interrupting their own conversation. However, if the avatar of terminal 1 does not speak, the generation unit 32 stops generating control data 22 from actions that are likely to attract attention for the avatar of terminal 1, and generates control data 22 from the action data 21. This makes it possible for each avatar to behave naturally in the virtual space.
[0038] The receiving unit 34 receives the other avatar control data 23 from the other party's terminal 2. The receiving unit 34 sequentially receives the other avatar control data 23 from the other party's terminal 2 and causes the playback unit 35 to play the other avatar control data. The receiving unit 34 may also receive utterance data from the other party's terminal 2 that indicates the content to be spoken by the other avatar.
[0039] The playback unit 35 plays back the control data 22 and the other avatar control data 23 to generate video data in which multiple avatars act in a virtual space. The playback unit 35 causes an avatar corresponding to the user of terminal 1 to act in the virtual space in accordance with the control data 22. The playback unit 35 also causes a second avatar corresponding to the user of the other terminal 2 to act in the virtual space in accordance with the other avatar control data 23. The playback unit 35 may display text data, generate audio data from the text data, or play the audio data in accordance with utterance data indicating what the avatar should say. The video data is displayed on the output device 12.
[0040] When a predetermined motion of an attempt to speak is detected in the motion data 21, the generation unit 32 generates control data 22 indicating a motion that makes it easy for the avatar to speak. The reproduction unit 35 also makes the avatar move in the virtual space in accordance with the control data 22 indicating a motion that makes it easy for the avatar to speak.
[0041] The processing of the terminal 1 will be described with reference to Fig. 3. The processing shown in Fig. 3 is an example and is not limited to this.
[0042] In step S1, the terminal 1 determines from the action data 21 whether or not a movement of making a statement is detected.
[0043] If detected, the terminal 1 generates the control data 22 indicating a motion that makes it easy for the avatar to speak in step S2. If not detected, the terminal 1 generates the control data 22 from the motion data 21 in step S3.
[0044] In step S4, terminal 1 plays back the control data 22 generated in step S2 or step S3. At this time, terminal 1 also references other avatar control data 23 received from the other terminal 2 and utterance data input from the user or the other terminal 2, and plays back the state of each avatar acting in the virtual space.
[0045] The processing in the communication system 5 will be described with reference to Fig. 4. The processing shown in Fig. 4 is an example, and the present invention is not limited to this.
[0046] In step S51, the terminal 1 generates control data 22 from the motion data 21. In step S52, the counterpart terminal 2 generates control data (other avatar control data 23) from the motion data input by the user of the counterpart terminal 2.
[0047] In step S53, the control data 22 generated in terminal 1 is transmitted to the other terminal 2, and the control data generated in the other terminal 2 is transmitted to terminal 1. In steps S54 and S55, terminal 1 and the other terminal 2 play video data showing each avatar acting in the virtual space in accordance with the control data generated in their own terminals and the control data generated in the other terminals.
[0048] In step S56, terminal 1 detects a movement in which the avatar is about to speak in the movement data 21. In step S57, terminal 1 generates control data 22 indicating a movement that makes it easy for the avatar to speak. In step S58, if the counterpart terminal 2 does not detect a movement in which the avatar is about to speak in the movement data 21, it generates control data from the movement data.
[0049] In step S59, the control data 22 generated on terminal 1 is transmitted to the other terminal 2, and the control data generated on the other terminal 2 is transmitted to terminal 1. In steps S60 and S61, terminal 1 and the other terminal 2 play video data of their avatars moving around in the virtual space in accordance with the control data generated on their own terminals and the control data generated on the other terminals. Here, the avatar on terminal 1 performs an action that is easy to make a statement in the virtual space played on terminal 1 and the other terminal 2, regardless of the action data 21 input by the user of terminal 1.
[0050] In an embodiment of the present invention, when a motion indicating an attempt to speak is detected in the motion data 21 input by the user of terminal 1, control data 22 is generated that indicates a motion that makes it easy for the user's avatar to speak. The generated control data 22 is played back on terminal 1 and the other terminal 2, and the participants in the conversation recognize that the avatar is about to speak. This creates an environment in which it is easy for the avatar of terminal 1 to speak, such as when other avatars stop speaking, and the user of terminal 1 has an opportunity for their avatar to speak easily. In this way, terminal 1 can create an environment in which it is easy for users participating in a conversation in virtual space to speak.
[0051] (Variation) In the embodiment of the present invention, the movement of an avatar attempting to speak is controlled to create an environment in which the avatar feels comfortable speaking. In the modified example, the movement of avatars other than the avatar attempting to speak is controlled to create an environment in which the avatar feels comfortable speaking.
[0052] A terminal 1a according to a modified example will be described with reference to Fig. 5. The counterpart terminal 2a also has the same configuration as the terminal 1a.
[0053] The terminal 1a according to the modification is different from the terminal 1 shown in Fig. 2 in that it includes instruction data 24m and 24n. Furthermore, the processes of the generating unit 32a, transmitting unit 33a, and receiving unit 34a in the terminal 1a are different.
[0054] The instruction data 24m indicates that the user corresponding to the terminal 1a is attempting to speak. The instruction data 24m is generated by the generation unit 32a and transmitted to the partner terminal 2a by the transmission unit 33a.
[0055] The instruction data 24n indicates that an avatar corresponding to a user corresponding to the remote terminal 2a is about to make a statement. The instruction data 24n is transmitted to the remote terminal 2a by the receiving unit 34a. The instruction data 24n is processed by the generating unit 32a.
[0056] When the generation unit 32a receives instruction data 24n indicating that an avatar corresponding to the user of the other terminal 2a is about to speak, the generation unit 32a generates control data 22 indicating an action of focusing on the avatar corresponding to the user of the other terminal 2a. The action of focusing on the avatar corresponding to the user of the other terminal 2a is an action such as turning toward the avatar corresponding to the user of the terminal 1a in the direction of the avatar that is about to speak. The generation unit 32a may also refer to the action data 21 input by the user to generate the control data 22.
[0057] Furthermore, when the generation unit 32a detects a predetermined motion indicating that the avatar corresponding to the user is about to speak in the motion data of the user, the generation unit 32a generates instruction data 24m indicating that the avatar corresponding to the user is about to speak. By notifying the other terminal 2a that the avatar corresponding to the user of the terminal 1a is about to speak, the generation unit 32a prompts the avatar of the other terminal 2a to perform an action that will make it easier for the avatar of the terminal 1a to speak.
[0058] The playback unit 35 plays back the control data 22 and the other avatar control data 23 to generate video data showing multiple avatars moving around in a virtual space. When the user of terminal 1a attempts to speak, the playback unit 35 plays back video data that focuses the attention of the other avatars on the avatar corresponding to the user of terminal 1a. This makes it easier for the user of terminal 1a to speak.
[0059] The processing of the terminal 1a will be described with reference to Fig. 6. The processing shown in Fig. 6 is an example and is not limiting.
[0060] In step S101, the terminal 1a determines from the action data 21 whether or not a motion of making a statement is detected.
[0061] If detected, in step S102 terminal 1a generates instruction data 24m indicating that the user is about to make a statement and transmits it to the counterpart terminal 2a. Furthermore, in step S103 terminal 1a generates control data 22 from the action data 21. If not detected in step S101 and if instruction data 24n is not received in step S104, terminal 1a generates control data 22 from the action data 21 in step S103. If instruction data 24n is received in step S104, terminal 1a generates control data 22 in step S105 that focuses on the avatar corresponding to the user of counterpart terminal 2a that transmitted the instruction data 24n.
[0062] In step S106, terminal 1a plays back the control data 22 generated in step S103 or step S105. At this time, terminal 1a also references other avatar control data 23 received from counterpart terminal 2a and utterance data input from the user or counterpart terminal 2a, and plays back the state of each avatar acting in the virtual space.
[0063] The processing in the communication system 5a (not shown) will be described with reference to Fig. 7. The processing shown in Fig. 7 is an example, and the present invention is not limited to this.
[0064] In step S151, terminal 1a generates control data 22 from the motion data 21. In step S152, counterpart terminal 2a generates control data (other avatar control data 23) from the motion data input by the user of counterpart terminal 2a.
[0065] In step S153, the control data 22 generated by terminal 1a is transmitted to counterpart terminal 2a, and the control data generated by counterpart terminal 2a is transmitted to terminal 1a. In steps S154 and S155, terminal 1a and counterpart terminal 2a play video data of each avatar acting in a virtual space in accordance with the control data generated by their own terminals and the control data generated by the other terminals.
[0066] In step S156, terminal 1a detects a movement in which the avatar is about to speak in the movement data 21. In step S157, terminal 1a generates instruction data 24m indicating that the avatar is about to speak. In step S158, terminal 1a generates control data 22 from the movement data 21. In step S159, the counterpart terminal 2a generates control data that focuses on the avatar corresponding to terminal 1a that sent the instruction data 24m.
[0067] In step S160, the control data 22 generated by terminal 1a is transmitted to the other terminal 2a, and the control data generated by the other terminal 2a is transmitted to terminal 1a. In steps S161 and S162, terminal 1a and the other terminal 2a play video data of their avatars acting in virtual space in accordance with the control data generated by their own terminals and the control data generated by the other terminals. Here, the avatar of the other terminal 2a performs an action that draws attention to the avatar of terminal 1a in the virtual space played back on each of terminal 1a and the other terminal 2a, regardless of the action data input by the user of the other terminal 2a.
[0068] In this modified example, when a motion indicating that the user is about to speak is detected in the motion data 21 input by the user of terminal 1a, terminal 1a transmits instruction data 24m indicating that the user is about to speak to the other terminal 2a. The other terminal 2a generates control data 22 indicating a motion to draw attention to the avatar of terminal 1a. The generated control data 22 is played on terminal 1 and the other terminal 2a, and the participants in the conversation recognize that the avatar of terminal 1a is attracting attention and about to speak. This creates an environment in which the avatar of terminal 1a can easily speak, and the user of terminal 1a can obtain an opportunity for the avatar to easily speak. In this way, terminal 1a can create an environment in which users participating in a conversation in virtual space can easily speak.
[0069] The terminal 1 of the present embodiment described above is, for example, a general-purpose computer system including a CPU (Central Processing Unit, processor) 901, a memory 902, a storage 903 (HDD: Hard Disk Drive, SSD: Solid State Drive), a communication device 904, an input device 905, and an output device 906. In this computer system, the CPU 901 executes a program loaded on the memory 902, thereby realizing each function of the terminal 1.
[0070] The terminal 1 may be implemented by one computer or by multiple computers, or may be a virtual machine implemented on a computer.
[0071] The program of terminal 1 can be stored on a computer-readable recording medium such as an HDD, SSD, USB (Universal Serial Bus) memory, CD (Compact Disc), or DVD (Digital Versatile Disc), or can be distributed via a network.
[0072] The present invention is not limited to the above-described embodiment, and various modifications are possible within the scope of the present invention. [Explanation of symbols]
[0073] 1 device 2. Remote device 3 Server 4 Network 5. Communication Systems 21 Operational Data 22 Control Data 23 Other avatar control data 24 Instruction Data 31 Acquisition Department 32 Generation part 33 Transmitter 34 Receiving unit 35 Playback Department 901 CPU 902 memory 903 Storage 904 Communication equipment 11,905 Input device 12,906 output devices
Claims
1. an acquisition unit that acquires motion data based on motions of users participating in a conversation in a virtual space; a generating unit that generates control data for an avatar corresponding to the user from the user's motion data, and, when a predetermined motion of the user intending to make a statement is detected in the user's motion data, replaces at least a part of the user's motion data to generate control data indicating a motion that makes it easier for the avatar to make a statement; a transmitter for transmitting the control data; a reproduction unit that causes the avatar to act in a virtual space in accordance with the control data; A terminal comprising:
2. an acquisition unit that acquires motion data of users participating in a conversation in a virtual space; a generation unit that generates control data for an avatar corresponding to the user from the user's motion data, and that, when a predetermined motion of the user intending to make a statement is detected in the user's motion data, generates control data indicating a motion that makes it easy for the avatar to make a statement; a transmitter for transmitting the control data; a reproduction unit that causes the avatar to act in a virtual space in accordance with the control data; Equipped with If the avatar does not make a statement after transmitting control data indicating a motion that makes it easy for the avatar to make a statement, the generation unit generates control data from the motion data. Terminal.
3. an acquisition unit that acquires motion data of users participating in a conversation in a virtual space; a generating unit that generates control data for an avatar corresponding to the user from the user's motion data, and, when a predetermined motion of the user intending to make a statement is detected in the user's motion data, replaces at least a part of the user's motion data to generate control data indicating a motion that makes it easier for the avatar to make a statement; a transmitter for transmitting the control data; a reproduction unit that causes the avatar to act in a virtual space in accordance with the control data; Equipped with When a predetermined action of making a statement is detected in the action data, the generation unit generates control data indicating a motion that makes it easy for the avatar to make a statement for a portion of the avatar related to a motion that makes it easy for the avatar to make a statement; The reproduction unit causes the avatar to act in a virtual space in accordance with control data indicating an action that makes it easy for the avatar to make a statement. Terminal.
4. an acquisition unit that acquires motion data of users participating in a conversation in a virtual space; a generation unit that generates control data for an avatar corresponding to the user from the user's motion data, and that, when a predetermined motion of the user intending to make a statement is detected in the user's motion data, generates control data indicating a motion that makes it easy for the avatar to make a statement; a transmitter for transmitting the control data; a reproduction unit that causes the avatar to act in a virtual space in accordance with the control data; Equipped with When receiving instruction data indicating that an avatar corresponding to a user of another terminal is about to make a statement, the generation unit replaces at least a portion of the user's motion data to generate control data indicating a motion that focuses on the avatar corresponding to the user of the other terminal. Terminal.
5. A step in which a computer acquires motion data based on motions of users participating in a conversation in a virtual space; the computer generates control data for an avatar corresponding to the user from the user's motion data, and when a predetermined motion of the user intending to make a statement is detected in the user's motion data, replaces at least a part of the user's motion data to generate control data indicating a motion that makes it easy for the avatar to make a statement; the computer transmitting the control data; a step in which the computer causes the avatar to act in a virtual space in accordance with the control data; A processing method comprising:
6. a first terminal used by a first user who participates in a conversation in a virtual space; a second terminal connected to the first terminal and used by a second user who will participate in the conversation; The first terminal an acquisition unit that acquires motion data based on the motion of the first user; a generating unit that generates control data of a first avatar corresponding to the first user from the motion data of the first user, and when a predetermined motion indicating an attempt to speak is detected in the motion data of the first user, replaces at least a part of the motion data of the first user to generate control data indicating a motion that makes it easy for the first avatar to speak; a transmitter that transmits the control data to the second terminal; a receiving unit that receives other avatar control data from the second terminal; a reproduction unit that causes the first avatar to act in the virtual space in accordance with the control data and causes a second avatar corresponding to a second user to act in accordance with the other avatar control data; Communication system.
7. a first terminal used by a first user who participates in a conversation in a virtual space; a second terminal connected to the first terminal and used by a second user who will participate in the conversation; The first terminal an acquisition unit that acquires motion data of the first user; Control data for a first avatar corresponding to the first user is generated from the motion data of the first user, and when a predetermined motion indicating that the first avatar is about to speak is detected in the motion data of the first user, instruction data indicating that the first avatar is about to speak is generated. a generation unit; a transmitting unit that transmits the instruction data to a second terminal; The second terminal an acquisition unit that acquires motion data of the second user; a generating unit that generates control data of a second avatar corresponding to the second user from the motion data of the second user, and, upon receiving the instruction data, replaces at least a portion of the motion data of the second user with the control data of the second avatar, thereby generating control data of the second avatar indicating a motion focusing on the first avatar; a transmission unit that transmits control data of the second avatar to the first terminal; The first terminal further comprises: a reproduction unit that causes the first avatar to act in accordance with the control data of the first avatar and causes the second avatar to act in accordance with the control data of the second avatar in the virtual space; Communication system.
8. an acquisition unit that acquires motion data based on motions of users participating in a conversation in a virtual space; a generating unit that generates control data for an avatar corresponding to the user from the user's motion data, and, when a predetermined motion of the user intending to make a statement is detected in the user's motion data, replaces at least a part of the user's motion data to generate control data indicating a motion that makes it easier for the avatar to make a statement; a transmitter for transmitting the control data; a reproduction unit that causes the avatar to act in a virtual space in accordance with the control data; A program that makes a computer function as such.
9. an acquisition unit that acquires motion data based on motions of users participating in a conversation in a virtual space; a generating unit that generates control data for an avatar corresponding to the user from the user's motion data, and, when a predetermined motion of the user intending to make a statement is detected in the user's motion data, replaces at least a part of the user's motion data to generate control data indicating a motion that makes it easier for the avatar to make a statement; a transmitter for transmitting the control data; a reproduction unit that causes the avatar to act in a virtual space in accordance with the control data; A recording medium that stores a program that causes a computer to function as a
Citation Information
Patent Citations
Network communication system and method
JP2003308285A
Conference communication system, method, and program
JP2011061450A
In-call translation
JP2017525167A
Information processing system
JP2020080154A
Integrated input and output (i / o) for three-dimensional (3D) environment
JP2022133254A