Avatar generation system and communication system
The centralized management of multiple avatar generation APIs simplifies the process of creating personalized avatars by selecting the optimal service and automating the generation, addressing user inconvenience and inefficiency across different services.
Patent Information
- Application Number
- JP2024061896
- Authority / Receiving Office
- JP · JP
- Patent Type
- Applications
- Current Assignee / Owner
- Filing Date
- 2024-04-08
- Publication Date
- 2025-10-21
- Estimated Expiration
- 2044-04-08
AI Technical Summary
Users face inconvenience and inefficiency when using multiple avatar generation services due to differing specifications and capabilities, requiring trial-and-error to select appropriate services and understand parameter settings.
A system that centrally manages multiple avatar generation APIs, allowing users to easily generate avatars by inputting parameters, dynamically selects the optimal API, and automates the generation process, integrating facial expression, voice, and personality information to create personalized avatars.
Enables efficient and cost-effective avatar generation by reducing the need to understand individual service specifications, ensuring high-quality results, and improving user convenience through automated processes.
Smart Images

Figure 2025159404000001_ABST
Abstract
Description
[Technical Field]
[0001] The present invention relates to a system that centrally manages multiple APIs, selects an appropriate API in response to a request, and generates an avatar, and a communication system that uses the avatar generated by the system. [Background technology]
[0002] In recent years, with the development of computer graphics technology, many services are being provided that generate avatars according to user preferences. These services allow users to select parameters such as hairstyle, eye color, and clothing to obtain the avatar they desire, and input them into an avatar generation API to generate the avatar.
[0003] In this technical field, a method for creating and editing avatars and navigating an avatar selection interface (see Patent Document 1) and a faster and more efficient interface for creating and editing avatars (Patent Document 2) have been disclosed. [Prior art documents] [Patent documents]
[0004] [Patent Document 1] Japanese Patent Publication No. 2022-008470 [Patent Document 2] Japanese Patent Publication No. 2023-085356 Summary of the Invention [Problem to be solved by the invention]
[0005] However, there are many providers offering avatar generation services, and each has different avatar generation API specifications and parameter setting methods. As a result, when a user uses multiple avatar generation services, they need to understand the specifications of each service and set parameters, which can be inconvenient.
[0006] In addition, because the avatar generation capabilities of each avatar generation service differ, users must select an appropriate service to obtain the avatar they desire. However, it is not easy for users to understand the avatar generation capabilities of each service, which forces them to use the service on a trial-and-error basis.
[0007] The present invention has been made in view of the above-mentioned problems, and aims to provide a system that makes it easier to generate avatars. [Means for solving the problem]
[0008] According to the present invention, is obtained. [Effects of the Invention]
[0009] According to the present invention, the following effects are achieved.
[0010] By centrally managing multiple avatar generation APIs, users can easily generate avatars without having to worry about the specifications of each service.
[0011] By managing the avatar generation capabilities of each avatar generation API, it is possible to select the most suitable avatar generation API to generate an avatar according to the user's request. By utilizing the avatar generation history information, an avatar that matches the user's preferences can be generated.
[0012] By having a computer execute the avatar generation process, avatar generation can be automated, improving user convenience.
[0013] Centralized management allows us to understand the usage status of each avatar generation API, which can be used to improve services and provide new services.
[0014] By combining multiple avatar generation APIs, it is possible to generate avatars that cannot be achieved with a single service.
[0015] It reduces the cost of generating avatars. Compared to using multiple services individually, centralized management reduces API usage fees.
[0016] As described above, according to the present invention, a plurality of avatar generation services can be efficiently used, improving user convenience. [Brief explanation of the drawings]
[0017] [Figure 1] 1 is a diagram showing an example of the configuration of an avatar generation system (hereinafter referred to as "the system") according to an embodiment of the present invention. [Figure 2] FIG. 2 is a functional block diagram of the system of FIG. 1. [Figure 3] 2 is an example of a management screen provided to a user by the system avatar management server of FIG. 1. [Figure 4] FIG. 2 is a sequence diagram of the processing of the system of FIG. [Figure 5] FIG. 2 is a schematic diagram showing a fixed conversation registration function of a communication system that uses an avatar generated by the system of FIG. 1. [Figure 6] FIG. 2 is a schematic diagram showing a fixed conversation candidate registration function of a communication system that uses avatars generated by the system of FIG. 1. DETAILED DESCRIPTION OF THE INVENTION
[0018] The present invention will be described below by listing the contents of the embodiments. The present invention has the following configuration. [Item 1] An avatar generation system including at least a facial expression generation API providing server, a voice generation API providing server, a personality generation API providing server, and an avatar management server, The avatar management server a parameter receiving unit that receives input of facial expression parameters, voice parameters, and personality parameters from a user; a parameter sending unit that sends the facial expression parameters, voice parameters, and personality parameters to the corresponding API providing servers; a generated information acquisition unit that receives facial expression information, voice information, and personality information generated by each of the API providing servers; an avatar generation unit that generates an avatar by integrating the facial expression information, voice information, and personality information; An avatar generation system comprising: [Item 2] Item 1 is an avatar generation system according to the present invention, the avatar management server further includes an avatar storage unit that stores the generated avatar. Avatar generation system. [Item 3] The avatar generation system according to item 1 or 2, the avatar management server further comprises a preview unit that reads and displays the stored avatar in response to a preview request from the user; Avatar generation system. [Item 4] The avatar generation system according to any one of items 1 to 3, the avatar management server further comprises an avatar editing unit that edits the generated avatar based on an editing instruction from a user; Avatar generation system. [Item 5] Item 4: The avatar generation system according to any one of items 1 to 4, the avatar management server further includes an avatar sharing unit that shares the generated avatar with other users based on a sharing instruction from the user; Avatar generation system. [Item 6] Item 5. The avatar generation system according to any one of items 1 to 5, The avatar management server The system further includes an API selection unit that selects an optimal API providing server from each of the plurality of facial expression generation API providing servers, the plurality of voice generation API providing servers, and the plurality of personality generation API providing servers. Avatar generation system.
[0019] <Details of implementation form> Hereinafter, embodiments of the present invention will be described with reference to the drawings.
[0020] <Summary of the Invention> As shown in Figure 1, an avatar generation system (hereinafter sometimes referred to as "system") 1 according to an embodiment of the present invention relates to a system that generates an avatar by combining multiple APIs. In the following embodiment, we will explain system 1, which is composed of a facial expression generation server 30, a voice generation API providing server 31, and a personality generation API providing server 33, which respectively provide a facial expression generation API, a voice generation API, and a personality generation API, and an avatar management server 20 that integrates these APIs to generate an optimal avatar.
[0021] According to the present invention, a user can easily create an avatar by simply inputting facial expression, voice, and personality parameters on the platform system provided by the avatar management server 20. Furthermore, when there are multiple API providers, the optimal API can be dynamically selected, thereby improving the performance and efficiency of the system.
[0022] <Hardware configuration example> This system is configured to be able to communicate with each other via the Internet. Note that this system may be configured as a cloud / network type system provided by a designated business operator, or as an on-premise system operated independently within the company that adopts the system.
[0023] Each of the above-mentioned functional blocks can be configured by, for example, hardware provided in a server device (terminal device), a DSP (Digital Signal Processor), or software. For example, when configured by software, each of the above-mentioned functional blocks is actually configured with a CPU, RAM, ROM, etc. of a computer, and is realized by the operation of a program stored in a recording medium such as RAM, ROM, a hard disk, or a semiconductor memory.
[0024] <Network configuration> This system consists of a facial expression generation API providing server, a voice generation API providing server, a personality generation API providing server, and an avatar management server. Each server is an independent computer system and is connected to each other via a network so that they can communicate with each other. The avatar management server accepts requests from user terminals and sends and receives data to and from each API providing server. The avatar management server also has a database for storing the generated avatars.
[0025] <Avatar management server> As shown in FIG. 2, avatar management server 20 according to this embodiment has the following functions for configuring the system.
[0026] <Storage section> The storage unit 101 is implemented in a storage device such as a hard disk or SSD in the avatar management server. Data is managed using a relational database, NoSQL database, or the like. The storage unit 101 is configured to ensure data consistency and permanence, and to enable high-speed reading and writing.
[0027] The storage unit 101 is provided in the avatar management server and stores various data necessary for the operation of the system. The main data stored are as follows: Avatar data Save data such as the image, voice, personality, and actions of the generated avatar. Store the data received from the avatar storage unit in a format that is easy to search and reuse. · User data Save information about the users who use the system. This includes user IDs, passwords, personal settings, purchase histories, etc. · API data Save information about the servers that provide various APIs used by the system, such as the expression generation API, voice generation API, personality generation API, etc. This includes access information, authentication information, usage status, etc. of the API providing server. · Setting data Save various setting information for controlling the operation of System 1. This includes criteria for API selection, data storage format, security settings, etc. Also, the storage unit 101 appropriately reads and writes data from other functional departments of the avatar management server. For example, avatar data generated by the avatar generation unit may be saved in the storage unit 101, or API data used by the API selection unit may be read from the storage unit 101.
[0028] <Parameter reception unit> [[ID=第十九]] The parameter reception unit 102 receives the input of expression parameters, voice parameters, and personality parameters necessary for avatar generation from the user terminal. These parameters are used by the user to customize the expression, voice, and personality of the avatar. The parameter reception unit provides an interface with the user terminal, analyzes and validates the parameters input by the user, and makes them available within the system for subsequent processing.
[0029] [[ID=2二十三]] <API selection unit> The API selection unit 103 is responsible for selecting the optimal server from among multiple API servers available for facial expression generation, voice generation, and personality generation. Selection criteria include the API's response speed, the quality of the generated facial expressions, voice, and personality, the API's usage cost, and the API's reliability. The API selection unit may evaluate APIs based on these criteria and dynamically select the optimal API for each task. This allows the system to efficiently use resources and generate high-quality avatars.
[0030] <Parameter sending section> The parameter sending unit 104 sends corresponding parameters to the API providing server selected by the API selecting unit. Facial expression parameters are sent to the facial expression generation API, voice parameters to the voice generation API, and personality parameters to the personality generation API. The parameter sending unit converts the parameters according to the data format required by each API and sends them to the API providing server via the network. The parameter sending unit also monitors responses from the API providing server, and takes appropriate action if an error occurs.
[0031] <Generation information acquisition section> The generation information acquisition unit 105 receives facial expression information, voice information, and personality information returned from the facial expression generation API, voice generation API, and personality generation API. The generation information acquisition unit manages communication with each API and verifies the consistency of the received data. The received information is passed to the avatar generation unit and integrated.
[0032] <Avatar Generation> The avatar generation unit 106 generates the final avatar by integrating the facial expression, voice, and personality information received from the generation information acquisition unit. The avatar generation unit uses 3D modeling and rendering techniques to create a consistent avatar with facial expressions, voice, and personality. The generated avatar is displayed on the user's device and is also passed to the avatar storage unit for further processing.
[0033] <Avatar Storage Department> The avatar storage unit 107 stores the avatar generated by the avatar generation unit in a database. The avatar data includes information on expressions, voices, and personalities. The avatar storage unit stores the avatar data in the database in a structured format that is easy to search and update. The stored avatar is used for the user to reuse or edit the avatar later.
[0034] <Avatar reading unit> The avatar reading unit 108 reads out the stored avatar from the database in response to an avatar playback request from the user terminal. When the user requests the playback of an avatar, the avatar reading unit acquires the avatar data from the database and transmits it to the user terminal. The avatar reading unit manages the connection to the database and realizes efficient data search and access.
[0035] <Avatar editing unit> The avatar editing unit 109 provides a function to edit the generated avatar based on an editing instruction from the user terminal. The user transmits an editing instruction to change the expression, voice, personality, etc. of the avatar. The avatar editing unit receives these instructions and modifies the avatar data. The edited avatar is stored in the database and transmitted to the user terminal.
[0036] <Avatar sharing unit> The avatar sharing unit 110 provides a function to share the generated avatar with other users based on a sharing instruction from the user terminal. The user can transmit the avatar to other users or share it on social media. The avatar sharing unit receives the user's sharing instruction, converts the avatar data into an appropriate format, and transmits it to the specified destination. In addition, the avatar sharing unit performs privacy settings and permission management to prevent unauthorized sharing of the avatar.
[0037] <Configuration of the API providing server> The API providing server according to this embodiment has the following functions according to their respective characteristics.
[0038] <Facial expression generation processing section> The facial expression generation API server generates facial expressions for the avatar based on the facial expression parameters received from the avatar management server. The facial expression parameters include the degree of eye opening, mouth shape, and eyebrow position. The facial expression generation API uses these parameters as input to generate facial expressions in the form of a 3D model or 2D image. The generated facial expression information is sent back to the avatar management server.
[0039] <Speech generation processing unit> The voice generation API server generates the voice of the avatar based on the voice parameters received from the avatar management server. Voice parameters include pitch, speed, intonation, etc. The voice generation API uses these parameters to generate the voice of the avatar using text-to-speech or voice synthesis technology. The generated voice information is sent back to the avatar management server.
[0040] <Personality generation processing section> The personality generation API server generates an avatar's personality based on the personality parameters received from the avatar management server. Personality parameters include cheerfulness, positivity, friendliness, etc. The personality generation API uses these parameters as input to generate an avatar's personality using natural language processing and dialogue system technology. The generated personality information is used to control the avatar's dialogue and behavior and is sent back to the avatar management server.
[0041] In addition to the API providing server described above, the following APIs may be used in combination or independently. ·Movement generation API Generates avatar body movements and gestures. Actions such as walking, running, jumping, and waving can be generated naturally, resulting in more realistic and lifelike avatars. ·Facial expression recognition API The system recognizes the user's facial expressions and reflects them on the avatar. When the user smiles or gets angry in front of the camera, the avatar will also reproduce the same facial expression. This allows for emotional synchronization between the user and the avatar. Speech Recognition API It recognizes the user's voice and reflects it in the avatar's responses and actions. When the user speaks, the avatar responds appropriately or takes the requested action. This allows for interactive communication with the avatar. Natural Language Processing API Realize natural conversation with avatars. Understand what the user says and generate appropriate responses based on the context. This allows for realistic interactions with avatars. ·Sentiment analysis API Emotions are analyzed from text and voice and reflected in the avatar's facial expressions and speaking style. The avatar reads the emotion in the user's words and responds empathetically. This allows for more natural and empathetic interactions with the avatar. Rendering API Render your avatar in high-quality 3D graphics. Avatars are smoothly animated in real time and displayed in high resolution, enhancing visual immersion. Physics Simulation API The movement of avatar hair and clothing is physically simulated, allowing hair to flutter in the wind and clothing to sway with movement, further enhancing the realism of avatars. Virtual Background API Place avatars in various virtual environments. Create any background, such as a room, a city, or a natural environment, and blend the avatar seamlessly into it. This allows avatars to be used in a variety of scenes. Accessory generation API Users can dress up their avatars with hats, glasses, accessories, etc. to suit their preferences, which increases the variety of avatar appearances. Animation control API Control avatar behavior. You can apply animations to avatars and customize their movements through the API, allowing you to fine-tune the behavior of your avatar.
[0042] By combining these APIs, it is possible to create more expressive and attractive avatars. The avatar's appearance, movement, dialogue, emotional expression, etc. can be enhanced in multiple ways, deepening interaction with the user.
[0043] <Processing flow> The flow of processing in the system 1 according to this embodiment will be described with reference to FIGS.
[0044] When a user logs in and accesses the avatar management server 20, the management screen shown in Fig. 3 is displayed. As shown in the figure, the management screen displays an API selection area G and an avatar management area A. In the API selection area G, buttons Btn for inputting parameters for various APIs are displayed. When each button Btn is selected, a parameter input screen for the corresponding QPI is displayed.
[0045] When the user inputs facial expression parameters, voice parameters, and personality parameters into each parameter input screen (SQ001), the parameter receiving unit of the avatar management server 20 receives the parameters, and the API selecting unit selects the most suitable API providing server (SQ002).
[0046] Next, the parameter sending unit sends the corresponding parameters to the selected API providing server. That is, the avatar management server 20 sends the parameters necessary to generate a facial expression to the facial expression generation API providing server 30 (SQ003). The parameters here include not only whether or not to generate the facial expression, but also items for adjusting the generation process, such as instructions, numerical values, degree, frequency, intensity, start and end times, entered by the user. The technology provided by the facial expression generation API according to this embodiment is lip-sync technology. Specifically, the user uploads an image to which they wish to apply lip-sync to the avatar management server. The uploaded image is sent to the selected facial expression generation API. The facial expression generation API performs the following processes on the received image. Detect faces in an image and extract facial features (mouth, eyes, nose, etc.). Analyzes the specified audio data and extracts audio features (phonemes, pitch, volume, etc.). Lip sync is generated by moving facial feature points over time based on the characteristics of the voice. - Generates an image with lip sync applied and returns it to the avatar management server. This technology allows users to apply lip synchronization to still images, making it possible to generate moving images that make it appear as if the images are speaking.
[0047] Similarly, the avatar management server 20 transmits the parameters required to generate facial expressions to the voice generation API providing server 31 (SQ005). The technology provided by the voice generation API providing server according to this embodiment allows a user to upload a voice sample of about several minutes, and the technology learns the features of that person's voice from the voice sample, enabling all voices to be generated in the same voice as that person. Specifically, the user uploads a voice sample of the person for whom voice generation is desired via the avatar management server. The uploaded voice sample is sent to the selected voice generation API providing server. The voice generation API performs the following processing on the received voice sample. Analyzes voice samples and extracts the characteristics of a person's voice (timbre, pitch, intonation, etc.). The extracted features are used to train a speech synthesis model to reproduce the person's voice. Using a trained speech synthesis model, input any text and generate a voice reading the text in the person's voice. The generated voice data is returned to the avatar management server. Such technology allows users to create a speech synthesis model that reproduces the voice of a particular person from a short voice sample.
[0048] Similarly, the avatar management server 20 transmits parameters necessary for generating the avatar's personality to the personality generation API providing server 32 (SQ007). The technology provided by the personality generation API providing server according to this embodiment provides technology for generating the personality, tone of voice, etc. of the generated avatar when the generated avatar responds via chat or voice using a generative AI. Specifically, the user specifies the personality, tone of voice, etc. that they wish to assign to the avatar as a prompt input via the avatar management server. The specified prompt is sent to the selected personality generation API. The personality generation API performs the following processing based on the received prompt input. Analyze prompt input and extract characteristics such as specified personality and tone of voice. The extracted features are used to generate a personality model to define the avatar's personality and tone of voice. By applying the generated personality model to generative AI, chat and voice responses can be based on that personality and tone of voice. The generated personality model is returned to the avatar management server. This technology allows users to freely define the personality and tone of their avatar through simple prompt input. The personality model generated by this technology is seamlessly integrated with generative AI, allowing for consistent responses based on the personality and tone of voice, providing a more immersive experience for users.
[0049] In this way, facial expression information, voice information, and personality information are generated based on the parameters received by each API provider server. Each piece of generated information is returned to the avatar management server (SQ004, SQ006, SQ008).
[0050] The generation information acquisition unit of the avatar management server receives the generation information from each API providing server, and the avatar generation unit integrates the received information to generate an avatar (SQ009).
[0051] The avatar management server stores the generated avatar in a database and provides it to the user terminal 10 (SQ010). When a user requests to play, edit, or share an avatar, the avatar reading unit, avatar editing unit, and avatar sharing unit can perform the corresponding processing.
[0052] As described above, the present invention is a system that centrally manages multiple avatar generation APIs and selects the optimal API in response to a user request to generate an avatar. Specifically, the management system of the present invention is capable of communicating with avatar generation APIs provided by multiple providers and selects the optimal API in response to a user request for avatar generation. The selected API is then assigned parameters based on the user's request and a request is made to generate an avatar. The generated avatar is acquired by the management system and provided to the user. This series of processes allows the user to obtain a desired avatar with simple operations without having to be aware of multiple avatar generation services.
[0053] According to this invention, users can easily generate avatars through a single management system, without having to understand the specifications of multiple avatar generation services individually. Furthermore, the management system understands the avatar generation capabilities of each API and selects the optimal API according to the user's requirements, allowing for efficient generation of high-quality avatars. Furthermore, when adding a new avatar generation API, it can be provided to users simply by establishing a connection with the management system, ensuring the flexibility and scalability of the system.
[0054] <Communication System> The avatars generated by the avatar generation system according to the above-described embodiment can be implemented as the following communication system: That is, the communication system according to the present embodiment includes a statement receiving unit that receives statements from users via chat, an answer information acquisition unit that communicates with an answer generation API providing server that generates answers in response to the statements and acquires answer information for the statements, and an avatar statement generation unit that generates statements for the generated avatar based on the answer information.
[0055] The comment receiving unit provides a chat interface between the user and the avatar, and receives text or voice comments from the user. The received comments are sent to the answer information acquisition unit.
[0056] The answer information acquisition unit communicates with the answer generation API providing server to generate an answer in response to the received utterance. The answer generation API providing server uses a large-scale language model to generate a natural answer to the utterance. The generated answer information is sent to the avatar utterance generation unit via the answer information acquisition unit.
[0057] The avatar utterance generator generates a utterance for the avatar generated by the avatar generation system based on the response information. The generated utterance is converted into text or voice and presented to the user via a chat interface.
[0058] The avatar utterance generator can also generate facial expressions and movements for the avatar based on the response information and display them in sync with the utterance, thereby achieving more realistic and natural communication with the avatar.
[0059] According to the communication system of this embodiment, it is possible to have natural conversations with users using avatars generated by the avatar generation system. By utilizing advanced language processing technology provided by the answer generation API server, the avatars can generate appropriate and natural answers to user comments, realizing realistic communication.
[0060] <Fixed Conversation Registration Function> As shown in Figure 5, the communication system may also include a question and answer registration unit that registers predetermined questions and answers, and a question and answer determination unit that sends the registered answer directly to the avatar statement generation unit when a user's comment matches the question. This allows a pre-prepared answer to be returned without going through the generative AI, preventing unintended answers and enabling faster response times.
[0061] The question and answer registration unit has a database for registering predetermined question contents and answers. For example, the following question and answer pairs are registered in this database. Questions asked: What is your favorite animal? Answer: I like dogs. Questions asked: What is your favorite food? Answer: It's an apple. The question and answer determination unit analyzes the content of the user's statement received by the statement receiving unit and compares it with the question registered in the question and answer registration unit. If the content of the statement matches the registered question or can be interpreted as being located, the question and answer determination unit transmits the corresponding answer registered in the question and answer registration unit directly to the avatar statement generation unit without going through the answer information acquisition unit.
[0062] For example, if a user says, "What is your favorite animal?", the question and answer determination unit detects that the content of this statement matches the question content registered in the question and answer registration unit. Then, it obtains the corresponding answer, "I like dogs," and sends it to the avatar statement generation unit.
[0063] The avatar statement generator generates a statement for the avatar using the answer received from the question and answer determination unit. In this case, the prepared answer is used as the avatar's statement without going through the answer generation API server.
[0064] By registering questions and answers in advance, the avatar can provide appropriate answers without using generative AI. This prevents unintended answers, shortens the processing time for answer generation, and enables high-speed response.
[0065] <Fixed conversation candidate registration function> 6, the system may further include a fixed conversation registration unit that registers the user's utterance content received by the utterance receiving unit and the avatar's utterance content generated by the avatar utterance generation unit as a question and an answer in the question and answer registration unit based on an instruction from the user. This allows the user to register the question and answer at any time as a "fixed conversation" while communicating with an avatar using generative AI.
[0066] For example, if a user asks, "What are some recommended tourist spots?" and the avatar replies, "My recommendation is Kamakura. The hydrangeas at Hasedera Temple are very beautiful," and the user likes this exchange, they can register it in the fixed conversation registration section, so that the same answer will be returned for the same question in the future.
[0067] The fixed conversation registration unit has the function of registering the user's statement content received by the statement receiving unit and the avatar's statement content generated by the avatar statement generation unit as question content and answer content in the question and answer registration unit based on instructions from the user.
[0068] When a user is having a conversation with an avatar and wants to register the question and answer at that time as a fixed conversation, the user performs a predetermined operation (for example, clicking the "Register this conversation as a fixed conversation" button). When the fixed conversation registration unit detects this operation, it obtains the user's most recent statement and the avatar's statement and registers them in the question and answer registration unit.
[0069] For example, suppose the following exchange took place: User: What food do you dislike? Avatar: A pear. If the user likes this conversation and operates the register button as a fixed conversation, the fixed conversation registration unit registers the user's statement as a question and the avatar's statement as an answer in the question and answer registration unit.
[0070] Once this registration is complete, from then on, when the user utters, "What is your least favorite food?", the question and answer determination unit will detect that the content of this utterance matches the question registered in the question and answer registration unit, and will send the registered answer to the avatar utterance generation unit. As a result, the avatar will always respond with, "Pears."
[0071] In this way, by providing a fixed conversation registration unit, users can fix their favorite exchanges from natural conversations with the generative AI, realizing reproducible communication. This makes it possible to customize communication with the avatar according to the user's preferences.
[0072] <Industrial application fields> The avatar generation system according to the embodiment of the present invention described above can be applied to avatar generation in various fields, such as games, entertainment, education, and communication. In particular, by combining multiple APIs, high-quality avatars with different elements such as facial expressions, voices, and personalities can be efficiently generated, which can greatly contribute to the development and provision of services that utilize avatars. Furthermore, the API selection function can generate the optimal avatar depending on the situation, which can also contribute to improving user satisfaction.
[0073] The above-described embodiment is merely an example for facilitating understanding of the present invention, and is not intended to limit the present invention. The present invention can be modified and improved without departing from the spirit thereof, and it goes without saying that the present invention includes equivalents thereof. [Explanation of symbols]
[0074] 10 Legal Affairs Officer Terminal 20. Client terminal 100 Legal consultation support server
Claims
1. An avatar generation system including at least a facial expression generation API providing server, a voice generation API providing server, a personality generation API providing server, and an avatar management server, The avatar management server a parameter receiving unit that receives input of facial expression parameters, voice parameters, and personality parameters from a user; a parameter sending unit that sends the facial expression parameters, voice parameters, and personality parameters to the corresponding API providing servers; a generated information acquisition unit that receives facial expression information, voice information, and personality information generated by each of the API providing servers; an avatar generation unit that generates an avatar by integrating the facial expression information, voice information, and personality information; An avatar generation system comprising:
2. 10. The avatar generation system according to claim 1, the avatar management server further includes an avatar storage unit that stores the generated avatar. Avatar generation system.
3. The avatar generation system according to claim 1 or claim 2, the avatar management server further comprises a preview unit that reads and displays the stored avatar in response to a preview request from the user; Avatar generation system.
4. 4. The avatar generation system according to claim 1, wherein: the avatar management server further comprises an avatar editing unit that edits the generated avatar based on an editing instruction from a user; Avatar generation system.
5. 5. The avatar generation system according to claim 1, wherein: the avatar management server further includes an avatar sharing unit that shares the generated avatar with other users based on a sharing instruction from the user; Avatar generation system.
6. 6. The avatar generation system according to claim 1, The avatar management server further comprising an API selection unit that selects an optimal API providing server from each of the plurality of facial expression generation API providing servers, the plurality of voice generation API providing servers, and the plurality of personality generation API providing servers; Avatar generation system.
7. A communication system for having a conversation with an avatar generated by the avatar generation system of any one of claims 1 to 6, a comment receiving unit that receives comments from users via chat; an answer information acquisition unit that communicates with an answer generation API providing server that generates an answer in response to the comment and acquires answer information for the comment; an avatar comment generation unit that generates a comment for the generated avatar based on the answer information; A communication system comprising:
8. 8. The communication system according to claim 7, a question and answer registration unit that accepts registration of predetermined question contents and answers; a question and answer determination unit that, when the content of the statement received by the statement receiving unit matches the content of the question registered in the question and answer registration unit, transmits the content of the answer registered in the question and answer registration unit to the avatar statement generation unit without going through the answer information acquisition unit; A communication system comprising:
9. 9. The communication system according to claim 8, The system further includes a fixed conversation registration unit that registers questions and answers in the question and answer registration unit based on instructions from a user, The fixed conversation registration unit registering the content of the user's comment received by the comment receiving unit and the content of the avatar's comment generated by the avatar comment generating unit as a question and an answer in the question and answer registration unit in accordance with an instruction from the user; Communication system.
Citation Information
Patent Citations
Avatar creation user interface
JP2022008470A
Avatar creation user interface
JP2023085356A
Cited By
SEMICONDUCTOR LASER ELEMENT
DE102025139515A1