system

The system allows users to create, share, and rate video content based on their preferences, addressing the limitations of conventional systems by enabling personalized content generation and suggestions.

JP2026037267APending Publication Date: 2026-03-06SOFTBANK GROUP CORP
View PDF 1 Cites 0 Cited by

Patent Information

Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
Filing Date
2024-08-21
Publication Date
2026-03-06

AI Technical Summary

Technical Problem

Conventional video entertainment systems lack the ability for users to create original video works based on their desired genre or theme, share these creations with others, collect ratings, and receive personalized content suggestions that match their preferences.

Method used

A system that includes a user interface for input, automatic video content generation based on user preferences, storage and distribution of generated content, user rating mechanisms, and personalized content suggestions based on user profiles.

Benefits of technology

Enables users to create, share, and receive ratings for their video content, while receiving personalized recommendations tailored to their preferences, enhancing user engagement and satisfaction.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2026037267000001_ABST
    Figure 2026037267000001_ABST
Patent Text Reader

Abstract

Provide a system. A user interface means for accepting user input; means for identifying a preferred genre or theme based on the input; means for automatically generating video content based on the genre or theme; a means for storing the generated video content; means for playing or distributing the stored video content to a user; means for accepting user ratings; means for updating the genre rankings based on the evaluations; A system including:
Need to check novelty before this filing date? Find Prior Art

Description

[Technical Field]

[0001] The technology of the present disclosure relates to a system. [Background technology]

[0002] Patent document 1 discloses a persona chatbot control method performed by at least one processor, the method including the steps of receiving a user utterance, adding the user utterance to a prompt including an instruction sentence related to a description of the chatbot character, encoding the prompt, and inputting the encoded prompt into a language model to generate a chatbot utterance in response to the user utterance. [Prior art documents] [Patent documents]

[0003] [Patent Document 1] Japanese Patent Publication No. 2022-180282 Summary of the Invention [Problem to be solved by the invention]

[0004] In conventional video entertainment, it was difficult for users to create their own original video works based on their desired genre or theme. It was also not easy to share the created video works with other users, collect their ratings, and create rankings. Furthermore, there was a lack of systems that could suggest personalized content based on user preferences. This meant that users could not enjoy video entertainment that perfectly matched their preferences. [Means for solving the problem]

[0005] To solve the above problems, the system of the present invention includes the following means: a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to users, a means for accepting user ratings, and a means for updating genre rankings based on the ratings. The system also includes a means for uploading the automatically generated video content to a posting platform and a means for sharing the uploaded video content with other users. The system further includes a means for creating and saving a user profile and a means for making personalized content suggestions based on the profile, allowing users to enjoy video entertainment that perfectly matches their preferences.

[0006] The "user interface means" is a means for a user to input information to the system, and provides an interface such as a menu screen or a selection screen.

[0007] The "means for specifying a genre or theme" is a means for determining and setting a desired genre or theme based on information input by the user.

[0008] The "means for automatically generating video content" refers to a means for automatically creating video content according to a specified genre or theme.

[0009] The "means for storing" is a means for storing the generated video content in a storage device so that it can be accessed later.

[0010] The "means for playing or distributing" refers to means for playing back the stored video content so that the user can view it, or distributing it to other devices.

[0011] The "means for accepting user evaluations" refers to means for accepting evaluations made by users after viewing video content.

[0012] The "means for updating the genre rankings" is a means for updating the rating rankings for video content in each genre based on the received ratings.

[0013] The "means for uploading to a posting platform" refers to a means for uploading the generated video content to an online posting platform.

[0014] The "means for sharing with other users" refers to a means for allowing other users to view and rate the uploaded video content.

[0015] The "means for creating and saving a user profile" refers to a means for recording and saving information such as a user's preferences and viewing history.

[0016] The "means for making personalized content suggestions" is a means for suggesting optimal video content to a user based on a stored user profile. [Brief explanation of the drawings]

[0017] [Figure 1] 1 is a conceptual diagram showing an example of the configuration of a data processing system according to a first embodiment. [Figure 2] 1 is a conceptual diagram showing an example of main functions of a data processing device and a smart device according to a first embodiment. [Figure 3] FIG. 10 is a conceptual diagram showing an example of the configuration of a data processing system according to a second embodiment. [Figure 4] FIG. 10 is a conceptual diagram showing an example of main functions of a data processing device and smart glasses according to a second embodiment. [Figure 5] FIG. 10 is a conceptual diagram showing an example of the configuration of a data processing system according to a third embodiment. [Figure 6]FIG. 11 is a conceptual diagram showing an example of main functions of a data processing device and a headset-type terminal according to a third embodiment. [Figure 7] FIG. 10 is a conceptual diagram showing an example of the configuration of a data processing system according to a fourth embodiment. [Figure 8] FIG. 10 is a conceptual diagram showing an example of main functions of a data processing device and a robot according to a fourth embodiment. [Figure 9] 1 shows an emotion map onto which multiple emotions are mapped. [Figure 10] 1 shows an emotion map onto which multiple emotions are mapped. [Figure 11] FIG. 3 is a sequence diagram illustrating a processing flow of the data processing system according to the first embodiment. [Figure 12] FIG. 10 is a sequence diagram showing the flow of processing in the data processing system in Application Example 1. [Figure 13] FIG. 10 is a sequence diagram showing the flow of processing in the data processing system according to the second embodiment when an emotion engine is combined. [Figure 14] FIG. 10 is a sequence diagram showing the flow of processing in the data processing system in Application Example 2 when an emotion engine is combined. DETAILED DESCRIPTION OF THE INVENTION

[0018] An example of an embodiment of a system according to the technology of the present disclosure will be described below with reference to the accompanying drawings.

[0019] First, the terms used in the following description will be explained.

[0020] In the following embodiments, a coded processor (hereinafter simply referred to as a "processor") may be a single arithmetic device or a combination of multiple arithmetic devices. Furthermore, a processor may be a single type of arithmetic device or a combination of multiple types of arithmetic devices. Examples of arithmetic devices include a CPU (Central Processing Unit), a GPU (Graphics Processing Unit), a GPGPU (General-Purpose computing on Graphics Processing Units), and an APU (Accelerated Processing Unit).

[0021] In the following embodiments, a coded RAM (Random Access Memory) is a memory in which information is temporarily stored and is used as a working memory by a processor.

[0022] In the following embodiments, the coded storage is one or more non-volatile storage devices that store various programs, various parameters, etc. Examples of non-volatile storage devices include flash memory (SSD (Solid State Drive)), magnetic disks (e.g., hard disks), and magnetic tapes.

[0023] In the following embodiments, a communication I / F (Interface) with a symbol is an interface including a communication processor, an antenna, etc. The communication I / F controls communication between multiple computers. Examples of communication standards applied to the communication I / F include wireless communication standards including 5G (5th Generation Mobile Communication System), Wi-Fi (registered trademark), Bluetooth (registered trademark), etc.

[0024] In the following embodiments, "A and / or B" is synonymous with "at least one of A and B." In other words, "A and / or B" means that it may be only A, only B, or a combination of A and B. Furthermore, in this specification, the same concept as "A and / or B" is also applied when three or more things are expressed connected by "and / or."

[0025] [First embodiment]

[0026] FIG. 1 shows an example of the configuration of a data processing system 10 according to the first embodiment.

[0027] 1, a data processing system 10 includes a data processing device 12 and a smart device 14. An example of the data processing device 12 is a server.

[0028] The data processing device 12 includes a computer 22, a database 24, and a communication I / F 26. The computer 22 is an example of a "computer" according to the technology of the present disclosure. The computer 22 includes a processor 28, a RAM 30, and a storage 32. The processor 28, the RAM 30, and the storage 32 are connected to a bus 34. The database 24 and the communication I / F 26 are also connected to the bus 34. The communication I / F 26 is connected to a network 54. Examples of the network 54 include a WAN (Wide Area Network) and / or a LAN (Local Area Network).

[0029] The smart device 14 includes a computer 36, a reception device 38, an output device 40, a camera 42, and a communication I / F 44. The computer 36 includes a processor 46, a RAM 48, and a storage 50. The processor 46, the RAM 48, and the storage 50 are connected to a bus 52. The reception device 38, the output device 40, and the camera 42 are also connected to the bus 52.

[0030] The reception device 38 includes a touch panel 38A, a microphone 38B, and the like, and receives user input. The touch panel 38A detects contact with an indicator (for example, a pen or a finger) to receive user input by the touch of the indicator. The microphone 38B detects the user's voice to receive user input by voice. The control unit 46A transmits data indicating the user input received by the touch panel 38A and the microphone 38B to the data processing device 12. In the data processing device 12, the specific processing unit 290 acquires the data indicating the user input.

[0031] The output device 40 includes a display 40A and a speaker 40B, and presents data to the user 20 by outputting the data in a form of expression that the user 20 can perceive (for example, audio and / or text). The display 40A displays visible information such as text and images in accordance with instructions from the processor 46. The speaker 40B outputs audio in accordance with instructions from the processor 46. The camera 42 is a compact digital camera equipped with an optical system including a lens, aperture, and shutter, and an imaging element such as a CMOS (Complementary Metal-Oxide-Semiconductor) image sensor or a CCD (Charge Coupled Device) image sensor.

[0032] The communication I / F 44 is connected to a network 54. The communication I / Fs 44 and 26 control the exchange of various information between the processor 46 and the processor 28 via the network 54.

[0033] FIG. 2 shows an example of the main functions of the data processing device 12 and the smart device 14.

[0034] 2, in the data processing device 12, a specific process is performed by the processor 28. A specific processing program 56 is stored in the storage 32. The specific processing program 56 is an example of a "program" according to the technology of the present disclosure. The processor 28 reads the specific processing program 56 from the storage 32 and executes the read specific processing program 56 on the RAM 30. The specific process is realized by the processor 28 operating as a specific processing unit 290 in accordance with the specific processing program 56 executed on the RAM 30.

[0035] The storage 32 stores a data generation model 58 and an emotion identification model 59. The data generation model 58 and the emotion identification model 59 are used by the identification processing unit 290.

[0036] In the smart device 14, the processor 46 performs the reception output process. The storage 50 stores a reception output program 60. The reception output program 60 is used in conjunction with the specific processing program 56 by the data processing system 10. The processor 46 reads the reception output program 60 from the storage 50 and executes the read reception output program 60 on the RAM 48. The reception output process is realized by the processor 46 operating as the control unit 46A in accordance with the reception output program 60 executed on the RAM 48.

[0037] Next, a description will be given of the specific processing performed by the specific processing unit 290 of the data processing device 12. In the following description, the data processing device 12 will be referred to as a "server" and the smart device 14 will be referred to as a "terminal."

[0038] This system allows users to create original video content based on their preferences, share it with other users, and receive ratings. Specific embodiments of each component of the system will be described below.

[0039] User Interface Means

[0040] Device:

[0041] The terminal provides the interface through which the user enters input into the system. The user can access their account by opening the application and entering their credentials on the login screen. On first use, the user is prompted to select their preferred genre and theme.

[0042] A means of identifying a genre or theme

[0043] server:

[0044] The system receives and analyzes information about genres and themes entered by the user to identify genres and themes based on the user's preferences. For example, if the user selects "action" and "fantasy," it sets parameters for generating video content related to these genres.

[0045] A means of automatically generating video content

[0046] server:

[0047] The video generation AI module is activated and automatically generates video content based on the genre and theme selected by the user. The AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if the user selects "action + fantasy," a video containing an episode of a warrior fighting a dragon will be generated.

[0048] Means of preservation

[0049] server:

[0050] The generated video content is temporarily saved and then saved in a dedicated folder for the user, allowing the user to access the generated video at any time.

[0051] Means of playback or distribution

[0052] Device:

[0053] The generated video content can be streamed from the server to the device or downloaded directly to the device, allowing users to watch the video in real time on their own devices.

[0054] A means of accepting user ratings

[0055] Device:

[0056] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[0057] server:

[0058] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[0059] A way to update genre rankings

[0060] server:

[0061] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[0062] How to upload to a publishing platform

[0063] Device:

[0064] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[0065] server:

[0066] Store uploaded video content on the platform and make it accessible to other users.

[0067] How to share with other users

[0068] server:

[0069] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[0070] A means to create and store user profiles

[0071] server:

[0072] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[0073] A means of making personalized content suggestions

[0074] server:

[0075] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[0076] The above is a detailed description of the system in the "Description of Embodiments."

[0077] The processing flow will be explained below.

[0078] Step 1:

[0079] Device:

[0080] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[0081] Step 2:

[0082] server:

[0083] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[0084] Step 3:

[0085] Device:

[0086] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[0087] Step 4:

[0088] Device:

[0089] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[0090] Step 5:

[0091] server:

[0092] The received request is analyzed and the necessary parameters are set for the AI ​​model, for example, setting video generation based on "action" and "fantasy."

[0093] Step 6:

[0094] server:

[0095] The video generation AI module is activated to generate a scenario, construct a scene, and synthesize audio. Specifically, it automatically generates a video containing an episode of a warrior fighting a dragon.

[0096] Step 7:

[0097] server:

[0098] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[0099] Step 8:

[0100] server:

[0101] The terminal is notified that the video content has been saved.

[0102] Step 9:

[0103] Device:

[0104] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[0105] Step 10:

[0106] Device:

[0107] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[0108] Step 11:

[0109] User:

[0110] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[0111] Step 12:

[0112] Device:

[0113] After the user enters their rating and comments, the data is sent to the server.

[0114] Step 13:

[0115] server:

[0116] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[0117] Step 14:

[0118] User:

[0119] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[0120] Step 15:

[0121] Device:

[0122] The upload procedure for the video content is carried out and an upload request is sent to the server.

[0123] Step 16:

[0124] server:

[0125] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[0126] These are the specific processing steps involved in generating and sharing a video work based on a user request.

[0127] Example 1

[0128] Next, a description will be given of Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the smart device 14 will be referred to as a "terminal."

[0129] In the creation and sharing of digital media content, personalized content delivery based on user preferences is required, but existing systems do not effectively generate, evaluate, and share content that reflects user preferences. Furthermore, the accuracy of recommendation systems and ranking updates based on user ratings also needs to be improved.

[0130] The specific processing by the specific processing unit 290 of the data processing device 12 in the first embodiment is realized by the following means.

[0131] In this invention, the server includes a user interface means for accepting user input, a means for identifying preferred categories or topics based on the input, a means for automatically generating digital media content based on the categories or topics, a means for saving the generated digital media content, a means for playing or distributing the saved digital media content to the user, a means for accepting user ratings, and a means for updating category rankings based on the ratings, thereby enabling efficient generation, rating, sharing, and ranking updates of personalized digital media content based on user preferences.

[0132] The "user interface means" is a device or program that provides an interface for a user to input information to the system.

[0133] A "category" is a concept that refers to a genre or topic of digital media content selected by a user.

[0134] A "topic" is a concept that refers to a specific theme or subject of digital media content selected by a user.

[0135] "Digital media content" refers to media data in digital format, such as video, audio, and images, that is generated based on user preferences.

[0136] An "automatic generation means" is a device or program that automatically creates digital media content based on category or topic information entered by a user.

[0137] A "storage means" is a device or program for temporary and long-term storage of the generated digital media content.

[0138] A "delivery means" is a device or program that provides stored digital media content to a user's terminal in a streaming or downloadable format.

[0139] The "means for receiving ratings" is a device or program that collects rating information on digital media content from users.

[0140] The "means for updating the ranking by category" is a device or program that updates the ranking of digital media content for each category based on evaluation information collected from users.

[0141] A "sharing platform" is an online service or website for sharing user-generated digital media content with other users.

[0142] MODE FOR CARRYING OUT THE INVENTION

[0143] The system automates the creation, storage, distribution, rating, and sharing of digital media content. It uses user input to identify preferred categories and topics, and then uses generative AI models to automatically generate content. A specific implementation of this system is described below.

[0144] 1. User Interface Methods

[0145] (Terminal)

[0146] The terminal provides an application as a user interface for the user to input data to the system. The user accesses the system by starting this application and entering a user name and password on the login screen.

[0147] 2. How to identify your favorite categories and topics

[0148] (Terminal)

[0149] On first use, the device will display an interface for setting user preferences, where users can select their preferred categories and topics, for example, "action" and "fantasy."

[0150] (server)

[0151] The server receives and analyzes this selection information and stores it in a database. This data is used to generate content using the AI ​​model described below.

[0152] 3. Content Generation Methods

[0153] (server)

[0154] The server generates prompt sentences based on the user's preference information and sends them to a generative AI model (e.g., OpenAI (registered trademark) GPT-4 (registered trademark), a video generation model created by DeepMind).

[0155] 4. Video content generation

[0156] (server)

[0157] The server's generative AI model automatically generates video content based on the input prompt. For example, for a user who likes "action" and "fantasy," a video containing a story about a warrior fighting a dragon is generated. Video generation involves the processes of scenario generation, scene construction, and voice synthesis.

[0158] 5. How content is stored

[0159] (server)

[0160] The server temporarily stores the generated video content, and then stores it in the user's dedicated folder. The video files are managed using a database management system (e.g., MySQL (registered trademark), PostgreSQL).

[0161] 6. Content playback and distribution methods

[0162] (Terminal)

[0163] The device allows users to play or download stored video content, often using streaming technology and a media player (e.g., VLC, QuickTime).

[0164] 7. Method of accepting evaluations

[0165] (Terminal)

[0166] After watching the video, the device displays a user rating input interface, allowing users to enter star ratings and comments.

[0167] (server)

[0168] The server receives user ratings and stores them in a database, which is used to update rankings for each category.

[0169] 8. How to update rankings by category

[0170] (server)

[0171] The server updates the rankings for each category based on the collected rating information, and then makes this ranking information available to other users, who can then recommend highly rated content.

[0172] 9. How to share content

[0173] (Terminal)

[0174] The terminal provides an interface for uploading the generated digital media content to a sharing platform.

[0175] (server)

[0176] The server manages the uploaded content and makes it accessible to other users, generating shareable links for easy sharing via social media or email.

[0177] 10. How to create and store user profiles

[0178] (server)

[0179] A profile is created and saved based on the user's preferences and viewing history, allowing for personalized content suggestions the next time the user uses the service.

[0180] 11. Means of personalized content recommendations

[0181] (server)

[0182] Based on a saved user profile, the app suggests digital media content that is most suitable for the user, for example, if a user previously indicated that they like "action," the app will suggest the latest action movies.

[0183] Examples of concrete examples and prompts

[0184] Examples:

[0185] The user selects "action" and "fantasy" and enters the prompt "Generate a movie about a warrior fighting a dragon." Based on this prompt, the server's generative AI model automatically generates video content and saves it in the user's dedicated folder. This video content can be viewed on the device and shared with other users.

[0186] Example prompt sentence:

[0187] "Generate footage of action scenes in fantasy movies"

[0188] "Suggest your next watchlist based on your favorite movie genres."

[0189] This system allows users to easily create, save, rate, and share content tailored to their preferences. This specific embodiment will clarify how the invention actually works.

[0190] The flow of the identification process in the first embodiment will be described with reference to FIG.

[0191] Step 1:

[0192] (User):

[0193] The user launches the application and enters their username and password on the login screen.

[0194] input:

[0195] Username, Password

[0196] output:

[0197] Authentication Request

[0198] (Device):

[0199] The terminal accepts user input and sends an authentication request to the server.

[0200] Specific behavior:

[0201] Displaying the login screen

[0202] Obtaining the entered username and password

[0203] Sending an authentication request to the server

[0204] Step 2:

[0205] (server):

[0206] The server checks the received authentication request against its database and returns the authentication result.

[0207] input:

[0208] Authentication Request

[0209] output:

[0210] Authentication result (success / failure)

[0211] (Device):

[0212] The terminal performs screen transitions based on the authentication results.

[0213] Specific behavior:

[0214] Receiving the authentication result

[0215] If authentication is successful, the main screen is displayed. If authentication fails, an error message is displayed.

[0216] Step 3:

[0217] (User):

[0218] On first use, users make selections in an interface that allows them to choose their preferred categories and topics.

[0219] input:

[0220] Category selection (e.g., "Action"), Topic selection (e.g., "Fantasy")

[0221] output:

[0222] Selected Data

[0223] (Device):

[0224] The terminal transmits the selected data to the server.

[0225] Specific behavior:

[0226] Receiving user selections

[0227] Sending Selected Data to the Server

[0228] Step 4:

[0229] (server):

[0230] The server stores the received selection data in a database.

[0231] input:

[0232] Selected Data

[0233] output:

[0234] Results saved in the database

[0235] Specific behavior:

[0236] Analysis of selected data

[0237] Saving to the database

[0238] Step 5:

[0239] (User):

[0240] The user inputs the content of the video they want to generate on the prompt sentence input screen.

[0241] input:

[0242] Prompt text (e.g., "Generate a movie of a warrior fighting a dragon")

[0243] output:

[0244] Prompt Data

[0245] (Device):

[0246] The terminal sends the prompt data to the server.

[0247] Specific behavior:

[0248] Receiving a prompt

[0249] Sending prompt data to the server

[0250] Step 6:

[0251] (server):

[0252] The server inputs the prompt data into the generative AI model and begins generating video content.

[0253] input:

[0254] Prompt Data

[0255] output:

[0256] Generated video content

[0257] Specific behavior:

[0258] Parsing prompt data

[0259] Input to generative AI models

[0260] Video content generation

[0261] Step 7:

[0262] (server):

[0263] The server temporarily stores the generated video content, and then stores it in a folder dedicated to the user.

[0264] input:

[0265] Generated video content

[0266] output:

[0267] Saved results in user folder

[0268] Specific behavior:

[0269] Temporary storage of video content

[0270] Move to user-specific folder

[0271] Step 8:

[0272] (Device):

[0273] The user plays or downloads the stored video content.

[0274] input:

[0275] User requests (playback, download)

[0276] output:

[0277] Playback and download of video content

[0278] Specific behavior:

[0279] Click the play button to stream

[0280] Click the download button to save the video to your device

[0281] Step 9:

[0282] (User):

[0283] After viewing the video, the user inputs a rating.

[0284] input:

[0285] Rating data (stars, comments)

[0286] output:

[0287] Evaluation data submission

[0288] (Device):

[0289] The terminal transmits the evaluation data to the server.

[0290] Specific behavior:

[0291] View the rating interface

[0292] Receiving and sending evaluation data to the server

[0293] Step 10:

[0294] (server):

[0295] The server stores the received rating data in a database and updates the rankings by category.

[0296] input:

[0297] Evaluation Data

[0298] output:

[0299] Updated ranking data

[0300] Specific behavior:

[0301] Analysis of evaluation data

[0302] Saving to a database

[0303] Category ranking updates

[0304] Step 11:

[0305] (User):

[0306] Users upload the generated video content to a sharing platform.

[0307] input:

[0308] Upload Request

[0309] output:

[0310] Upload completion notification

[0311] (Device):

[0312] Pressing the upload button sends the video content to the server.

[0313] Specific behavior:

[0314] Click the upload button

[0315] Sending video content to the server

[0316] Step 12:

[0317] (server):

[0318] The server manages the uploaded content, sets it up so that other users can access it, and generates a share link so that it can be shared via social media or email.

[0319] input:

[0320] Uploaded content

[0321] output:

[0322] Shared Links

[0323] Specific behavior:

[0324] Organizing uploaded content

[0325] Generate shareable links for social media and email

[0326] Step 13:

[0327] (server):

[0328] The server creates and stores a profile based on the user's preferences and viewing history.

[0329] input:

[0330] User preferences and viewing history

[0331] output:

[0332] User Profile

[0333] Specific behavior:

[0334] User data analysis

[0335] Creating and Saving a Profile

[0336] Step 14:

[0337] (server):

[0338] The server makes personalized content suggestions based on the stored profile.

[0339] input:

[0340] User Profile

[0341] output:

[0342] Suggested Content List

[0343] Specific behavior:

[0344] Viewing Profiles

[0345] Selection of proposed content

[0346] The above are the specific steps of the program processing of this system.

[0347] (Application example 1)

[0348] Next, a description will be given of Application Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the smart device 14 will be referred to as a "terminal."

[0349] Conventional video content generation and distribution systems lack the functionality to allow users to easily create original video content based on their preferences and share it with other users. Furthermore, the lack of personalized content suggestions and rating systems hinders the improvement of the user experience.

[0350] The specific processing by the specific processing unit 290 of the data processing device 12 in the application example 1 is realized by the following means.

[0351] In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre rankings based on the ratings, a means for creating and saving a user profile, a means for suggesting personalized content based on the profile, a means for a user to log in and select a genre using an application installed on a smartphone or smart glasses, a means for saving and sharing the generated video in cloud storage, and a means for generating video using an AI model by inputting a prompt. This allows users to easily generate video content according to their preferences, share it with other users, receive ratings, and improve the accuracy of content suggestions from the next time onwards.

[0352] The "user interface means" is an interface that allows a user to input information into the system, and is a function that allows operation through a smartphone or smart glasses, for example.

[0353] The "means for identifying a favorite genre or theme" is a function of the system that analyzes and identifies a user's favorite genre or theme based on information input by the user.

[0354] The "means for automatically generating video content" is an AI module for generating video content based on the genre or theme selected by the user.

[0355] The "means for saving the generated video content" is a function for temporarily or permanently storing the generated video in a recording device.

[0356] The "means for playing or distributing stored video content to users" is a function for providing stored video so that users can view it.

[0357] The "means for accepting user ratings" is a function including an interface for users to input ratings and comments on video content.

[0358] The "means for updating genre rankings" is a function for automatically updating genre rankings based on user ratings.

[0359] "Means for creating and saving user profiles" refers to a function that creates individual profiles based on the user's preferences and viewing history and saves them in a database.

[0360] The "means for making personalized content suggestions" is a function that suggests the most suitable video content to a user based on a saved user profile.

[0361] "Application installed on a smartphone or smart glasses" refers to software for a mobile device that allows a user to access the video content generation system.

[0362] "Means for saving and sharing the generated video in cloud storage" refers to a function for saving the generated video file to a remote server via the Internet and providing a sharing link from there.

[0363] "Means for generating images using an AI model by inputting a prompt sentence" refers to a function in which a user inputs a specific instruction sentence, and the AI ​​model generates an image based on that instruction.

[0364] This invention is a system that enables users to use mobile devices such as smartphones or smart glasses to create original video content based on their preferences, share it with other users, and receive ratings.

[0365] User Interface Means

[0366] The device is equipped with an interface that allows users to input information into the system. Users can access their account by opening the application on their smartphone or smart glasses and entering their authentication information on the login screen. On first use, a screen appears where users can select their preferred genre (e.g., "action" or "fantasy") and theme.

[0367] A means of identifying a genre or theme

[0368] The server receives the genre and theme information entered by the user, analyzes it, and identifies the user's preferences. For example, if the user selects "thriller" and "mystery," parameters for generating video content related to these genres are set.

[0369] A means of automatically generating video content

[0370] The server launches a video generation AI model and automatically generates video content based on the genre and theme selected by the user. During this process, the AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if a user selects "Thriller + Mystery," a short video of a detective solving a mysterious case will be generated. The AI ​​model used here utilizes machine learning frameworks such as TENSORFLOW (registered trademark) and PyTorch.

[0371] Storage and playback methods

[0372] The server temporarily stores the generated video content in cloud storage such as Google Cloud Storage, and then saves it in the user's dedicated folder. The user can then view the video in real time on their smartphone or smart glasses.

[0373] How to receive ratings

[0374] The device provides an interface for users to input ratings after watching videos. Users can enter ratings from 1 to 5 stars and comments. The server records the received ratings and comments in a database such as Firebase and analyzes them to update the rankings by genre.

[0375] How to update the rankings

[0376] The server automatically updates the rankings of video content by genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[0377] How to store and share content

[0378] The generated video content is stored in cloud storage and a sharing link is generated, allowing users to easily share their work via social media or email.

[0379] User profiles and personalized content suggestions

[0380] The server creates and stores a profile based on the user's preferences and viewing history. The next time the user logs in, personalized content suggestions are made based on this profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[0381] Examples of concrete examples and prompts

[0382] After a user logs in, they can select their preferred genre, create, watch, and rate a short video. The following prompts are used:

[0383] Input: "Generate a short film that mixes horror and fantasy genres. My Firebase UID is abcdef."

[0384] Output: "Generating movie... The URL for the generated movie is https: / / storage.googleapis.com / abcdef / generated_video.mp4."

[0385] This system allows users to easily create and enjoy video content according to their preferences.

[0386] The flow of the specific processing in the application example 1 will be described with reference to FIG.

[0387] Step 1:

[0388] A user launches an application on their smartphone or smart glasses and accesses the login screen. The input is the user's authentication information, and the output is the authentication success / failure result. The server validates the user's authentication information using Firebase Auth and returns the user ID if authentication is successful.

[0389] Step 2:

[0390] After logging in, the user selects their preferred genre and theme. The input is the genre and theme selected by the user, and the output is the data of the selected genre and theme. The terminal sends the selected genre and theme to the server.

[0391] Step 3:

[0392] The server analyzes the received genre and theme information and sets the parameters of the AI ​​model. The input is genre and theme data, and the output is the parameter settings of the AI ​​model. The server uses TensorFlow or PyTorch to initialize the AI ​​model and set the settings based on the received data.

[0393] Step 4:

[0394] The server launches the configured AI model and generates video content. The input is the AI ​​model's parameters and generation instructions (prompt text), and the output is the generated video file. The AI ​​model generates a scenario, constructs scenes, and synthesizes audio. In this step, if the user inputs, for example, "Please generate a short film that combines the horror and fantasy genres," a video of the corresponding scenario will be automatically generated.

[0395] Step 5:

[0396] The server saves the generated video file in cloud storage (e.g., Google Cloud Storage). The input is the generated video file, and the output is the cloud storage URL. The server uploads the video file to cloud storage and generates a URL link for the saved file.

[0397] Step 6:

[0398] The server provides the user with a URL for the generated video. The input is the URL of the cloud storage, and the output is the URL sent to the user's device. The server then notifies the smartphone or smart glasses of the generated URL, allowing the user to access the video.

[0399] Step 7:

[0400] Users watch videos and rate them through a rating input screen. The input is the user's rating information (e.g., a rating from 1 to 5 stars, comments), and the output is the rating data. The device accepts the user's rating and sends it to the server.

[0401] Step 8:

[0402] The server records the received rating data and updates the genre rankings. The input is the rating data received from users, and the output is the updated ranking information. The server saves the rating data in Firebase or MongoDB and automatically updates the rankings based on it.

[0403] Step 9:

[0404] The server creates and stores user profiles. The input is the user's selection history and rating data, and the output is a user profile. The server creates a profile based on the viewing history and rating information and stores it in a database.

[0405] Step 10:

[0406] The server will then suggest personalized content based on the saved user profile from the next time onwards. The input is the user profile, and the output is the suggested content information. The server analyzes the user profile, generates information to suggest appropriate video content, and provides it to the user.

[0407] Through these steps, users can efficiently create, view, and rate video content according to their preferences, improving the user experience.

[0408] Furthermore, an emotion engine that estimates the user's emotion may be combined. That is, the identification processing unit 290 may estimate the user's emotion using the emotion identification model 59 and perform identification processing using the user's emotion.

[0409] This system not only generates original video content based on the user's preferences, but also recognizes the user's emotions and proposes and generates content. Specific embodiments of each component of the system are described below.

[0410] User Interface Means

[0411] Device:

[0412] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening an application and entering their credentials on a login screen.

[0413] A means of identifying a genre or theme

[0414] server:

[0415] The genre and theme information input by the user or the emotion information recognized by the emotion engine means is analyzed to identify a genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the appropriate video content is identified.

[0416] Emotion Engine Means

[0417] Device:

[0418] The terminal captures the user's facial expressions and voice in real time using a built-in camera and microphone, and the emotion engine means analyzes the captured images. The emotion engine means recognizes emotions from the user's facial expressions and tone of voice, and transmits the information to a server.

[0419] server:

[0420] The emotion data sent from the emotion engine means is received and analyzed. Based on the analysis results, the system recommends genres and themes that are most suitable for the user.

[0421] A means of automatically generating video content

[0422] server:

[0423] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[0424] Means of preservation

[0425] server:

[0426] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[0427] Means of playback or distribution

[0428] Device:

[0429] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[0430] A means of accepting user ratings

[0431] Device:

[0432] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[0433] server:

[0434] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[0435] A way to update genre rankings

[0436] server:

[0437] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[0438] How to upload to a publishing platform

[0439] Device:

[0440] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[0441] server:

[0442] Uploaded video content is stored on the posting platform and made accessible to other users.

[0443] How to share with other users

[0444] server:

[0445] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[0446] A means to create and store user profiles

[0447] server:

[0448] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[0449] A means of making personalized content suggestions

[0450] server:

[0451] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[0452] The above is a detailed description of the system in the "Description of Embodiments."

[0453] The processing flow will be explained below.

[0454] Step 1:

[0455] Device:

[0456] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[0457] Step 2:

[0458] server:

[0459] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[0460] Step 3:

[0461] Device:

[0462] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[0463] Step 4:

[0464] Device:

[0465] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[0466] Step 5:

[0467] Device:

[0468] The user selects the Enable Emotion Engine option in the settings screen, which prepares the device's built-in camera and microphone to collect emotion data.

[0469] Step 6:

[0470] Device:

[0471] While watching a video or requesting a new video, the device captures the user's facial expressions and voice through the camera and microphone, analyzes the emotional data in real time, and temporarily stores this data.

[0472] Step 7:

[0473] server:

[0474] Emotion data transmitted from the terminal is received and analyzed, and the emotion engine means identifies the user's emotion (e.g., excitement, joy, sadness).

[0475] Step 8:

[0476] server:

[0477] Based on the analysis results, the system recommends genres and themes that suit the user's current emotions. For example, if the user is recognized as "excited," the system will identify "action" and "thriller" genres.

[0478] Step 9:

[0479] server:

[0480] Based on the user's selection and recognized emotions, the video generation AI module is activated to generate scenarios, construct scenes, and synthesize audio.

[0481] Specifically, it automatically generates footage including episodes of warriors fighting dragons.

[0482] Step 10:

[0483] server:

[0484] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[0485] Step 11:

[0486] server:

[0487] The terminal is notified that the video content has been saved.

[0488] Step 12:

[0489] Device:

[0490] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[0491] Step 13:

[0492] Device:

[0493] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[0494] Step 14:

[0495] User:

[0496] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[0497] Step 15:

[0498] Device:

[0499] After the user enters their rating and comments, the data is sent to the server.

[0500] Step 16:

[0501] server:

[0502] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[0503] Step 17:

[0504] User:

[0505] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[0506] Step 18:

[0507] Device:

[0508] The upload procedure for the video content is carried out and an upload request is sent to the server.

[0509] Step 19:

[0510] server:

[0511] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[0512] These are the specific processing steps for generating and sharing a video work based on the user's requests and emotions.

[0513] Example 2

[0514] Next, a description will be given of Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the smart device 14 will be referred to as a "terminal."

[0515] Conventional video content generation systems have had difficulty generating video content that fully reflects a user's preferences and momentary emotions. Furthermore, sharing the generated content with other users is time-consuming, and it is difficult to reflect the user's ratings in the system. Furthermore, they lacked functionality for making personalized suggestions based on user profiles, leaving a need for an improved user experience.

[0516] The identification process by the identification processing unit 290 of the data processing device 12 in Example 2 is realized by the following means. In this invention, the server includes an information terminal means for accepting user input, a data analysis means for identifying a preferred genre or theme based on the input, an emotion recognition means for capturing the user's facial expressions and voice in real time using a built-in camera and microphone and analyzing their emotions, a means including a generative AI model for automatically generating video content based on the genre, theme, and emotion data, a data storage means for saving the generated video content, a playback / distribution means for playing or distributing the saved video content to users, an evaluation input means for accepting user ratings, and a ranking update means for updating genre rankings based on the ratings. This enables the generation and sharing of personalized video content that reflects the user's preferences and emotions, and further enables appropriate suggestions based on the user profile.

[0517] The "information terminal means for accepting user input" is a means for providing an interface for the user to operate or give instructions to the system.

[0518] The "data analysis means for identifying a preferred genre or theme based on the input" is a means for analyzing the information input by the user and identifying a genre or theme according to the user's preferences.

[0519] "Emotion recognition means that uses a built-in camera and microphone to capture the user's facial expressions and voice in real time and analyze their emotions" refers to a means of recognizing the user's emotional state by using the camera and microphone attached to the device to photograph and record the user's facial expressions and voice, and analyzing that data.

[0520] "Means including a generative AI model that automatically generates video content based on the genre, theme, and emotional data" refers to means including an artificial intelligence model that receives a specified genre, theme, and emotional data as input and generates video content based on that input.

[0521] The "data storage means for storing the generated video content" is a means for temporarily storing the generated video content and for long-term storage as required.

[0522] "Playback and distribution means for playing or distributing the stored video content to users" refers to means for playing or distributing the stored video content in a form that can be viewed by users.

[0523] The "evaluation input means for accepting user evaluations" is a means for accepting and recording user evaluations and feedback on video content.

[0524] The "ranking update means for updating the genre rankings based on the ratings" is a means for updating the rankings of video content by genre to the latest status based on the rating data received from users.

[0525] This system generates original video content based on the user's preferences, and can also suggest and generate content by recognizing the user's emotions. Specific embodiments of each component of the system will be described below.

[0526] User Interface Means

[0527] Device:

[0528] A device provides an interface for users to enter input into the system. Users can access their account by opening an application and entering their credentials on a login screen. Devices include smartphones, tablets, and PCs.

[0529] A means of identifying a genre or theme

[0530] server:

[0531] The server analyzes the genre and theme information entered by the user or the emotion information recognized by the emotion recognition means to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the server identifies the appropriate video content.

[0532] emotion recognition means

[0533] Device:

[0534] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine. The device recognizes emotions from the user's facial expressions and tone of voice, and sends the information to a server.

[0535] server:

[0536] The server receives and analyzes the emotion data sent from the emotion recognition means, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[0537] A means of automatically generating video content

[0538] server:

[0539] The server then activates the generative AI model to automatically generate video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion recognition engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[0540] Means of preservation

[0541] server:

[0542] The server temporarily stores the generated video content, then saves it in the user's dedicated folder. Information about the saved video content is recorded in a database.

[0543] Means of playback or distribution

[0544] Device:

[0545] The generated video content can be streamed from the server to the terminal or downloaded directly by the user for viewing.

[0546] A means of accepting user ratings

[0547] Device:

[0548] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[0549] server:

[0550] The server records and analyzes the received ratings and comments in a database, which then updates the genre rankings.

[0551] A way to update genre rankings

[0552] server:

[0553] The server updates the rankings of video content in each genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[0554] How to upload to a publishing platform

[0555] Device:

[0556] The device provides a button for uploading user-generated video content to a posting platform, allowing users to share their creations with others.

[0557] server:

[0558] The server stores the uploaded video content on the posting platform and makes it accessible to other users.

[0559] How to share with other users

[0560] server:

[0561] The server manages the posted video content so that other users can view and rate it, and also generates a sharing link for users to use on social media or by email, making it easy to share.

[0562] A means to create and store user profiles

[0563] server:

[0564] The server creates and stores a profile based on the user's preferences and viewing history, which enables personalized video content to be suggested the next time the user uses the service.

[0565] A means of making personalized content suggestions

[0566] server:

[0567] The server then suggests the most suitable video content for the user based on the stored user profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[0568] Examples and prompts

[0569] Examples:

[0570] User Action: The user logs in, selects the genres "Action" and "Sci-Fi," and has a happy look on their face.

[0571] Server processing: The server receives the emotion data of "fun" from the emotion engine and automatically generates video content that includes action and science fiction elements.

[0572] Playback: The generated video content is streamed to the device.

[0573] Generate AI model prompt:

[0574] "The user likes the 'action' and 'sci-fi' genres, and their current emotion is 'fun'. Based on this condition, please generate a sci-fi video that contains exciting action scenes."

[0575] Through such a system, it becomes possible to provide users with video content that suits their mood and preferences at the time.

[0576] The flow of the identification process in the second embodiment will be described with reference to FIG.

[0577] Step 1:

[0578] User: The user opens the application on their device and enters their authentication information (username and password) on the login screen. This input includes text data such as username and password.

[0579] Terminal: The terminal sends the entered authentication information to the server using a secure communication protocol (e.g. HTTPS).

[0580] Server: The server receives the authentication information and authenticates it by comparing the username and password against the user information in its database.

[0581] Result: If authentication is successful, the server generates a session ID for the user and sends it to the terminal. The terminal displays the session ID and notifies the user that the login was successful.

[0582] Step 2:

[0583] User: After logging in, the user selects their preferred genre and theme from the displayed interface. This input includes text data such as "action" and "fantasy" as genre and theme.

[0584] Device: The device sends the selected information to the server, which is sent in JSON format.

[0585] Server: The server analyzes the received genre and theme information and stores it in a database, which is then used for further processing.

[0586] Step 3:

[0587] Device: The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time. This data includes image data and audio data.

[0588] Device: Passes the captured data to the emotion recognition engine.

[0589] Server: The server receives and analyzes the emotion data sent from the emotion recognition engine. Image processing algorithms and voice recognition algorithms are used for this analysis. The analysis result is the user's emotional state (e.g., "excited").

[0590] Step 4:

[0591] Server: The server takes genre, theme, and emotion data as input and runs the generative AI model. This input includes text data and emotion data.

[0592] Generative AI model: The generative AI model automatically generates video content based on this data. The output is a video content file.

[0593] Specific action: Send a prompt to the generative AI model saying, "The user likes the 'action' and 'fantasy' genres, and their current emotion is 'excitement.' Based on these conditions, please generate a video that includes exciting action scenes."

[0594] Step 5:

[0595] Server: The server temporarily stores the generated video content and then saves it in the user's dedicated folder. The server records the path and metadata of the saved file (e.g., creation date and time, user ID) in the database.

[0596] Device: A notification will appear on your device when the save is complete.

[0597] Step 6:

[0598] Device: The user plays or downloads the stored video content. When the user clicks the "Play" button, the device receives the video stream from the server and displays it in the video player. When the user clicks the "Download" button, the device saves the video file to local storage.

[0599] Step 7:

[0600] Terminal: After watching the video, an interface is displayed for the user to input a rating. The user can input a rating from 1 to 5 stars and a comment. This input includes numerical data and text data.

[0601] Device: The device sends the evaluation data to the server.

[0602] Server: The server receives the rating data and records it in a database. This data is used to update the rankings.

[0603] Step 8:

[0604] Server: The server updates the rankings of video content in each genre based on the received rating data. New rankings are generated and stored in the database.

[0605] On the device: The latest information is displayed so that users can view ranking information.

[0606] Step 9:

[0607] Terminal: Provides a button to upload user-generated video content to a posting platform. The user clicks the "Upload" button and selects the posting platform.

[0608] Server: The server receives the video file and transfers it to the designated uploading platform. Once uploaded, a public link is generated.

[0609] On your device: The user will see a notification that the upload is complete and a public link.

[0610] Step 10:

[0611] Server: Manages the posted video content so that other users can view and rate it. Also generates sharing links for social media and email, allowing users to easily share. Shared links are generated and sent to social media and email.

[0612] Step 11:

[0613] Server: The server creates and stores a profile based on the user's preferences and viewing history, including genre preferences and viewing history.

[0614] Specific operation: Collects user viewing history and genre preference information and records it in a database as a profile.

[0615] Step 12:

[0616] Server: The server uses a recommendation algorithm to suggest the most suitable video content for a user based on the stored user profile.

[0617] On the device: A screen will appear where the user can view a list of suggested content.

[0618] (Application example 2)

[0619] Next, a description will be given of Application Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the smart device 14 will be referred to as a "terminal."

[0620] Current content distribution services struggle to automatically generate and deliver personalized video content based on users' interests and emotions. Because they are unable to generate or suggest content that reflects a user's temporary emotional state in real time, it is difficult to provide users with an optimal entertainment experience. Furthermore, there is a lack of systems that effectively utilize user ratings as feedback and reflect them in rankings. Technology that can solve these problems and provide a more personalized user experience is needed.

[0621] The identification process by the identification processing unit 290 of the data processing device 12 in Application Example 2 is realized by the following means. In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre-specific rankings based on the ratings, a means for recognizing emotions from the user's facial expressions and voice, and a means for using a generative AI model for generating video content based on the recognized emotion information. This enables the generation and distribution of personalized video content based on the user's emotions, and also enables the updating of genre-specific rankings to reflect the user's ratings.

[0622] The "user interface means" is a means for providing an interface for a user to input to the system.

[0623] The "means for identifying a genre or theme" is a means for identifying a genre or theme that the user likes based on the user's input or emotional information.

[0624] The "means for automatically generating video content" is a means for automatically generating video content based on a genre or theme selected by a user or a recognized emotion.

[0625] The "means for saving" is a means for temporarily saving the generated video content and then saving it in a location accessible to the user.

[0626] "Means for playing or distributing" refers to means for allowing users to stream or download stored video content.

[0627] The "means for receiving a rating" is a means for allowing a user to input a rating for the video content that the user has viewed, and for the rating to be received within the system.

[0628] The "means for updating the genre rankings" is a means for updating the rankings of video content in each genre based on user ratings.

[0629] The "means for recognizing emotions" refers to a means for recognizing emotions from the user's facial expressions and tone of voice and analyzing that information.

[0630] A "generative AI model" is an artificial intelligence model used to generate video content based on input emotional information and prompts.

[0631] This invention relates to a system that automatically generates and distributes original video content based on user input and emotional information. The system includes a user interface, a means for identifying genres and themes, a means for recognizing emotions, a means for using a generative AI model, a means for automatically generating video content, a means for saving, a means for playing or distributing, a means for accepting ratings, and a means for updating genre-specific rankings. Each component of this system is described in detail below.

[0632] User Interface Means

[0633] Device:

[0634] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening the smartphone application and entering their credentials on the login screen.

[0635] A means of identifying a genre or theme

[0636] server:

[0637] The system analyzes the genre and theme information entered by the user, or the emotional information acquired by the emotion recognition means, to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," it will identify video content that matches that.

[0638] A means of recognizing emotions

[0639] Device:

[0640] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine, which recognizes emotions from the user's facial expressions and tone of voice and sends the information to a server.

[0641] server:

[0642] It receives and analyzes the emotional data sent from the emotion engine, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[0643] A means of automatically generating video content

[0644] server:

[0645] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[0646] Means of preservation

[0647] server:

[0648] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[0649] Means of playback or distribution

[0650] Device:

[0651] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[0652] A means of accepting user ratings

[0653] Device:

[0654] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[0655] server:

[0656] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[0657] A way to update genre rankings

[0658] server:

[0659] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[0660] Specific examples

[0661] For example, if a user expresses the emotion "happy," the emotion recognition engine sends that emotion to the server, which then sends prompts like "comedy" or "adventure" to the generative AI model. In this way, video content that best suits the user's current emotional state is generated and delivered.

[0662] Example prompt sentence:

[0663] "Emotion: Happy, Genre: Comedy, Adventure"

[0664] The above system enables the generation and distribution of personalized video content based on the user's emotions, improving the user's entertainment experience.

[0665] The flow of the specific processing in the application example 2 will be described with reference to FIG.

[0666] Step 1:

[0667] The terminal accepts the user's login. The user enters authentication information (user name and password) using the user interface and sends it to the server. Based on this input data, the server authenticates the user, and if authentication is successful, obtains the user's profile information and sends it to the terminal.

[0668] Step 2:

[0669] The terminal accepts genre and theme selections from the user. The user uses the interface to select their preferred genre (e.g., action, comedy) or theme (e.g., fantasy, suspense), and sends the input data to the server. The server then analyzes the user's preferences based on this data.

[0670] Step 3:

[0671] The device captures the user's facial expressions and voice using a built-in camera and microphone. The captured data is processed to recognize emotions. Specifically, facial expression data is analyzed using OpenCV, and voice data is processed using TensorFlow / Keras. The recognized emotion data is then sent to a server.

[0672] Step 4:

[0673] The server generates a prompt based on the received genre and theme selection data and emotional data. This prompt includes information such as "Emotion: happy, Genre: comedy, adventure." This prompt is then input into a generative AI model to automatically generate video content.

[0674] Step 5:

[0675] The server temporarily stores the generated video content and stores it in a dedicated folder for each user. Metadata related to the content storage is also recorded in the database.

[0676] Step 6:

[0677] The terminal allows the playback or download of video content stored by the server, and for the user to view the content, streaming playback is initiated or the content is downloaded to the terminal.

[0678] Step 7:

[0679] The terminal provides an interface for users to input ratings after viewing. Users input ratings (1 to 5 stars) and text comments, and then send the input data to the server. The server records the ratings and comments in a database.

[0680] Step 8:

[0681] The server updates the genre rankings based on the received rating data. In the ranking update process, it calculates the average rating score for each piece of content and generates new ranking information. This ranking information is saved in a database and made public to other users.

[0682] This enables the generation and distribution of personalized video content based on user emotions and the updating of rankings based on user ratings.

[0683] The specific processing unit 290 transmits the result of the specific processing to the smart device 14. In the smart device 14, the control unit 46A causes the output device 40 to output the result of the specific processing. The microphone 38B acquires audio indicating a user input regarding the result of the specific processing. The control unit 46A transmits audio data indicating the user input acquired by the microphone 38B to the data processing device 12. In the data processing device 12, the specific processing unit 290 acquires the audio data.

[0684] The data generation model 58 is a so-called generative AI (Artificial Intelligence). An example of the data generation model 58 is ChatGPT (registered trademark) (Internet search engine).<URL: https: / / openai.com / blog / chatgpt> ), Gemini (registered trademark) (Internet search <url: https: gemini.google.com ?hl="ja">) and other generation AIs. The data generation model 58 is obtained by performing deep learning on a neural network. A prompt including an instruction is input to the data generation model 58, and inference data such as voice data indicating voice, text data indicating text, and image data indicating an image is also input. The data generation model 58 performs inference on the input inference data in accordance with the instruction indicated by the prompt, and outputs the inference result in a data format such as voice data and text data. Here, inference refers to, for example, analysis, classification, prediction, and / or summarization.

[0685] In the above embodiment, an example in which the specific process is performed by the data processing device 12 has been given, but the technology of the present disclosure is not limited to this, and the specific process may be performed by the smart device 14.

[0686] [Second embodiment]

[0687] FIG. 3 shows an example of the configuration of a data processing system 210 according to the second embodiment.

[0688] 3, the data processing system 210 includes the data processing device 12 and smart glasses 214. An example of the data processing device 12 is a server.

[0689] The data processing device 12 includes a computer 22, a database 24, and a communication I / F 26. The computer 22 is an example of a "computer" according to the technology of the present disclosure. The computer 22 includes a processor 28, a RAM 30, and a storage 32. The processor 28, the RAM 30, and the storage 32 are connected to a bus 34. The database 24 and the communication I / F 26 are also connected to the bus 34. The communication I / F 26 is connected to a network 54. Examples of the network 54 include a WAN (Wide Area Network) and / or a LAN (Local Area Network).

[0690] The smart glasses 214 include a computer 36, a microphone 238, a speaker 240, a camera 42, and a communication I / F 44. The computer 36 includes a processor 46, a RAM 48, and a storage 50. The processor 46, the RAM 48, and the storage 50 are connected to a bus 52. The microphone 238, the speaker 240, and the camera 42 are also connected to the bus 52.

[0691] The microphone 238 receives instructions and the like from the user 20 by receiving voice uttered by the user 20. The microphone 238 captures the voice uttered by the user 20, converts the captured voice into audio data, and outputs it to the processor 46. The speaker 240 outputs audio in accordance with instructions from the processor 46.

[0692] Camera 42 is a small digital camera equipped with an optical system including a lens, aperture, and shutter, and an imaging element such as a CMOS (Complementary Metal-Oxide-Semiconductor) image sensor or a CCD (Charge Coupled Device) image sensor, and captures images of the surroundings of user 20 (for example, an imaging range defined by an angle of view equivalent to the field of vision of a typical healthy person).

[0693] The communication I / F 44 is connected to a network 54. The communication I / Fs 44 and 26 are responsible for the exchange of various information between the processor 46 and the processor 28 via the network 54. The exchange of various information between the processor 46 and the processor 28 using the communication I / Fs 44 and 26 is carried out in a secure state.

[0694] Fig. 4 shows an example of the main functions of the data processing device 12 and the smart glasses 214. As shown in Fig. 4, in the data processing device 12, a specific process is performed by the processor 28. A specific process program 56 is stored in the storage 32.

[0695] The specific processing program 56 is an example of a "program" according to the technology of the present disclosure. The processor 28 reads the specific processing program 56 from the storage 32 and executes the read specific processing program 56 on the RAM 30. The specific processing is realized by the processor 28 operating as a specific processing unit 290 in accordance with the specific processing program 56 executed on the RAM 30.

[0696] The storage 32 stores a data generation model 58 and an emotion identification model 59. The data generation model 58 and the emotion identification model 59 are used by the identification processing unit 290.

[0697] In the smart glasses 214, the reception output process is performed by the processor 46. A reception output program 60 is stored in the storage 50. The processor 46 reads the reception output program 60 from the storage 50 and executes the read reception output program 60 on the RAM 48. The reception output process is realized by the processor 46 operating as the control unit 46A in accordance with the reception output program 60 executed on the RAM 48.

[0698] Next, a description will be given of the identification process performed by the identification processing unit 290 of the data processing device 12. In the following description, the data processing device 12 will be referred to as the "server" and the smart glasses 214 will be referred to as the "terminal."

[0699] This system allows users to create original video content based on their preferences, share it with other users, and receive ratings. Specific embodiments of each component of the system will be described below.

[0700] User Interface Means

[0701] Device:

[0702] The terminal provides the interface through which the user enters input into the system. The user can access their account by opening the application and entering their credentials on the login screen. On first use, the user is prompted to select their preferred genre and theme.

[0703] A means of identifying a genre or theme

[0704] server:

[0705] The system receives and analyzes information about genres and themes entered by the user to identify genres and themes based on the user's preferences. For example, if the user selects "action" and "fantasy," it sets parameters for generating video content related to these genres.

[0706] A means of automatically generating video content

[0707] server:

[0708] The video generation AI module is activated and automatically generates video content based on the genre and theme selected by the user. The AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if the user selects "action + fantasy," a video containing an episode of a warrior fighting a dragon will be generated.

[0709] Means of preservation

[0710] server:

[0711] The generated video content is temporarily saved and then saved in a dedicated folder for the user, allowing the user to access the generated video at any time.

[0712] Means of playback or distribution

[0713] Device:

[0714] The generated video content can be streamed from the server to the device or downloaded directly to the device, allowing users to watch the video in real time on their own devices.

[0715] A means of accepting user ratings

[0716] Device:

[0717] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[0718] server:

[0719] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[0720] A way to update genre rankings

[0721] server:

[0722] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[0723] How to upload to a publishing platform

[0724] Device:

[0725] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[0726] server:

[0727] Store uploaded video content on the platform and make it accessible to other users.

[0728] How to share with other users

[0729] server:

[0730] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[0731] A means to create and store user profiles

[0732] server:

[0733] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[0734] A means of making personalized content suggestions

[0735] server:

[0736] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[0737] The above is a detailed description of the system in the "Description of Embodiments."

[0738] The processing flow will be explained below.

[0739] Step 1:

[0740] Device:

[0741] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[0742] Step 2:

[0743] server:

[0744] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[0745] Step 3:

[0746] Device:

[0747] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[0748] Step 4:

[0749] Device:

[0750] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[0751] Step 5:

[0752] server:

[0753] The received request is analyzed and the necessary parameters are set for the AI ​​model, for example, setting video generation based on "action" and "fantasy."

[0754] Step 6:

[0755] server:

[0756] The video generation AI module is activated to generate a scenario, construct a scene, and synthesize audio. Specifically, it automatically generates a video containing an episode of a warrior fighting a dragon.

[0757] Step 7:

[0758] server:

[0759] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[0760] Step 8:

[0761] server:

[0762] The terminal is notified that the video content has been saved.

[0763] Step 9:

[0764] Device:

[0765] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[0766] Step 10:

[0767] Device:

[0768] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[0769] Step 11:

[0770] User:

[0771] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[0772] Step 12:

[0773] Device:

[0774] After the user enters their rating and comments, the data is sent to the server.

[0775] Step 13:

[0776] server:

[0777] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[0778] Step 14:

[0779] User:

[0780] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[0781] Step 15:

[0782] Device:

[0783] The upload procedure for the video content is carried out and an upload request is sent to the server.

[0784] Step 16:

[0785] server:

[0786] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[0787] These are the specific processing steps involved in generating and sharing a video work based on a user request.

[0788] Example 1

[0789] Next, a description will be given of Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the smart glasses 214 will be referred to as a "terminal."

[0790] In the creation and sharing of digital media content, personalized content delivery based on user preferences is required, but existing systems do not effectively generate, evaluate, and share content that reflects user preferences. Furthermore, the accuracy of recommendation systems and ranking updates based on user ratings also needs to be improved.

[0791] The specific processing by the specific processing unit 290 of the data processing device 12 in the first embodiment is realized by the following means.

[0792] In this invention, the server includes a user interface means for accepting user input, a means for identifying preferred categories or topics based on the input, a means for automatically generating digital media content based on the categories or topics, a means for saving the generated digital media content, a means for playing or distributing the saved digital media content to the user, a means for accepting user ratings, and a means for updating category rankings based on the ratings, thereby enabling efficient generation, rating, sharing, and ranking updates of personalized digital media content based on user preferences.

[0793] The "user interface means" is a device or program that provides an interface for a user to input information to the system.

[0794] A "category" is a concept that refers to a genre or topic of digital media content selected by a user.

[0795] A "topic" is a concept that refers to a specific theme or subject of digital media content selected by a user.

[0796] "Digital media content" refers to media data in digital format, such as video, audio, and images, that is generated based on user preferences.

[0797] An "automatic generation means" is a device or program that automatically creates digital media content based on category or topic information entered by a user.

[0798] A "storage means" is a device or program for temporary and long-term storage of the generated digital media content.

[0799] A "delivery means" is a device or program that provides stored digital media content to a user's terminal in a streaming or downloadable format.

[0800] The "means for receiving ratings" is a device or program that collects rating information on digital media content from users.

[0801] The "means for updating the ranking by category" is a device or program that updates the ranking of digital media content for each category based on evaluation information collected from users.

[0802] A "sharing platform" is an online service or website for sharing user-generated digital media content with other users.

[0803] MODE FOR CARRYING OUT THE INVENTION

[0804] The system automates the creation, storage, distribution, rating, and sharing of digital media content. It uses user input to identify preferred categories and topics, and then uses generative AI models to automatically generate content. A specific implementation of this system is described below.

[0805] 1. User Interface Methods

[0806] (Terminal)

[0807] The terminal provides an application as a user interface for the user to input data to the system. The user accesses the system by starting this application and entering a user name and password on the login screen.

[0808] 2. How to identify your favorite categories and topics

[0809] (Terminal)

[0810] On first use, the device will display an interface for setting user preferences, where users can select their preferred categories and topics, for example, "action" and "fantasy."

[0811] (server)

[0812] The server receives and analyzes this selection information and stores it in a database. This data is used to generate content using the AI ​​model described below.

[0813] 3. Content Generation Methods

[0814] (server)

[0815] The server generates prompts based on the user's preferences and sends them to a generative AI model (e.g., OpenAI GPT-4, DeepMind's video generation model).

[0816] 4. Video content generation

[0817] (server)

[0818] The server's generative AI model automatically generates video content based on the input prompt. For example, for a user who likes "action" and "fantasy," a video containing a story about a warrior fighting a dragon is generated. Video generation involves the processes of scenario generation, scene construction, and voice synthesis.

[0819] 5. How content is stored

[0820] (server)

[0821] The server temporarily stores the generated video content and then saves it in the user's dedicated folder. The video files are managed using a database management system (e.g., MySQL, PostgreSQL).

[0822] 6. Content playback and distribution methods

[0823] (Terminal)

[0824] The device allows users to play or download stored video content, often using streaming technology and a media player (e.g., VLC, QuickTime).

[0825] 7. Method of accepting evaluations

[0826] (Terminal)

[0827] After watching the video, the device displays a user rating input interface, allowing users to enter star ratings and comments.

[0828] (server)

[0829] The server receives user ratings and stores them in a database, which is used to update rankings for each category.

[0830] 8. How to update rankings by category

[0831] (server)

[0832] The server updates the rankings for each category based on the collected rating information, and then makes this ranking information available to other users, who can then recommend highly rated content.

[0833] 9. How to share content

[0834] (Terminal)

[0835] The terminal provides an interface for uploading the generated digital media content to a sharing platform.

[0836] (server)

[0837] The server manages the uploaded content and makes it accessible to other users, generating shareable links for easy sharing via social media or email.

[0838] 10. How to create and store user profiles

[0839] (server)

[0840] A profile is created and saved based on the user's preferences and viewing history, allowing for personalized content suggestions the next time the user uses the service.

[0841] 11. Means of personalized content recommendations

[0842] (server)

[0843] Based on a saved user profile, the app suggests digital media content that is most suitable for the user, for example, if a user previously indicated that they like "action," the app will suggest the latest action movies.

[0844] Examples of concrete examples and prompts

[0845] Examples:

[0846] The user selects "action" and "fantasy" and enters the prompt "Generate a movie about a warrior fighting a dragon." Based on this prompt, the server's generative AI model automatically generates video content and saves it in the user's dedicated folder. This video content can be viewed on the device and shared with other users.

[0847] Example prompt sentence:

[0848] "Generate footage of action scenes in fantasy movies"

[0849] "Suggest your next watchlist based on your favorite movie genres."

[0850] This system allows users to easily create, save, rate, and share content tailored to their preferences. This specific embodiment will clarify how the invention actually works.

[0851] The flow of the identification process in the first embodiment will be described with reference to FIG.

[0852] Step 1:

[0853] (User):

[0854] The user launches the application and enters their username and password on the login screen.

[0855] input:

[0856] Username, Password

[0857] output:

[0858] Authentication Request

[0859] (Device):

[0860] The terminal accepts user input and sends an authentication request to the server.

[0861] Specific behavior:

[0862] Displaying the login screen

[0863] Obtaining the entered username and password

[0864] Sending an authentication request to the server

[0865] Step 2:

[0866] (server):

[0867] The server checks the received authentication request against its database and returns the authentication result.

[0868] input:

[0869] Authentication Request

[0870] output:

[0871] Authentication result (success / failure)

[0872] (Device):

[0873] The terminal performs screen transitions based on the authentication results.

[0874] Specific behavior:

[0875] Receiving the authentication result

[0876] If authentication is successful, the main screen is displayed. If authentication fails, an error message is displayed.

[0877] Step 3:

[0878] (User):

[0879] On first use, users make selections in an interface that allows them to choose their preferred categories and topics.

[0880] input:

[0881] Category selection (e.g., "Action"), Topic selection (e.g., "Fantasy")

[0882] output:

[0883] Selected Data

[0884] (Device):

[0885] The terminal transmits the selected data to the server.

[0886] Specific behavior:

[0887] Receiving user selections

[0888] Sending Selected Data to the Server

[0889] Step 4:

[0890] (server):

[0891] The server stores the received selection data in a database.

[0892] input:

[0893] Selected Data

[0894] output:

[0895] Results saved in the database

[0896] Specific behavior:

[0897] Analysis of selected data

[0898] Saving to the database

[0899] Step 5:

[0900] (User):

[0901] The user inputs the content of the video they want to generate on the prompt sentence input screen.

[0902] input:

[0903] Prompt text (e.g., "Generate a movie of a warrior fighting a dragon")

[0904] output:

[0905] Prompt Data

[0906] (Device):

[0907] The terminal sends the prompt data to the server.

[0908] Specific behavior:

[0909] Receiving a prompt

[0910] Sending prompt data to the server

[0911] Step 6:

[0912] (server):

[0913] The server inputs the prompt data into the generative AI model and begins generating video content.

[0914] input:

[0915] Prompt Data

[0916] output:

[0917] Generated video content

[0918] Specific behavior:

[0919] Parsing prompt data

[0920] Input to generative AI models

[0921] Video content generation

[0922] Step 7:

[0923] (server):

[0924] The server temporarily stores the generated video content, and then stores it in a folder dedicated to the user.

[0925] input:

[0926] Generated video content

[0927] output:

[0928] Saved results in user folder

[0929] Specific behavior:

[0930] Temporary storage of video content

[0931] Move to user-specific folder

[0932] Step 8:

[0933] (Device):

[0934] The user plays or downloads the stored video content.

[0935] input:

[0936] User requests (playback, download)

[0937] output:

[0938] Playback and download of video content

[0939] Specific behavior:

[0940] Click the play button to stream

[0941] Click the download button to save the video to your device

[0942] Step 9:

[0943] (User):

[0944] After viewing the video, the user inputs a rating.

[0945] input:

[0946] Rating data (stars, comments)

[0947] output:

[0948] Evaluation data submission

[0949] (Device):

[0950] The terminal transmits the evaluation data to the server.

[0951] Specific behavior:

[0952] View the rating interface

[0953] Receiving and sending evaluation data to the server

[0954] Step 10:

[0955] (server):

[0956] The server stores the received rating data in a database and updates the rankings by category.

[0957] input:

[0958] Evaluation Data

[0959] output:

[0960] Updated ranking data

[0961] Specific behavior:

[0962] Analysis of evaluation data

[0963] Saving to a database

[0964] Category ranking updates

[0965] Step 11:

[0966] (User):

[0967] Users upload the generated video content to a sharing platform.

[0968] input:

[0969] Upload Request

[0970] output:

[0971] Upload completion notification

[0972] (Device):

[0973] Pressing the upload button sends the video content to the server.

[0974] Specific behavior:

[0975] Click the upload button

[0976] Sending video content to the server

[0977] Step 12:

[0978] (server):

[0979] The server manages the uploaded content, sets it up so that other users can access it, and generates a share link so that it can be shared via social media or email.

[0980] input:

[0981] Uploaded content

[0982] output:

[0983] Shared Links

[0984] Specific behavior:

[0985] Organizing uploaded content

[0986] Generate shareable links for social media and email

[0987] Step 13:

[0988] (server):

[0989] The server creates and stores a profile based on the user's preferences and viewing history.

[0990] input:

[0991] User preferences and viewing history

[0992] output:

[0993] User Profile

[0994] Specific behavior:

[0995] User data analysis

[0996] Creating and Saving a Profile

[0997] Step 14:

[0998] (server):

[0999] The server makes personalized content suggestions based on the stored profile.

[1000] input:

[1001] User Profile

[1002] output:

[1003] Suggested Content List

[1004] Specific behavior:

[1005] Viewing Profiles

[1006] Selection of proposed content

[1007] The above are the specific steps of the program processing of this system.

[1008] (Application example 1)

[1009] Next, a description will be given of Application Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the smart glasses 214 will be referred to as a "terminal."

[1010] Conventional video content generation and distribution systems lack the functionality to allow users to easily create original video content based on their preferences and share it with other users. Furthermore, the lack of personalized content suggestions and rating systems hinders the improvement of the user experience.

[1011] The specific processing by the specific processing unit 290 of the data processing device 12 in the application example 1 is realized by the following means.

[1012] In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre rankings based on the ratings, a means for creating and saving a user profile, a means for suggesting personalized content based on the profile, a means for a user to log in and select a genre using an application installed on a smartphone or smart glasses, a means for saving and sharing the generated video in cloud storage, and a means for generating video using an AI model by inputting a prompt. This allows users to easily generate video content according to their preferences, share it with other users, receive ratings, and improve the accuracy of content suggestions from the next time onwards.

[1013] The "user interface means" is an interface that allows a user to input information into the system, and is a function that allows operation through a smartphone or smart glasses, for example.

[1014] The "means for identifying a favorite genre or theme" is a function of the system that analyzes and identifies a user's favorite genre or theme based on information input by the user.

[1015] The "means for automatically generating video content" is an AI module for generating video content based on the genre or theme selected by the user.

[1016] The "means for saving the generated video content" is a function for temporarily or permanently storing the generated video in a recording device.

[1017] The "means for playing or distributing stored video content to users" is a function for providing stored video so that users can view it.

[1018] The "means for accepting user ratings" is a function including an interface for users to input ratings and comments on video content.

[1019] The "means for updating genre rankings" is a function for automatically updating genre rankings based on user ratings.

[1020] "Means for creating and saving user profiles" refers to a function that creates individual profiles based on the user's preferences and viewing history and saves them in a database.

[1021] The "means for making personalized content suggestions" is a function that suggests the most suitable video content to a user based on a saved user profile.

[1022] "Application installed on a smartphone or smart glasses" refers to software for a mobile device that allows a user to access the video content generation system.

[1023] "Means for saving and sharing the generated video in cloud storage" refers to a function for saving the generated video file to a remote server via the Internet and providing a sharing link from there.

[1024] "Means for generating images using an AI model by inputting a prompt sentence" refers to a function in which a user inputs a specific instruction sentence, and the AI ​​model generates an image based on that instruction.

[1025] This invention is a system that enables users to use mobile devices such as smartphones or smart glasses to create original video content based on their preferences, share it with other users, and receive ratings.

[1026] User Interface Means

[1027] The device is equipped with an interface that allows users to input information into the system. Users can access their account by opening the application on their smartphone or smart glasses and entering their authentication information on the login screen. On first use, a screen appears where users can select their preferred genre (e.g., "action" or "fantasy") and theme.

[1028] A means of identifying a genre or theme

[1029] The server receives the genre and theme information entered by the user, analyzes it, and identifies the user's preferences. For example, if the user selects "thriller" and "mystery," parameters for generating video content related to these genres are set.

[1030] A means of automatically generating video content

[1031] The server launches a video generation AI model and automatically generates video content based on the genre and theme selected by the user. The AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if a user selects "Thriller + Mystery," a short video of a detective solving a mysterious case will be generated. The AI ​​model used here utilizes machine learning frameworks such as TensorFlow and PyTorch.

[1032] Storage and playback methods

[1033] The server temporarily stores the generated video content in cloud storage such as Google Cloud Storage, and then saves it in the user's dedicated folder. Users can then view the video in real time on their smartphones or smart glasses.

[1034] How to receive ratings

[1035] The device provides an interface for users to input ratings after watching videos. Users can enter ratings from 1 to 5 stars and comments. The server records the received ratings and comments in a database such as Firebase and analyzes them to update the rankings by genre.

[1036] How to update the rankings

[1037] The server automatically updates the rankings of video content by genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1038] How to store and share content

[1039] The generated video content is stored in cloud storage and a sharing link is generated, allowing users to easily share their work via social media or email.

[1040] User profiles and personalized content suggestions

[1041] The server creates and stores a profile based on the user's preferences and viewing history. The next time the user logs in, personalized content suggestions are made based on this profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[1042] Examples of concrete examples and prompts

[1043] After a user logs in, they can select their preferred genre, create, watch, and rate a short video. The following prompts are used:

[1044] Input: "Generate a short film that mixes horror and fantasy genres. My Firebase UID is abcdef."

[1045] Output: "Generating movie... The URL for the generated movie is https: / / storage.googleapis.com / abcdef / generated_video.mp4."

[1046] This system allows users to easily create and enjoy video content according to their preferences.

[1047] The flow of the specific processing in the application example 1 will be described with reference to FIG.

[1048] Step 1:

[1049] A user launches an application on their smartphone or smart glasses and accesses the login screen. The input is the user's authentication information, and the output is the authentication success / failure result. The server validates the user's authentication information using Firebase Auth and returns the user ID if authentication is successful.

[1050] Step 2:

[1051] After logging in, the user selects their preferred genre and theme. The input is the genre and theme selected by the user, and the output is the data of the selected genre and theme. The terminal sends the selected genre and theme to the server.

[1052] Step 3:

[1053] The server analyzes the received genre and theme information and sets the parameters of the AI ​​model. The input is genre and theme data, and the output is the parameter settings of the AI ​​model. The server uses TensorFlow or PyTorch to initialize the AI ​​model and set the settings based on the received data.

[1054] Step 4:

[1055] The server launches the configured AI model and generates video content. The input is the AI ​​model's parameters and generation instructions (prompt text), and the output is the generated video file. The AI ​​model generates a scenario, constructs scenes, and synthesizes audio. In this step, if the user inputs, for example, "Please generate a short film that combines the horror and fantasy genres," a video of the corresponding scenario will be automatically generated.

[1056] Step 5:

[1057] The server saves the generated video file in cloud storage (e.g., Google Cloud Storage). The input is the generated video file, and the output is the cloud storage URL. The server uploads the video file to cloud storage and generates a URL link for the saved file.

[1058] Step 6:

[1059] The server provides the user with a URL for the generated video. The input is the URL of the cloud storage, and the output is the URL sent to the user's device. The server then notifies the smartphone or smart glasses of the generated URL, allowing the user to access the video.

[1060] Step 7:

[1061] Users watch videos and rate them through a rating input screen. The input is the user's rating information (e.g., a rating from 1 to 5 stars, comments), and the output is the rating data. The device accepts the user's rating and sends it to the server.

[1062] Step 8:

[1063] The server records the received rating data and updates the genre rankings. The input is the rating data received from users, and the output is the updated ranking information. The server saves the rating data in Firebase or MongoDB and automatically updates the rankings based on it.

[1064] Step 9:

[1065] The server creates and stores user profiles. The input is the user's selection history and rating data, and the output is a user profile. The server creates a profile based on the viewing history and rating information and stores it in a database.

[1066] Step 10:

[1067] The server will then suggest personalized content based on the saved user profile from the next time onwards. The input is the user profile, and the output is the suggested content information. The server analyzes the user profile, generates information to suggest appropriate video content, and provides it to the user.

[1068] Through these steps, users can efficiently create, view, and rate video content according to their preferences, improving the user experience.

[1069] Furthermore, an emotion engine that estimates the user's emotion may be further combined. That is, the identification processing unit 290 may estimate the user's emotion using the emotion identification model 59, and perform identification processing using the user's emotion.

[1070] This system not only generates original video content based on the user's preferences, but also recognizes the user's emotions and proposes and generates content. Specific embodiments of each component of the system are described below.

[1071] User Interface Means

[1072] Device:

[1073] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening an application and entering their credentials on a login screen.

[1074] A means of identifying a genre or theme

[1075] server:

[1076] The genre and theme information input by the user or the emotion information recognized by the emotion engine means is analyzed to identify a genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the appropriate video content is identified.

[1077] Emotion Engine Means

[1078] Device:

[1079] The terminal captures the user's facial expressions and voice in real time using a built-in camera and microphone, and the emotion engine means analyzes the captured images. The emotion engine means recognizes emotions from the user's facial expressions and tone of voice, and transmits the information to a server.

[1080] server:

[1081] The emotion data sent from the emotion engine means is received and analyzed. Based on the analysis results, the system recommends genres and themes that are most suitable for the user.

[1082] A means of automatically generating video content

[1083] server:

[1084] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[1085] Means of preservation

[1086] server:

[1087] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1088] Means of playback or distribution

[1089] Device:

[1090] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[1091] A means of accepting user ratings

[1092] Device:

[1093] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[1094] server:

[1095] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[1096] A way to update genre rankings

[1097] server:

[1098] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1099] How to upload to a publishing platform

[1100] Device:

[1101] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[1102] server:

[1103] Uploaded video content is stored on the posting platform and made accessible to other users.

[1104] How to share with other users

[1105] server:

[1106] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[1107] A means to create and store user profiles

[1108] server:

[1109] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[1110] A means of making personalized content suggestions

[1111] server:

[1112] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[1113] The above is a detailed description of the system in the "Description of Embodiments."

[1114] The processing flow will be explained below.

[1115] Step 1:

[1116] Device:

[1117] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[1118] Step 2:

[1119] server:

[1120] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[1121] Step 3:

[1122] Device:

[1123] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[1124] Step 4:

[1125] Device:

[1126] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[1127] Step 5:

[1128] Device:

[1129] The user selects the Enable Emotion Engine option in the settings screen, which prepares the device's built-in camera and microphone to collect emotion data.

[1130] Step 6:

[1131] Device:

[1132] While watching a video or requesting a new video, the device captures the user's facial expressions and voice through the camera and microphone, analyzes the emotional data in real time, and temporarily stores this data.

[1133] Step 7:

[1134] server:

[1135] Emotion data transmitted from the terminal is received and analyzed, and the emotion engine means identifies the user's emotion (e.g., excitement, joy, sadness).

[1136] Step 8:

[1137] server:

[1138] Based on the analysis results, the system recommends genres and themes that suit the user's current emotions. For example, if the user is recognized as "excited," the system will identify "action" and "thriller" genres.

[1139] Step 9:

[1140] server:

[1141] Based on the user's selection and recognized emotions, the video generation AI module is activated to generate scenarios, construct scenes, and synthesize audio.

[1142] Specifically, it automatically generates footage including episodes of warriors fighting dragons.

[1143] Step 10:

[1144] server:

[1145] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1146] Step 11:

[1147] server:

[1148] The terminal is notified that the video content has been saved.

[1149] Step 12:

[1150] Device:

[1151] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[1152] Step 13:

[1153] Device:

[1154] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[1155] Step 14:

[1156] User:

[1157] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[1158] Step 15:

[1159] Device:

[1160] After the user enters their rating and comments, the data is sent to the server.

[1161] Step 16:

[1162] server:

[1163] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[1164] Step 17:

[1165] User:

[1166] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[1167] Step 18:

[1168] Device:

[1169] The upload procedure for the video content is carried out and an upload request is sent to the server.

[1170] Step 19:

[1171] server:

[1172] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[1173] These are the specific processing steps for generating and sharing a video work based on the user's requests and emotions.

[1174] Example 2

[1175] Next, a description will be given of Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the smart glasses 214 will be referred to as a "terminal."

[1176] Conventional video content generation systems have had difficulty generating video content that fully reflects a user's preferences and momentary emotions. Furthermore, sharing the generated content with other users is time-consuming, and it is difficult to reflect the user's ratings in the system. Furthermore, they lacked functionality for making personalized suggestions based on user profiles, leaving a need for an improved user experience.

[1177] The identification process by the identification processing unit 290 of the data processing device 12 in Example 2 is realized by the following means. In this invention, the server includes an information terminal means for accepting user input, a data analysis means for identifying a preferred genre or theme based on the input, an emotion recognition means for capturing the user's facial expressions and voice in real time using a built-in camera and microphone and analyzing their emotions, a means including a generative AI model for automatically generating video content based on the genre, theme, and emotion data, a data storage means for saving the generated video content, a playback / distribution means for playing or distributing the saved video content to users, an evaluation input means for accepting user ratings, and a ranking update means for updating genre rankings based on the ratings. This enables the generation and sharing of personalized video content that reflects the user's preferences and emotions, and further enables appropriate suggestions based on the user profile.

[1178] The "information terminal means for accepting user input" is a means for providing an interface for the user to operate or give instructions to the system.

[1179] The "data analysis means for identifying a preferred genre or theme based on the input" is a means for analyzing the information input by the user and identifying a genre or theme according to the user's preferences.

[1180] "Emotion recognition means that uses a built-in camera and microphone to capture the user's facial expressions and voice in real time and analyze their emotions" refers to a means of recognizing the user's emotional state by using the camera and microphone attached to the device to photograph and record the user's facial expressions and voice, and analyzing that data.

[1181] "Means including a generative AI model that automatically generates video content based on the genre, theme, and emotional data" refers to means including an artificial intelligence model that receives a specified genre, theme, and emotional data as input and generates video content based on that input.

[1182] The "data storage means for storing the generated video content" is a means for temporarily storing the generated video content and for long-term storage as required.

[1183] "Playback and distribution means for playing or distributing the stored video content to users" refers to means for playing or distributing the stored video content in a form that can be viewed by users.

[1184] The "evaluation input means for accepting user evaluations" is a means for accepting and recording user evaluations and feedback on video content.

[1185] The "ranking update means for updating the genre rankings based on the ratings" is a means for updating the rankings of video content by genre to the latest status based on the rating data received from users.

[1186] This system generates original video content based on the user's preferences, and can also suggest and generate content by recognizing the user's emotions. Specific embodiments of each component of the system will be described below.

[1187] User Interface Means

[1188] Device:

[1189] A device provides an interface for users to enter input into the system. Users can access their account by opening an application and entering their credentials on a login screen. Devices include smartphones, tablets, and PCs.

[1190] A means of identifying a genre or theme

[1191] server:

[1192] The server analyzes the genre and theme information entered by the user or the emotion information recognized by the emotion recognition means to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the server identifies the appropriate video content.

[1193] emotion recognition means

[1194] Device:

[1195] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine. The device recognizes emotions from the user's facial expressions and tone of voice, and sends the information to a server.

[1196] server:

[1197] The server receives and analyzes the emotion data sent from the emotion recognition means, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[1198] A means of automatically generating video content

[1199] server:

[1200] The server then activates the generative AI model to automatically generate video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion recognition engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[1201] Means of preservation

[1202] server:

[1203] The server temporarily stores the generated video content, then saves it in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1204] Means of playback or distribution

[1205] Device:

[1206] The generated video content can be streamed from the server to the terminal or downloaded directly by the user for viewing.

[1207] A means of accepting user ratings

[1208] Device:

[1209] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[1210] server:

[1211] The server records and analyzes the received ratings and comments in a database, which then updates the genre rankings.

[1212] A way to update genre rankings

[1213] server:

[1214] The server updates the rankings of video content in each genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1215] How to upload to a publishing platform

[1216] Device:

[1217] The device provides a button for uploading user-generated video content to a posting platform, allowing users to share their creations with others.

[1218] server:

[1219] The server stores the uploaded video content on the posting platform and makes it accessible to other users.

[1220] How to share with other users

[1221] server:

[1222] The server manages the posted video content so that other users can view and rate it, and also generates a sharing link for users to use on social media or by email, making it easy to share.

[1223] A means to create and store user profiles

[1224] server:

[1225] The server creates and stores a profile based on the user's preferences and viewing history, which enables personalized video content to be suggested the next time the user uses the service.

[1226] A means of making personalized content suggestions

[1227] server:

[1228] The server then suggests the most suitable video content for the user based on the stored user profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[1229] Examples and prompts

[1230] Examples:

[1231] User Action: The user logs in, selects the genres "Action" and "Sci-Fi," and has a happy look on their face.

[1232] Server processing: The server receives the emotion data of "fun" from the emotion engine and automatically generates video content that includes action and science fiction elements.

[1233] Playback: The generated video content is streamed to the device.

[1234] Generate AI model prompt:

[1235] "The user likes the 'action' and 'sci-fi' genres, and their current emotion is 'fun'. Based on this condition, please generate a sci-fi video that contains exciting action scenes."

[1236] Through such a system, it becomes possible to provide users with video content that suits their mood and preferences at the time.

[1237] The flow of the identification process in the second embodiment will be described with reference to FIG.

[1238] Step 1:

[1239] User: The user opens the application on their device and enters their authentication information (username and password) on the login screen. This input includes text data such as username and password.

[1240] Terminal: The terminal sends the entered authentication information to the server using a secure communication protocol (e.g. HTTPS).

[1241] Server: The server receives the authentication information and authenticates it by comparing the username and password against the user information in its database.

[1242] Result: If authentication is successful, the server generates a session ID for the user and sends it to the terminal. The terminal displays the session ID and notifies the user that the login was successful.

[1243] Step 2:

[1244] User: After logging in, the user selects their preferred genre and theme from the displayed interface. This input includes text data such as "action" and "fantasy" as genre and theme.

[1245] Device: The device sends the selected information to the server, which is sent in JSON format.

[1246] Server: The server analyzes the received genre and theme information and stores it in a database, which is then used for further processing.

[1247] Step 3:

[1248] Device: The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time. This data includes image data and audio data.

[1249] Device: Passes the captured data to the emotion recognition engine.

[1250] Server: The server receives and analyzes the emotion data sent from the emotion recognition engine. Image processing algorithms and voice recognition algorithms are used for this analysis. The analysis result is the user's emotional state (e.g., "excited").

[1251] Step 4:

[1252] Server: The server takes genre, theme, and emotion data as input and runs the generative AI model. This input includes text data and emotion data.

[1253] Generative AI model: The generative AI model automatically generates video content based on this data. The output is a video content file.

[1254] Specific action: Send a prompt to the generative AI model saying, "The user likes the 'action' and 'fantasy' genres, and their current emotion is 'excitement.' Based on these conditions, please generate a video that includes exciting action scenes."

[1255] Step 5:

[1256] Server: The server temporarily stores the generated video content and then saves it in the user's dedicated folder. The server records the path and metadata of the saved file (e.g., creation date and time, user ID) in the database.

[1257] Device: A notification will appear on your device when the save is complete.

[1258] Step 6:

[1259] Device: The user plays or downloads the stored video content. When the user clicks the "Play" button, the device receives the video stream from the server and displays it in the video player. When the user clicks the "Download" button, the device saves the video file to local storage.

[1260] Step 7:

[1261] Terminal: After watching the video, an interface is displayed for the user to input a rating. The user can input a rating from 1 to 5 stars and a comment. This input includes numerical data and text data.

[1262] Device: The device sends the evaluation data to the server.

[1263] Server: The server receives the rating data and records it in a database. This data is used to update the rankings.

[1264] Step 8:

[1265] Server: The server updates the rankings of video content in each genre based on the received rating data. New rankings are generated and stored in the database.

[1266] On the device: The latest information is displayed so that users can view ranking information.

[1267] Step 9:

[1268] Terminal: Provides a button to upload user-generated video content to a posting platform. The user clicks the "Upload" button and selects the posting platform.

[1269] Server: The server receives the video file and transfers it to the designated uploading platform. Once uploaded, a public link is generated.

[1270] On your device: The user will see a notification that the upload is complete and a public link.

[1271] Step 10:

[1272] Server: Manages the posted video content so that other users can view and rate it. Also generates sharing links for social media and email, allowing users to easily share. Shared links are generated and sent to social media and email.

[1273] Step 11:

[1274] Server: The server creates and stores a profile based on the user's preferences and viewing history, including genre preferences and viewing history.

[1275] Specific operation: Collects user viewing history and genre preference information and records it in a database as a profile.

[1276] Step 12:

[1277] Server: The server uses a recommendation algorithm to suggest the most suitable video content for a user based on the stored user profile.

[1278] On the device: A screen will appear where the user can view a list of suggested content.

[1279] (Application example 2)

[1280] Next, a description will be given of Application Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the smart glasses 214 will be referred to as a "terminal."

[1281] Current content distribution services struggle to automatically generate and deliver personalized video content based on users' interests and emotions. Because they are unable to generate or suggest content that reflects a user's temporary emotional state in real time, it is difficult to provide users with an optimal entertainment experience. Furthermore, there is a lack of systems that effectively utilize user ratings as feedback and reflect them in rankings. Technology that can solve these problems and provide a more personalized user experience is needed.

[1282] The identification process by the identification processing unit 290 of the data processing device 12 in Application Example 2 is realized by the following means. In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre-specific rankings based on the ratings, a means for recognizing emotions from the user's facial expressions and voice, and a means for using a generative AI model for generating video content based on the recognized emotion information. This enables the generation and distribution of personalized video content based on the user's emotions, and also enables the updating of genre-specific rankings to reflect the user's ratings.

[1283] The "user interface means" is a means for providing an interface for a user to input to the system.

[1284] The "means for identifying a genre or theme" is a means for identifying a genre or theme that the user likes based on the user's input or emotional information.

[1285] The "means for automatically generating video content" is a means for automatically generating video content based on a genre or theme selected by a user or a recognized emotion.

[1286] The "means for saving" is a means for temporarily saving the generated video content and then saving it in a location accessible to the user.

[1287] "Means for playing or distributing" refers to means for allowing users to stream or download stored video content.

[1288] The "means for receiving a rating" is a means for allowing a user to input a rating for the video content that the user has viewed, and for the rating to be received within the system.

[1289] The "means for updating the genre rankings" is a means for updating the rankings of video content in each genre based on user ratings.

[1290] The "means for recognizing emotions" refers to a means for recognizing emotions from the user's facial expressions and tone of voice and analyzing that information.

[1291] A "generative AI model" is an artificial intelligence model used to generate video content based on input emotional information and prompts.

[1292] This invention relates to a system that automatically generates and distributes original video content based on user input and emotional information. The system includes a user interface, a means for identifying genres and themes, a means for recognizing emotions, a means for using a generative AI model, a means for automatically generating video content, a means for saving, a means for playing or distributing, a means for accepting ratings, and a means for updating genre-specific rankings. Each component of this system is described in detail below.

[1293] User Interface Means

[1294] Device:

[1295] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening the smartphone application and entering their credentials on the login screen.

[1296] A means of identifying a genre or theme

[1297] server:

[1298] The system analyzes the genre and theme information entered by the user, or the emotional information acquired by the emotion recognition means, to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," it will identify video content that matches that.

[1299] A means of recognizing emotions

[1300] Device:

[1301] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine, which recognizes emotions from the user's facial expressions and tone of voice and sends the information to a server.

[1302] server:

[1303] It receives and analyzes the emotional data sent from the emotion engine, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[1304] A means of automatically generating video content

[1305] server:

[1306] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[1307] Means of preservation

[1308] server:

[1309] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1310] Means of playback or distribution

[1311] Device:

[1312] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[1313] A means of accepting user ratings

[1314] Device:

[1315] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[1316] server:

[1317] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[1318] A way to update genre rankings

[1319] server:

[1320] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1321] Specific examples

[1322] For example, if a user expresses the emotion "happy," the emotion recognition engine sends that emotion to the server, which then sends prompts like "comedy" or "adventure" to the generative AI model. In this way, video content that best suits the user's current emotional state is generated and delivered.

[1323] Example prompt sentence:

[1324] "Emotion: Happy, Genre: Comedy, Adventure"

[1325] The above system enables the generation and distribution of personalized video content based on the user's emotions, improving the user's entertainment experience.

[1326] The flow of the specific processing in the application example 2 will be described with reference to FIG.

[1327] Step 1:

[1328] The terminal accepts the user's login. The user enters authentication information (user name and password) using the user interface and sends it to the server. Based on this input data, the server authenticates the user, and if authentication is successful, obtains the user's profile information and sends it to the terminal.

[1329] Step 2:

[1330] The terminal accepts genre and theme selections from the user. The user uses the interface to select their preferred genre (e.g., action, comedy) or theme (e.g., fantasy, suspense), and sends the input data to the server. The server then analyzes the user's preferences based on this data.

[1331] Step 3:

[1332] The device captures the user's facial expressions and voice using a built-in camera and microphone. The captured data is processed to recognize emotions. Specifically, facial expression data is analyzed using OpenCV, and voice data is processed using TensorFlow / Keras. The recognized emotion data is then sent to a server.

[1333] Step 4:

[1334] The server generates a prompt based on the received genre and theme selection data and emotional data. This prompt includes information such as "Emotion: happy, Genre: comedy, adventure." This prompt is then input into a generative AI model to automatically generate video content.

[1335] Step 5:

[1336] The server temporarily stores the generated video content and stores it in a dedicated folder for each user. Metadata related to the content storage is also recorded in the database.

[1337] Step 6:

[1338] The terminal allows the playback or download of video content stored by the server, and for the user to view the content, streaming playback is initiated or the content is downloaded to the terminal.

[1339] Step 7:

[1340] The terminal provides an interface for users to input ratings after viewing. Users input ratings (1 to 5 stars) and text comments, and then send the input data to the server. The server records the ratings and comments in a database.

[1341] Step 8:

[1342] The server updates the genre rankings based on the received rating data. In the ranking update process, it calculates the average rating score for each piece of content and generates new ranking information. This ranking information is saved in a database and made public to other users.

[1343] This enables the generation and distribution of personalized video content based on user emotions and the updating of rankings based on user ratings.

[1344] The specific processing unit 290 transmits the result of the specific processing to the smart glasses 214. In the smart glasses 214, the control unit 46A causes the speaker 240 to output the result of the specific processing. The microphone 238 acquires audio indicating a user input regarding the result of the specific processing. The control unit 46A transmits audio data indicating the user input acquired by the microphone 238 to the data processing device 12. In the data processing device 12, the specific processing unit 290 acquires the audio data.

[1345] The data generation model 58 is a so-called generative AI (Artificial Intelligence). An example of the data generation model 58 is ChatGPT (Internet Search<URL: https: / / openai.com / blog / chatgpt> ), Gemini (Internet search <url: https: gemini.google.com ?hl="ja">) and other generation AIs. The data generation model 58 is obtained by performing deep learning on a neural network. A prompt including an instruction is input to the data generation model 58, and inference data such as voice data indicating voice, text data indicating text, and image data indicating an image is also input. The data generation model 58 performs inference on the input inference data in accordance with the instruction indicated by the prompt, and outputs the inference result in a data format such as voice data and text data. Here, inference refers to, for example, analysis, classification, prediction, and / or summarization.

[1346] In the above embodiment, an example in which the specific processing is performed by the data processing device 12 has been given, but the technology of the present disclosure is not limited to this, and the specific processing may be performed by the smart glasses 214.

[1347] [Third embodiment]

[1348] FIG. 5 shows an example of the configuration of a data processing system 310 according to the third embodiment.

[1349] 5, the data processing system 310 includes the data processing device 12 and a headset terminal 314. An example of the data processing device 12 is a server.

[1350] The data processing device 12 includes a computer 22, a database 24, and a communication I / F 26. The computer 22 is an example of a "computer" according to the technology of the present disclosure. The computer 22 includes a processor 28, a RAM 30, and a storage 32. The processor 28, the RAM 30, and the storage 32 are connected to a bus 34. The database 24 and the communication I / F 26 are also connected to the bus 34. The communication I / F 26 is connected to a network 54. Examples of the network 54 include a WAN (Wide Area Network) and / or a LAN (Local Area Network).

[1351] The headset type terminal 314 includes a computer 36, a microphone 238, a speaker 240, a camera 42, a communication I / F 44, and a display 343. The computer 36 includes a processor 46, a RAM 48, and a storage 50. The processor 46, the RAM 48, and the storage 50 are connected to a bus 52. The microphone 238, the speaker 240, the camera 42, and the display 343 are also connected to the bus 52.

[1352] The microphone 238 receives instructions and the like from the user 20 by receiving voice uttered by the user 20. The microphone 238 captures the voice uttered by the user 20, converts the captured voice into audio data, and outputs it to the processor 46. The speaker 240 outputs audio in accordance with instructions from the processor 46.

[1353] Camera 42 is a small digital camera equipped with an optical system including a lens, aperture, and shutter, and an imaging element such as a CMOS (Complementary Metal-Oxide-Semiconductor) image sensor or a CCD (Charge Coupled Device) image sensor, and captures images of the surroundings of user 20 (for example, an imaging range defined by an angle of view equivalent to the field of vision of a typical healthy person).

[1354] The communication I / F 44 is connected to a network 54. The communication I / Fs 44 and 26 are responsible for the exchange of various information between the processor 46 and the processor 28 via the network 54. The exchange of various information between the processor 46 and the processor 28 using the communication I / Fs 44 and 26 is carried out in a secure state.

[1355] Fig. 6 shows an example of the main functions of the data processing device 12 and the headset type terminal 314. As shown in Fig. 6, in the data processing device 12, a specific process is performed by the processor 28. A specific process program 56 is stored in the storage 32.

[1356] The specific processing program 56 is an example of a "program" according to the technology of the present disclosure. The processor 28 reads the specific processing program 56 from the storage 32 and executes the read specific processing program 56 on the RAM 30. The specific processing is realized by the processor 28 operating as a specific processing unit 290 in accordance with the specific processing program 56 executed on the RAM 30.

[1357] The storage 32 stores a data generation model 58 and an emotion identification model 59. The data generation model 58 and the emotion identification model 59 are used by the identification processing unit 290.

[1358] In the headset type terminal 314, a reception output process is performed by the processor 46. A reception output program 60 is stored in the storage 50. The processor 46 reads the reception output program 60 from the storage 50 and executes the read reception output program 60 on the RAM 48. The reception output process is realized by the processor 46 operating as the control unit 46A in accordance with the reception output program 60 executed on the RAM 48.

[1359] Next, a description will be given of the identification process performed by the identification processing unit 290 of the data processing device 12. In the following description, the data processing device 12 will be referred to as the "server" and the headset type terminal 314 will be referred to as the "terminal."

[1360] This system allows users to create original video content based on their preferences, share it with other users, and receive ratings. Specific embodiments of each component of the system will be described below.

[1361] User Interface Means

[1362] Device:

[1363] The terminal provides the interface through which the user enters input into the system. The user can access their account by opening the application and entering their credentials on the login screen. On first use, the user is prompted to select their preferred genre and theme.

[1364] A means of identifying a genre or theme

[1365] server:

[1366] The system receives and analyzes information about genres and themes entered by the user to identify genres and themes based on the user's preferences. For example, if the user selects "action" and "fantasy," it sets parameters for generating video content related to these genres.

[1367] A means of automatically generating video content

[1368] server:

[1369] The video generation AI module is activated and automatically generates video content based on the genre and theme selected by the user. The AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if the user selects "action + fantasy," a video containing an episode of a warrior fighting a dragon will be generated.

[1370] Means of preservation

[1371] server:

[1372] The generated video content is temporarily saved and then saved in a dedicated folder for the user, allowing the user to access the generated video at any time.

[1373] Means of playback or distribution

[1374] Device:

[1375] The generated video content can be streamed from the server to the device or downloaded directly to the device, allowing users to watch the video in real time on their own devices.

[1376] A means of accepting user ratings

[1377] Device:

[1378] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[1379] server:

[1380] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[1381] A way to update genre rankings

[1382] server:

[1383] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1384] How to upload to a publishing platform

[1385] Device:

[1386] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[1387] server:

[1388] Store uploaded video content on the platform and make it accessible to other users.

[1389] How to share with other users

[1390] server:

[1391] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[1392] A means to create and store user profiles

[1393] server:

[1394] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[1395] A means of making personalized content suggestions

[1396] server:

[1397] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[1398] The above is a detailed description of the system in the "Description of Embodiments."

[1399] The processing flow will be explained below.

[1400] Step 1:

[1401] Device:

[1402] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[1403] Step 2:

[1404] server:

[1405] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[1406] Step 3:

[1407] Device:

[1408] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[1409] Step 4:

[1410] Device:

[1411] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[1412] Step 5:

[1413] server:

[1414] The received request is analyzed and the necessary parameters are set for the AI ​​model, for example, setting video generation based on "action" and "fantasy."

[1415] Step 6:

[1416] server:

[1417] The video generation AI module is activated to generate a scenario, construct a scene, and synthesize audio. Specifically, it automatically generates a video containing an episode of a warrior fighting a dragon.

[1418] Step 7:

[1419] server:

[1420] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1421] Step 8:

[1422] server:

[1423] The terminal is notified that the video content has been saved.

[1424] Step 9:

[1425] Device:

[1426] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[1427] Step 10:

[1428] Device:

[1429] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[1430] Step 11:

[1431] User:

[1432] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[1433] Step 12:

[1434] Device:

[1435] After the user enters their rating and comments, the data is sent to the server.

[1436] Step 13:

[1437] server:

[1438] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[1439] Step 14:

[1440] User:

[1441] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[1442] Step 15:

[1443] Device:

[1444] The upload procedure for the video content is carried out and an upload request is sent to the server.

[1445] Step 16:

[1446] server:

[1447] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[1448] These are the specific processing steps involved in generating and sharing a video work based on a user request.

[1449] Example 1

[1450] Next, a description will be given of Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the headset type terminal 314 will be referred to as a "terminal."

[1451] In the creation and sharing of digital media content, personalized content delivery based on user preferences is required, but existing systems do not effectively generate, evaluate, and share content that reflects user preferences. Furthermore, the accuracy of recommendation systems and ranking updates based on user ratings also needs to be improved.

[1452] The specific processing by the specific processing unit 290 of the data processing device 12 in the first embodiment is realized by the following means.

[1453] In this invention, the server includes a user interface means for accepting user input, a means for identifying preferred categories or topics based on the input, a means for automatically generating digital media content based on the categories or topics, a means for saving the generated digital media content, a means for playing or distributing the saved digital media content to the user, a means for accepting user ratings, and a means for updating category rankings based on the ratings, thereby enabling efficient generation, rating, sharing, and ranking updates of personalized digital media content based on user preferences.

[1454] The "user interface means" is a device or program that provides an interface for a user to input information to the system.

[1455] A "category" is a concept that refers to a genre or topic of digital media content selected by a user.

[1456] A "topic" is a concept that refers to a specific theme or subject of digital media content selected by a user.

[1457] "Digital media content" refers to media data in digital format, such as video, audio, and images, that is generated based on user preferences.

[1458] An "automatic generation means" is a device or program that automatically creates digital media content based on category or topic information entered by a user.

[1459] A "storage means" is a device or program for temporary and long-term storage of the generated digital media content.

[1460] A "delivery means" is a device or program that provides stored digital media content to a user's terminal in a streaming or downloadable format.

[1461] The "means for receiving ratings" is a device or program that collects rating information on digital media content from users.

[1462] The "means for updating the ranking by category" is a device or program that updates the ranking of digital media content for each category based on evaluation information collected from users.

[1463] A "sharing platform" is an online service or website for sharing user-generated digital media content with other users.

[1464] MODE FOR CARRYING OUT THE INVENTION

[1465] The system automates the creation, storage, distribution, rating, and sharing of digital media content. It uses user input to identify preferred categories and topics, and then uses generative AI models to automatically generate content. A specific implementation of this system is described below.

[1466] 1. User Interface Methods

[1467] (Terminal)

[1468] The terminal provides an application as a user interface for the user to input data to the system. The user accesses the system by starting this application and entering a user name and password on the login screen.

[1469] 2. How to identify your favorite categories and topics

[1470] (Terminal)

[1471] On first use, the device will display an interface for setting user preferences, where users can select their preferred categories and topics, for example, "action" and "fantasy."

[1472] (server)

[1473] The server receives and analyzes this selection information and stores it in a database. This data is used to generate content using the AI ​​model described below.

[1474] 3. Content Generation Methods

[1475] (server)

[1476] The server generates prompts based on the user's preferences and sends them to a generative AI model (e.g., OpenAI GPT-4, DeepMind's video generation model).

[1477] 4. Video content generation

[1478] (server)

[1479] The server's generative AI model automatically generates video content based on the input prompt. For example, for a user who likes "action" and "fantasy," a video containing a story about a warrior fighting a dragon is generated. Video generation involves the processes of scenario generation, scene construction, and voice synthesis.

[1480] 5. How content is stored

[1481] (server)

[1482] The server temporarily stores the generated video content and then saves it in the user's dedicated folder. The video files are managed using a database management system (e.g., MySQL, PostgreSQL).

[1483] 6. Content playback and distribution methods

[1484] (Terminal)

[1485] The device allows users to play or download stored video content, often using streaming technology and a media player (e.g., VLC, QuickTime).

[1486] 7. Method of accepting evaluations

[1487] (Terminal)

[1488] After watching the video, the device displays a user rating input interface, allowing users to enter star ratings and comments.

[1489] (server)

[1490] The server receives user ratings and stores them in a database, which is used to update rankings for each category.

[1491] 8. How to update rankings by category

[1492] (server)

[1493] The server updates the rankings for each category based on the collected rating information, and then makes this ranking information available to other users, who can then recommend highly rated content.

[1494] 9. How to share content

[1495] (Terminal)

[1496] The terminal provides an interface for uploading the generated digital media content to a sharing platform.

[1497] (server)

[1498] The server manages the uploaded content and makes it accessible to other users, generating shareable links for easy sharing via social media or email.

[1499] 10. How to create and store user profiles

[1500] (server)

[1501] A profile is created and saved based on the user's preferences and viewing history, allowing for personalized content suggestions the next time the user uses the service.

[1502] 11. Means of personalized content recommendations

[1503] (server)

[1504] Based on a saved user profile, the app suggests digital media content that is most suitable for the user, for example, if a user previously indicated that they like "action," the app will suggest the latest action movies.

[1505] Examples of concrete examples and prompts

[1506] Examples:

[1507] The user selects "action" and "fantasy" and enters the prompt "Generate a movie about a warrior fighting a dragon." Based on this prompt, the server's generative AI model automatically generates video content and saves it in the user's dedicated folder. This video content can be viewed on the device and shared with other users.

[1508] Example prompt sentence:

[1509] "Generate footage of action scenes in fantasy movies"

[1510] "Suggest your next watchlist based on your favorite movie genres."

[1511] This system allows users to easily create, save, rate, and share content tailored to their preferences. This specific embodiment will clarify how the invention actually works.

[1512] The flow of the identification process in the first embodiment will be described with reference to FIG.

[1513] Step 1:

[1514] (User):

[1515] The user launches the application and enters their username and password on the login screen.

[1516] input:

[1517] Username, Password

[1518] output:

[1519] Authentication Request

[1520] (Device):

[1521] The terminal accepts user input and sends an authentication request to the server.

[1522] Specific behavior:

[1523] Displaying the login screen

[1524] Obtaining the entered username and password

[1525] Sending an authentication request to the server

[1526] Step 2:

[1527] (server):

[1528] The server checks the received authentication request against its database and returns the authentication result.

[1529] input:

[1530] Authentication Request

[1531] output:

[1532] Authentication result (success / failure)

[1533] (Device):

[1534] The terminal performs screen transitions based on the authentication results.

[1535] Specific behavior:

[1536] Receiving the authentication result

[1537] If authentication is successful, the main screen is displayed. If authentication fails, an error message is displayed.

[1538] Step 3:

[1539] (User):

[1540] On first use, users make selections in an interface that allows them to choose their preferred categories and topics.

[1541] input:

[1542] Category selection (e.g., "Action"), Topic selection (e.g., "Fantasy")

[1543] output:

[1544] Selected Data

[1545] (Device):

[1546] The terminal transmits the selected data to the server.

[1547] Specific behavior:

[1548] Receiving user selections

[1549] Sending Selected Data to the Server

[1550] Step 4:

[1551] (server):

[1552] The server stores the received selection data in a database.

[1553] input:

[1554] Selected Data

[1555] output:

[1556] Results saved in the database

[1557] Specific behavior:

[1558] Analysis of selected data

[1559] Saving to the database

[1560] Step 5:

[1561] (User):

[1562] The user inputs the content of the video they want to generate on the prompt sentence input screen.

[1563] input:

[1564] Prompt text (e.g., "Generate a movie of a warrior fighting a dragon")

[1565] output:

[1566] Prompt Data

[1567] (Device):

[1568] The terminal sends the prompt data to the server.

[1569] Specific behavior:

[1570] Receiving a prompt

[1571] Sending prompt data to the server

[1572] Step 6:

[1573] (server):

[1574] The server inputs the prompt data into the generative AI model and begins generating video content.

[1575] input:

[1576] Prompt Data

[1577] output:

[1578] Generated video content

[1579] Specific behavior:

[1580] Parsing prompt data

[1581] Input to generative AI models

[1582] Video content generation

[1583] Step 7:

[1584] (server):

[1585] The server temporarily stores the generated video content, and then stores it in a folder dedicated to the user.

[1586] input:

[1587] Generated video content

[1588] output:

[1589] Saved results in user folder

[1590] Specific behavior:

[1591] Temporary storage of video content

[1592] Move to user-specific folder

[1593] Step 8:

[1594] (Device):

[1595] The user plays or downloads the stored video content.

[1596] input:

[1597] User requests (playback, download)

[1598] output:

[1599] Playback and download of video content

[1600] Specific behavior:

[1601] Click the play button to stream

[1602] Click the download button to save the video to your device

[1603] Step 9:

[1604] (User):

[1605] After viewing the video, the user inputs a rating.

[1606] input:

[1607] Rating data (stars, comments)

[1608] output:

[1609] Evaluation data submission

[1610] (Device):

[1611] The terminal transmits the evaluation data to the server.

[1612] Specific behavior:

[1613] View the rating interface

[1614] Receiving and sending evaluation data to the server

[1615] Step 10:

[1616] (server):

[1617] The server stores the received rating data in a database and updates the rankings by category.

[1618] input:

[1619] Evaluation Data

[1620] output:

[1621] Updated ranking data

[1622] Specific behavior:

[1623] Analysis of evaluation data

[1624] Saving to a database

[1625] Category ranking updates

[1626] Step 11:

[1627] (User):

[1628] Users upload the generated video content to a sharing platform.

[1629] input:

[1630] Upload Request

[1631] output:

[1632] Upload completion notification

[1633] (Device):

[1634] Pressing the upload button sends the video content to the server.

[1635] Specific behavior:

[1636] Click the upload button

[1637] Sending video content to the server

[1638] Step 12:

[1639] (server):

[1640] The server manages the uploaded content, sets it up so that other users can access it, and generates a share link so that it can be shared via social media or email.

[1641] input:

[1642] Uploaded content

[1643] output:

[1644] Shared Links

[1645] Specific behavior:

[1646] Organizing uploaded content

[1647] Generate shareable links for social media and email

[1648] Step 13:

[1649] (server):

[1650] The server creates and stores a profile based on the user's preferences and viewing history.

[1651] input:

[1652] User preferences and viewing history

[1653] output:

[1654] User Profile

[1655] Specific behavior:

[1656] User data analysis

[1657] Creating and Saving a Profile

[1658] Step 14:

[1659] (server):

[1660] The server makes personalized content suggestions based on the stored profile.

[1661] input:

[1662] User Profile

[1663] output:

[1664] Suggested Content List

[1665] Specific behavior:

[1666] Viewing Profiles

[1667] Selection of proposed content

[1668] The above are the specific steps of the program processing of this system.

[1669] (Application example 1)

[1670] Next, a description will be given of Application Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the headset type terminal 314 will be referred to as a "terminal."

[1671] Conventional video content generation and distribution systems lack the functionality to allow users to easily create original video content based on their preferences and share it with other users. Furthermore, the lack of personalized content suggestions and rating systems hinders the improvement of the user experience.

[1672] The specific processing by the specific processing unit 290 of the data processing device 12 in the application example 1 is realized by the following means.

[1673] In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre rankings based on the ratings, a means for creating and saving a user profile, a means for suggesting personalized content based on the profile, a means for a user to log in and select a genre using an application installed on a smartphone or smart glasses, a means for saving and sharing the generated video in cloud storage, and a means for generating video using an AI model by inputting a prompt. This allows users to easily generate video content according to their preferences, share it with other users, receive ratings, and improve the accuracy of content suggestions from the next time onwards.

[1674] The "user interface means" is an interface that allows a user to input information into the system, and is a function that allows operation through a smartphone or smart glasses, for example.

[1675] The "means for identifying a favorite genre or theme" is a function of the system that analyzes and identifies a user's favorite genre or theme based on information input by the user.

[1676] The "means for automatically generating video content" is an AI module for generating video content based on the genre or theme selected by the user.

[1677] The "means for saving the generated video content" is a function for temporarily or permanently storing the generated video in a recording device.

[1678] The "means for playing or distributing stored video content to users" is a function for providing stored video so that users can view it.

[1679] The "means for accepting user ratings" is a function including an interface for users to input ratings and comments on video content.

[1680] The "means for updating genre rankings" is a function for automatically updating genre rankings based on user ratings.

[1681] "Means for creating and saving user profiles" refers to a function that creates individual profiles based on the user's preferences and viewing history and saves them in a database.

[1682] The "means for making personalized content suggestions" is a function that suggests the most suitable video content to a user based on a saved user profile.

[1683] "Application installed on a smartphone or smart glasses" refers to software for a mobile device that allows a user to access the video content generation system.

[1684] "Means for saving and sharing the generated video in cloud storage" refers to a function for saving the generated video file to a remote server via the Internet and providing a sharing link from there.

[1685] "Means for generating images using an AI model by inputting a prompt sentence" refers to a function in which a user inputs a specific instruction sentence, and the AI ​​model generates an image based on that instruction.

[1686] This invention is a system that enables users to use mobile devices such as smartphones or smart glasses to create original video content based on their preferences, share it with other users, and receive ratings.

[1687] User Interface Means

[1688] The device is equipped with an interface that allows users to input information into the system. Users can access their account by opening the application on their smartphone or smart glasses and entering their authentication information on the login screen. On first use, a screen appears where users can select their preferred genre (e.g., "action" or "fantasy") and theme.

[1689] A means of identifying a genre or theme

[1690] The server receives the genre and theme information entered by the user, analyzes it, and identifies the user's preferences. For example, if the user selects "thriller" and "mystery," parameters for generating video content related to these genres are set.

[1691] A means of automatically generating video content

[1692] The server launches a video generation AI model and automatically generates video content based on the genre and theme selected by the user. The AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if a user selects "Thriller + Mystery," a short video of a detective solving a mysterious case will be generated. The AI ​​model used here utilizes machine learning frameworks such as TensorFlow and PyTorch.

[1693] Storage and playback methods

[1694] The server temporarily stores the generated video content in cloud storage such as Google Cloud Storage, and then saves it in the user's dedicated folder. Users can then view the video in real time on their smartphones or smart glasses.

[1695] How to receive ratings

[1696] The device provides an interface for users to input ratings after watching videos. Users can enter ratings from 1 to 5 stars and comments. The server records the received ratings and comments in a database such as Firebase and analyzes them to update the rankings by genre.

[1697] How to update the rankings

[1698] The server automatically updates the rankings of video content by genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1699] How to store and share content

[1700] The generated video content is stored in cloud storage and a sharing link is generated, allowing users to easily share their work via social media or email.

[1701] User profiles and personalized content suggestions

[1702] The server creates and stores a profile based on the user's preferences and viewing history. The next time the user logs in, personalized content suggestions are made based on this profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[1703] Examples of concrete examples and prompts

[1704] After a user logs in, they can select their preferred genre, create, watch, and rate a short video. The following prompts are used:

[1705] Input: "Generate a short film that mixes horror and fantasy genres. My Firebase UID is abcdef."

[1706] Output: "Generating movie... The URL for the generated movie is https: / / storage.googleapis.com / abcdef / generated_video.mp4."

[1707] This system allows users to easily create and enjoy video content according to their preferences.

[1708] The flow of the specific processing in the application example 1 will be described with reference to FIG.

[1709] Step 1:

[1710] A user launches an application on their smartphone or smart glasses and accesses the login screen. The input is the user's authentication information, and the output is the authentication success / failure result. The server validates the user's authentication information using Firebase Auth and returns the user ID if authentication is successful.

[1711] Step 2:

[1712] After logging in, the user selects their preferred genre and theme. The input is the genre and theme selected by the user, and the output is the data of the selected genre and theme. The terminal sends the selected genre and theme to the server.

[1713] Step 3:

[1714] The server analyzes the received genre and theme information and sets the parameters of the AI ​​model. The input is genre and theme data, and the output is the parameter settings of the AI ​​model. The server uses TensorFlow or PyTorch to initialize the AI ​​model and set the settings based on the received data.

[1715] Step 4:

[1716] The server launches the configured AI model and generates video content. The input is the AI ​​model's parameters and generation instructions (prompt text), and the output is the generated video file. The AI ​​model generates a scenario, constructs scenes, and synthesizes audio. In this step, if the user inputs, for example, "Please generate a short film that combines the horror and fantasy genres," a video of the corresponding scenario will be automatically generated.

[1717] Step 5:

[1718] The server saves the generated video file in cloud storage (e.g., Google Cloud Storage). The input is the generated video file, and the output is the cloud storage URL. The server uploads the video file to cloud storage and generates a URL link for the saved file.

[1719] Step 6:

[1720] The server provides the user with a URL for the generated video. The input is the URL of the cloud storage, and the output is the URL sent to the user's device. The server then notifies the smartphone or smart glasses of the generated URL, allowing the user to access the video.

[1721] Step 7:

[1722] Users watch videos and rate them through a rating input screen. The input is the user's rating information (e.g., a rating from 1 to 5 stars, comments), and the output is the rating data. The device accepts the user's rating and sends it to the server.

[1723] Step 8:

[1724] The server records the received rating data and updates the genre rankings. The input is the rating data received from users, and the output is the updated ranking information. The server saves the rating data in Firebase or MongoDB and automatically updates the rankings based on it.

[1725] Step 9:

[1726] The server creates and stores user profiles. The input is the user's selection history and rating data, and the output is a user profile. The server creates a profile based on the viewing history and rating information and stores it in a database.

[1727] Step 10:

[1728] The server will then suggest personalized content based on the saved user profile from the next time onwards. The input is the user profile, and the output is the suggested content information. The server analyzes the user profile, generates information to suggest appropriate video content, and provides it to the user.

[1729] Through these steps, users can efficiently create, view, and rate video content according to their preferences, improving the user experience.

[1730] Furthermore, an emotion engine that estimates the user's emotion may be further combined. That is, the identification processing unit 290 may estimate the user's emotion using the emotion identification model 59, and perform identification processing using the user's emotion.

[1731] This system not only generates original video content based on the user's preferences, but also recognizes the user's emotions and proposes and generates content. Specific embodiments of each component of the system are described below.

[1732] User Interface Means

[1733] Device:

[1734] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening an application and entering their credentials on a login screen.

[1735] A means of identifying a genre or theme

[1736] server:

[1737] The genre and theme information input by the user or the emotion information recognized by the emotion engine means is analyzed to identify a genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the appropriate video content is identified.

[1738] Emotion Engine Means

[1739] Device:

[1740] The terminal captures the user's facial expressions and voice in real time using a built-in camera and microphone, and the emotion engine means analyzes the captured images. The emotion engine means recognizes emotions from the user's facial expressions and tone of voice, and transmits the information to a server.

[1741] server:

[1742] The emotion data sent from the emotion engine means is received and analyzed. Based on the analysis results, the system recommends genres and themes that are most suitable for the user.

[1743] A means of automatically generating video content

[1744] server:

[1745] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[1746] Means of preservation

[1747] server:

[1748] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1749] Means of playback or distribution

[1750] Device:

[1751] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[1752] A means of accepting user ratings

[1753] Device:

[1754] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[1755] server:

[1756] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[1757] A way to update genre rankings

[1758] server:

[1759] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1760] How to upload to a publishing platform

[1761] Device:

[1762] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[1763] server:

[1764] Uploaded video content is stored on the posting platform and made accessible to other users.

[1765] How to share with other users

[1766] server:

[1767] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[1768] A means to create and store user profiles

[1769] server:

[1770] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[1771] A means of making personalized content suggestions

[1772] server:

[1773] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[1774] The above is a detailed description of the system in the "Description of Embodiments."

[1775] The processing flow will be explained below.

[1776] Step 1:

[1777] Device:

[1778] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[1779] Step 2:

[1780] server:

[1781] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[1782] Step 3:

[1783] Device:

[1784] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[1785] Step 4:

[1786] Device:

[1787] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[1788] Step 5:

[1789] Device:

[1790] The user selects the Enable Emotion Engine option in the settings screen, which prepares the device's built-in camera and microphone to collect emotion data.

[1791] Step 6:

[1792] Device:

[1793] While watching a video or requesting a new video, the device captures the user's facial expressions and voice through the camera and microphone, analyzes the emotional data in real time, and temporarily stores this data.

[1794] Step 7:

[1795] server:

[1796] Emotion data transmitted from the terminal is received and analyzed, and the emotion engine means identifies the user's emotion (e.g., excitement, joy, sadness).

[1797] Step 8:

[1798] server:

[1799] Based on the analysis results, the system recommends genres and themes that suit the user's current emotions. For example, if the user is recognized as "excited," the system will identify "action" and "thriller" genres.

[1800] Step 9:

[1801] server:

[1802] Based on the user's selection and recognized emotions, the video generation AI module is activated to generate scenarios, construct scenes, and synthesize audio.

[1803] Specifically, it automatically generates footage including episodes of warriors fighting dragons.

[1804] Step 10:

[1805] server:

[1806] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1807] Step 11:

[1808] server:

[1809] The terminal is notified that the video content has been saved.

[1810] Step 12:

[1811] Device:

[1812] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[1813] Step 13:

[1814] Device:

[1815] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[1816] Step 14:

[1817] User:

[1818] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[1819] Step 15:

[1820] Device:

[1821] After the user enters their rating and comments, the data is sent to the server.

[1822] Step 16:

[1823] server:

[1824] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[1825] Step 17:

[1826] User:

[1827] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[1828] Step 18:

[1829] Device:

[1830] The upload procedure for the video content is carried out and an upload request is sent to the server.

[1831] Step 19:

[1832] server:

[1833] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[1834] These are the specific processing steps for generating and sharing a video work based on the user's requests and emotions.

[1835] Example 2

[1836] Next, a description will be given of Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the headset type terminal 314 will be referred to as a "terminal."

[1837] Conventional video content generation systems have had difficulty generating video content that fully reflects a user's preferences and momentary emotions. Furthermore, sharing the generated content with other users is time-consuming, and it is difficult to reflect the user's ratings in the system. Furthermore, they lacked functionality for making personalized suggestions based on user profiles, leaving a need for an improved user experience.

[1838] The identification process by the identification processing unit 290 of the data processing device 12 in Example 2 is realized by the following means. In this invention, the server includes an information terminal means for accepting user input, a data analysis means for identifying a preferred genre or theme based on the input, an emotion recognition means for capturing the user's facial expressions and voice in real time using a built-in camera and microphone and analyzing their emotions, a means including a generative AI model for automatically generating video content based on the genre, theme, and emotion data, a data storage means for saving the generated video content, a playback / distribution means for playing or distributing the saved video content to users, an evaluation input means for accepting user ratings, and a ranking update means for updating genre rankings based on the ratings. This enables the generation and sharing of personalized video content that reflects the user's preferences and emotions, and further enables appropriate suggestions based on the user profile.

[1839] The "information terminal means for accepting user input" is a means for providing an interface for the user to operate or give instructions to the system.

[1840] The "data analysis means for identifying a preferred genre or theme based on the input" is a means for analyzing the information input by the user and identifying a genre or theme according to the user's preferences.

[1841] "Emotion recognition means that uses a built-in camera and microphone to capture the user's facial expressions and voice in real time and analyze their emotions" refers to a means of recognizing the user's emotional state by using the camera and microphone attached to the device to photograph and record the user's facial expressions and voice, and analyzing that data.

[1842] "Means including a generative AI model that automatically generates video content based on the genre, theme, and emotional data" refers to means including an artificial intelligence model that receives a specified genre, theme, and emotional data as input and generates video content based on that input.

[1843] The "data storage means for storing the generated video content" is a means for temporarily storing the generated video content and for long-term storage as required.

[1844] "Playback and distribution means for playing or distributing the stored video content to users" refers to means for playing or distributing the stored video content in a form that can be viewed by users.

[1845] The "evaluation input means for accepting user evaluations" is a means for accepting and recording user evaluations and feedback on video content.

[1846] The "ranking update means for updating the genre rankings based on the ratings" is a means for updating the rankings of video content by genre to the latest status based on the rating data received from users.

[1847] This system generates original video content based on the user's preferences, and can also suggest and generate content by recognizing the user's emotions. Specific embodiments of each component of the system will be described below.

[1848] User Interface Means

[1849] Device:

[1850] A device provides an interface for users to enter input into the system. Users can access their account by opening an application and entering their credentials on a login screen. Devices include smartphones, tablets, and PCs.

[1851] A means of identifying a genre or theme

[1852] server:

[1853] The server analyzes the genre and theme information entered by the user or the emotion information recognized by the emotion recognition means to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the server identifies the appropriate video content.

[1854] emotion recognition means

[1855] Device:

[1856] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine. The device recognizes emotions from the user's facial expressions and tone of voice, and sends the information to a server.

[1857] server:

[1858] The server receives and analyzes the emotion data sent from the emotion recognition means, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[1859] A means of automatically generating video content

[1860] server:

[1861] The server then activates the generative AI model to automatically generate video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion recognition engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[1862] Means of preservation

[1863] server:

[1864] The server temporarily stores the generated video content, then saves it in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1865] Means of playback or distribution

[1866] Device:

[1867] The generated video content can be streamed from the server to the terminal or downloaded directly by the user for viewing.

[1868] A means of accepting user ratings

[1869] Device:

[1870] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[1871] server:

[1872] The server records and analyzes the received ratings and comments in a database, which then updates the genre rankings.

[1873] A way to update genre rankings

[1874] server:

[1875] The server updates the rankings of video content in each genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1876] How to upload to a publishing platform

[1877] Device:

[1878] The device provides a button for uploading user-generated video content to a posting platform, allowing users to share their creations with others.

[1879] server:

[1880] The server stores the uploaded video content on the posting platform and makes it accessible to other users.

[1881] How to share with other users

[1882] server:

[1883] The server manages the posted video content so that other users can view and rate it, and also generates a sharing link for users to use on social media or by email, making it easy to share.

[1884] A means to create and store user profiles

[1885] server:

[1886] The server creates and stores a profile based on the user's preferences and viewing history, which enables personalized video content to be suggested the next time the user uses the service.

[1887] A means of making personalized content suggestions

[1888] server:

[1889] The server then suggests the most suitable video content for the user based on the stored user profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[1890] Examples and prompts

[1891] Examples:

[1892] User Action: The user logs in, selects the genres "Action" and "Sci-Fi," and has a happy look on their face.

[1893] Server processing: The server receives the emotion data of "fun" from the emotion engine and automatically generates video content that includes action and science fiction elements.

[1894] Playback: The generated video content is streamed to the device.

[1895] Generate AI model prompt:

[1896] "The user likes the 'action' and 'sci-fi' genres, and their current emotion is 'fun'. Based on this condition, please generate a sci-fi video that contains exciting action scenes."

[1897] Through such a system, it becomes possible to provide users with video content that suits their mood and preferences at the time.

[1898] The flow of the identification process in the second embodiment will be described with reference to FIG.

[1899] Step 1:

[1900] User: The user opens the application on their device and enters their authentication information (username and password) on the login screen. This input includes text data such as username and password.

[1901] Terminal: The terminal sends the entered authentication information to the server using a secure communication protocol (e.g. HTTPS).

[1902] Server: The server receives the authentication information and authenticates it by comparing the username and password against the user information in its database.

[1903] Result: If authentication is successful, the server generates a session ID for the user and sends it to the terminal. The terminal displays the session ID and notifies the user that the login was successful.

[1904] Step 2:

[1905] User: After logging in, the user selects their preferred genre and theme from the displayed interface. This input includes text data such as "action" and "fantasy" as genre and theme.

[1906] Device: The device sends the selected information to the server, which is sent in JSON format.

[1907] Server: The server analyzes the received genre and theme information and stores it in a database, which is then used for further processing.

[1908] Step 3:

[1909] Device: The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time. This data includes image data and audio data.

[1910] Device: Passes the captured data to the emotion recognition engine.

[1911] Server: The server receives and analyzes the emotion data sent from the emotion recognition engine. Image processing algorithms and voice recognition algorithms are used for this analysis. The analysis result is the user's emotional state (e.g., "excited").

[1912] Step 4:

[1913] Server: The server takes genre, theme, and emotion data as input and runs the generative AI model. This input includes text data and emotion data.

[1914] Generative AI model: The generative AI model automatically generates video content based on this data. The output is a video content file.

[1915] Specific action: Send a prompt to the generative AI model saying, "The user likes the 'action' and 'fantasy' genres, and their current emotion is 'excitement.' Based on these conditions, please generate a video that includes exciting action scenes."

[1916] Step 5:

[1917] Server: The server temporarily stores the generated video content and then saves it in the user's dedicated folder. The server records the path and metadata of the saved file (e.g., creation date and time, user ID) in the database.

[1918] Device: A notification will appear on your device when the save is complete.

[1919] Step 6:

[1920] Device: The user plays or downloads the stored video content. When the user clicks the "Play" button, the device receives the video stream from the server and displays it in the video player. When the user clicks the "Download" button, the device saves the video file to local storage.

[1921] Step 7:

[1922] Terminal: After watching the video, an interface is displayed for the user to input a rating. The user can input a rating from 1 to 5 stars and a comment. This input includes numerical data and text data.

[1923] Device: The device sends the evaluation data to the server.

[1924] Server: The server receives the rating data and records it in a database. This data is used to update the rankings.

[1925] Step 8:

[1926] Server: The server updates the rankings of video content in each genre based on the received rating data. New rankings are generated and stored in the database.

[1927] On the device: The latest information is displayed so that users can view ranking information.

[1928] Step 9:

[1929] Terminal: Provides a button to upload user-generated video content to a posting platform. The user clicks the "Upload" button and selects the posting platform.

[1930] Server: The server receives the video file and transfers it to the designated uploading platform. Once uploaded, a public link is generated.

[1931] On your device: The user will see a notification that the upload is complete and a public link.

[1932] Step 10:

[1933] Server: Manages the posted video content so that other users can view and rate it. Also generates sharing links for social media and email, allowing users to easily share. Shared links are generated and sent to social media and email.

[1934] Step 11:

[1935] Server: The server creates and stores a profile based on the user's preferences and viewing history, including genre preferences and viewing history.

[1936] Specific operation: Collects user viewing history and genre preference information and records it in a database as a profile.

[1937] Step 12:

[1938] Server: The server uses a recommendation algorithm to suggest the most suitable video content for a user based on the stored user profile.

[1939] On the device: A screen will appear where the user can view a list of suggested content.

[1940] (Application example 2)

[1941] Next, a description will be given of Application Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the headset type terminal 314 will be referred to as a "terminal."

[1942] Current content distribution services struggle to automatically generate and deliver personalized video content based on users' interests and emotions. Because they are unable to generate or suggest content that reflects a user's temporary emotional state in real time, it is difficult to provide users with an optimal entertainment experience. Furthermore, there is a lack of systems that effectively utilize user ratings as feedback and reflect them in rankings. Technology that can solve these problems and provide a more personalized user experience is needed.

[1943] The identification process by the identification processing unit 290 of the data processing device 12 in Application Example 2 is realized by the following means. In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre-specific rankings based on the ratings, a means for recognizing emotions from the user's facial expressions and voice, and a means for using a generative AI model for generating video content based on the recognized emotion information. This enables the generation and distribution of personalized video content based on the user's emotions, and also enables the updating of genre-specific rankings to reflect the user's ratings.

[1944] The "user interface means" is a means for providing an interface for a user to input to the system.

[1945] The "means for identifying a genre or theme" is a means for identifying a genre or theme that the user likes based on the user's input or emotional information.

[1946] The "means for automatically generating video content" is a means for automatically generating video content based on a genre or theme selected by a user or a recognized emotion.

[1947] The "means for saving" is a means for temporarily saving the generated video content and then saving it in a location accessible to the user.

[1948] "Means for playing or distributing" refers to means for allowing users to stream or download stored video content.

[1949] The "means for receiving a rating" is a means for allowing a user to input a rating for the video content that the user has viewed, and for the rating to be received within the system.

[1950] The "means for updating the genre rankings" is a means for updating the rankings of video content in each genre based on user ratings.

[1951] The "means for recognizing emotions" refers to a means for recognizing emotions from the user's facial expressions and tone of voice and analyzing that information.

[1952] A "generative AI model" is an artificial intelligence model used to generate video content based on input emotional information and prompts.

[1953] This invention relates to a system that automatically generates and distributes original video content based on user input and emotional information. The system includes a user interface, a means for identifying genres and themes, a means for recognizing emotions, a means for using a generative AI model, a means for automatically generating video content, a means for saving, a means for playing or distributing, a means for accepting ratings, and a means for updating genre-specific rankings. Each component of this system is described in detail below.

[1954] User Interface Means

[1955] Device:

[1956] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening the smartphone application and entering their credentials on the login screen.

[1957] A means of identifying a genre or theme

[1958] server:

[1959] The system analyzes the genre and theme information entered by the user, or the emotional information acquired by the emotion recognition means, to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," it will identify video content that matches that.

[1960] A means of recognizing emotions

[1961] Device:

[1962] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine, which recognizes emotions from the user's facial expressions and tone of voice and sends the information to a server.

[1963] server:

[1964] It receives and analyzes the emotional data sent from the emotion engine, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[1965] A means of automatically generating video content

[1966] server:

[1967] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[1968] Means of preservation

[1969] server:

[1970] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[1971] Means of playback or distribution

[1972] Device:

[1973] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[1974] A means of accepting user ratings

[1975] Device:

[1976] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[1977] server:

[1978] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[1979] A way to update genre rankings

[1980] server:

[1981] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[1982] Specific examples

[1983] For example, if a user expresses the emotion "happy," the emotion recognition engine sends that emotion to the server, which then sends prompts like "comedy" or "adventure" to the generative AI model. In this way, video content that best suits the user's current emotional state is generated and delivered.

[1984] Example prompt sentence:

[1985] "Emotion: Happy, Genre: Comedy, Adventure"

[1986] The above system enables the generation and distribution of personalized video content based on the user's emotions, improving the user's entertainment experience.

[1987] The flow of the specific processing in the application example 2 will be described with reference to FIG.

[1988] Step 1:

[1989] The terminal accepts the user's login. The user enters authentication information (user name and password) using the user interface and sends it to the server. Based on this input data, the server authenticates the user, and if authentication is successful, obtains the user's profile information and sends it to the terminal.

[1990] Step 2:

[1991] The terminal accepts genre and theme selections from the user. The user uses the interface to select their preferred genre (e.g., action, comedy) or theme (e.g., fantasy, suspense), and sends the input data to the server. The server then analyzes the user's preferences based on this data.

[1992] Step 3:

[1993] The device captures the user's facial expressions and voice using a built-in camera and microphone. The captured data is processed to recognize emotions. Specifically, facial expression data is analyzed using OpenCV, and voice data is processed using TensorFlow / Keras. The recognized emotion data is then sent to a server.

[1994] Step 4:

[1995] The server generates a prompt based on the received genre and theme selection data and emotional data. This prompt includes information such as "Emotion: happy, Genre: comedy, adventure." This prompt is then input into a generative AI model to automatically generate video content.

[1996] Step 5:

[1997] The server temporarily stores the generated video content and stores it in a dedicated folder for each user. Metadata related to the content storage is also recorded in the database.

[1998] Step 6:

[1999] The terminal allows the playback or download of video content stored by the server, and for the user to view the content, streaming playback is initiated or the content is downloaded to the terminal.

[2000] Step 7:

[2001] The terminal provides an interface for users to input ratings after viewing. Users input ratings (1 to 5 stars) and text comments, and then send the input data to the server. The server records the ratings and comments in a database.

[2002] Step 8:

[2003] The server updates the genre rankings based on the received rating data. In the ranking update process, it calculates the average rating score for each piece of content and generates new ranking information. This ranking information is saved in a database and made public to other users.

[2004] This enables the generation and distribution of personalized video content based on user emotions and the updating of rankings based on user ratings.

[2005] The specific processing unit 290 transmits the result of the specific processing to the headset type terminal 314. In the headset type terminal 314, the control unit 46A causes the speaker 240 and the display 343 to output the result of the specific processing. The microphone 238 acquires audio indicating a user input regarding the result of the specific processing. The control unit 46A transmits audio data indicating the user input acquired by the microphone 238 to the data processing device 12. In the data processing device 12, the specific processing unit 290 acquires the audio data.

[2006] The data generation model 58 is a so-called generative AI (Artificial Intelligence). An example of the data generation model 58 is ChatGPT (Internet Search<URL: https: / / openai.com / blog / chatgpt> ), Gemini (Internet search <url: https: gemini.google.com ?hl="ja">) and other generation AIs. The data generation model 58 is obtained by performing deep learning on a neural network. A prompt including an instruction is input to the data generation model 58, and inference data such as voice data indicating voice, text data indicating text, and image data indicating an image is also input. The data generation model 58 performs inference on the input inference data in accordance with the instruction indicated by the prompt, and outputs the inference result in a data format such as voice data and text data. Here, inference refers to, for example, analysis, classification, prediction, and / or summarization.

[2007] In the above embodiment, an example was given in which the specific processing is performed by the data processing device 12, but the technology of the present disclosure is not limited to this, and the specific processing may be performed by the headset type terminal 314.

[2008] [Fourth embodiment]

[2009] FIG. 7 shows an example of the configuration of a data processing system 410 according to the fourth embodiment.

[2010] 7, a data processing system 410 includes a data processing device 12 and a robot 414. An example of the data processing device 12 is a server.

[2011] The data processing device 12 includes a computer 22, a database 24, and a communication I / F 26. The computer 22 is an example of a "computer" according to the technology of the present disclosure. The computer 22 includes a processor 28, a RAM 30, and a storage 32. The processor 28, the RAM 30, and the storage 32 are connected to a bus 34. The database 24 and the communication I / F 26 are also connected to the bus 34. The communication I / F 26 is connected to a network 54. Examples of the network 54 include a WAN (Wide Area Network) and / or a LAN (Local Area Network).

[2012] The robot 414 includes a computer 36, a microphone 238, a speaker 240, a camera 42, a communication I / F 44, and a control target 443. The computer 36 includes a processor 46, a RAM 48, and a storage 50. The processor 46, the RAM 48, and the storage 50 are connected to a bus 52. The microphone 238, the speaker 240, the camera 42, and the control target 443 are also connected to the bus 52.

[2013] The microphone 238 receives instructions and the like from the user 20 by receiving voice uttered by the user 20. The microphone 238 captures the voice uttered by the user 20, converts the captured voice into audio data, and outputs it to the processor 46. The speaker 240 outputs audio in accordance with instructions from the processor 46.

[2014] Camera 42 is a small digital camera equipped with an optical system including a lens, aperture, and shutter, and an imaging element such as a CMOS (Complementary Metal-Oxide-Semiconductor) image sensor or a CCD (Charge Coupled Device) image sensor, and captures images of the surroundings of user 20 (for example, an imaging range defined by an angle of view equivalent to the field of vision of a typical healthy person).

[2015] The communication I / F 44 is connected to a network 54. The communication I / Fs 44 and 26 are responsible for the exchange of various information between the processor 46 and the processor 28 via the network 54. The exchange of various information between the processor 46 and the processor 28 using the communication I / Fs 44 and 26 is carried out in a secure state.

[2016] The control object 443 includes a display device, LEDs in the eyes, and motors for driving the arms, hands, and feet. The posture and gestures of the robot 414 are controlled by controlling the motors of the arms, hands, and feet. Some of the emotions of the robot 414 can be expressed by controlling these motors. In addition, the facial expressions of the robot 414 can also be expressed by controlling the light emission state of the LEDs in the eyes of the robot 414.

[2017] Fig. 8 shows an example of the main functions of the data processing device 12 and the robot 414. As shown in Fig. 8, in the data processing device 12, a specific process is performed by the processor 28. A specific process program 56 is stored in the storage 32.

[2018] The specific processing program 56 is an example of a "program" according to the technology of the present disclosure. The processor 28 reads the specific processing program 56 from the storage 32 and executes the read specific processing program 56 on the RAM 30. The specific processing is realized by the processor 28 operating as a specific processing unit 290 in accordance with the specific processing program 56 executed on the RAM 30.

[2019] The storage 32 stores a data generation model 58 and an emotion identification model 59. The data generation model 58 and the emotion identification model 59 are used by the identification processing unit 290.

[2020] In the robot 414, the processor 46 performs the reception output process. A reception output program 60 is stored in the storage 50. The processor 46 reads the reception output program 60 from the storage 50 and executes the read reception output program 60 on the RAM 48. The reception output process is realized by the processor 46 operating as the control unit 46A in accordance with the reception output program 60 executed on the RAM 48.

[2021] Next, a description will be given of the specific processing performed by the specific processing unit 290 of the data processing device 12. In the following description, the data processing device 12 will be referred to as a "server" and the robot 414 will be referred to as a "terminal."

[2022] This system allows users to create original video content based on their preferences, share it with other users, and receive ratings. Specific embodiments of each component of the system will be described below.

[2023] User Interface Means

[2024] Device:

[2025] The terminal provides the interface through which the user enters input into the system. The user can access their account by opening the application and entering their credentials on the login screen. On first use, the user is prompted to select their preferred genre and theme.

[2026] A means of identifying a genre or theme

[2027] server:

[2028] The system receives and analyzes information about genres and themes entered by the user to identify genres and themes based on the user's preferences. For example, if the user selects "action" and "fantasy," it sets parameters for generating video content related to these genres.

[2029] A means of automatically generating video content

[2030] server:

[2031] The video generation AI module is activated and automatically generates video content based on the genre and theme selected by the user. The AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if the user selects "action + fantasy," a video containing an episode of a warrior fighting a dragon will be generated.

[2032] Means of preservation

[2033] server:

[2034] The generated video content is temporarily saved and then saved in a dedicated folder for the user, allowing the user to access the generated video at any time.

[2035] Means of playback or distribution

[2036] Device:

[2037] The generated video content can be streamed from the server to the device or downloaded directly to the device, allowing users to watch the video in real time on their own devices.

[2038] A means of accepting user ratings

[2039] Device:

[2040] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[2041] server:

[2042] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[2043] A way to update genre rankings

[2044] server:

[2045] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[2046] How to upload to a publishing platform

[2047] Device:

[2048] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[2049] server:

[2050] Store uploaded video content on the platform and make it accessible to other users.

[2051] How to share with other users

[2052] server:

[2053] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[2054] A means to create and store user profiles

[2055] server:

[2056] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[2057] A means of making personalized content suggestions

[2058] server:

[2059] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[2060] The above is a detailed description of the system in the "Description of Embodiments."

[2061] The processing flow will be explained below.

[2062] Step 1:

[2063] Device:

[2064] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[2065] Step 2:

[2066] server:

[2067] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[2068] Step 3:

[2069] Device:

[2070] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[2071] Step 4:

[2072] Device:

[2073] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[2074] Step 5:

[2075] server:

[2076] The received request is analyzed and the necessary parameters are set for the AI ​​model, for example, setting video generation based on "action" and "fantasy."

[2077] Step 6:

[2078] server:

[2079] The video generation AI module is activated to generate a scenario, construct a scene, and synthesize audio. Specifically, it automatically generates a video containing an episode of a warrior fighting a dragon.

[2080] Step 7:

[2081] server:

[2082] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[2083] Step 8:

[2084] server:

[2085] The terminal is notified that the video content has been saved.

[2086] Step 9:

[2087] Device:

[2088] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[2089] Step 10:

[2090] Device:

[2091] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[2092] Step 11:

[2093] User:

[2094] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[2095] Step 12:

[2096] Device:

[2097] After the user enters their rating and comments, the data is sent to the server.

[2098] Step 13:

[2099] server:

[2100] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[2101] Step 14:

[2102] User:

[2103] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[2104] Step 15:

[2105] Device:

[2106] The upload procedure for the video content is carried out and an upload request is sent to the server.

[2107] Step 16:

[2108] server:

[2109] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[2110] These are the specific processing steps involved in generating and sharing a video work based on a user request.

[2111] Example 1

[2112] Next, a description will be given of Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the robot 414 will be referred to as a "terminal."

[2113] In the creation and sharing of digital media content, personalized content delivery based on user preferences is required, but existing systems do not effectively generate, evaluate, and share content that reflects user preferences. Furthermore, the accuracy of recommendation systems and ranking updates based on user ratings also needs to be improved.

[2114] The specific processing by the specific processing unit 290 of the data processing device 12 in the first embodiment is realized by the following means.

[2115] In this invention, the server includes a user interface means for accepting user input, a means for identifying preferred categories or topics based on the input, a means for automatically generating digital media content based on the categories or topics, a means for saving the generated digital media content, a means for playing or distributing the saved digital media content to the user, a means for accepting user ratings, and a means for updating category rankings based on the ratings, thereby enabling efficient generation, rating, sharing, and ranking updates of personalized digital media content based on user preferences.

[2116] The "user interface means" is a device or program that provides an interface for a user to input information to the system.

[2117] A "category" is a concept that refers to a genre or topic of digital media content selected by a user.

[2118] A "topic" is a concept that refers to a specific theme or subject of digital media content selected by a user.

[2119] "Digital media content" refers to media data in digital format, such as video, audio, and images, that is generated based on user preferences.

[2120] An "automatic generation means" is a device or program that automatically creates digital media content based on category or topic information entered by a user.

[2121] A "storage means" is a device or program for temporary and long-term storage of the generated digital media content.

[2122] A "delivery means" is a device or program that provides stored digital media content to a user's terminal in a streaming or downloadable format.

[2123] The "means for receiving ratings" is a device or program that collects rating information on digital media content from users.

[2124] The "means for updating the ranking by category" is a device or program that updates the ranking of digital media content for each category based on evaluation information collected from users.

[2125] A "sharing platform" is an online service or website for sharing user-generated digital media content with other users.

[2126] MODE FOR CARRYING OUT THE INVENTION

[2127] The system automates the creation, storage, distribution, rating, and sharing of digital media content. It uses user input to identify preferred categories and topics, and then uses generative AI models to automatically generate content. A specific implementation of this system is described below.

[2128] 1. User Interface Methods

[2129] (Terminal)

[2130] The terminal provides an application as a user interface for the user to input data to the system. The user accesses the system by starting this application and entering a user name and password on the login screen.

[2131] 2. How to identify your favorite categories and topics

[2132] (Terminal)

[2133] On first use, the device will display an interface for setting user preferences, where users can select their preferred categories and topics, for example, "action" and "fantasy."

[2134] (server)

[2135] The server receives and analyzes this selection information and stores it in a database. This data is used to generate content using the AI ​​model described below.

[2136] 3. Content Generation Methods

[2137] (server)

[2138] The server generates prompts based on the user's preferences and sends them to a generative AI model (e.g., OpenAI GPT-4, DeepMind's video generation model).

[2139] 4. Video content generation

[2140] (server)

[2141] The server's generative AI model automatically generates video content based on the input prompt. For example, for a user who likes "action" and "fantasy," a video containing a story about a warrior fighting a dragon is generated. Video generation involves the processes of scenario generation, scene construction, and voice synthesis.

[2142] 5. How content is stored

[2143] (server)

[2144] The server temporarily stores the generated video content and then saves it in the user's dedicated folder. The video files are managed using a database management system (e.g., MySQL, PostgreSQL).

[2145] 6. Content playback and distribution methods

[2146] (Terminal)

[2147] The device allows users to play or download stored video content, often using streaming technology and a media player (e.g., VLC, QuickTime).

[2148] 7. Method of accepting evaluations

[2149] (Terminal)

[2150] After watching the video, the device displays a user rating input interface, allowing users to enter star ratings and comments.

[2151] (server)

[2152] The server receives user ratings and stores them in a database, which is used to update rankings for each category.

[2153] 8. How to update rankings by category

[2154] (server)

[2155] The server updates the rankings for each category based on the collected rating information, and then makes this ranking information available to other users, who can then recommend highly rated content.

[2156] 9. How to share content

[2157] (Terminal)

[2158] The terminal provides an interface for uploading the generated digital media content to a sharing platform.

[2159] (server)

[2160] The server manages the uploaded content and makes it accessible to other users, generating shareable links for easy sharing via social media or email.

[2161] 10. How to create and store user profiles

[2162] (server)

[2163] A profile is created and saved based on the user's preferences and viewing history, allowing for personalized content suggestions the next time the user uses the service.

[2164] 11. Means of personalized content recommendations

[2165] (server)

[2166] Based on a saved user profile, the app suggests digital media content that is most suitable for the user, for example, if a user previously indicated that they like "action," the app will suggest the latest action movies.

[2167] Examples of concrete examples and prompts

[2168] Examples:

[2169] The user selects "action" and "fantasy" and enters the prompt "Generate a movie about a warrior fighting a dragon." Based on this prompt, the server's generative AI model automatically generates video content and saves it in the user's dedicated folder. This video content can be viewed on the device and shared with other users.

[2170] Example prompt sentence:

[2171] "Generate footage of action scenes in fantasy movies"

[2172] "Suggest your next watchlist based on your favorite movie genres."

[2173] This system allows users to easily create, save, rate, and share content tailored to their preferences. This specific embodiment will clarify how the invention actually works.

[2174] The flow of the identification process in the first embodiment will be described with reference to FIG.

[2175] Step 1:

[2176] (User):

[2177] The user launches the application and enters their username and password on the login screen.

[2178] input:

[2179] Username, Password

[2180] output:

[2181] Authentication Request

[2182] (Device):

[2183] The terminal accepts user input and sends an authentication request to the server.

[2184] Specific behavior:

[2185] Displaying the login screen

[2186] Obtaining the entered username and password

[2187] Sending an authentication request to the server

[2188] Step 2:

[2189] (server):

[2190] The server checks the received authentication request against its database and returns the authentication result.

[2191] input:

[2192] Authentication Request

[2193] output:

[2194] Authentication result (success / failure)

[2195] (Device):

[2196] The terminal performs screen transitions based on the authentication results.

[2197] Specific behavior:

[2198] Receiving the authentication result

[2199] If authentication is successful, the main screen is displayed. If authentication fails, an error message is displayed.

[2200] Step 3:

[2201] (User):

[2202] On first use, users make selections in an interface that allows them to choose their preferred categories and topics.

[2203] input:

[2204] Category selection (e.g., "Action"), Topic selection (e.g., "Fantasy")

[2205] output:

[2206] Selected Data

[2207] (Device):

[2208] The terminal transmits the selected data to the server.

[2209] Specific behavior:

[2210] Receiving user selections

[2211] Sending Selected Data to the Server

[2212] Step 4:

[2213] (server):

[2214] The server stores the received selection data in a database.

[2215] input:

[2216] Selected Data

[2217] output:

[2218] Results saved in the database

[2219] Specific behavior:

[2220] Analysis of selected data

[2221] Saving to the database

[2222] Step 5:

[2223] (User):

[2224] The user inputs the content of the video they want to generate on the prompt sentence input screen.

[2225] input:

[2226] Prompt text (e.g., "Generate a movie of a warrior fighting a dragon")

[2227] output:

[2228] Prompt Data

[2229] (Device):

[2230] The terminal sends the prompt data to the server.

[2231] Specific behavior:

[2232] Receiving a prompt

[2233] Sending prompt data to the server

[2234] Step 6:

[2235] (server):

[2236] The server inputs the prompt data into the generative AI model and begins generating video content.

[2237] input:

[2238] Prompt Data

[2239] output:

[2240] Generated video content

[2241] Specific behavior:

[2242] Parsing prompt data

[2243] Input to generative AI models

[2244] Video content generation

[2245] Step 7:

[2246] (server):

[2247] The server temporarily stores the generated video content, and then stores it in a folder dedicated to the user.

[2248] input:

[2249] Generated video content

[2250] output:

[2251] Saved results in user folder

[2252] Specific behavior:

[2253] Temporary storage of video content

[2254] Move to user-specific folder

[2255] Step 8:

[2256] (Device):

[2257] The user plays or downloads the stored video content.

[2258] input:

[2259] User requests (playback, download)

[2260] output:

[2261] Playback and download of video content

[2262] Specific behavior:

[2263] Click the play button to stream

[2264] Click the download button to save the video to your device

[2265] Step 9:

[2266] (User):

[2267] After viewing the video, the user inputs a rating.

[2268] input:

[2269] Rating data (stars, comments)

[2270] output:

[2271] Evaluation data submission

[2272] (Device):

[2273] The terminal transmits the evaluation data to the server.

[2274] Specific behavior:

[2275] View the rating interface

[2276] Receiving and sending evaluation data to the server

[2277] Step 10:

[2278] (server):

[2279] The server stores the received rating data in a database and updates the rankings by category.

[2280] input:

[2281] Evaluation Data

[2282] output:

[2283] Updated ranking data

[2284] Specific behavior:

[2285] Analysis of evaluation data

[2286] Saving to a database

[2287] Category ranking updates

[2288] Step 11:

[2289] (User):

[2290] Users upload the generated video content to a sharing platform.

[2291] input:

[2292] Upload Request

[2293] output:

[2294] Upload completion notification

[2295] (Device):

[2296] Pressing the upload button sends the video content to the server.

[2297] Specific behavior:

[2298] Click the upload button

[2299] Sending video content to the server

[2300] Step 12:

[2301] (server):

[2302] The server manages the uploaded content, sets it up so that other users can access it, and generates a share link so that it can be shared via social media or email.

[2303] input:

[2304] Uploaded content

[2305] output:

[2306] Shared Links

[2307] Specific behavior:

[2308] Organizing uploaded content

[2309] Generate shareable links for social media and email

[2310] Step 13:

[2311] (server):

[2312] The server creates and stores a profile based on the user's preferences and viewing history.

[2313] input:

[2314] User preferences and viewing history

[2315] output:

[2316] User Profile

[2317] Specific behavior:

[2318] User data analysis

[2319] Creating and Saving a Profile

[2320] Step 14:

[2321] (server):

[2322] The server makes personalized content suggestions based on the stored profile.

[2323] input:

[2324] User Profile

[2325] output:

[2326] Suggested Content List

[2327] Specific behavior:

[2328] Viewing Profiles

[2329] Selection of proposed content

[2330] The above are the specific steps of the program processing of this system.

[2331] (Application example 1)

[2332] Next, a description will be given of Application Example 1. In the following description, the data processing device 12 will be referred to as a "server" and the robot 414 will be referred to as a "terminal."

[2333] Conventional video content generation and distribution systems lack the functionality to allow users to easily create original video content based on their preferences and share it with other users. Furthermore, the lack of personalized content suggestions and rating systems hinders the improvement of the user experience.

[2334] The specific processing by the specific processing unit 290 of the data processing device 12 in the application example 1 is realized by the following means.

[2335] In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre rankings based on the ratings, a means for creating and saving a user profile, a means for suggesting personalized content based on the profile, a means for a user to log in and select a genre using an application installed on a smartphone or smart glasses, a means for saving and sharing the generated video in cloud storage, and a means for generating video using an AI model by inputting a prompt. This allows users to easily generate video content according to their preferences, share it with other users, receive ratings, and improve the accuracy of content suggestions from the next time onwards.

[2336] The "user interface means" is an interface that allows a user to input information into the system, and is a function that allows operation through a smartphone or smart glasses, for example.

[2337] The "means for identifying a favorite genre or theme" is a function of the system that analyzes and identifies a user's favorite genre or theme based on information input by the user.

[2338] The "means for automatically generating video content" is an AI module for generating video content based on the genre or theme selected by the user.

[2339] The "means for saving the generated video content" is a function for temporarily or permanently storing the generated video in a recording device.

[2340] The "means for playing or distributing stored video content to users" is a function for providing stored video so that users can view it.

[2341] The "means for accepting user ratings" is a function including an interface for users to input ratings and comments on video content.

[2342] The "means for updating genre rankings" is a function for automatically updating genre rankings based on user ratings.

[2343] "Means for creating and saving user profiles" refers to a function that creates individual profiles based on the user's preferences and viewing history and saves them in a database.

[2344] The "means for making personalized content suggestions" is a function that suggests the most suitable video content to a user based on a saved user profile.

[2345] "Application installed on a smartphone or smart glasses" refers to software for a mobile device that allows a user to access the video content generation system.

[2346] "Means for saving and sharing the generated video in cloud storage" refers to a function for saving the generated video file to a remote server via the Internet and providing a sharing link from there.

[2347] "Means for generating images using an AI model by inputting a prompt sentence" refers to a function in which a user inputs a specific instruction sentence, and the AI ​​model generates an image based on that instruction.

[2348] This invention is a system that enables users to use mobile devices such as smartphones or smart glasses to create original video content based on their preferences, share it with other users, and receive ratings.

[2349] User Interface Means

[2350] The device is equipped with an interface that allows users to input information into the system. Users can access their account by opening the application on their smartphone or smart glasses and entering their authentication information on the login screen. On first use, a screen appears where users can select their preferred genre (e.g., "action" or "fantasy") and theme.

[2351] A means of identifying a genre or theme

[2352] The server receives the genre and theme information entered by the user, analyzes it, and identifies the user's preferences. For example, if the user selects "thriller" and "mystery," parameters for generating video content related to these genres are set.

[2353] A means of automatically generating video content

[2354] The server launches a video generation AI model and automatically generates video content based on the genre and theme selected by the user. The AI ​​generates the scenario, constructs the scene, and synthesizes the audio. For example, if a user selects "Thriller + Mystery," a short video of a detective solving a mysterious case will be generated. The AI ​​model used here utilizes machine learning frameworks such as TensorFlow and PyTorch.

[2355] Storage and playback methods

[2356] The server temporarily stores the generated video content in cloud storage such as Google Cloud Storage, and then saves it in the user's dedicated folder. Users can then view the video in real time on their smartphones or smart glasses.

[2357] How to receive ratings

[2358] The device provides an interface for users to input ratings after watching videos. Users can enter ratings from 1 to 5 stars and comments. The server records the received ratings and comments in a database such as Firebase and analyzes them to update the rankings by genre.

[2359] How to update the rankings

[2360] The server automatically updates the rankings of video content by genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[2361] How to store and share content

[2362] The generated video content is stored in cloud storage and a sharing link is generated, allowing users to easily share their work via social media or email.

[2363] User profiles and personalized content suggestions

[2364] The server creates and stores a profile based on the user's preferences and viewing history. The next time the user logs in, personalized content suggestions are made based on this profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[2365] Examples of concrete examples and prompts

[2366] After a user logs in, they can select their preferred genre, create, watch, and rate a short video. The following prompts are used:

[2367] Input: "Generate a short film that mixes horror and fantasy genres. My Firebase UID is abcdef."

[2368] Output: "Generating movie... The URL for the generated movie is https: / / storage.googleapis.com / abcdef / generated_video.mp4."

[2369] This system allows users to easily create and enjoy video content according to their preferences.

[2370] The flow of the specific processing in the application example 1 will be described with reference to FIG.

[2371] Step 1:

[2372] A user launches an application on their smartphone or smart glasses and accesses the login screen. The input is the user's authentication information, and the output is the authentication success / failure result. The server validates the user's authentication information using Firebase Auth and returns the user ID if authentication is successful.

[2373] Step 2:

[2374] After logging in, the user selects their preferred genre and theme. The input is the genre and theme selected by the user, and the output is the data of the selected genre and theme. The terminal sends the selected genre and theme to the server.

[2375] Step 3:

[2376] The server analyzes the received genre and theme information and sets the parameters of the AI ​​model. The input is genre and theme data, and the output is the parameter settings of the AI ​​model. The server uses TensorFlow or PyTorch to initialize the AI ​​model and set the settings based on the received data.

[2377] Step 4:

[2378] The server launches the configured AI model and generates video content. The input is the AI ​​model's parameters and generation instructions (prompt text), and the output is the generated video file. The AI ​​model generates a scenario, constructs scenes, and synthesizes audio. In this step, if the user inputs, for example, "Please generate a short film that combines the horror and fantasy genres," a video of the corresponding scenario will be automatically generated.

[2379] Step 5:

[2380] The server saves the generated video file in cloud storage (e.g., Google Cloud Storage). The input is the generated video file, and the output is the cloud storage URL. The server uploads the video file to cloud storage and generates a URL link for the saved file.

[2381] Step 6:

[2382] The server provides the user with a URL for the generated video. The input is the URL of the cloud storage, and the output is the URL sent to the user's device. The server then notifies the smartphone or smart glasses of the generated URL, allowing the user to access the video.

[2383] Step 7:

[2384] Users watch videos and rate them through a rating input screen. The input is the user's rating information (e.g., a rating from 1 to 5 stars, comments), and the output is the rating data. The device accepts the user's rating and sends it to the server.

[2385] Step 8:

[2386] The server records the received rating data and updates the genre rankings. The input is the rating data received from users, and the output is the updated ranking information. The server saves the rating data in Firebase or MongoDB and automatically updates the rankings based on it.

[2387] Step 9:

[2388] The server creates and stores user profiles. The input is the user's selection history and rating data, and the output is a user profile. The server creates a profile based on the viewing history and rating information and stores it in a database.

[2389] Step 10:

[2390] The server will then suggest personalized content based on the saved user profile from the next time onwards. The input is the user profile, and the output is the suggested content information. The server analyzes the user profile, generates information to suggest appropriate video content, and provides it to the user.

[2391] Through these steps, users can efficiently create, view, and rate video content according to their preferences, improving the user experience.

[2392] Furthermore, an emotion engine that estimates the user's emotion may be further combined. That is, the identification processing unit 290 may estimate the user's emotion using the emotion identification model 59, and perform identification processing using the user's emotion.

[2393] This system not only generates original video content based on the user's preferences, but also recognizes the user's emotions and proposes and generates content. Specific embodiments of each component of the system are described below.

[2394] User Interface Means

[2395] Device:

[2396] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening an application and entering their credentials on a login screen.

[2397] A means of identifying a genre or theme

[2398] server:

[2399] The genre and theme information input by the user or the emotion information recognized by the emotion engine means is analyzed to identify a genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the appropriate video content is identified.

[2400] Emotion Engine Means

[2401] Device:

[2402] The terminal captures the user's facial expressions and voice in real time using a built-in camera and microphone, and the emotion engine means analyzes the captured images. The emotion engine means recognizes emotions from the user's facial expressions and tone of voice, and transmits the information to a server.

[2403] server:

[2404] The emotion data sent from the emotion engine means is received and analyzed. Based on the analysis results, the system recommends genres and themes that are most suitable for the user.

[2405] A means of automatically generating video content

[2406] server:

[2407] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[2408] Means of preservation

[2409] server:

[2410] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[2411] Means of playback or distribution

[2412] Device:

[2413] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[2414] A means of accepting user ratings

[2415] Device:

[2416] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[2417] server:

[2418] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[2419] A way to update genre rankings

[2420] server:

[2421] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[2422] How to upload to a publishing platform

[2423] Device:

[2424] It provides a button to upload user-generated video content to a publishing platform, allowing users to share their creations with others.

[2425] server:

[2426] Uploaded video content is stored on the posting platform and made accessible to other users.

[2427] How to share with other users

[2428] server:

[2429] It manages posted video content so that other users can view and rate it, and also generates sharing links for social media and email, making it easy for users to share.

[2430] A means to create and store user profiles

[2431] server:

[2432] A profile is created and saved based on the user's preferences and viewing history, allowing personalized video content to be suggested the next time the service is used.

[2433] A means of making personalized content suggestions

[2434] server:

[2435] Based on the saved user profile, the app suggests the most suitable video content for the user. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[2436] The above is a detailed description of the system in the "Description of Embodiments."

[2437] The processing flow will be explained below.

[2438] Step 1:

[2439] Device:

[2440] The user launches the application on the device and enters their authentication information (user ID and password) on the login screen.

[2441] Step 2:

[2442] server:

[2443] The login information is received and the user is authenticated. If authentication is successful, the user's profile information (preferred genres and viewing history) is retrieved from the database and sent to the device.

[2444] Step 3:

[2445] Device:

[2446] The user selects a preferred genre or theme on the genre selection screen. For example, the user selects "action" and "fantasy."

[2447] Step 4:

[2448] Device:

[2449] When the user presses the video generation request button, information on the selected genre and theme is sent to the server.

[2450] Step 5:

[2451] Device:

[2452] The user selects the Enable Emotion Engine option in the settings screen, which prepares the device's built-in camera and microphone to collect emotion data.

[2453] Step 6:

[2454] Device:

[2455] While watching a video or requesting a new video, the device captures the user's facial expressions and voice through the camera and microphone, analyzes the emotional data in real time, and temporarily stores this data.

[2456] Step 7:

[2457] server:

[2458] Emotion data transmitted from the terminal is received and analyzed, and the emotion engine means identifies the user's emotion (e.g., excitement, joy, sadness).

[2459] Step 8:

[2460] server:

[2461] Based on the analysis results, the system recommends genres and themes that suit the user's current emotions. For example, if the user is recognized as "excited," the system will identify "action" and "thriller" genres.

[2462] Step 9:

[2463] server:

[2464] Based on the user's selection and recognized emotions, the video generation AI module is activated to generate scenarios, construct scenes, and synthesize audio.

[2465] Specifically, it automatically generates footage including episodes of warriors fighting dragons.

[2466] Step 10:

[2467] server:

[2468] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[2469] Step 11:

[2470] server:

[2471] The terminal is notified that the video content has been saved.

[2472] Step 12:

[2473] Device:

[2474] The user will be notified that the video has been generated and can select either the streaming play button or the download button.

[2475] Step 13:

[2476] Device:

[2477] If you select streaming playback, the video will be played in real time from the server. If you select download, the video data will be downloaded and saved to your device.

[2478] Step 14:

[2479] User:

[2480] The generated video content is viewed, and after viewing is complete, a rating (1 to 5 stars) and comments are entered on the rating screen.

[2481] Step 15:

[2482] Device:

[2483] After the user enters their rating and comments, the data is sent to the server.

[2484] Step 16:

[2485] server:

[2486] The received ratings and comments are recorded in a database and analyzed, which then updates the rankings by genre.

[2487] Step 17:

[2488] User:

[2489] When a user wants to share video content with others, they press a button to upload it to a posting platform.

[2490] Step 18:

[2491] Device:

[2492] The upload procedure for the video content is carried out and an upload request is sent to the server.

[2493] Step 19:

[2494] server:

[2495] Uploaded video content is stored on the posting platform so that other users can view and rate it, and a share link is generated so that users can share it via social media or email.

[2496] These are the specific processing steps for generating and sharing a video work based on the user's requests and emotions.

[2497] Example 2

[2498] Next, a description will be given of Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the robot 414 will be referred to as a "terminal."

[2499] Conventional video content generation systems have had difficulty generating video content that fully reflects a user's preferences and momentary emotions. Furthermore, sharing the generated content with other users is time-consuming, and it is difficult to reflect the user's ratings in the system. Furthermore, they lacked functionality for making personalized suggestions based on user profiles, leaving a need for an improved user experience.

[2500] The identification process by the identification processing unit 290 of the data processing device 12 in Example 2 is realized by the following means. In this invention, the server includes an information terminal means for accepting user input, a data analysis means for identifying a preferred genre or theme based on the input, an emotion recognition means for capturing the user's facial expressions and voice in real time using a built-in camera and microphone and analyzing their emotions, a means including a generative AI model for automatically generating video content based on the genre, theme, and emotion data, a data storage means for saving the generated video content, a playback / distribution means for playing or distributing the saved video content to users, an evaluation input means for accepting user ratings, and a ranking update means for updating genre rankings based on the ratings. This enables the generation and sharing of personalized video content that reflects the user's preferences and emotions, and further enables appropriate suggestions based on the user profile.

[2501] The "information terminal means for accepting user input" is a means for providing an interface for the user to operate or give instructions to the system.

[2502] The "data analysis means for identifying a preferred genre or theme based on the input" is a means for analyzing the information input by the user and identifying a genre or theme according to the user's preferences.

[2503] "Emotion recognition means that uses a built-in camera and microphone to capture the user's facial expressions and voice in real time and analyze their emotions" refers to a means of recognizing the user's emotional state by using the camera and microphone attached to the device to photograph and record the user's facial expressions and voice, and analyzing that data.

[2504] "Means including a generative AI model that automatically generates video content based on the genre, theme, and emotional data" refers to means including an artificial intelligence model that receives a specified genre, theme, and emotional data as input and generates video content based on that input.

[2505] The "data storage means for storing the generated video content" is a means for temporarily storing the generated video content and for long-term storage as required.

[2506] "Playback and distribution means for playing or distributing the stored video content to users" refers to means for playing or distributing the stored video content in a form that can be viewed by users.

[2507] The "evaluation input means for accepting user evaluations" is a means for accepting and recording user evaluations and feedback on video content.

[2508] The "ranking update means for updating the genre rankings based on the ratings" is a means for updating the rankings of video content by genre to the latest status based on the rating data received from users.

[2509] This system generates original video content based on the user's preferences, and can also suggest and generate content by recognizing the user's emotions. Specific embodiments of each component of the system will be described below.

[2510] User Interface Means

[2511] Device:

[2512] A device provides an interface for users to enter input into the system. Users can access their account by opening an application and entering their credentials on a login screen. Devices include smartphones, tablets, and PCs.

[2513] A means of identifying a genre or theme

[2514] server:

[2515] The server analyzes the genre and theme information entered by the user or the emotion information recognized by the emotion recognition means to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," the server identifies the appropriate video content.

[2516] emotion recognition means

[2517] Device:

[2518] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine. The device recognizes emotions from the user's facial expressions and tone of voice, and sends the information to a server.

[2519] server:

[2520] The server receives and analyzes the emotion data sent from the emotion recognition means, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[2521] A means of automatically generating video content

[2522] server:

[2523] The server then activates the generative AI model to automatically generate video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion recognition engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[2524] Means of preservation

[2525] server:

[2526] The server temporarily stores the generated video content, then saves it in the user's dedicated folder. Information about the saved video content is recorded in a database.

[2527] Means of playback or distribution

[2528] Device:

[2529] The generated video content can be streamed from the server to the terminal or downloaded directly by the user for viewing.

[2530] A means of accepting user ratings

[2531] Device:

[2532] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[2533] server:

[2534] The server records and analyzes the received ratings and comments in a database, which then updates the genre rankings.

[2535] A way to update genre rankings

[2536] server:

[2537] The server updates the rankings of video content in each genre based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[2538] How to upload to a publishing platform

[2539] Device:

[2540] The device provides a button for uploading user-generated video content to a posting platform, allowing users to share their creations with others.

[2541] server:

[2542] The server stores the uploaded video content on the posting platform and makes it accessible to other users.

[2543] How to share with other users

[2544] server:

[2545] The server manages the posted video content so that other users can view and rate it, and also generates a sharing link for users to use on social media or by email, making it easy to share.

[2546] A means to create and store user profiles

[2547] server:

[2548] The server creates and stores a profile based on the user's preferences and viewing history, which enables personalized video content to be suggested the next time the user uses the service.

[2549] A means of making personalized content suggestions

[2550] server:

[2551] The server then suggests the most suitable video content for the user based on the stored user profile. For example, if a user previously indicated that they like "action," new action movies will be suggested.

[2552] Examples and prompts

[2553] Examples:

[2554] User Action: The user logs in, selects the genres "Action" and "Sci-Fi," and has a happy look on their face.

[2555] Server processing: The server receives the emotion data of "fun" from the emotion engine and automatically generates video content that includes action and science fiction elements.

[2556] Playback: The generated video content is streamed to the device.

[2557] Generate AI model prompt:

[2558] "The user likes the 'action' and 'sci-fi' genres, and their current emotion is 'fun'. Based on this condition, please generate a sci-fi video that contains exciting action scenes."

[2559] Through such a system, it becomes possible to provide users with video content that suits their mood and preferences at the time.

[2560] The flow of the identification process in the second embodiment will be described with reference to FIG.

[2561] Step 1:

[2562] User: The user opens the application on their device and enters their authentication information (username and password) on the login screen. This input includes text data such as username and password.

[2563] Terminal: The terminal sends the entered authentication information to the server using a secure communication protocol (e.g. HTTPS).

[2564] Server: The server receives the authentication information and authenticates it by comparing the username and password against the user information in its database.

[2565] Result: If authentication is successful, the server generates a session ID for the user and sends it to the terminal. The terminal displays the session ID and notifies the user that the login was successful.

[2566] Step 2:

[2567] User: After logging in, the user selects their preferred genre and theme from the displayed interface. This input includes text data such as "action" and "fantasy" as genre and theme.

[2568] Device: The device sends the selected information to the server, which is sent in JSON format.

[2569] Server: The server analyzes the received genre and theme information and stores it in a database, which is then used for further processing.

[2570] Step 3:

[2571] Device: The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time. This data includes image data and audio data.

[2572] Device: Passes the captured data to the emotion recognition engine.

[2573] Server: The server receives and analyzes the emotion data sent from the emotion recognition engine. Image processing algorithms and voice recognition algorithms are used for this analysis. The analysis result is the user's emotional state (e.g., "excited").

[2574] Step 4:

[2575] Server: The server takes genre, theme, and emotion data as input and runs the generative AI model. This input includes text data and emotion data.

[2576] Generative AI model: The generative AI model automatically generates video content based on this data. The output is a video content file.

[2577] Specific action: Send a prompt to the generative AI model saying, "The user likes the 'action' and 'fantasy' genres, and their current emotion is 'excitement.' Based on these conditions, please generate a video that includes exciting action scenes."

[2578] Step 5:

[2579] Server: The server temporarily stores the generated video content and then saves it in the user's dedicated folder. The server records the path and metadata of the saved file (e.g., creation date and time, user ID) in the database.

[2580] Device: A notification will appear on your device when the save is complete.

[2581] Step 6:

[2582] Device: The user plays or downloads the stored video content. When the user clicks the "Play" button, the device receives the video stream from the server and displays it in the video player. When the user clicks the "Download" button, the device saves the video file to local storage.

[2583] Step 7:

[2584] Terminal: After watching the video, an interface is displayed for the user to input a rating. The user can input a rating from 1 to 5 stars and a comment. This input includes numerical data and text data.

[2585] Device: The device sends the evaluation data to the server.

[2586] Server: The server receives the rating data and records it in a database. This data is used to update the rankings.

[2587] Step 8:

[2588] Server: The server updates the rankings of video content in each genre based on the received rating data. New rankings are generated and stored in the database.

[2589] On the device: The latest information is displayed so that users can view ranking information.

[2590] Step 9:

[2591] Terminal: Provides a button to upload user-generated video content to a posting platform. The user clicks the "Upload" button and selects the posting platform.

[2592] Server: The server receives the video file and transfers it to the designated uploading platform. Once uploaded, a public link is generated.

[2593] On your device: The user will see a notification that the upload is complete and a public link.

[2594] Step 10:

[2595] Server: Manages the posted video content so that other users can view and rate it. Also generates sharing links for social media and email, allowing users to easily share. Shared links are generated and sent to social media and email.

[2596] Step 11:

[2597] Server: The server creates and stores a profile based on the user's preferences and viewing history, including genre preferences and viewing history.

[2598] Specific operation: Collects user viewing history and genre preference information and records it in a database as a profile.

[2599] Step 12:

[2600] Server: The server uses a recommendation algorithm to suggest the most suitable video content for a user based on the stored user profile.

[2601] On the device: A screen will appear where the user can view a list of suggested content.

[2602] (Application example 2)

[2603] Next, a description will be given of Application Example 2. In the following description, the data processing device 12 will be referred to as a "server" and the robot 414 will be referred to as a "terminal."

[2604] Current content distribution services struggle to automatically generate and deliver personalized video content based on users' interests and emotions. Because they are unable to generate or suggest content that reflects a user's temporary emotional state in real time, it is difficult to provide users with an optimal entertainment experience. Furthermore, there is a lack of systems that effectively utilize user ratings as feedback and reflect them in rankings. Technology that can solve these problems and provide a more personalized user experience is needed.

[2605] The identification process by the identification processing unit 290 of the data processing device 12 in Application Example 2 is realized by the following means. In this invention, the server includes a user interface means for accepting user input, a means for identifying a preferred genre or theme based on the input, a means for automatically generating video content based on the genre or theme, a means for saving the generated video content, a means for playing or distributing the saved video content to the user, a means for accepting user ratings, a means for updating genre-specific rankings based on the ratings, a means for recognizing emotions from the user's facial expressions and voice, and a means for using a generative AI model for generating video content based on the recognized emotion information. This enables the generation and distribution of personalized video content based on the user's emotions, and also enables the updating of genre-specific rankings to reflect the user's ratings.

[2606] The "user interface means" is a means for providing an interface for a user to input to the system.

[2607] The "means for identifying a genre or theme" is a means for identifying a genre or theme that the user likes based on the user's input or emotional information.

[2608] The "means for automatically generating video content" is a means for automatically generating video content based on a genre or theme selected by a user or a recognized emotion.

[2609] The "means for saving" is a means for temporarily saving the generated video content and then saving it in a location accessible to the user.

[2610] "Means for playing or distributing" refers to means for allowing users to stream or download stored video content.

[2611] The "means for receiving a rating" is a means for allowing a user to input a rating for the video content that the user has viewed, and for the rating to be received within the system.

[2612] The "means for updating the genre rankings" is a means for updating the rankings of video content in each genre based on user ratings.

[2613] The "means for recognizing emotions" refers to a means for recognizing emotions from the user's facial expressions and tone of voice and analyzing that information.

[2614] A "generative AI model" is an artificial intelligence model used to generate video content based on input emotional information and prompts.

[2615] This invention relates to a system that automatically generates and distributes original video content based on user input and emotional information. The system includes a user interface, a means for identifying genres and themes, a means for recognizing emotions, a means for using a generative AI model, a means for automatically generating video content, a means for saving, a means for playing or distributing, a means for accepting ratings, and a means for updating genre-specific rankings. Each component of this system is described in detail below.

[2616] User Interface Means

[2617] Device:

[2618] The terminal provides an interface for users to enter input into the system. Users can access their accounts by opening the smartphone application and entering their credentials on the login screen.

[2619] A means of identifying a genre or theme

[2620] server:

[2621] The system analyzes the genre and theme information entered by the user, or the emotional information acquired by the emotion recognition means, to identify the genre and theme based on the user's preferences and mood at the time. For example, if the user selects "action" and "fantasy," or if the emotion engine recognizes the user's emotion as "excitement," it will identify video content that matches that.

[2622] A means of recognizing emotions

[2623] Device:

[2624] The device uses a built-in camera and microphone to capture the user's facial expressions and voice in real time, which are then analyzed by an emotion recognition engine, which recognizes emotions from the user's facial expressions and tone of voice and sends the information to a server.

[2625] server:

[2626] It receives and analyzes the emotional data sent from the emotion engine, and based on the analysis results, recommends genres and themes that are most suitable for the user.

[2627] A means of automatically generating video content

[2628] server:

[2629] The video generation AI module is activated and automatically generates video content based on the user's selected genre or theme, or the user's recognized emotion. For example, if the emotion engine recognizes the user's emotion as "excitement," it will generate a video with many action scenes.

[2630] Means of preservation

[2631] server:

[2632] The generated video content is temporarily saved, and then saved in the user's dedicated folder. Information about the saved video content is recorded in a database.

[2633] Means of playback or distribution

[2634] Device:

[2635] The generated video content can be streamed from the server to the terminal or downloaded directly for viewing.

[2636] A means of accepting user ratings

[2637] Device:

[2638] After watching the video, the user is presented with an interface to enter a rating, for example, a 1-5 star rating or a comment.

[2639] server:

[2640] The ratings and comments received are recorded in a database and analyzed, which then updates the rankings by genre.

[2641] A way to update genre rankings

[2642] server:

[2643] The rankings of video content in each genre are updated based on user ratings. The ranking information is made public to other users, and highly rated works are recommended to other users.

[2644] Specific examples

[2645] For example, if a user expresses the emotion "happy," the emotion recognition engine sends that emotion to the server, which then sends prompts like "comedy" or "adventure" to the generative AI model. In this way, video content that best suits the user's current emotional state is generated and delivered.

[2646] Example prompt sentence:

[2647] "Emotion: Happy, Genre: Comedy, Adventure"

[2648] The above system enables the generation and distribution of personalized video content based on the user's emotions, improving the user's entertainment experience.

[2649] The flow of the specific processing in the application example 2 will be described with reference to FIG.

[2650] Step 1:

[2651] The terminal accepts the user's login. The user enters authentication information (user name and password) using the user interface and sends it to the server. Based on this input data, the server authenticates the user, and if authentication is successful, obtains the user's profile information and sends it to the terminal.

[2652] Step 2:

[2653] The terminal accepts genre and theme selections from the user. The user uses the interface to select their preferred genre (e.g., action, comedy) or theme (e.g., fantasy, suspense), and sends the input data to the server. The server then analyzes the user's preferences based on this data.

[2654] Step 3:

[2655] The device captures the user's facial expressions and voice using a built-in camera and microphone. The captured data is processed to recognize emotions. Specifically, facial expression data is analyzed using OpenCV, and voice data is processed using TensorFlow / Keras. The recognized emotion data is then sent to a server.

[2656] Step 4:

[2657] The server generates a prompt based on the received genre and theme selection data and emotional data. This prompt includes information such as "Emotion: happy, Genre: comedy, adventure." This prompt is then input into a generative AI model to automatically generate video content.

[2658] Step 5:

[2659] The server temporarily stores the generated video content and stores it in a dedicated folder for each user. Metadata related to the content storage is also recorded in the database.

[2660] Step 6:

[2661] The terminal allows the playback or download of video content stored by the server, and for the user to view the content, streaming playback is initiated or the content is downloaded to the terminal.

[2662] Step 7:

[2663] The terminal provides an interface for users to input ratings after viewing. Users input ratings (1 to 5 stars) and text comments, and then send the input data to the server. The server records the ratings and comments in a database.

[2664] Step 8:

[2665] The server updates the genre rankings based on the received rating data. In the ranking update process, it calculates the average rating score for each piece of content and generates new ranking information. This ranking information is saved in a database and made public to other users.

[2666] This enables the generation and distribution of personalized video content based on user emotions and the updating of rankings based on user ratings.

[2667] The specific processing unit 290 transmits the result of the specific processing to the robot 414. In the robot 414, the control unit 46A causes the speaker 240 and the control target 443 to output the result of the specific processing. The microphone 238 acquires voice indicating a user input regarding the result of the specific processing. The control unit 46A transmits voice data indicating the user input acquired by the microphone 238 to the data processing device 12. In the data processing device 12, the specific processing unit 290 acquires the voice data.

[2668] The data generation model 58 is a so-called generative AI (Artificial Intelligence). An example of the data generation model 58 is ChatGPT (Internet Search<URL: https: / / openai.com / blog / chatgpt> ), Gemini (Internet search <url: https: gemini.google.com ?hl="ja">) and other generation AIs. The data generation model 58 is obtained by performing deep learning on a neural network. A prompt including an instruction is input to the data generation model 58, and inference data such as voice data indicating voice, text data indicating text, and image data indicating an image is also input. The data generation model 58 performs inference on the input inference data in accordance with the instruction indicated by the prompt, and outputs the inference result in a data format such as voice data and text data. Here, inference refers to, for example, analysis, classification, prediction, and / or summarization.

[2669] In the above embodiment, an example in which the specific processing is performed by the data processing device 12 has been given, but the technology of the present disclosure is not limited to this, and the specific processing may be performed by the robot 414.

[2670] The emotion identification model 59 as an emotion engine may determine the user's emotion according to a specific mapping. Specifically, the emotion identification model 59 may determine the user's emotion according to an emotion map (see FIG. 9), which is a specific mapping. Similarly, the emotion identification model 59 may determine the robot's emotion, and the identification processing unit 290 may perform identification processing using the robot's emotion.

[2671] FIG. 9 is a diagram illustrating an emotion map 400 on which multiple emotions are mapped. In the emotion map 400, emotions are arranged in concentric circles radiating from the center. Emotions closer to the center of the concentric circles are more primitive. Emotions representing states and actions arising from a state of mind are arranged on the outer edges of the concentric circles. The concept of emotion includes both affect and mental states. Emotions generally generated from reactions occurring in the brain are arranged on the left side of the concentric circles. Emotions generally induced by situational judgment are arranged on the right side of the concentric circles. Emotions generally generated from reactions occurring in the brain and induced by situational judgment are arranged on the upper and lower sides of the concentric circles. Furthermore, the emotion of "pleasure" is arranged on the upper side of the concentric circles, and the emotion of "discomfort" is arranged on the lower side. In this way, in the emotion map 400, multiple emotions are mapped based on the structure by which emotions are generated, and emotions that tend to occur simultaneously are mapped close to each other.

[2672] These emotions are distributed in the 3 o'clock direction on emotion map 400, and typically fluctuate between relief and anxiety. In the right half of emotion map 400, situational awareness dominates over internal sensations, resulting in a sense of calm.

[2673] The inside of emotion map 400 represents what is going on in the mind, and the outside of emotion map 400 represents behavior, so the further you go outside emotion map 400, the more visible the emotions become (the more they are expressed in behavior).

[2674] Human emotions are based on various balances, such as posture and blood sugar levels. When these balances deviate from the ideal, a state of discomfort is indicated, and when they approach the ideal, a state of pleasure is indicated. Emotions can also be created for robots, automobiles, and motorcycles, based on various balances, such as posture and remaining battery life. When these balances deviate from the ideal, a state of discomfort is indicated, and when they approach the ideal, a state of pleasure is indicated. An emotion map can be generated, for example, based on Dr. Mitsuyoshi's emotion map (Research on Voice Emotion Recognition and Emotional Brain Physi...

Claims

1. a user interface means for accepting user input; means for identifying a preferred genre or theme based on the input; means for automatically generating video content based on the genre or theme; a means for storing the generated video content; means for playing or distributing the stored video content to a user; means for accepting user ratings; means for updating the genre rankings based on the evaluations; A system including:

2. means for uploading the automatically generated video content to a posting platform; means for sharing the uploaded video content with other users; The system of claim 1 further comprising:

3. means for creating and storing user profiles; means for making personalized content suggestions based on said profile; The system of claim 1 further comprising:

Citation Information

Patent Citations

  • Persona chatbot control method and system

    JP2022180282A