Content platform and profit sharing method based on artificial intelligence secondary content production
The AI-powered content creation platform addresses the inefficiencies and copyright issues in individual content creation by generating secondary content and implementing a fair revenue sharing method, enhancing the creation and consumption of content while protecting original creators' rights.
Patent Information
- Application Number
- PCT/KR2024/013429
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2023-12-07
- Filing Date
- 2024-09-05
- Publication Date
- 2025-06-12
AI Technical Summary
The existing content creation process is labor-intensive and requires professional tools, limiting the ability of individual creators to produce high-quality content efficiently. Additionally, the rise of individual content creators has led to increased cases of copyright infringement and unfair profit distribution.
A text-based artificial intelligence content creation service platform that utilizes AI programs to generate secondary content such as text, audio, images, and videos from original text data, while implementing a revenue sharing method among original authors, secondary creators, sharers, and the platform.
The platform facilitates efficient creation and consumption of high-quality secondary content, protects original authors' copyrights, and ensures fair profit distribution among creators and sharers, thereby promoting individual content creation and reducing copyright infringement.
Smart Images

Figure KR2024013429_12062025_PF_FP_ABST
Abstract
Description
A content platform and revenue sharing method based on AI-powered secondary content creation.
[0001] The present invention relates to a content platform and a profit sharing method based on artificial intelligence secondary content production, and more specifically, to a content platform service and a profit sharing method that produces secondary content (text, audiobook, webtoon, video, etc. modified from the original work) using an artificial intelligence content production service provided by a platform based on one's own or another's text creation, and distributes profits among the original author, secondary creator, sharer, and platform by advertising the same through selling or sharing it with other users.
[0002] While creative work has long been the domain of professionals, advancements in information and communication technology, improved communication performance, and shifts in content consumption culture have led to its gradual expansion into the realm of the individual. As a result, the era of the individual creator, where anyone can become a content producer, has arrived.
[0003] However, in order to create content, users must go through the entire process from ideation to sketching, image search, and editing. Currently, to create content, especially videos with motion, they must use professional editors such as Photoshop, Illustrator, and Premiere, or use these professional editors.
[0004] Even if you use a professional editor, it takes a lot of time to create the final content, and especially in the case of videos, the workload is very high because each image frame must be edited and then stitched together.
[0005] In addition, with the rapid increase in the number of individual content creators, cases of copyright infringement are increasing, such as unauthorized copying, reprinting, redistribution, and plagiarism of others' content, or the creation of secondary creations without the consent of the original creator to generate revenue.
[0006] [Prior Art Literature]
[0007] [Patent Document]
[0008] (Patent Document 1) Korean Patent Publication No. 10-2021-0107473 (September 1, 2021)
[0009] (Patent Document 2) Korean Patent Publication No. 10-2496362 (February 1, 2023)
[0010] (Patent Document 3) Korean Patent Publication No. 10-2459775 (October 24, 2022)
[0011] The technical task to be achieved by the present invention is to promote the creation and consumption of artificial intelligence secondary content such as text, audio, images, and videos by utilizing one's own or another's text, to protect the copyright of the original author, and to provide a means to generate fair profits for the original author, secondary content creators, and sharers.
[0012] The problems to be solved by the present invention are not limited to the problems mentioned above, and other problems not mentioned will be clearly understood by those skilled in the art from the description below.
[0013] A text-based artificial intelligence content creation service platform according to an embodiment of the present invention for achieving the above-described task includes a communication interface unit for receiving text data for creation of content using an artificial intelligence (AI) program from a terminal device of a user or another user, and a control unit for applying the received text data to an artificial intelligence (AI) program based on type information related to a type of content creation selected by the user terminal device, and for creating and outputting content corresponding to the type information.
[0014] The control unit can generate the content from the text data as new text (Text to Text), text to audio, text to image, or text to video based on the type information.
[0015] The above control unit can generate the new modified text content by executing the artificial intelligence language model as the artificial intelligence program, generate audio content by executing the TTS (Text to Speech) model, and generate the image and video content by executing the image and video generation model, respectively.
[0016] The above control unit can set a consumer sales price for the generated content by type of the generated content on the user terminal device and the original creator's terminal device.
[0017] The above control unit can process charging of a preset cost when consuming the generated content, including purchase of the generated content, from another user terminal device, when requesting to view artificial intelligence content from the user terminal device.
[0018] The above control unit may distribute a portion of the advertising revenue based on the advertising performance generated on the SNS channel only to the user terminal device (or service account) of the user who first shares the created content on the SNS (Social Network Service) after creating a review.
[0019] In addition, a method for sharing revenue according to an embodiment of the present invention for achieving the above-mentioned task may include a step of distinguishing between the original author and secondary creator of a text, determining a revenue distribution ratio among the original author, secondary creator, and platform based on the information, and distributing revenue according to the consumer's payment.
[0020] Regarding advertising revenue for the platform on which the above content is provided, a step may be included in which a portion of the advertising revenue is distributed to the first sharer of the content created within the platform based on sharing and advertising performance.
[0021] In addition, a method for operating a text-based artificial intelligence content service device according to an embodiment of the present invention for achieving the above-mentioned task includes a step in which a communication interface unit receives text data for generating content using an artificial intelligence (AI) program from a user terminal device, and a step in which a control unit applies the received text data to an artificial intelligence (AI) program based on type information related to a type of content generation selected by the user terminal device and generates and outputs content corresponding to the type information.
[0022] The step of generating and outputting the above content can generate the content from text data as new text (TTT), text as audio (TTA), text as image (TTI), or text as video (TTV) based on the above type information.
[0023] The step of generating and outputting the above content may include generating the content of the new text by executing the TTS model as the artificial intelligence program, generating the content of the image by executing the image generation model, and generating the content of the video by executing the video generation model.
[0024] The above driving method may further include a step of processing a charge for a preset cost when consuming the generated content, including purchasing the generated content, from another user terminal device, when the control unit requests viewing of artificial intelligence content from the user terminal device for the generated content.
[0025] The above driving method may further include a step of distributing a portion of the revenue generated from the SNS channel to a user terminal device (or service account) of a user who first shares the content created by the control unit on SNS after creating a review.
[0026] According to an embodiment of the present invention, texts such as texts or e-books provided by the user or other users can be recombined at the scenario level to reflect the intention of the secondary content creator, and thus, it can be utilized as a basic technology for automatically generating audiobooks, audio dramas, webtoons, and animations based on books, thereby promoting the production and consumption of secondary content.
[0027] In addition, an embodiment of the present invention can extract individual speaker identification, keywords that can produce sounds, and keywords that can identify spaces for synthesizing voices and sounds based on the content of a text, thereby producing high-quality audiobooks, video sound sources, etc.
[0028] Furthermore, the embodiment of the present invention can secure diversity of content by generating new AI content of text, audio, image or video based on text when a general user or a company such as a publisher requests content creation, and can also generate revenue for the original creator, secondary creator, sharer, etc. by enabling consumption or sharing of the generated content in connection with other users online.
[0029] FIG. 1 is a diagram illustrating a text-based artificial intelligence content service system according to an embodiment of the present invention.
[0030] FIG. 2 is a diagram for explaining a content platform service based on artificial intelligence secondary content production according to an embodiment of the present invention.
[0031] Figure 3 is a block diagram illustrating a detailed structure of the artificial intelligence content service device of Figure 1.
[0032] Figure 4 is a diagram showing the content creation process of the artificial intelligence content service device of Figure 1.
[0033] Figure 5 is a flowchart showing the operation process of the artificial intelligence content service platform of Figure 2.
[0034] The embodiments of the present invention described below are provided to more clearly explain the present invention to a person having ordinary skill in the art, and the scope of the present invention is not limited by the following embodiments, and the following embodiments may be modified in various other forms.
[0035] The terminology used herein is used to describe particular embodiments and is not intended to limit the present invention. The singular forms used herein may include the plural forms unless the context clearly dictates otherwise. In addition, the terms "comprise" and / or "comprising" used herein specify the presence of a stated feature, step, number, operation, element, element, and / or group thereof, but do not exclude the presence or addition of one or more other features, steps, numbers, operations, elements, elements, and / or groups thereof. In addition, the term "connection" used herein not only means that certain elements are directly connected, but also includes a concept that indirectly connects elements by interposing another element between them.
[0036] In addition, when it is said in this specification that a certain element is located "on" another element, this includes not only cases where a certain element is in contact with another element, but also cases where another element exists between the two elements. The term "and / or" as used in this specification includes any one of the listed items and any and all combinations of one or more of them. In addition, terms of degree such as "about", "substantially", etc. as used in this specification are used to mean a range of or close to the numerical value or degree, taking into account inherent manufacturing and material tolerances, and are used to prevent infringers from unfairly using the disclosure that mentions exact or absolute numbers provided to help the understanding of this specification.
[0037] Hereinafter, embodiments of the present invention will be described in detail with reference to the attached drawings. The sizes and thicknesses of areas or parts depicted in the attached drawings may be somewhat exaggerated for clarity and convenience of explanation. Like reference numbers designate like components throughout the detailed description.
[0038] FIG. 1 is a drawing showing a text-based artificial intelligence content service system according to an embodiment of the present invention, and FIG. 2 is a drawing briefly explaining a content creation process of the artificial intelligence content service device of FIG. 1.
[0039] As illustrated in FIG. 1, a text-based artificial intelligence content service system (hereinafter, artificial intelligence content service system) (90) according to an embodiment of the present invention includes part or all of a user terminal device (100), a communication network (110), an artificial intelligence content service device (120), and a third-party device (130).
[0040] Here, “including some or all” means that some components, such as a third-party device (130), are omitted so that the artificial intelligence content service system (90) is configured, or some or all of the components constituting the artificial intelligence content service device (120) can be integrated into a network device (e.g., a wireless switching device, etc.) constituting a communication network (110), etc. In order to help a sufficient understanding of the invention, it is explained as including all.
[0041] The user terminal device (100) may refer to a terminal device of users that performs an operation to generate AI content, i.e., content created by an AI program or model based on user input, i.e., text input by a user on a service screen or text such as an e-book, and to consume and share the generated content in connection with other users. More specifically, the user terminal device (100) may include various types of terminal devices that allow users to generate desired content based on text using the service of the AI content service device (120), consume content created by the AI content service device (120) or uploaded to a service, and further share content created in various forms with other users by connecting to external SNS, etc. Here, various forms refer to forms that convert text into (new) text, text into audio, text into image, or text into video, and content may refer to audiobooks, webtoons, animations / movies, games, etc. created in such forms.
[0042] The user terminal device (100) may include not only PC-based terminal devices such as desktop computers or laptop computers, but also mobile-based terminal devices such as smartphones, tablet PCs, and wearable devices worn by users on their wrists, etc. Furthermore, it may further include various devices such as smart TVs capable of wired and wireless Internet access. Of course, it may also include artificial intelligence (AI) speakers that operate in conjunction with smart TVs, etc.
[0043] For example, the artificial intelligence content service device (120) according to an embodiment of the present invention may receive and store text data related to a novel, and then generate it in the form of an audio file and store it in the DB (120a) of FIG. 1. Here, the audio file form may refer to a dictionary operation for automatically converting into a scenario when the 'Convert Scenario' button displayed on the service screen provided by the artificial intelligence content service device (120) in the user terminal device (100) of FIG. 1 is selected, and automatically generating audio data when the 'Listen to Audiobook' button is selected. In other words, this may be a data construction operation for providing a service according to an embodiment of the present invention. Of course, the generation of an audio file may correspond to a part of the content generated by the artificial intelligence content service device (120).
[0044] Looking further into the content consumption of users using user terminal devices (100), content consumption (or purchase or viewing, etc.) can be divided into text content and artificial intelligence (AI)-generated content. Accordingly, text content may be provided free of charge or may incur a fee of 100 won when using services such as viewing user-uploaded content text. Furthermore, for paid text viewing, such as e-books, a fee of 100 to 300 won per 5,000 characters may be charged, for example, and this fee may be determined by the author, allowing users to consume content for a fee. Of course, in the case of AI-generated content, viewing AI content created from user-uploaded text or AI content created from paid text, such as e-books, may be possible, and a certain fee may be paid for this. The creation, consumption, and sharing of content will be discussed in more detail later when describing the AI content service device (120) of FIG. 1.
[0045] The communication network (110) can be configured in various forms. The communication network (110) can include both wired and wireless communication networks. For example, a wired or wireless Internet network can be used or linked as the communication network (110). Here, the wired network includes an Internet network such as a cable network or a public switched telephone network (PSTN), and the wireless communication network includes a CDMA, WCDMA, GSM, EPC (Evolved Packet Core), LTE (Long Term Evolution), Wibro network, etc. Of course, the communication network (110) according to the embodiment of the present invention is not limited thereto, and can be used as an access network of a next-generation mobile communication system to be implemented in the future, for example, a cloud computing network in a cloud computing environment, a 5G network, a 6G network, etc. For example, if the communication network (110) is a wired communication network, an access point within the communication network can connect to a telephone exchange, etc., but if it is a wireless communication network, data can be processed by connecting to an SGSN or GGSN (Gateway GPRS Support Node) operated by a communication company, or data can be processed by connecting to various relays such as a BTS (Base Transceiver Station), NodeB, or e-NodeB.
[0046] The communication network (110) may include an access point. The access point may include a small base station, such as a femto or pico base station, which is often installed inside a building. Here, the femto or pico base station may be classified according to the maximum number of user terminal devices (100) or third-party devices (130) of FIG. 1 that can be connected to the small base station. Of course, the communication network (110) may include a short-range communication module for performing short-range communication, such as Zigbee or Wi-Fi, with the user terminal devices (100) or third-party devices (130). The access point may use TCP / IP or RTSP (Real-Time Streaming Protocol) for wireless communication. Here, short-range communication can be performed using various standards such as Bluetooth, Zigbee, infrared (IrDA), radio frequency (RF) such as ultra-high frequency (UHF) and very high frequency (VHF), and ultra-wideband (UWB) in addition to Wi-Fi. Accordingly, the access point can extract the location of the data packet, designate the best communication path for the extracted location, and transmit the data packet along the designated communication path to the next device, such as the artificial intelligence content service device (120). The access point can share multiple lines in a general network environment, and may include, for example, a router, a repeater, and a repeater.
[0047] The artificial intelligence content service device (120) may include, for example, a cloud server, and may be configured to include a DB (120a) linked to the server. The artificial intelligence content service device (120) may generate content by applying an artificial intelligence program (or model) based on text input by a user or text of an e-book in order to provide a service according to an embodiment of the present invention in conjunction with the user terminal device (100) and the third-party device (130) of FIG. 1, and may also perform an operation to consume the generated content, etc., and share it with other users of the user terminal device (100). In this process, the artificial intelligence content service device (120) may receive text data using, for example, a GUI service screen or a messenger-based artificial intelligence chatbot such as ChatGPT, and thereby generate and provide text, audio, images, or videos desired by users of the user terminal device (100).
[0048] To be more specific, the AI content service device (120) can generate content using the AI engine of the AI program when text is input from the user terminal device (100). In the process, the AI content service device (120) can perform operations such as charging. For example, in relation to AI content generation, at least one operation among Text to Text, Text to Audio, Text to Image, and Text to Video can be performed. Of course, the AI content service device (120) can also utilize LLM (Large Language Model) or the like to generate AI content.
[0049] First, during the AI content generation process of Text to Text, changes such as changes to the main character's name, speech pattern, ending, and narrator's perspective can occur. For simple changes, such as the main character's name, a fee of 100 won per 5,000 characters may be charged. Payment can be made by recharging the original amount and then deducting the amount from the recharged amount, or various payment methods, such as card payment or mobile phone payment, may be available at the time of service use. Furthermore, for more complex changes, such as speech pattern, ending, and narrator's perspective, a fee of 1,000 won per 5,000 characters may be charged. Of course, the cost setting can vary depending on the service provider's intent, and thus the embodiments of the present invention will not be specifically limited to any one method. However, service costs may differ between simple and complex changes.
[0050] The AI content service device (120) can generate audio dramas with various castings for AI content creation from Text to Audio. It differs from existing Text to Speech (TTS) in that it includes AI automatic character casting and emotional analysis, rather than simple reading. For example, generating only a character narration voice may incur a charge of 3,000 won per minute. Furthermore, generating audio for a character + narration + background music + sound effects may incur a charge of 5,000 won per minute.
[0051] In addition, the artificial intelligence content service device (120) generates an image as a scene of a 'webtoon' or 'illustration' by considering the characters and background of the text when text is input for AI content generation of Text to Image. For example, the input text can be analyzed to summarize the content of the text and an image of a scene can be generated based on the summary. For example, if the content of the text is analyzed and summarized into 4~5 main contents, 4~5 images can be generated based on this. In addition, in case of generating only characters, 10 images can be generated per sentence, and a charge of 3,000 won can be incurred. In case of generating characters and backgrounds, 10 images can be generated per sentence, and a charge of 5,000 won can be incurred.
[0052] In the case of Text to Video, the AI content service device (120) generates an animated video by considering the text's characters, story, background, etc. when the text is input. For moving videos with a fixed background, character expressions, movements, etc., a charge of 5,000 won per minute may be incurred, and for videos with both a moving background and character, a charge of 30,000 won per minute may be incurred.
[0053] Meanwhile, the AI content service device (120) according to an embodiment of the present invention can perform operations for content consumption. For example, content consumption can be divided into text content consumption and AI-generated content consumption. In the case of text content, viewing user-uploaded text may be provided free of charge or a small fee of 100 won may be charged. In addition, in the case of viewing paid text such as e-books, a fee of 100 to 300 won per 5,000 characters may be charged. Since this fee can be determined by the author, the AI content service device (120) can set the cost (information) for charging in conjunction with the author's user terminal device (100), a user terminal device (100) such as a computer operated by a publisher, or a third-party device (130). In addition, in the case of AI-generated content, viewing AI content created from user-uploaded text may be provided free of charge or a small fee of 200 won may be charged. Additionally, for viewing AI content created with paid text, such as e-books, a fee of 200 to 500 won may be charged.
[0054] Revenue distribution related to content consumption through the AI content service device (120) can be achieved in various ways. In other words, revenue distribution can be achieved in various forms depending on whether the content provider is a general user or a publisher. Of course, revenue distribution can also be achieved for service providers who provide services through the AI content service device (120) during this process. To this end, the AI content service device (120) can enter into a contract for revenue distribution, and in the process, can communicate with the user terminal device (100) or a third-party device (130). In the case of text / contexts uploaded by users, revenue can be distributed at a ratio of 5:5 between the uploading user and the service provider (e.g., Oddbooks). In the case of publisher text such as e-books, revenue can be distributed at a ratio of 9:1 between the publisher and the service provider. In the case of secondary AI content such as e-books, revenue can be distributed at a ratio of 6:2:2 between the publisher, the producing user, and the service provider.
[0055] In addition, the AI content service device (120) according to an embodiment of the present invention can perform operations for content sharing. In other words, the AI content service device (120) can share content by linking with a third-party device (130) that provides a social networking service (SNS) such as Instagram or Facebook (e.g., API linkage, etc.), and can also generate advertising revenue in the process, which can be distributed to sharers. For example, if an advertiser generates revenue of 10, the service provider's margin can be 5, or 50%, and the remaining 5, or 50%, of the revenue can be distributed to SNS sharers. Advertising revenue can be generated by attracting advertisements for web novels, webtoons, animations, and goods, and fees can be collected from advertisers. In addition, the main banner, side banner, recommended works, and related works of Sulpia can be the targets of revenue generation. In the case of content sharing, the AI content service device (120) can also operate in a structure similar to YouTube monetization. For example, you can generate revenue of 1 to 2 won per view, and of course, this is not a fixed unit price and can fluctuate depending on advertising revenue. When sharing free content or paid content review + link on external SNS, revenue can be recognized to the original sharer (or original author) of the link post. For example, in the case of revenue distribution of 1 won per repost or like, whichever is greater, on X (formerly Twitter), if User A reposts "This novel is really fun!" (link) and gets 1,000 likes, the revenue can be calculated as 200 → 1,000 won, and if User B reposts "User A. This novel is really fun~ (link)", the original sharer X can be calculated as 0 won. Of course, revenue distribution according to sharing can be set in various ways depending on the service policy, so the embodiment of the present invention will not be particularly limited to any one form.
[0056] The AI content service device (120) according to an embodiment of the present invention, as shown in FIG. 2, can perform operations for generating data, scenarios, and final content using AI text analysis technology for content creation (S200 to S260). As previously described, the AI content service device (120) can generate various types of content, such as audiobooks, webtoons, animation / movies, and games, by modifying data such as characters, titles, and personalities, as well as scenarios such as character casting and character-title matching, based on the service request, when a desired service is selected and text is input through a service screen provided to the user terminal device (100) through the service screen. Furthermore, the AI content service device (120) can generate and provide various types of content, such as audiobooks, webtoons, animation / movies, and games, based on the input text. In this process, a TTS model may be executed for audiobook creation. Furthermore, an image generation model may be executed for webtoon content, and a video generation model may be executed for animation / movie content. A 3D object generation model may be executed for games. It can be assumed that an AI model specialized (or trained) for each content creation is used. In other words, the artificial intelligence content service device (120) according to an embodiment of the present invention can include multiple artificial intelligence models for generating various types of AI content and execute each artificial intelligence model. Each artificial intelligence model is an artificial intelligence model, i.e., a program, specialized for generating different types of content, and each artificial intelligence model may have different software (SW) modules or algorithms for performing detailed operations, but above all, the type of learning data for generating each AI content may be different.
[0057] Let's take a look at the creation of a new text (TTT) using text. For example, the artificial intelligence content service device (120) according to an embodiment of the present invention can be equipped with an artificial intelligence model such as a Large Language Model (LLM) and perform an operation to create an audiobook by utilizing the same. The artificial intelligence content service device (120) analyzes text data received from a third-party device (130) or pre-stored text data to identify who the speaker of the dialogue within the text is from a text (e.g., a set of sentences), and can also extract sound keywords (e.g., verbs, onomatopoeia, etc.) and background keywords (e.g., temporal background, spatial background) within the text, and can generate an audiobook by combining the two. Of course, in this process, the artificial intelligence content service device (120) can analyze the text data by utilizing an artificial intelligence module such as LLM and synthesize voice data based on the analysis results to generate an audiobook. In this process, the artificial intelligence content service device (120) may operate so that a developer or administrator sets a command through a prompt and data analysis or voice data synthesis is performed accordingly.
[0058] Text data such as novels may include text for narration and dialogue-related lines. Of course, since text related to dialogue can be marked with double quotation marks, it is entirely possible to distinguish lines through this. To generate an audiobook, the AI content service device (120) first separates the text and dialogue, and then identifies who the dialogue belongs to, so that it can synthesize speech with a specific voice. To this end, the AI content service device (120) can identify lines using an AI model that identifies the relationship between words and sentences, or more precisely, the relationship between objects and sentences within words. Alternatively, it is entirely possible to determine characters or characters based on their relationships with surrounding dialogue.
[0059] More specifically, the artificial intelligence content service device (120) can extract character names from sentences, convert them into data, and manage them. In addition, for each sentence, the correlation between the "character name data" and the sentence can be inferred (or predicted), and if there is no valid character, the character name "Unknown" can be selected. In addition, the artificial intelligence content service device (120) can learn in advance the target sentence, surrounding sentences (e.g., n sentences before and after the target sentence), and the character name spoken in the target sentence as training data for the artificial intelligence model. Through this, the artificial intelligence content service device (120) can generate an artificial intelligence model that receives sentences and surrounding sentences as input and infers the character name, or execute the pre-generated model.
[0060] In addition, the artificial intelligence content service device (120) can classify keywords that can make sounds (e.g., verbs: knock, open, hit, drop, etc. / onomatopoeia: whirrrik, swoosh, boom, tak, thump, etc.) within the text or dialogue that constitutes text data through data analysis, and can specifically identify verbs and onomatopoeia related to sounds. To this end, the artificial intelligence content service device (120) can infer the relationship between related sound keywords and sentences. The artificial intelligence content service device (120) can perform an operation to extract background keywords, and such backgrounds can be used when synthesizing character voices or voices related to sound keywords. In order to extract background keywords from the text, the artificial intelligence content service device (120) can operate with a concept such as a subset of the intersection of named entity recognition and semantic role determination methods in a broad sense. In other words, the artificial intelligence content service device (120) according to an embodiment of the present invention can infer the relationship between a sentence and a background keyword with a single operation in this part. For this purpose, LLM can be utilized. Currently, many semantic role determination (SRL) and named entity recognition (NER) models capable of identifying words within sentences have been developed, and these can be utilized in embodiments of the present invention.
[0061] The AI content service device (120) according to an embodiment of the present invention can separate verbs, adverbs, and nouns from a sentence by utilizing a morphological analysis model for keyword identification. Many Korean morphological classification models capable of separating morphemes from sentences, as well as semantic role determination (SRL) and named entity recognition (NFR) models capable of identifying words within a sentence, have been developed and are widely available on the market. Therefore, it is entirely possible to utilize such analysis models in the embodiment of the present invention. The AI content service device (120) can use the model to determine and extract the sound relevance of separated verbs and adverbs, and can also determine the spatial background based on nouns. In the process, the AI content service device (120) can learn data in advance, such as target sentences, keywords of the target sentences, and keyword types, as training data, i.e., learning data. Through this, the AI content service device (120) can generate an AI model that receives a sentence as input and infers key keywords. Alternatively, the pre-generated AI model can be executed.
[0062] According to the above configuration and operation, the AI content service device (120) according to the embodiment of the present invention can produce high-quality audiobooks by extracting individual speaker identification for voice and sound synthesis based on the content of the text, keywords that can produce sounds, and keywords that can identify spaces. As a result, since the book can be recombined at the scenario level, audiobooks, audio dramas, webtoons, and animations based on the book can be automatically generated. Of course, the above description was only in detail related to the generation of audiobooks using input text data, but the generation of audio dramas and webtoons may not be significantly different from the generation of audiobooks in that they also utilize AI models.
[0063] The third-party device (130) may refer to a computer of an administrator who operates the service of the artificial intelligence content service device (120) of FIG. 1, but may also include a server of a program developer who develops and installs a program for providing services to the artificial intelligence content service device (120), or a computer used by a developer. For example, if the third-party device (130) is a server or computer of a developer or a developer, it may be connected to the artificial intelligence content service device (120) of FIG. 1 and performed. Typically, it may include an operation for generating a prompt. In the case of a prompt, it may mean that the developer provides a kind of guideline or inputs setting information so that an operation is performed accordingly. In this way, the third-party device (130) can essentially operate as a single device for generating audiobooks, etc. However, in the embodiment of the present invention, since the artificial intelligence content service device (120) is an expensive device, it is preferable to perform content creation, consumption, and sharing operations according to the embodiment of the present invention on the artificial intelligence content service device (120) under the assumption that it can be installed with an artificial intelligence program such as LLM.
[0064] In addition to the above, the user terminal device (100), communication network (110), artificial intelligence content service device (120), and third-party device (130) of FIG. 1 can perform various operations, and since related contents will be continuously covered later, detailed contents will be replaced with those contents.
[0065] Figure 3 is a block diagram illustrating a detailed structure of the artificial intelligence content service device of Figure 1.
[0066] As illustrated in FIG. 3, the artificial intelligence content service device (120) according to an embodiment of the present invention includes part or all of a communication interface unit (300), a control unit (310), an artificial intelligence content service unit (320), and a storage unit (330).
[0067] Here, “including some or all” means that some components, such as the storage unit (330), may be omitted to configure the artificial intelligence content service device (120) of FIG. 1, or some components, such as the artificial intelligence content service unit (320), may be integrated into other components, such as the control unit (310), and is explained as including all to help sufficient understanding of the invention.
[0068] The communication interface unit (300) can communicate with the user terminal device (100) and the third-party device (130) of FIG. 1, respectively. During the communication process, the communication interface unit (300) can perform various operations, such as modulation / demodulation, muxing / demuxing, encoding / decoding, encryption / decryption, and scaling to convert resolution, which are obvious to those skilled in the art and thus will not be further described.
[0069] The communication interface unit (300) can communicate with the third-party device (130) of FIG. 1 to perform a data construction operation so that various types of content, including audiobooks, can be generated based on a service according to an embodiment of the present invention, that is, text provided by a user terminal device (100) or a company representative such as a publisher through the user terminal device (100) or the third-party device (130). More precisely, an operation for learning (or training) an artificial intelligence program such as LLM can be performed. For example, in an embodiment of the present invention, an operation for training each artificial intelligence model can be performed to generate a new type of text, such as content that changes the name or speech of the main character in a text such as a novel, audio, image, and even video, based on text data, and the communication interface unit (300) can be involved in such an operation under the control of the control unit (310).
[0070] In addition, the communication interface unit (300) can perform operations related to content creation, consumption, and sharing, etc., upon completion of the construction operation of an artificial intelligence program for providing a service according to an embodiment of the present invention. For example, a user of the user terminal device (100) of FIG. 1 can request a service and input text data for creating desired content on a service screen. At this time, the user can select type information regarding the desired type of content to be created, such as whether the content is new text, audio, image content, or video content, on the service screen. Accordingly, the communication interface unit (300) can execute an artificial intelligence program (e.g., a generative AI program, etc.) on the text data provided by the user to create and provide the type of content desired by the user. Of course, the communication interface unit (300) can also be involved in these operations.
[0071] The control unit (310) is responsible for the overall control operations of the communication interface unit (300), the control unit (310), the artificial intelligence content service unit (320), and the storage unit (330) of FIG. 3. For example, the control unit (310) may perform operations for online consumption and sharing of the generated content when generating data, scenarios, and final content using AI text analysis technology, thereby distributing profits accordingly. Of course, for this purpose, the control unit (310) executes a program installed in the artificial intelligence content service unit (320), and the program may be included and operated within a program such as LLM, and may also be operated in conjunction with an artificial intelligence program such as LLM. For example, let's assume that the control unit (310) requests the generation of an image such as a webtoon while providing text data necessary for generating AI content from the user terminal device (100). In this case, the control unit (310) can temporarily store text data and type information related to content creation in the storage unit (330) and then retrieve them to provide them to the artificial intelligence content service unit (320). Accordingly, the artificial intelligence content service unit (320) can use an image creation model related to image creation based on the type information to create an image to be used in webtoons, etc. and provide the image to the control unit (310), and the control unit (310) can control the communication of the communication interface unit (300) to provide the image to the user terminal device (100).
[0072] In addition, if a user of the user terminal device (100) wants to consume, for example, sell, image content such as webtoons that he or she created so that other users can view them, the control unit (310) can request consumption through the service screen so that subscriptions or sales can be made. Accordingly, the control unit (310) can upload content created by users of the user terminal device (100) through the main screen, sub-screen, or menu screen of the online platform provided by the artificial intelligence content service device (120) of FIG. 1. For example, if an e-book is created using an artificial intelligence program according to an embodiment of the present invention, more precisely, using the TTS model as shown in FIG. 2, if the author of the e-book requests paid text viewing and sets the cost related to the viewing in the range of, for example, 100 to 300 won, a charging operation can be performed on the user terminal device (100) of the user who has viewed the content accordingly. Of course, during this process, the control unit (310) can perform data processing in conjunction with the artificial intelligence content service device (320). Here, data processing may include an operation to perform charging based on the identification information of a user terminal device (100) when a certain content is consumed on a user terminal device (100) and to store and manage the data by systematically classifying it by user in the DB (120a) of FIG. 1.
[0073] The artificial intelligence content service unit (320) can generate artificial intelligence content based on text input by users of the user terminal device (100), of course, the text can be text written directly by the users, and also generate revenue by consuming (e.g., viewing, selling) and sharing it on SNS channels, etc., as well as text in the form of an e-book. Of course, for this purpose, the artificial intelligence content service unit (320) can perform operations to generate content such as audiobooks, webtoons, animations / movies, etc. by connecting data related to characters, titles, personalities, etc. and neural networks of scenarios such as character casting, character-title matching, and character-sentence matching, as illustrated in FIG. 2, and executing the TTS model, image generation model, and video generation model installed therein.
[0074] In other words, the artificial intelligence content service unit (320) can generate various types of content using pre-entered text data at the request of general users or relevant parties such as publishers. As shown in FIG. 2, corresponding TTS models, image generation models, and video generation models can be executed for content generation such as audiobooks, webtoon images, and movie videos. However, since the neural networks of data such as characters and titles and scenarios such as character casting are interconnected through various paths, it can be seen that countless types of content can be generated through this. Of course, even if the user of the user terminal device (100) provides the same text content and requests the generation of audiobook content, the content that can be generated can vary depending on the character, title, and in the case of a scenario, character casting, as shown in FIG. 2. As shown in FIG. 2, the character casting scenario is interconnected with the character, title, personality, gender, emotion, background, etc. Accordingly, when creating content, it is possible to specify and provide data on characters, etc., and information related to scenarios such as character casting, etc. from the user terminal device (100), but it is also possible to automatically set each artificial intelligence model to create content based on text data provided by the user.
[0075] In addition, the AI content service unit (320) can perform operations to enable other users to purchase and view the generated content, and can also operate to share content and links by connecting to external SNS, etc. For example, if a user of a user terminal device (100) requests the creation of desired content and then requests consumption, or even sale, of the content, the AI content service unit (320) can provide the content through a sales menu, such as the main screen or sub-screen. In addition, for example, in the case of webtoons, the AI content service unit (320) can allow others to purchase and view the content, and charge for this process. For example, if a monthly fee has been paid for a specific webtoon, the content can be viewed only during the corresponding period of time by visiting the site. Of course, even when downloading webtoon content from the user terminal device (100), the content can be automatically deleted from the memory of the user terminal device (100) after a specified period of time based on payment. For example, an SW module, such as a timer, can be provided in the form of firmware, etc., to enable automatic deletion. The artificial intelligence content service department (320) can participate in actions related to consumption of content generated by users, etc.
[0076] Furthermore, the AI content service unit (320) may perform an operation to distribute a portion of the revenue generated from the above-mentioned parasitic content to the creator of the content or to the user who posted a comment and shared it on the SNS channel when the generated content is shared on a social networking service (SNS) channel or the like, when revenue is generated on the relevant channel. For example, the AI content service unit (320) may perform the operation when User A creates content and requests sale or consumption. Accordingly, User B may purchase and view User A's content, and if, in the process, he / she posts a comment and shares the content on a social networking service (SNS) channel such as Facebook, a portion of the revenue generated on the relevant channel may be distributed to User B. Of course, it may be desirable not to distribute revenue to users who have liked the content on the social networking service (SNS). Revenue distribution may be implemented in various ways depending on the intention of the service provider. Above all, the AI content service unit (320) according to an embodiment of the present invention may engage in an operation to generate content such as text, images, and videos by applying an AI program based on text, sell the content to others, and share the content on a social networking service (SNS) channel or the like, thereby generating revenue.
[0077] The storage unit (330) can temporarily store various types of data or information processed under the control of the control unit (310). Here, since the terms information and data are used interchangeably in practice, the concept of such terms will not be particularly limited. For example, the content of text provided by the user terminal device (100) can be called text data. Of course, content can also be used as a term that includes the meaning of the text. On the other hand, content related to whether the text data is to be generated as an image or a video, etc., can be referred to as information. For example, when text data and type information for content generation to generate the text data as image content are provided by the user terminal device (100), the storage unit (330) can temporarily store the text data and then retrieve it and provide it to the artificial intelligence content service unit (320) for use in content generation.
[0078] In addition to the above, the communication interface unit (300), control unit (310), artificial intelligence content service unit (320), and storage unit (330) of FIG. 3 can perform various operations, and other detailed information has been sufficiently explained above, so it will be replaced with those contents.
[0079] According to an embodiment of the present invention, the communication interface unit (300), the control unit (310), the artificial intelligence content service unit (320), and the storage unit (330) of FIG. 3 are configured as physically separate hardware modules, but each module may store and execute software for performing the above operations. However, the software is a collection of software modules, and each module may be formed of hardware, so there is no particular limitation on the configuration, such as software or hardware. For example, the storage unit (330) may be hardware, such as storage or memory. However, since it is also possible to store information (repository) in software, there is no particular limitation on the above.
[0080] Meanwhile, as another embodiment of the present invention, the control unit (310) may include a CPU and a memory, and may be formed as a single chip. The CPU may include a control circuit, an operation unit (ALU), a command interpretation unit, and a registry, and the memory may include a RAM. The control circuit may perform a control operation, the operation unit may perform an operation of binary bit information, and the command interpretation unit may perform an operation of converting a high-level language into machine language and vice versa, including an interpreter or a compiler, and the registry may be involved in software data storage. According to the above configuration, for example, at the initial stage of the operation of the artificial intelligence content service device (120) of FIG. 1, a program stored in the artificial intelligence content service unit (320) may be copied and loaded into memory, i.e., RAM, and then executed, thereby rapidly increasing the data operation processing speed. In the case of a deep learning model, it may be loaded into the GPU memory instead of the RAM and executed by accelerating the execution speed using the GPU.
[0081] Figure 4 is a diagram showing the content creation process of the artificial intelligence content service device of Figure 1.
[0082] For convenience of explanation, referring to FIG. 4 together with FIG. 1, according to an embodiment of the present invention, for example, a user terminal device (100) of FIG. 1 may be a device used by a general user or a person concerned such as a publisher, and may receive text and information on the type of content to be created, i.e., content type information (S400). For example, when the user terminal device (100) requests a service from the artificial intelligence content service device (120) of FIG. 1 after subscribing to a service, a first area for entering text data for creating content or retrieving data such as an e-book may be displayed on the service screen, and a second area located around the first area may be displayed, and a service screen may be provided in the second area so that the type of content to be created can be selected. Since the UI screen may be created in various ways, it will not be particularly limited to any one form.
[0083] The AI content service device (120) of FIG. 1 can analyze text data when it receives it and perform a metadata generation operation. Here, the metadata generation operation can refer to various information extracted from the text data analysis results, such as the type of novel, characters, and background, and information based on such information. In other words, various types of information specified in the embodiments of the present invention can be extracted from the text data.
[0084] In addition, the AI content service device (120) can select and execute an AI model corresponding to the type information regarding content creation selected by the user (S420). In an embodiment of the present invention, an audiobook creation model for creating audiobook content using text data, a webtoon creation model for creating images such as webtoons using text data, and a video creation model for producing video content such as dramas using text data can be selected and executed (S430, S440, S450).
[0085] For example, when an audiobook creation model is selected (S430), the AI content service device (120) can create audiobook content (S434) through detailed operations such as audiobook script creation (S431), AI voice automatic casting and sound source creation (S432), background music and sound effect creation (S433), and voice and background music synchronization (S434), as shown in FIG. 4. Audiobook creation has already been sufficiently explained above, so I will replace that with the explanations therein. However, to briefly reiterate, when a reader purchases and selects (specifies the start / end) all / part of a romance web novel text (prior author consultation, platform profit sharing agreement completed) for which he or she wants to modify the content, and inputs the selected text, the AI content service device (120) generates data such as web novel character traits, narration, background music, and sound effects from the text. Based on the generated data, an appropriate AI VOICE is cast and assigned to a TTS model to generate a voice sound source. Then, among the generated data, background music and sound effects are searched for in the database or newly synthesized. The audio sources from the voice source generation and background music synthesis stages are combined with synchronization to create the final audiobook content.
[0086] In addition, when a webtoon creation model is selected (S440), the artificial intelligence content service device (120) can create and output webtoon content through detailed operations such as webtoon storyboard creation (S441), image creation by cut (S442), speech bags and other text creation by cut (S443), and image text layer integration (S444). For example, in the case of a webtoon, 8 to 9 images can be created based on the content of the text data. The reader purchases and selects (specifies the start / end) all / part of the romance web novel text (e.g., prior author consultation, platform profit sharing contract completion) for which he or she wants to modify the content. Then, the selected text is input. Then, the artificial intelligence content service device (120) creates data such as web novel character characteristics, narration, background music, and sound effects from the text. It performs operations for collecting and creating additional character appearance description data for webtoon creation (e.g., hair color, eye color, clothing, accessories, posture, expression, etc.). Then, based on the generated data, prompts for each image generation model are generated. Of course, text description prompt styles that generate images well may vary depending on the model, such as midjourney and stable diffusion. Images are generated for each cut by assigning them to the corresponding image generation AI model. Using the dialogue for each cut, speech bubble dialogue and background dialogue images are generated as GIFs (Graphics Interchange Format). An integrated image is created with each layer of the images generated in the cut-by-cut image generation and GIF generation steps. Then, by seamlessly connecting the cuts, it is generated as a single, long webtoon image that can be scrolled.
[0087] Furthermore, when a video creation model is selected (S450), the artificial intelligence content service device (120) generates and outputs video content (S455) through detailed operations of creating a scenario for each video scene (S451), creating multiple 30-second videos for each scene (S452), creating multiple sound sources for each scene (S453), and synchronizing 30-second videos and sound sources (S454). When a reader purchases and selects (specifies start / end) all / part of a romance web novel text (e.g., prior author consultation, platform profit sharing agreement completed) for which he or she wants to modify the content, the artificial intelligence content service device (120) receives the selected text and generates data such as web novel character traits, narration, background music, and sound effects from the text. It can perform operations of collecting and generating additional character appearance description data for webtoon creation (e.g., hair color, eye color, clothing, accessories, posture, expression, etc.). And it generates text - scene description (e.g., character, background, sound effect description) that will make up a 30-second video. Generate dialogue, background music, and sound effects suitable for the video. Then, synchronize and integrate the video and audio generated from the text-scene description generation and dialogue, sound effects, etc.
[0088] In addition to the above, the user terminal device (100), communication network (110), artificial intelligence content service device (120), and third-party device (130) of FIG. 1 can perform various operations, and other detailed information has been sufficiently explained above, so it will be replaced with those contents.
[0089] Figure 5 is a flowchart showing the operation process of the artificial intelligence content service device of Figure 1.
[0090] For convenience of explanation, referring to FIG. 5 together with FIG. 1, an artificial intelligence content service device (120) according to an embodiment of the present invention receives text data for creating content using an artificial intelligence program installed in the artificial intelligence content service device (120) from an external device such as a user terminal device (100) used by a general user or a publisher or other relevant party (S500). In the process of receiving text data, type information related to the type of content desired by the user, i.e., the type, can be received together. Of course, it is clear that the user can be distinguished by receiving device identification information such as the device ID of the user terminal device (100).
[0091] In addition, the artificial intelligence content service device (120) can generate and output content corresponding to the type information by using an artificial intelligence program, such as a TTS model, an image generation model for generating images such as webtoons, or a video generation model for generating videos such as dramas or movies, based on the type information related to content generation selected from an external device such as a user terminal device (100) (S510). Of course, before this process, the artificial intelligence content service device (120) can perform an operation of analyzing the text data to generate metadata. By analyzing the text data through the LLM artificial intelligence model, the contents of text data such as novels can be analyzed and various metadata related to the background, characters, etc. contained therein can be extracted. For example, if the background is known, a corresponding sound effect can be inserted when generating content, or it can be inserted as a background image or background video when generating an image or video.
[0092] Above all, the artificial intelligence content service device (120) according to an embodiment of the present invention does not simply generate content requested by users, but when a user of a user terminal device (100) requests the sale of content such as an e-book or webtoon that he or she has created so that other users can view the content he or she has created, the sale can be made on the main screen or sub-screen, etc., so that the content can be sold to other users according to the request. In addition, when content created by users is shared through an SNS channel, etc., the artificial intelligence content service device (120) can, according to a preset method, distribute a designated percentage of revenue to the first user who creates a comment on the created content and then shares it on the SNS channel, for example, when revenue is generated on the channel.
[0093] More specifically, the artificial intelligence content service device (120) determines revenue distribution based on the identity of the original text author and the content creator (S520), allows the user to select whether to advertise the generated content, receives the advertising performance from the first sharer of the content (S530), and distributes revenue periodically based on the revenue distribution ratio between each user and the platform (S540).
[0094] In addition to the above, the artificial intelligence content service device (120) of FIG. 1 can perform various operations, and other detailed information has been sufficiently explained above, so we will replace it with that information.
[0095] Even though all components constituting the embodiments of the present invention have been described as being combined or operating in combination, the present invention is not necessarily limited to such embodiments. That is, within the scope of the present invention, all of the components may be selectively combined and operated one or more times. In addition, although all of the components may be implemented as individual hardware, some or all of the components may be selectively combined and implemented as a computer program having program modules that perform some or all of the functions of the combined hardware in one or more pieces. The codes and code segments constituting the computer program will be readily inferred by those skilled in the art. Such a computer program may be stored in a non-transitory computer-readable storage medium and read and executed by a computer, thereby implementing the embodiments of the present invention.
[0096] Here, the non-transitory readable storage medium refers to a medium that permanently stores data and can be read by a device, rather than a medium that stores data for a short period of time, such as a register, cache, or memory. Specifically, the above-described programs may be stored and provided on a non-transitory readable storage medium, such as a CD, DVD, hard disk, Blu-ray disc, USB, memory card, or ROM.
[0097] Although embodiments of the present invention have been described with reference to the attached drawings, those skilled in the art will appreciate that the present invention can be implemented in other specific forms without altering the technical spirit or essential characteristics of the present invention. Therefore, the embodiments described above should be understood to be illustrative in all respects and not restrictive.
[0098] [Explanation of symbols]
[0099] 100: User terminal device 110: Communication network
[0100] 120: Artificial Intelligence Content Service Device 130: Third-Party Device
[0101] 300: Communication interface section 310: Control section
[0102] 320: Artificial Intelligence Content Service Department 330: Storage Department
Claims
1. A communication interface unit that receives text creations owned by the user or others related to the original work from a user terminal device to create secondary content that is different from the original work using an artificial intelligence (AI) program; and A control unit that applies the data of the received text creation to an artificial intelligence (AI) program based on type information related to the type of secondary content creation selected from the user terminal device to generate and output the secondary content corresponding to the type information; including, The above control unit sets the consumer sales price for each type of secondary content created in the user terminal device and the original author's terminal device related to the original work for the secondary content created, The above control unit processes the preset cost for the generated secondary content when consuming the generated secondary content, including purchasing the generated secondary content from another user terminal device, when requesting to view AI content on the user terminal device for the generated secondary content. The above control unit is an artificial intelligence content service device for a content platform based on artificial intelligence secondary content production, which distributes a portion of the advertising revenue generated from the SNS (Social Network Service) channel to the user terminal device of the user who first shares the created secondary content on the SNS (Social Network Service) after creating a review, according to a predetermined revenue distribution ratio.
2. In paragraph 1, The above control unit is an artificial intelligence content service device for a content platform based on artificial intelligence secondary content production, which generates the content from the text data as new text (Text to Text), text to audio (Text to Audio), text to image (Text to Image), or text to video (Text to Video) based on the type information.
3. In paragraph 2, The above control unit is an artificial intelligence content service device for a content platform based on artificial intelligence secondary content production, which executes an artificial intelligence language model as the artificial intelligence program to generate the new transformed text content, executes a TTS (Text to Speech) model to generate audio content, and executes an image and video generation model to generate the image and video content, respectively.
4. A step in which the communication interface unit receives a text creation work owned by the user or another person related to the original work from the user terminal device to create a modified secondary content different from the original work using an artificial intelligence (AI) program; and A step in which the control unit applies the data of the received text creation to an artificial intelligence (AI) program based on type information related to the type of secondary content creation selected from the user terminal device, thereby generating and outputting the secondary content corresponding to the type information; including, A step for the control unit to set a consumer sales price for each type of secondary content created in the user terminal device and the original author's terminal device related to the original work for the secondary content created; The step of the above control unit processing a preset cost for the generated secondary content when consuming the generated secondary content, including purchasing the generated secondary content from another user terminal device, when requesting to view artificial intelligence content on the user terminal device for the generated secondary content; and A step of distributing a portion of the advertising revenue generated from the SNS channel to the user terminal device of the user who first shares the secondary content generated by the above control unit on SNS after creating a review, according to a predetermined revenue distribution ratio; A method for operating an artificial intelligence content service device for a content platform based on artificial intelligence secondary content production, which further includes:
5. In paragraph 4, The steps for creating and outputting the above content are: A method for operating an artificial intelligence content service device for a content platform based on artificial intelligence secondary content production, which generates the content from the text data as new text (TTT), text as audio (TTA), text as image (TTI), or text as video (TTV) based on the above type information.
6. In paragraph 5, The steps for creating and outputting the above content are: A method for operating an artificial intelligence content service device for a content platform based on artificial intelligence secondary content production, which generates new transformed text content by executing an artificial intelligence language model as the artificial intelligence program, generates audio content by executing a TTS (Text to Speech) model, and generates image and video content by executing an image and video generation model, respectively.
Citation Information
Patent Citations
URL-based advertisement platform system and providing method thereof
KR1020160097987A
System and method for providing social service based on video contens
KR1020180067977A
Method, device and system for automatically processing creation of web book based on web novel using artificial intelligence model
KR102586799B1
Content Platform and Profit Sharing Method Based on Artificial Intelligence Secondary Content Production
KR102693274B1
Information processing system, information processing method, and program
WO2020153193A1