Method, device, equipment, storage medium and program product for content generation
By presenting input components associated with the area content in the canvas and triggering media content generation, the problem of the inability to simultaneously present input materials and generated results in existing tools is solved, improving operational efficiency and user experience.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- BEIJING ZITIAO NETWORK TECH CO LTD
- Filing Date
- 2026-03-20
- Publication Date
- 2026-06-19
Smart Images

Figure CN122239992A_ABST
Abstract
Description
Technical Field
[0001] The examples in this article generally relate to the field of computers, and in particular to methods, apparatuses, electronic devices, computer-readable storage media, and computer program products for content generation. Background Technology
[0002] With the rapid development of computer technology, a variety of content generation tools have emerged on the market, such as applications for generating media content like images and videos. These tools can quickly generate referable or editable media content based on user input, significantly improving the efficiency of creating media content or works containing media content. Summary of the Invention
[0003] In a first aspect, a method for content generation is provided. The method includes: receiving a first operation indicating selection of a first region of a canvas; presenting an input component at an associated location in the first region, wherein the state of the input component is related to content in the first region; and, in response to receiving a second operation via the input component, presenting at least one piece of media content in a second region of the canvas.
[0004] In a second aspect, an apparatus for content generation is provided. The apparatus includes: a region selection module, an input component rendering module, and a media content rendering module, wherein the region selection module is configured to receive a first operation indicating the selection of a first region of a canvas; the input component rendering module is configured to render an input component at an associated location in the first region, wherein the state of the input component is related to content in the first region; and the media content rendering module is configured to render at least one piece of media content in a second region of the canvas in response to receiving a second operation via the input component.
[0005] In a third aspect, an electronic device is provided. The device includes at least one processor; and at least one memory coupled to the at least one processor and storing instructions for execution by the at least one processor. When executed by the at least one processor, the instructions cause the device to perform the method of the first aspect.
[0006] In a fourth aspect, a computer-readable storage medium is provided. The computer-readable storage medium stores computer-executable instructions that can be executed by a processor to implement the method of the first aspect.
[0007] In a fifth aspect, a computer program product is provided, which is tangibly stored in a computer storage medium and includes computer-executable instructions that, when executed by a device, cause the device to perform the method of the first aspect.
[0008] This method can improve the efficiency of content generation.
[0009] It should be understood that the content described in this section is not intended to limit the key or important features of the examples in this article, nor is it intended to restrict the scope of the solution. Other features will become readily apparent from the following description. Attached Figure Description
[0010] The above and other features, advantages, and aspects of the various examples herein will become more apparent when taken in conjunction with the accompanying drawings and the following detailed description. In the accompanying drawings, the same or similar reference numerals denote the same or similar elements, wherein: Figure 1 A schematic diagram of the example environment is shown; Figures 2A to 2E Example interfaces for some scenarios are shown; Figures 3A to 3D Example interfaces for some scenarios are shown; Figures 4A to 4E Example interfaces for some scenarios are shown; Figures 5A to 5C Example interfaces for some scenarios are shown; Figures 6A to 6B Example interfaces for some scenarios are shown; Figure 7 The flowchart illustrates an example process for content generation in several scenarios; Figure 8 Schematic block diagrams of example devices for content generation in several scenarios are shown; and Figure 9 A block diagram of an electronic device capable of implementing multiple illustrative scenarios is shown. Detailed Implementation
[0011] The examples in this document will now be described in more detail with reference to the accompanying drawings. While some examples are shown in the drawings, it should be understood that solutions can be implemented in various forms and should not be construed as limited to the examples presented herein. Rather, these examples are provided to provide a more thorough and complete understanding of the solutions. It should be understood that the drawings and examples in this document are for illustrative purposes only and are not intended to limit the scope of protection of the solutions.
[0012] It should be noted that the headings of any section / subsection provided herein are not restrictive. Various examples are described throughout this document, and examples of any type may be included under any section / subsection. Furthermore, examples described in any section / subsection may be combined in any way with any other examples described in the same section / subsection and / or different sections / subsections.
[0013] In the description of the examples in this document, the term "including" and similar terms should be understood as open inclusion, i.e., "including but not limited to". The term "based on" should be understood as "at least partially based on". The term "an example" or "the example" should be understood as "at least one example". The term "some examples" should be understood as "at least some examples". Other explicit and implicit definitions may also be included below. The terms "first", "second", etc., may refer to different or the same objects. Other explicit and implicit definitions may also be included below.
[0014] The examples in this document may involve user data, data acquisition, and / or use. All of these aspects comply with relevant laws, regulations, and rules. In the examples, all data collection, acquisition, processing, manipulation, forwarding, and use are conducted with the user's knowledge and confirmation. Accordingly, when implementing each example, the type, scope of use, and usage scenarios of any data or information that may be involved should be communicated to the user and their authorization obtained through appropriate means, in accordance with relevant laws and regulations. The specific methods of notification and / or authorization can vary depending on the actual situation and application scenario; the scope of the solution is not limited in this regard.
[0015] In this manual and the sample solutions, any processing of personal information will be conducted only under legal grounds (such as obtaining the consent of the data subject or being necessary for the performance of a contract) and will only be carried out within the scope stipulated or agreed upon. A user's refusal to process personal information beyond what is necessary for basic functions will not affect the user's use of basic functions.
[0016] In this paper, "virtual object" refers to an object capable of autonomous control based on a machine learning model. A virtual object can, for example, make decisions and autonomously execute actions based on a machine learning model to achieve a preset goal or complete a preset task. In some cases, a virtual object can also be called an intelligent system, which can be, for example, an automated program that understands the user's intent and can use models or invoke tools to complete various types of tasks. Examples of virtual objects include, but are not limited to, agents, bots, chatbots, digital avatars, intelligent customer service representatives, and digital assistants. Alternatively, a virtual object can also be an intelligent role implemented based on a machine learning model. A "virtual object" can, for example, process user requests based on generative models (e.g., language models, multimodal models) to perform a specified type of task. In some cases, a virtual object can also refer to a virtual account, which may have a corresponding avatar or nickname.
[0017] As mentioned above, with the rapid development of computer technology, a variety of content generation tools have emerged on the market, such as applications for generating media content like images and videos. These content generation tools can quickly generate referable or editable media content based on user input, significantly improving the efficiency of creating media content or works containing media content.
[0018] However, in existing products, some applications can only present the final generated result on one interface, while the input materials and other content cannot be presented at the same time, which affects the user experience; some applications require switching interfaces to view different information such as the materials before generation and the media content after generation, resulting in cumbersome operation and low information presentation efficiency.
[0019] In related technologies, operations related to content generation need to be performed in a separate panel or window, which requires users to frequently switch between operation windows or operation areas, resulting in low operation efficiency.
[0020] Here, a content generation scheme is proposed. The scheme includes: receiving a first operation, the first operation indicating the selection of a first region of a canvas; presenting an input component at an associated location in the first region, wherein the state of the input component is related to the content in the first region; and in response to receiving a second operation via the input component, presenting at least one piece of media content in a second region of the canvas.
[0021] In this way, an input component can be presented at the associated position of the selected first area in the canvas, and the state of the input component can be associated with the content of the selected first area. At least one piece of media content can also be generated through the input component, thereby improving the visual relevance between the input control and the selected first area, improving the efficiency of input operations, and thus effectively improving the efficiency of media content generation. The presentation state of the input component can also be enriched based on the content of the selected area. Furthermore, by presenting at least one piece of generated media content in the second area of the canvas, the content in the first area is retained, and at least one piece of media content and the content in the first area are presented simultaneously in the same canvas.
[0022] The following describes various examples of this scheme in further detail with reference to the accompanying drawings.
[0023] Example Environment Figure 1 A schematic diagram of example environment 100 is shown. (e.g.) Figure 1 As shown, example environment 100 may include electronic device 110.
[0024] In this example environment 100, electronic device 110 may run an application 120 that supports content generation. Application 120 may be any suitable type of application for content generation, including but not limited to: content generation or creation applications, work publishing applications, or other suitable applications. User 140 may interact with application 120 via electronic device 110 and / or its attached devices.
[0025] exist Figure 1 In environment 100, if application 120 is active, electronic device 110 can use application 120 to present interface 150 for supporting content generation.
[0026] In some cases, electronic device 110 communicates with server 130 to provide services to application 120. Electronic device 110 can be any type of mobile terminal, fixed terminal, or portable terminal, including mobile phones, desktop computers, laptop computers, notebook computers, netbook computers, tablet computers, media computers, multimedia tablets, handheld computers, portable gaming terminals, VR / AR devices, personal communication system (PCS) devices, personal navigation devices, personal digital assistants (PDAs), audio / video players, digital cameras / camcorders, positioning devices, television receivers, radio receivers, e-book devices, gaming devices, or any combination of the foregoing, including accessories and peripherals of these devices or any combination thereof. In some cases, electronic device 110 can also support any type of user-facing interface (such as "wearable" circuitry).
[0027] Server 130 can be a standalone physical server, a server cluster or distributed system composed of multiple physical servers, or a cloud server providing basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communication, middleware services, domain name services, security services, content delivery networks, and big data and artificial intelligence platforms. Server 130 may include, for example, computing systems / servers such as mainframes, edge computing nodes, computing devices in a cloud environment, etc. Server 130 can provide backend services for applications 120 that support content generation in electronic devices 110.
[0028] A communication connection can be established between server 130 and electronic device 110. This communication connection can be established via wired or wireless means. The communication connection can include, but is not limited to, Bluetooth, mobile network, Universal Serial Bus (USB), and Wireless Fidelity (WiFi) connections. In some cases, server 130 and electronic device 110 can exchange signaling information through their communication connection.
[0029] It should be understood that the structure and function of the various elements in environment 100 are described for illustrative purposes only and do not imply any limitation on the scope of the scheme.
[0030] The following description of the example will continue with reference to the accompanying drawings.
[0031] Example Interaction Figures 2A to 2E Example interfaces 200A to 200E are shown, generated based on content from various scenarios. Interfaces 200A to 200E can, for example, be generated by... Figure 1 The electronic device 110 shown is provided. As an example, interfaces 200A to 200E can be interactive interfaces associated with application 120.
[0032] refer to Figure 2A As shown, in interface 200A, electronic device 110 can display a portion of canvas 210. As an example, canvas 210 can be a boundless virtual area (e.g., a planar area), and canvas 210 can also be referred to as an infinite canvas.
[0033] In some examples, the user 140 can switch the display area or display ratio of the canvas 210 in the interface 200A by at least one method such as sliding, panning, or inputting a scale. As an example, the electronic device 110 can provide an adjustment control 210a in the interface 200A. Through the adjustment control 210a, the electronic device 110 can receive the display ratio input by the user and update the display area of the canvas 210 in the interface 200A based on this display ratio. Thus, the electronic device 110 can update the display ratio of the content in the canvas 210 presented in the interface 200A.
[0034] Through interface 200A, user 140 can perform at least one interactive operation on canvas 210. As an example, such interactive operations may include, but are not limited to: drawing lines or shapes with a virtual pen, inputting text, erasing, selecting, and confirming.
[0035] In some cases, the electronic device 110 can receive a first operation via the canvas 210. As an example, such a first operation could indicate the selection of a first area 211 of the canvas 210.
[0036] In some cases, the electronic device 110 can determine the corresponding first operation and the first region 211 corresponding to the first operation based on the selection operation of at least one content (such as text content, image content, identification element, etc.) in the first region 211.
[0037] In some cases, the electronic device 110 may also determine the corresponding first region 211 based on the received drag operation (e.g., drawing a rectangle). As an example, such a first region 211 may correspond to a blank area or a non-blank area including at least one element, without limitation.
[0038] In some examples, in interface 200A, electronic device 110 can provide a set of controls 220, via a first control 221 in the set of controls 220, electronic device 110 can receive drag operations to determine the position and extent of the first region 211 by using the line between the start and end points of the drag operation as the diagonal of the first region 211.
[0039] In some examples, the electronic device 110 can also accept at least one click operation via a second control 222 in a set of controls 220 to determine the corresponding first region 211. As an example, in response to a trigger operation on the second control 222, the electronic device 110 can present a set of candidate specification information (e.g., multiple aspect ratios). In response to the selection of a first specification information from the set of candidate specification information, the electronic device 110 can determine the specification information corresponding to the first region 211. Then, in response to receiving a click operation on a first position in the canvas 210, the electronic device 110 can determine the position of the first region 211 in the canvas 210 based on the first position and the first specification information. For example, the electronic device 110 can use the coordinates of the first position as the center coordinates or corner coordinates (e.g., the upper left corner or lower left corner) of the first region, and then perform a diffusion calculation from the first position based on the first specification information to determine the area range of the first region 211 in the canvas 210, i.e., the position information of the first region 211.
[0040] In some cases, in response to determining the location information of the first region 211, the electronic device 110 may present the bounding box 211a of the first region 211 in the canvas 210 to visually represent the location of the first region 211 and its area extent in the canvas 210.
[0041] In some examples, the electronic device 110 may also display the specification information 211b of the first region 211 (e.g., aspect ratio information, which may be displayed as 300×400) at an associated location of the bounding box 211a. As an example, the electronic device 110 may display the specification information 211b at the upper right or lower right corner of the bounding box 211a.
[0042] In some examples, the electronic device 110 can also receive a switching operation via specification information 211b. For example, in response to a trigger operation on specification information 211b, the electronic device 110 can present a set of candidate specification information or switch specification information 211b to an editable state; then, in response to the selection of a second specification information from the set of candidate specification information or receiving an edit operation that modifies specification information 211b to the second specification information, the electronic device 110 can switch specification information 211b to the value corresponding to the second specification information.
[0043] In some cases, in response to receiving the first operation described above, the electronic device 110 may present the input component 230 (which may also be referred to as the first input component) at an associated location in the first region 211. As an example, such an associated location may be below or above the first region 211. For example, the electronic device 110 may present the input component 230 at a position spaced from the lower boundary of the first region 211 by a predetermined distance.
[0044] In some cases, the presentation state (or display style) of the input component 230 may be related to the content in the selected first area 211. As an example, in response to the inclusion of at least one element in the first area 211, the electronic device 110 may present a first state of the input component 230. Such a first state could be, for example, a collapsed or thumbnailed state of the input component 230. For instance, associated with such a first state, the electronic device 110 may present a confirmation control in the input component 230 to trigger a generation request associated with the content in the first area 211 via the confirmation control. For example, such a generation request could instruct the generation of at least one piece of media content (e.g., image content) based on the content in the first area 211.
[0045] In some examples, such as Figure 2A As shown, in response to the first region 211 being a blank region, the electronic device 110 can present a second state of the input component 230. As an example, such a second state could be an expanded state of the input component 230. For example, such an expanded state could indicate the input area of the input component 230, for example... Figure 2A The input area 231 is shown. Through the input area 231, the electronic device 110 can receive at least one input from the user 140 and display it in the input area 231. For example, such input may include… Figure 2B The input text shown is 241.
[0046] Comprehensive reference Figure 2A and Figure 2B As shown, corresponding to the second state, the electronic device 110 can also present an upload control 232 in the input component 230. Through the upload control 232, the electronic device 110 can receive at least one piece of reference content. As an example, such at least one piece of reference content may include image content, video content, document content, etc.
[0047] Taking the receipt of the first reference content as an example, the electronic device 110 can display the indication information 242 corresponding to the first reference content in the input component 230.
[0048] In some cases, such indication information 242 at least describes the type of the first reference content. As an example, such indication information 242 may include at least one type of identification information such as an image identifier or a label identifier corresponding to the first reference content.
[0049] In the input component 230, the electronic device 110 may also provide a confirmation control 233. Via the confirmation control 233, the electronic device 110 may receive a second operation that indicates a generation request for generating at least one type of media content.
[0050] In this way, an input component can be presented at the associated position of the selected first area in the canvas, and the state of the input component can be associated with the content of the selected first area. At least one piece of media content can also be generated through the input component, thereby improving the visual relevance between the input control and the selected first area, improving the efficiency of input operations, and thus effectively improving the efficiency of media content generation. The presentation state of the input component can also be enriched based on the content of the selected area. Furthermore, by presenting at least one piece of generated media content in the second area of the canvas, the content in the first area is retained, and at least one piece of media content and the content in the first area are presented simultaneously in the same canvas.
[0051] Comprehensive reference Figure 2B and Figure 2C As shown, taking the receipt of input text 241 via input component 230 as an example, in response to the triggering operation of confirmation control 233, electronic device 110 can receive a second operation and present at least one piece of media content (e.g., ...) in the second area of canvas 210. Figure 2C (At least one piece of media content 250 shown). As an example, such at least one piece of media content 250 is generated based on input text 241.
[0052] As an example, such a second region can be a region different from the first region 211, or it can be a region that includes the first region 211. In some examples, the electronic device 110 can determine the second region at an associated location (e.g., adjacent to the right, below, etc.) of the first region 211. For example, if the first region 211 includes at least one piece of content, the second region can be an adjacent region of the first region 211. In some examples, such as when the first region 211 is a blank area, the electronic device 110 can determine that the second region is an adjacent region of the first region 211, or it can determine that the second region includes or covers the first region 211. For example, the electronic device 110 can use a corner point of the first region 211 as one of the corner points of the second region, and determine the region of the second region in the canvas 210 based on the number and specifications of at least one piece of media content 250.
[0053] In some examples, after receiving at least one piece of reference content via input component 230, in response to a triggering operation of confirmation control 233, electronic device 110 may present at least one piece of media content 250, which is generated based on at least one piece of reference content. For example, such at least one piece of media content 250 may be based on at least one element (e.g., content object, visual effect, etc.) in the at least one piece of reference content.
[0054] In some examples, in response to receiving input text 241 and at least one reference content via input component 230, based on a trigger operation of confirmation control 233, electronic device 110 can present at least one corresponding media content 250, such at least one media content 250 being generated based on input text 241 and at least one reference content.
[0055] In some cases, the electronic device 110 may also provide at least one configuration control 234 in the input component 230. As an example, via at least one configuration control 234, the electronic device 110 may configure at least one generation parameter associated with at least one media content. For example, such generation parameters may indicate the number of at least one media content, the model or algorithm on which the at least one media content is generated, etc.
[0056] refer to Figure 2D As shown, after presenting at least one piece of media content 250, the electronic device 110 can receive a selection operation for any one of the at least one piece of media content 250 to trigger viewing, editing or modifying the media content.
[0057] As an example, in response to a triggering operation on media content 251, electronic device 110 may display a prompt message 261 at an associated location of media content 251 to indicate that an editing operation on media content 251 can be triggered. For example, in response to the triggering of prompt message 261, electronic device 110 may switch prompt message 261 to input component 230 to receive input content and trigger an adjustment request for media content 251. As an example, in response to receiving such an adjustment request, electronic device 110 may switch media content 251 to the adjusted media content, or, while retaining media content 251, display the adjusted media content at an associated location (e.g., below or to the right) of media content 251.
[0058] In some cases, in response to a triggering operation on media content 251, electronic device 110 may present an add control 262 at an associated location of media content 251. As an example, add control 262 may be configured to trigger the generation of additional media content.
[0059] Comprehensive reference Figure 2D and Figure 2E As shown, in response to the triggering of the add control 262, the electronic device 110 can present supplementary media content 270 at an associated location (e.g., the right side) of the media content 251. As an example, the basis for generating such supplementary media content may be at least partially the same as the basis for generating at least one piece of media content 250. For example, at least one piece of media content 250 is generated based on input text 241 and at least one reference content. In response to the triggering of the add control 262, the electronic device 110 can use a model to generate supplementary media content 270 based on the input text 241 and at least one reference content, and present the generated supplementary media content 270 at the associated location of the media content 250. For example, the electronic device 110 can switch the add control 262 to present supplementary media content 270.
[0060] Figures 3A to 3D Example interfaces 300A to 300D are shown, generated based on content from various scenarios. Interfaces 300A to 300D can, for example, be generated by... Figure 1 The electronic device 110 shown is provided. As an example, interfaces 300A to 300D can be interactive interfaces associated with application 120.
[0061] refer to Figure 3A As shown, in interface 300A, electronic device 110 can display a portion of canvas 310. As an example, canvas 310 can be a boundless virtual area (e.g., a planar area), and canvas 310 can also be referred to as an infinite canvas.
[0062] In some cases, the electronic device 110 can receive a first operation via the canvas 310. For example, such a first operation could indicate the selection of a first area 311 of the canvas 310.
[0063] As an example, electronic device 110 can receive the first operation by selecting an element or by clicking on an element. For instance, in response to a selection operation on elements 321 and 322 in canvas 310, electronic device 110 can determine the area range corresponding to the selected first region 311 based on the positions of elements 321 and 322. In some cases, element 321 may also be referred to as the first element, and element 322 may also be referred to as the second element.
[0064] As an example, such elements 321 and 322 may include at least one of the following: image elements, text elements, graphic elements, identifier elements (such as identifier lines or labels, etc.).
[0065] In some examples, in response to receiving a first operation and determining the area range corresponding to the first region 311, the electronic device 110 can present the bounding box 311a of the first region 311 in the canvas 310. In this way, the area range of the first region 311 can be visually presented, enriching the visual effects associated with the canvas 310.
[0066] In some cases, in response to receiving a selection operation on element 321 and / or element 322, electronic device 110 may also present a set of controls 323 at an associated location or associated direction (e.g., above) of the first region 311 to trigger at least one editing operation on the selected element.
[0067] As an example, such editing operations include, but are not limited to, at least one of the following: enhancing image quality, regenerating, blending operations, triggering to join a session, picking objects, generating videos, etc.
[0068] In some cases, in response to receiving the first operation described above, the electronic device 110 may present the input component 330 (which may also be referred to as the first input component) at an associated location in the first region 311. As an example, such an associated location may be below or above the first region 311. For example, the electronic device 110 may present the input component 330 at a predetermined distance from the lower boundary of the first region 311.
[0069] In some examples, the presentation state of input component 330 may be related to the content in the selected area 311. As an example, in response to the selected first area 311 including at least one element, such as elements 321 and 322, the electronic device 110 may present a first state of input component 330, for example, indicating that input component 330 is in a collapsed or thumbnailed state. As an example, corresponding to such a first state, the electronic device 110 may present a prompt message 331 in input component 330, which may be configured to trigger a switch of input component 330 from the first state to a second state. For example, such a second state may indicate that input component 330 is in an expanded state. As an example, prompt message 331 may be presented as "Describe an idea".
[0070] In some examples, in response to a triggering operation on prompt message 331 or input component 330, electronic device 110 can transfer input component 330 from... Figure 3A The first state shown is switched to Figure 3B The second state is shown.
[0071] Reference Figure 3B As shown, in response to the selection of elements 321 and 322, the electronic device 110 may correspond to a second state and present indication information 341 (e.g., also referred to as first indication information) and indication information 342 (e.g., also referred to as second indication information) in the input component 330.
[0072] As an example, instruction information 341 corresponds to element 321, and instruction information 342 corresponds to element 322. The information content corresponding to instruction information 341 and instruction information 342 can be the same or similar.
[0073] Taking instruction information 341 as an example, such instruction information 341 at least describes the type of element 321. As an example, instruction information 341 may include at least one type of identification information such as an image identifier or a label identifier corresponding to element 321. Such an image identifier may include at least one of the following: an icon corresponding to the element or at least one object within the element. Such a label identifier may be determined based on at least one of the following: the source of the element, the type of the element, the type of at least one object within the element, etc.
[0074] In some cases, the electronic device 110 may receive at least one interactive operation on the indication information (e.g., indication information 341 and indication information 342) presented in the input component 330. Such interactive operations may include, but are not limited to: click operations, delete operations, drag operations (e.g., moving the presentation order between different indication information), hover operations (e.g., hovering a mouse or finger over the presentation position of the indication information without clicking to trigger it), etc.
[0075] In some examples, continuing with the example of an interactive operation on indication information 341, in response to receiving an interactive operation associated with indication information 341, such as an interaction that instructs selection of indication information 341, electronic device 110 may display preview content 341a associated with indication information 341. As an example, such an interactive operation may also be referred to as a third operation, which may include a click or selection operation on indication information 341. As an example, such preview content 341a corresponds to element 321. For example, such preview content 341a may be a preview image corresponding to element 321 or at least one object within element 321.
[0076] In some examples, the electronic device 110 can also receive a deletion operation for the indication information 341. In response to receiving such a deletion operation, the electronic device 110 can stop displaying the indication information 341 in the input component 330 and switch the element 321 in the first area 311 from a selected state to an unselected state.
[0077] In some cases, electronic device 110 can receive input text 343 via the input area of input component 330. For example... Figure 3B As shown, the electronic device 110 can display the received input text 343 in the input area of the input component 330.
[0078] The input component 330 also includes a confirmation control 332. Through the confirmation control 332, the electronic device 110 can receive a generation request associated with the information presented in the input component 330.
[0079] In some cases, in response to receiving a trigger operation on the confirmation control 332, the electronic device 110 determines the presentation information in the input component 330. For example, if the input component 330 includes indication information 341, indication information 342, and input text 343, the electronic device 110 can determine the generation request indication corresponding to the trigger operation and generate at least one piece of media content based on elements 321, 322, and input text 343. For example, if the input component 330 only includes indication information 341 and indication information 342 when a trigger operation on the confirmation control 332 is received, the electronic device 110 can determine the generation request indication corresponding to the trigger operation and generate at least one piece of media content based on elements 321 and 322.
[0080] Comprehensive reference Figure 3B and Figure 3CAs shown, in response to a triggering operation of the confirmation control 332, the electronic device 110 can present at least one generated media content 350 in a second area of the canvas 310. As an example, such a second area can be a location area associated with the first area 311, or a location area associated with the selected elements 321 and 322. For example, such an associated location area can indicate a location area spaced at a preset distance in a preset direction (e.g., to the right or below).
[0081] In some cases, the electronic device 110 may also provide an interaction entry 360 in the interactive interface corresponding to the canvas 310 (e.g., interface 300A, interface 300B, or interface 300C). As an example, such an interaction entry 360 may be configured to trigger a dialogue interaction with a virtual object.
[0082] Comprehensive reference Figure 3C and Figure 3D As shown, in response to a trigger operation on the interaction entry 360, the electronic device 110 can present a session window 370 in the interface 300D. Such a session window 370 is used for dialogue interaction with virtual objects.
[0083] like Figure 3D As shown, the session window 370 may include an input component 371 (which may also be referred to as a second input component). Through the input component 371, the electronic device 110 can receive user input content, and in response to a confirmation operation of the input content, the electronic device 110 can present a session message corresponding to the input content in the session window 370, and can also present a response message for the session message in the session window 370.
[0084] In some examples, such a response message may indicate prompts for the input content, or it may indicate descriptive information about the generated results associated with the input content. As an example, such generated results may include media content generated at least based on the input content, and such generated results may be presented in canvas 310.
[0085] In some cases, in response to content being selected in canvas 310, such as element 321 in canvas 310 or at least one piece of media content 350 being selected, electronic device 110 may present indication information (e.g., referred to as third indication information) corresponding to the selected content in input component 371. As an example, such indication information at least describes the type of the selected content, such as image type, text type, etc.
[0086] When presenting such instruction information, the electronic device 110 can also receive a generation request or editing request associated with the selected content via the input component 371, and present at least one piece of media content obtained based on the generation request or editing request in the canvas 310.
[0087] In some cases, in response to via Figure 3B The input component 330 receives a generation request and presents at least one corresponding media content 350 in the canvas 310. The electronic device 110 can also present a session message 372 in the session window 370. As an example, such a session message 372 can describe the interactive operation received via the input component 330, or it can describe at least one media content 350 generated based on the interactive operation, without limitation.
[0088] Figures 4A to 4E Example interfaces 400A to 400E are shown, generated based on content from various scenarios. Interfaces 400A to 400E can, for example, be generated by... Figure 1 The electronic device 110 shown is provided. As an example, interfaces 400A to 400E can be interactive interfaces associated with application 120.
[0089] refer to Figure 4A As shown, in interface 400A, electronic device 110 can display a portion of canvas 410. As an example, canvas 410 can be a boundless virtual area (e.g., a planar area), and canvas 410 can also be referred to as an infinite canvas.
[0090] Through interface 400A, user 140 can perform at least one interactive operation on canvas 410. As an example, such interactive operations may include, but are not limited to: drawing lines or shapes with a virtual pen, inputting text, erasing, selecting, and confirming.
[0091] by Figure 4A Taking the interface 400A as an example, the electronic device 110 can present a set of controls 420 in the interactive interface 400A corresponding to the canvas 410, which are used to trigger and execute at least one interactive operation associated with the canvas 410.
[0092] In some cases, electronic device 110 can receive a tagging operation via canvas 410 and display tagging elements on canvas 410, such tagging elements corresponding to the received tagging operation.
[0093] As an example, a set of controls 420 may include a drawing control 421, which is triggered to add drawing elements such as lines or wireframes to the canvas 410. In some examples, in response to the drawing control 421 being triggered, the electronic device 110 may present a set of drawing options. For example, via such a set of drawing options, the user 140 may select the shape of the line to be drawn (e.g., straight line, curve, wireframe, etc.), the line thickness, the line color, the type of virtual brush (e.g., brush, pencil, pen, eraser, etc.).
[0094] Taking the selection of the curve option from a set of drawing options as an example, in response to receiving a drag operation (which can also be called a marking operation), the electronic device 110 can present the corresponding marking element in the canvas 410 based on the drag trajectory. As an example, such marking elements can correspond to the artwork created by the user 140 in the blank area of the canvas 410, or they can be marking information associated with existing content in the canvas 410, without limitation.
[0095] Taking an existing element 401 (also referred to as the first element) in canvas 410 as an example, electronic device 110 can present a marker element 402 in canvas 410 based on a received marker operation. Such a marker element 402 can be associated with element 401 or at least one object in element 401. For example, marker element 402 can be overlaid on element 401.
[0096] Taking the marking element 402 as an example of indicating the marking or selection of object 401a in element 401, the electronic device 110 can present the marking element 402 at the position in element 401 associated with object 401a, based on the received operation trajectory.
[0097] In some examples, in response to the presentation of marker element 402, electronic device 110 can also receive a click operation on marker element 402 to trigger switching marker element 402 to an editable state to adjust the line position, color, thickness or shape of the marker element 402.
[0098] In some cases, a set of controls 420 may also include a text control 422, which is triggered to add text content to the canvas 410. As an example, when the text control 422 is triggered, the electronic device 110 may receive a click or selection action on the canvas 410 and determine the presentation position of the text content based on such action; and in response to receiving the text content, the electronic device 110 may present the received text content at the determined presentation position. For example, such text content may be presented as a text element 403 (also referred to as a markup element).
[0099] In some examples, such a text element 403 may be associated with an element 401 or object 401a marked by a markup element 402. For example, text element 403 may indicate descriptive or modification information about the marked content in markup element 402.
[0100] In some examples, the electronic device 110 can also receive click or selection operations on the text element 403 to trigger the text element 403 to be rendered as editable, and receive editing operations on the text element 403 to adjust the text format, text content, etc. of the text element 403, and then render the adjusted text element 403.
[0101] In some cases, the electronic device 110 can also display a marker element 404 on the canvas 410 via a drawing operation received from the drawing control 421 to mark the relationship between the text element 403 and the marker element 402. As an example, such a marker element 404 can be displayed as a connecting line or an arrowed indicator line, without limitation.
[0102] In some cases, the electronic device 110 can also receive selection operations on element 401, marker element 402, text element 403, and marker element 404, and trigger adjustment requests for element 401 or content generation requests associated with the selected element via such selection operations. As an example, such selection operations may include point selection operations on each element, or box selection operations corresponding to drag operations, without limitation.
[0103] In some cases, in response to receiving a selection operation on element 401, marker element 402, text element 403, and marker element 404, electronic device 110 can determine a corresponding first region 411 based on the presentation position of the selected element in canvas 401. For example, electronic device 110 can determine the first region 411 based on the area range of all selected elements in canvas 410, such that the first region 411 at least includes the presentation position range of each selected element. In some examples, electronic device 110 can also determine the corresponding first region 411 based on the presentation area of element 401 in canvas 410, for example, the boundary of the first region 411 corresponds to the position boundary of element 401.
[0104] In some examples, in response to a received selection operation, the electronic device 110 may also determine a first region 411 based on the drag range of the selection operation, and determine the selected element based on the content within the coverage area of the first region 411.
[0105] In some examples, in response to the inclusion of element 401 in the first region 411, the electronic device 110 can render the bounding box 401b corresponding to element 401.
[0106] In some examples, in response to the inclusion of a marker element 402 in the first region 411, the electronic device 110 can present indication information 402a (also referred to as third indication information) on the canvas 410, such indication information 402 indicating that the marker element 402 is selected or in a selected state. For example, such indication information 402a can be presented as an indicator box, and the marker element 402 can be presented within the area corresponding to such an indicator box.
[0107] Additionally or alternatively, in response to the inclusion of text element 403 and marker element 404 in the first area 411, or in response to determining that text element 403 and marker element 404 are selected, electronic device 110 may also present instruction information 403a corresponding to text element 403 and instruction information 404a corresponding to marker element 404.
[0108] In some cases, in response to determining the first region 411, the electronic device 110 may present the input component 430 at an associated location in the first region 411 or the bounding box 401b. As an example, such an associated location may be below or above the first region 411. For instance, the electronic device 110 may present the input component 430 at a predetermined distance from the lower boundary of the first region 411.
[0109] In some cases, the presentation state (or display style) of the input component 430 may be related to the content in the selected first area 411. As an example, in response to the inclusion of at least one element in the first area 411, the electronic device 110 may present a first state of the input component 430. Such a first state could be, for example, a collapsed or thumbnailed state of the input component 430. As an example, corresponding to such a first state, the electronic device 110 may present a prompt message 431 in the input component 430. As an example, the prompt message 431 could be presented as "Describe your idea".
[0110] In some examples, in response to a triggering operation on the input component 430 or the prompt message 431, the electronic device 110 can switch the input component 430 from a first state to a second state. As an example, such a second state could indicate that the input component 430 is in an expanded state, for example... Figure 4C The input component 430 shown corresponds to the second state.
[0111] In some cases, in response to the inclusion of element 401 in the first region 411, the electronic device 110 may present indication information 440 in the input component 430. Such indication information 440 may at least describe the type of element 401. Taking element 401 as an image type as an example, the indication information 440 may include an image identifier corresponding to the image type or element 401, and may also include an identifier label associated with the image type or selection state of element 401.
[0112] In response to a triggering operation on the confirmation control 432 in the input component 430, the electronic device 110 can receive an interaction request associated with the selected element. For example, such an interaction request may include a request to generate media content based on the selected element. In some examples, in response to a triggering operation on the confirmation space 430, the electronic device 110 can switch the interface 400C to... Figure 4D The interface shown is 400D.
[0113] As shown in reference interface 400D, electronic device 110 can determine a second region 412 at an associated location of the first region 411 based on the amount of media content indicated by the generation request. As an example, electronic device 110 can determine the corresponding second region 412 at a position to the right or below the bounding box of the first region 411, at a preset distance.
[0114] In some cases, in response to determining the second region 412, the electronic device 110 may present a corresponding number of placeholders based on the amount of media content indicated by the generation request.
[0115] Taking a request to generate media content of 1 as an example, the electronic device 110 may present a placeholder identifier 450 in the second area 412. As an example, such a placeholder identifier 450 may be presented as a bounding box or a padding element corresponding to the size of the second area 412.
[0116] In some examples, in response to a triggering operation of the confirmation control 432, the electronic device 110 may stop displaying the indication information corresponding to the marker element 402, text element 403 and marker element 404 in the canvas 410, for example, stop displaying indication information 402a, indication information 403a and indication information 404a.
[0117] In some cases, in response to a triggering operation of the confirmation control 432, the electronic device 110 can also switch the input component 430 to a third state. Corresponding to the third state, the electronic device 110 can present the interactive control 433 in the input component 430. As an example, the interactive control 433 can be configured to trigger to pause or abort processing of the corresponding generation request. In some examples, in response to the interactive control 433 being triggered, the electronic device 110 can stop presenting the placeholder identifier 450.
[0118] In some cases, in response to the completion of the generation request, the electronic device 110 can switch the placeholder identifier 450 to the generated media content, for example... Figure 4E The media content 460 shown. As an example, in response to the confirmation control 432 being triggered, the marker element 402, text element 403, and marker element 404 in the first area 411 are selected, and such media content 460 is associated with the selected marker element. As an example, such media content 460 may be associated with at least one of the following: the content of the marker element, the positional relationship of the marker element to the first element, etc.
[0119] As an example, media content 460 can be associated with the text content of text element 403, or with the positional relationship between marker element 402 and element 401 or object 401a. For example, electronic device 110 can determine that object 401a in element 401 is selected based on the positional relationship between marker element 402 and element 401; it can also determine processing information corresponding to the selected object 401a based on the content of text element 403. For example, such processing information can instruct that object 401a be replaced with a reference object. Based on such processing information, electronic device 110 can process element 401 to generate corresponding media content 460, which is then presented in the second area 412.
[0120] Figures 5A to 5C Example interfaces 500A to 500C are shown, generated based on content from various scenarios. Interfaces 500A to 500C can, for example, be generated by... Figure 1 The electronic device 110 shown is provided. As an example, interfaces 500A to 500C can be interactive interfaces associated with application 120.
[0121] refer to Figure 5A As shown, in interface 500A, electronic device 110 can display a portion of canvas 510. As an example, canvas 510 can be a boundless virtual area (e.g., a planar area), and canvas 510 can also be referred to as an infinite canvas.
[0122] Through interface 500A, user 140 can perform at least one interactive operation on canvas 510. As an example, such interactive operations may include, but are not limited to: drawing lines or shapes with a virtual pen, inputting text, erasing, selecting, and confirming.
[0123] like Figure 5A As shown, the content in canvas 510 includes text content 520. In some cases, electronic device 110 may receive a selection operation (also referred to as a first operation) on the text content 520.
[0124] As an example, in response to a selection operation on text content 520, electronic device 110 can determine the corresponding first region 511 based on the position range of content 520 in canvas 510.
[0125] In some examples, in response to a selection operation on text content 520, electronic device 110 may render the bounding box 511a of the first region 511.
[0126] In some cases, in response to a selection operation on text content 520, electronic device 110 may also present input component 530 at an associated location in the first area 511. As an example, the state of input component 530 is related to the content type in the first area 511.
[0127] In some examples, in response to content (e.g., text content 520) being selected in the first area 511 and consisting only of text, the electronic device 110 may present the input component 530 corresponding to the first state. As an example, such an input component 530 corresponds to a collapsed state or a brief state. Figure 5A As shown, the electronic device 110 may provide a confirmation control 531 in the input component 530.
[0128] In some cases, when text content 520 is selected, in response to receiving a trigger operation (also known as a second operation) on confirmation control 531, electronic device 110 can determine the generation request indicated by the trigger operation. For example, such a generation request indicates the generation of at least one piece of media content based on text content 520.
[0129] As an example, see comprehensive reference Figure 5A and Figure 5B As shown, in response to a triggering operation of the confirmation control 531, the electronic device 110 can present media content 540 in the second area of the canvas 510, such media content 540 being generated based on text content 520.
[0130] As an example, such a second region can be determined based on the presentation position of the first region 511 or the text content 520. For example, the electronic device 110 can determine the starting position of the second region at a preset interval in a preset direction of the first region 511 or the text content 520.
[0131] Comprehensive reference Figure 5A and Figure 5CAs shown, in some cases, corresponding to the input component 530 presented in the first state, the electronic device 110 can also receive a trigger operation on the input component 530 to switch the presentation state of the input component 530. For example, in response to a trigger operation on the input component 530 in the first state in the interface 500A, the electronic device 110 can switch the input component 530 from the first state to... Figure 5C The second state is shown.
[0132] As an example, such a second state can indicate that the input component 530 is in an expanded state, for example, the input area of the input component 530 corresponds to the expanded state.
[0133] like Figure 5C As shown, corresponding to the second state, in response to the text content 520 being selected, the electronic device 110 can present instruction information 550 in the input component 530, which corresponds to the text content 520.
[0134] As an example, the instruction information 550 at least describes the type of the text content 520. For example, the instruction information 550 may include at least one type of identification information such as an image identifier or a label identifier corresponding to the text content 520. Such an image identifier can be a text identifier, such as an icon corresponding to the letter T. Such a label identifier can be a label identifier that is presented as "text".
[0135] In some cases, electronic device 110 may also receive input text via the input area of input component 530; and / or, via the upload control in input component 530, receive at least one piece of reference content. For example, in response to receiving reference content, electronic device 110 may also display indication information corresponding to the reference content in the input component.
[0136] In response to a triggering operation of a confirmation control in input component 530, electronic device 110 can determine a generation request corresponding to the triggering operation based on information presented in the input area of input component 530 (e.g., indicator information and / or input text). For example, if the input area of input component 530 includes indicator information 550 and input text, electronic device 110 can determine that the generation request indicates that at least one piece of media content be generated based on text content 520 and input text, and presented in a second area of canvas 510, for example... Figure 5B The media content shown is 540.
[0137] like Figure 5CAs shown, in some cases, in response to a selection operation on the text content 520, the electronic device 110 may also present a set of controls 560. As an example, such a set of controls 560 may indicate a trigger to perform at least one interactive operation on the text content 520. For example, such interactive operations include, but are not limited to: adjusting font size, adjusting font type, adjusting line spacing, adjusting font color, adjusting text background color, translating text content, etc.
[0138] Figures 6A to 6B Example interfaces 600A to 600B are shown, generated based on content from various scenarios. Interfaces 600A to 600B can, for example, be generated by... Figure 1 The electronic device 110 shown is provided. As an example, interfaces 600A to 600B can be interactive interfaces associated with application 120.
[0139] refer to Figure 6A As shown, in interface 600A, electronic device 110 can display a portion of canvas 610. As an example, canvas 610 can be a boundless virtual area (e.g., a planar area), and canvas 610 can also be referred to as an infinite canvas.
[0140] Through interface 600A, user 140 can perform at least one interactive operation on canvas 610. As an example, such interactive operations may include, but are not limited to: drawing lines or shapes with a virtual pen, inputting text, uploading media content such as images or videos, erasing, selecting, and confirming.
[0141] In some cases, electronic device 110 can receive at least one input content via interface 600A. As an example, such input content may include, but is not limited to: text content, image content, line content, graphic content, audio content, document content, etc. In response to receiving at least one input content, electronic device 110 can display the received input content on canvas 610.
[0142] As an example, in response to receiving media content 620 (which may also be referred to as first media content), electronic device 110 can present the media content 620 on canvas 610.
[0143] In some cases, the electronic device 110 can also receive content retrieval requests associated with the media content 620. As an example, such content retrieval requests can be triggered via an input component, via an interactive entry point associated with a virtual object in the interface, or via other controls in the interactive interface.
[0144] by Figure 6ATaking the scenario shown as an example, in interface 600A, electronic device 110 provides interactive control 630. In response to the triggering of interactive control 630, electronic device 110 can display an interactive panel, for example... Figure 6B The interactive panel 640 shown.
[0145] In some cases, in response to media content 620 being selected, electronic device 110 may receive a content acquisition request via interactive panel 640 and present at least one reference media content 641 in interactive panel 640, such at least one reference media content 641 being associated with the selected media content 620.
[0146] In some examples, at least one reference to media content 641 may be determined based on at least one piece of information such as objects, materials, or associated scenes in media content 620.
[0147] As an example, in response to media content 620 including object 625 (which may also be referred to as a first object), electronic device 110 can determine at least one reference media content 641 based on object 625.
[0148] In some cases, electronic device 110 can determine at least one reference media content 641 from a preset media set based on the association information of object 625. As an example, such association information may include, but is not limited to, at least one of the following: object type, physical characteristics (e.g., shape, color, material, etc.), application scenario, etc. As an example, a second object included in at least one reference media content 641 is related to object 625. For example, the second object may be of the same type as object 625, have a similar shape, or be applicable in the same scenario, etc.
[0149] For example, object 625 is a yellow shirt, and at least one of the objects in reference media content 641 may include a yellow shirt, a yellow jacket, a shirt with a similar pattern to object 625, etc.
[0150] In the interactive panel 640, the electronic device 110 may also provide an input control 642. Through the input control 642, the electronic device 110 may receive a content acquisition request or an update request for at least one piece of reference media content 641.
[0151] In some cases, in response to media content 620 being selected, electronic device 110 can display indication information 643 corresponding to media content 620 in input control 642. As an example, indication information 643 can indicate the content type of media content 620, or it can indicate the object 625 in media content 620 or the object type corresponding to object 625.
[0152] As an example, via input control 642, electronic device 110 can receive descriptive information input by user 140, determine the corresponding content acquisition request based on such descriptive information, and present at least one piece of reference media content acquired in interactive panel 640.
[0153] In some cases, the canvas in this article can be configured with a layer structure, which may include multiple layers. As an example, such multiple layers may correspond to different management dimensions.
[0154] In some examples, such multiple layers may include, but are not limited to: root node layer, region layer, page layer, element group layer, element layer, etc.
[0155] As an example, the root node layer can be configured as the top-level container of the canvas, with all content in the canvas attached to it. Region layers can be configured to organize multiple pages; they can be created manually or automatically generated by virtual objects such as smart assistants based on narrative content. Page layers can be configured as the basic display unit in the canvas, containing one or more element groups or just a single element. Element group layers, composed of multiple elements, can be configured to support overall operations, such as synchronously moving, scaling, locking, and hiding multiple elements within a group. Element layers are the most basic building blocks of the canvas, and supported element content can include, but is not limited to, images, text, graphics, stickers, videos, and documents.
[0156] In some cases, a region layer can contain a page layer, but not directly an element layer.
[0157] In some examples, a page layer cannot contain other page layers, but elements within a page can be grouped to form element groups.
[0158] In some examples, an element layer can belong to the page layer, or it can be located directly in the canvas root node layer and become a detached element.
[0159] In some cases, the canvas described in this article can be configured with a multi-level content organization structure to facilitate the management and manipulation of elements within the canvas at different granularities. As an example, such a multi-level structure could include a region layer, a page layer, and an element layer.
[0160] In some examples, region layers can be configured as logical aggregations of multiple page layers, suitable for displaying narrative content such as documents or storyboards. Page layers within a region layer can be configured to be arranged in the order of generation (e.g., from left to right or top to bottom). Corresponding to the region layer, the default number of page layers displayed per row or column can also be configured, for example, 5 page layers per row by default. Such region layers can be configured to support operations such as renaming, previewing, and merging.
[0161] In some examples, page layers can be configured as independent content units within the canvas, containing multiple element layers or element groups. Page layers can be configured to support background color settings, resizing, copying as an image, downloading, and other functions. Spacing between different page layers can also be configured, for example, a default spacing of 10px. In some examples, page layers can also be configured to support free dragging for layout adjustments.
[0162] In some examples, element layers can be configured as the most basic content units in the canvas, including images, text, graphics, stickers, etc. Element layers can be operated individually, or multiple elements can be grouped into an element group layer for unified management. Element layers can be configured to support operations such as locking, hiding, redrawing, text editing, image cutout, image enlargement, and high-resolution processing.
[0163] In some cases, the canvas in this article can also be configured with a hierarchical content download mechanism, which can provide differentiated download strategies based on different levels of the selected content. As examples, such differentiated download strategies could include single-selection download, multi-selection download, etc.
[0164] Taking the single-selection download strategy as an example, when a user selects a single element layer, the canvas can support downloading the single element as an image format and provide settings such as format, size, and quality; when a user selects a single element group layer, the electronic device 110 can merge all elements in the group into a single download and support flat export; when a user selects a single page layer, all content in the page layer is packaged into a single image for download.
[0165] In some examples, when a user selects a region layer, the canvas can support choosing the image format to download and package the entire region layer into a compressed file. As an example, such a compressed file could include thumbnails corresponding to the region layer, as well as images corresponding to all page layers. In some examples, when a user selects a region layer, it can also support selecting a file format to download the region layer's content as a file in that format. For example, electronic device 110 can automatically generate and output a file in that format according to the page layer order within the region layer.
[0166] Corresponding to the multi-select download strategy, when a user selects multiple element layers, the electronic device 110 can package all element layers into a compressed file; when a user selects multiple page layers, the electronic device 110 can package all page layers into a compressed file, with each page layer as a unit; when a user selects multiple region layers, the electronic device 110 can also package different region layers as different content units into a single compressed file.
[0167] In some examples, when a user selects multiple different levels of content, if the selected content includes area layers and the export format is an image, the electronic device 110 can package each area layer into a sub-file with its own independent compressed file format, and then package all the sub-files into a compressed file for output.
[0168] Example process Figure 7 A flowchart of an example process 700, generated based on certain scenarios, is shown. Process 700 can be implemented at electronic device 110. See below for reference. Figure 1 To describe process 700.
[0169] like Figure 7 As shown in box 710, electronic device 110 receives a first operation, which indicates the selection of a first area of the canvas.
[0170] In box 720, electronic device 110 presents an input component at an associated location in the first region, wherein the state of the input component is related to the content in the first region.
[0171] In frame 730, electronic device 110, in response to receiving a second operation via an input component, presents at least one piece of media content in a second area of the canvas.
[0172] In this way, an input component can be presented at the associated position of the selected first area in the canvas, and the state of the input component can be associated with the content of the selected first area. At least one piece of media content can also be generated through the input component, thereby improving the visual relevance between the input control and the selected first area, improving the efficiency of input operations, and thus effectively improving the efficiency of media content generation. The presentation state of the input component can also be enriched based on the content of the selected area. Furthermore, by presenting at least one piece of generated media content in the second area of the canvas, the content in the first area is retained, and at least one piece of media content and the content in the first area are presented simultaneously in the same canvas.
[0173] In some cases, process 700 further includes: in response to the first region including the first element, electronic device 110 presents first indication information in an input component, the first indication information corresponding to the first element.
[0174] In this way, the indicator information corresponding to the selected element can be displayed in the input component, thereby visually demonstrating the validity of the element being selected.
[0175] In some cases, the first instruction information at least describes the type of the first element.
[0176] In some cases, process 700 further includes: electronic device 110 receiving a third operation, the third operation instructing the selection of first instruction information; and, in association with the first instruction information, presenting preview content corresponding to the first element.
[0177] In this way, the corresponding preview content can be presented by selecting the instruction information, so as to ensure that the selected content can be previewed quickly and to ensure the effectiveness of the generated media content.
[0178] In some cases, process 700 further includes: electronic device 110 further includes a second element in response to the first region, and in the input component, presents second indication information corresponding to the second element.
[0179] In this way, corresponding instruction information can be presented in the input component for different elements in the first area, effectively ensuring the correspondence between the instruction information and the selected element, and visually demonstrating the validity of the element selection.
[0180] In some cases, process 700 further includes: electronic device 110 receiving a tagging operation; presenting a tagging element in a canvas, the tagging element corresponding to the tagging operation; and in response to the first area including the tagging element, presenting third indication information in the canvas, the third indication information indicating that the tagging element is selected.
[0181] In this way, marked elements can be presented on the canvas, and media content can be generated based on the selection of marked elements, thereby effectively ensuring the richness of the elements corresponding to the generated media content and enhancing the richness of the content generation process.
[0182] In some cases, at least one piece of media content is related to at least one of the following: the content of the tag element; the positional relationship between the tag element and the first element.
[0183] In some cases, at the associated location of the first region, the presentation of the input component includes: the electronic device 110 presenting a prompt message in response to the release of the first operation; and presenting the input component at the associated location in response to receiving a fourth operation associated with the prompt message.
[0184] In some cases, process 700 further includes: electronic device 110 responding to a first operation by presenting a bounding box in a canvas, the bounding box indicating the boundary of a first region, wherein prompts and input components are presented in a preset orientation of the bounding box.
[0185] In this way, the area range corresponding to the selected first region can be presented through the bounding box, and the presentation position of the prompts and input components can be determined based on the bounding box, thereby ensuring the visual relevance of the presented prompts and input components to the selected region.
[0186] In some cases, the input component includes an input area, and in response to receiving a second operation via the input component, presenting at least one piece of media content in the second area of the canvas includes: the electronic device 110 receiving input text via the input area; and in response to receiving a confirmation operation, presenting at least one piece of media content in the second area, the at least one piece of media content being generated based on the input text and the content in the first area.
[0187] In this way, media content can be generated based on the received input text and the content in the first area, effectively enhancing the richness of the interactive process associated with the canvas, enriching the materials on which the media content generation process depends, and ensuring the richness and effectiveness of the generated media content.
[0188] In some cases, the input component also includes an upload control, and process 700 further includes: electronic device 110 receiving at least one reference content via the upload control, wherein at least one media content is also generated based on at least one reference content.
[0189] In this way, at least one reference can be received to generate at least one piece of media content, further enhancing the richness of the materials on which the media content generation process depends and ensuring the effectiveness of the presented media content.
[0190] In some cases, the input component also includes a configuration control, and process 700 further includes: electronic device 110 configuring at least one generation parameter via the configuration control, wherein at least one media content is also generated based on at least one generation parameter.
[0191] In this way, at least one generation parameter can be configured for each content to generate at least one piece of media content, thereby ensuring the accuracy and effectiveness of the presented media content.
[0192] In some cases, process 700 further includes: electronic device 110 presenting first media content on a canvas; and in response to receiving a content acquisition request, presenting at least one reference media content associated with the first media content.
[0193] In this way, at least one related reference media content can be obtained based on the media content presented in the canvas, thereby increasing the richness of the ways to obtain media content through the canvas and improving the efficiency of obtaining reference media content.
[0194] In some cases, the first media content includes the first object, and the second object included in at least one reference media content is related to the first object.
[0195] In some cases, the input component is a first input component, and process 700 further includes: electronic device 110 presenting a session window in response to triggering an interactive entry in the canvas, the session window being used for dialogue interaction with a virtual object, the session window including a second input component; and in response to the first area including at least one element, presenting third instruction information in the second input component, the third instruction information corresponding to at least one element.
[0196] In this way, interactive sessions with virtual objects can be quickly triggered through interactive entry points on the canvas, thereby enriching the interactive methods associated with the canvas and enhancing the richness of the canvas-based interactive process.
[0197] In some cases, process 700 further includes: in response to receiving a second operation via a first input component, electronic device 110 presents a session message in a session window, the session message describing the second operation.
[0198] In this way, the operation information displayed on the canvas can be presented in the session window, thereby ensuring the traceability of operation information associated with the canvas.
[0199] Example devices and equipment A corresponding apparatus for implementing the above methods or processes is also provided. Figure 8 A schematic structural block diagram of an example device 800 for content generation under certain scenarios is shown. Device 800 may be implemented as or included in electronic device 110. The various modules / components in device 800 may be implemented by hardware, software, firmware, or any combination thereof.
[0200] like Figure 8 As shown, the device 800 includes: a region selection module 810, an input component presentation module 820, and a media content presentation module 830, wherein the region selection module 810 is configured to receive a first operation, the first operation indicating the selection of a first region of the canvas; the input component presentation module 820 is configured to present an input component at an associated location in the first region, wherein the state of the input component is related to the content in the first region; and the media content presentation module 830 is configured to present at least one piece of media content in a second region of the canvas in response to receiving a second operation via the input component.
[0201] In some cases, the device 800 further includes: a first information presentation module configured to, in response to a first area including a first element, present first indication information in an input component, the first indication information corresponding to the first element.
[0202] In some cases, the first instruction information at least describes the type of the first element.
[0203] In some cases, the device 800 also includes a preview content presentation module configured to: receive a third operation, the third operation instructing the selection of first instruction information; and, in association with the first instruction information, present preview content corresponding to a first element.
[0204] In some cases, the device 800 further includes a second information presentation module configured to, in response to the first area further including a second element, present second indication information in the input component, the second indication information corresponding to the second element.
[0205] In some cases, the device 800 also includes a module that performs the following processes: receiving a marking operation; presenting a marking element in a canvas, the marking element corresponding to the marking operation; and in response to the first area including the marking element, presenting third indication information in the canvas, the third indication information indicating that the marking element has been selected.
[0206] In some cases, at least one piece of media content is related to at least one of the following: the content of the tag element; the positional relationship between the tag element and the first element.
[0207] In some cases, the input component presentation module 820 is configured to: present a prompt message in response to a first release operation; and present the input component at the associated location in response to receiving a fourth operation associated with the prompt message.
[0208] In some cases, the device 800 further includes a bounding box rendering module configured to render a bounding box in a canvas in response to a first operation, the bounding box indicating the boundary of a first region, wherein the prompt and input components are presented in a preset orientation of the bounding box.
[0209] In some cases, the input component includes an input area, and the media content presentation module 830 is configured to: receive input text via the input area; and in response to receiving a confirmation operation, present at least one piece of media content in a second area, the at least one piece of media content being generated based on the input text and the content in the first area.
[0210] In some cases, the input component also includes an upload control, and the device 800 further includes a reference content receiving module configured to receive at least one piece of reference content via the upload control, wherein the at least one piece of media content is also generated based on the at least one piece of reference content.
[0211] In some cases, the input component also includes a configuration control, and the device 800 further includes a generation parameter configuration module configured to configure at least one generation parameter via the configuration control, wherein at least one media content is generated based on at least one generation parameter.
[0212] In some cases, the device 800 also includes a reference content presentation module configured to: present first media content in a canvas; and, in response to receiving a content acquisition request, present at least one reference media content associated with the first media content.
[0213] In some cases, the first media content includes the first object, and the second object included in at least one reference media content is related to the first object.
[0214] In some cases, the input component is a first input component, and the device 800 further includes a session window presentation module and a third information presentation module, wherein the session window presentation module is configured to: present a session window in response to triggering an interaction entry in the canvas, the session window being used for dialogue interaction with a virtual object, the session window including a second input component; and the third information presentation module is configured to: present third instruction information in the second input component in response to a first area including at least one element, the third instruction information corresponding to at least one element.
[0215] In some cases, the device 800 further includes a session message presentation module configured to, in response to receiving a second operation via a first input component, present a session message in a session window, the session message describing the second operation.
[0216] The modules included in device 800 can be implemented in various ways, including software, hardware, firmware, or any combination thereof. In some cases, one or more modules can be implemented using software and / or firmware, such as machine-executable instructions stored on a storage medium. In addition to or as an alternative to machine-executable instructions, some or all of the units in device 800 can be implemented at least partially by one or more hardware logic components. By way of example, and not limitation, exemplary types of hardware logic components that can be used include field-programmable gate arrays (FPGAs), application-specific integrated circuits (ASICs), application-specific standard parts (ASSPs), systems on a chip (SOCs), complex programmable logic devices (CPLDs), and so on.
[0217] Figure 9 A block diagram of an electronic device 900 in which one or more examples may be implemented is shown. It should be understood that... Figure 9 The electronic device 900 shown is merely exemplary and should not be construed as limiting the functionality and scope of the examples described herein. Figure 9 The illustrated electronic device 900 can be used to implement the electronic device 110 discussed above.
[0218] like Figure 9 As shown, electronic device 900 is in the form of a general-purpose electronic device. Components of electronic device 900 may include, but are not limited to, one or more processing units or processors 910, memory 920, storage devices 930, one or more communication units 940, one or more input devices 950, and one or more output devices 960. Processor 910 may be a physical or virtual processor and is capable of performing various processes according to programs stored in memory 920. In a multiprocessor system, multiple processors execute computer-executable instructions in parallel to improve the parallel processing capability of electronic device 900.
[0219] Electronic device 900 typically includes multiple computer storage media. Such media can be any accessible media that is accessible to electronic device 900, including but not limited to volatile and non-volatile media, removable and non-removable media. Memory 920 can be volatile memory (e.g., registers, cache, random access memory (RAM)), non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. Storage device 930 can be removable or non-removable media and can include machine-readable media, such as flash drives, disks, or any other media that can be used to store information and / or data and can be accessed within electronic device 900.
[0220] Electronic device 900 may further include additional removable / non-removable, volatile / non-volatile storage media. Although not explicitly stated... Figure 9 As shown, disk drives for reading from or writing to removable, non-volatile disks (e.g., "floppy disks") and optical disk drives for reading from or writing to removable, non-volatile optical disks can be provided. In these cases, each drive can be connected to a bus (not shown) via one or more data media interfaces. Memory 920 may include computer program product 925 having one or more program modules configured to perform various methods or actions of various examples.
[0221] The communication unit 940 enables communication with other electronic devices via a communication medium. Additionally, the functionality of the components of the electronic device 900 can be implemented using a single computing cluster or multiple computing machines capable of communicating via communication connections. Therefore, the electronic device 900 can operate in a networked environment using logical connections to one or more other servers, networked personal computers, or another network node.
[0222] Input device 950 can be one or more input devices, such as a mouse, keyboard, trackball, etc. Output device 960 can be one or more output devices, such as a monitor, speaker, printer, etc. Electronic device 900 can also communicate with one or more external devices (not shown) via communication unit 940 as needed. These external devices include storage devices, display devices, etc., and can communicate with one or more devices that enable user interaction with electronic device 900, or with any device that enables electronic device 900 to communicate with one or more other electronic devices (e.g., network card, modem, etc.). Such communication can be performed via input / output (I / O) interfaces (not shown).
[0223] A computer-readable storage medium is provided that stores computer-executable instructions thereon, wherein the computer-executable instructions are executed by a processor to implement the methods described above. A computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions, which are executed by a processor to implement the methods described above.
[0224] The flowcharts and / or block diagrams of the methods, apparatus, devices, and computer program products referred to herein describe various aspects. It should be understood that each block of the flowcharts and / or block diagrams, as well as combinations of blocks in the flowcharts and / or block diagrams, can be implemented by computer-readable program instructions.
[0225] These computer-readable program instructions can be provided to a processor of a general-purpose computer, a special-purpose computer, or other programmable data processing apparatus to produce a machine such that, when executed by the processor of the computer or other programmable data processing apparatus, they create means for implementing the functions / actions specified in one or more blocks of the flowchart and / or block diagram. These computer-readable program instructions can also be stored in a computer-readable storage medium that causes a computer, programmable data processing apparatus, and / or other device to operate in a particular manner; thus, the computer-readable medium storing the instructions comprises an article of manufacture that includes instructions for implementing aspects of the functions / actions specified in one or more blocks of the flowchart and / or block diagram.
[0226] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby causing the instructions that execute on the computer, other programmable data processing apparatus, or other device to perform the functions / actions specified in one or more boxes of a flowchart and / or block diagram.
[0227] The flowcharts and block diagrams in the accompanying figures illustrate the architecture, functionality, and operation of possible implementations of systems, methods, and computer program products under various scenarios. In this respect, each block in a flowchart or block diagram may represent a module, segment, or portion of an instruction, which contains one or more executable instructions for implementing the specified logical function. In some alternative implementations, the functions marked in the blocks may occur in a different order than those shown in the figures. For example, two consecutive blocks may actually be executed substantially in parallel, and they may sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, can be implemented using a dedicated hardware-based system that performs the specified function or action, or using a combination of dedicated hardware and computer instructions.
[0228] Various examples have been described above. The foregoing descriptions are exemplary and not exhaustive, nor are they limited to the disclosed implementations. Many modifications and variations will be apparent to those skilled in the art without departing from the scope and spirit of the described implementations. The terminology used herein is chosen to best explain the principles, practical applications, or improvements to technology in the market, or to enable others skilled in the art to understand the examples disclosed herein.
Claims
1. A content generation method, comprising: Receive a first operation, the first operation indicating the selection of a first area of the canvas; An input component is presented at an associated location in the first region, wherein the state of the input component is related to the content in the first region; as well as In response to receiving a second operation via the input component, at least one piece of media content is presented in a second area of the canvas.
2. The method according to claim 1, further comprising: In response to the first area including a first element, first indication information is presented in the input component, the first indication information corresponding to the first element.
3. The method of claim 2, wherein the first indication information at least describes the type of the first element.
4. The method according to claim 2, further comprising: Receive a third operation, wherein the third operation indicates that the first indication information is selected; as well as Associated with the first indication information, a preview content is presented, the preview content corresponding to the first element.
5. The method according to claim 2, further comprising: In response to the first area, a second element is also included, in which second indication information is presented in the input component, the second indication information corresponding to the second element.
6. The method according to claim 2, further comprising: Receive tag operation; In the canvas, marker elements are presented, and the marker elements correspond to the marker operations; as well as In response to the first area including the marker element, a third indication is presented in the canvas, the third indication indicating that the marker element is selected.
7. The method of claim 6, wherein the at least one media content is related to at least one of the following: The content of the marker element; The positional relationship between the marker element and the first element.
8. The method of claim 1, wherein presenting the input component at the associated location in the first region comprises: In response to the release of the first operation, a prompt message is displayed; as well as In response to receiving a fourth operation associated with the prompt information, the input component is presented at the associated location.
9. The method according to claim 8, further comprising: In response to the first operation, a bounding box is presented in the canvas, the bounding box indicating the boundary of the first region, wherein the prompt information and the input component are presented in a preset direction of the bounding box.
10. The method of claim 1, wherein the input component includes an input area, and the presentation of at least one piece of media content in a second area of the canvas in response to receiving a second operation via the input component comprises: Input text is received via the input area; as well as In response to receiving a confirmation operation, the at least one piece of media content is presented in the second area, the at least one piece of media content being generated based on the input text and the content in the first area.
11. The method of claim 10, wherein the input component further includes an upload control, and the method further includes: The upload control receives at least one piece of reference content, wherein the at least one piece of media content is also generated based on the at least one piece of reference content.
12. The method of claim 10, wherein the input component further includes a configuration control, and the method further includes: At least one generation parameter is configured via the configuration control, wherein the at least one media content is also generated based on the at least one generation parameter.
13. The method of claim 1, further comprising: The first media content is presented on the canvas; as well as In response to receiving a content retrieval request, at least one reference media content is presented, the at least one reference media content being associated with the first media content.
14. The method of claim 13, wherein the first media content includes a first object, and the second object included in the at least one reference media content is related to the first object.
15. The method of claim 1, wherein the input component is a first input component, and the method further comprises: In response to the triggering of an interactive entry point in the canvas, a session window is presented, the session window being used for dialogue interaction with virtual objects, the session window including a second input component; as well as In response to the first area including at least one element, third indication information is presented in the second input component, the third indication information corresponding to the at least one element.
16. The method of claim 15, further comprising: In response to receiving the second operation via the first input component, a session message describing the second operation is presented in the session window.
17. An apparatus for content generation, comprising: The region selection module is configured to receive a first operation, the first operation indicating the selection of a first region of the canvas; An input component presentation module is configured to present an input component at an associated location in the first region, wherein the state of the input component is related to the content in the first region; as well as The media content presentation module is configured to present at least one piece of media content in a second area of the canvas in response to receiving a second operation via the input component.
18. An electronic device comprising: At least one processor; as well as At least one memory coupled to the at least one processor and storing instructions for execution by the at least one processor, the instructions causing the electronic device to perform the method according to any one of claims 1 to 16 when executed by the at least one processor.
19. A computer-readable storage medium having stored thereon computer-executable instructions that can be executed by a processor to implement the method according to any one of claims 1 to 16.
20. A computer program product tangibly stored in a computer storage medium and comprising computer-executable instructions that, when executed by a device, cause the device to perform the method according to any one of claims 1 to 16.