User interfaces for generating automatically-generated content
The electronic device generates and integrates automatically-generated visual content using AI within a user interface, reducing user inputs and errors while optimizing computing resources.
Patent Information
- Application Number
- PCT/US2025/021893
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2024-10-20
- Filing Date
- 2025-03-27
- Publication Date
- 2025-10-16
AI Technical Summary
Existing electronic devices lack efficient methods for generating and integrating automatically-generated visual content without requiring additional user inputs or opening secondary applications, leading to errors and increased computing resources.
An electronic device generates and integrates automatically-generated visual content within a user interface of a first application using AI processes, allowing users to modify recognized concepts and display visual effects, such as HDR content, without navigating away from the first application.
This approach reduces user inputs, minimizes errors, and optimizes computing resources by enabling efficient generation and integration of visual content directly within the user interface, enhancing device operability and user interaction.
Smart Images

Figure US2025021893_16102025_PF_FP_ABST
Abstract
Description
Attorney Docket No.106842222440 (P65924WO1) USER INTERFACES FOR GENERATING AUTOMATICALLY- GENERATED CONTENT Cross-Reference to Related Applications
[0001] This application claims the benefit of U.S. Provisional Application No. 63 / 631,445, filed April 8, 2024, U.S. Provisional Application No.63 / 657,971, filed June 9, 2024, U.S. Provisional Application No.63 / 658,421, filed June 10, 2024, and U.S. Provisional Application No.63 / 709,506, filed October 20, 2024, the contents of which are herein incorporated by reference in their entireties for all purposes. Field of the Disclosure
[0002] This disclosure relates generally to an electronic device presenting user interfaces for generating automatically-generated content. Background of the Disclosure
[0003] User interaction with electronic devices has increased significantly in recent years. These devices can be devices such as computers, tablet computers, televisions, multimedia devices, or mobile devices. In some circumstances, users may wish to create generative images using reference images. The user may therefore desire efficient ways of generating automatically-generated content and displaying the automatically-generated content. Summary of the Disclosure
[0004] Providing efficient ways of displaying representations of recognized concepts of a prompt used to generate automatically-generated visual content allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors. Allowing a user to edit previously generated media allows the user to efficiently change recognized concepts of a previously generated media item, thereby reducing computing resources used by the electronic device, and also reduces erroneous inputs to the electronic device.
[0005] In some embodiments, an electronic device receives a prompt for generating an automatically-generated visual content. In some embodiments, the electronic device extracts recognized concepts from the prompt to be used to influence the generation of the -1- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) automatically-generated visual content. In some embodiments, while displaying the user interface including the recognized concepts, the electronic device receives one or more inputs to modify the recognized concepts. In some embodiments, the electronic device generates (e.g., using an AI process or a generative AI process) an automatically-generated visual content using the one or more recognized concepts.
[0006] In some embodiments, the electronic device adds an automatically-generated visual content to a content entry field of a first application, different than the automatically- generated visual media application, without opening the automatically-generated visual media application. In some embodiments, the electronic device displays a user interface within the first application including one or more previously generated automatically- generated visual content to be added to the content entry field. In some embodiments, the electronic device edits previously generated automatically-generated visual content and creates new automatically-generated visual content using reference media items while in the first application. In some embodiments, the electronic device detects an event corresponding to a respective functionality that outputs content generated based on an artificial intelligence (AI) model and displays the content with a visual effect that includes a visual characteristic associated with content generated based on an AI model. In some embodiments, the electronic device displays an animation indicative of an AI model content generation information, wherein the animation includes High Dynamic Range (HDR) content.
[0007] Displaying a user interface for inserting the automatically-generated visual content into the user interface of the first application allows a user to efficiently insert such visual media into the first application without opening a second application and / or navigating away from the first user interface, thereby reducing inputs needed to insert such visual media. Displaying a user interface for generating the automatically-generated visual content based on a reference media item in the user interface of the first application allows a user to generate an automatically-generated visual content for use with the first application without opening a second application and / or navigating away from the first application, thereby reducing inputs needed to generate such visual media. Displaying a user interface with a visual characteristic and / or animation to indicate that the content is based on AI generated content allows a user to visually assess when content was generated using an AI model.
[0008] Displaying a representation of second automatically-generated generative visual content based on a user selected prompt used to automatically generate first automatically-generated generative visual content in response to an input including a -2- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) movement component provides an intuitive way to display representations of different automatically-generated generative visual content efficient interactions between the user and the electronic device.
[0009] Restricting automatically-generated visual content generated in response to user inputs enhances the operability of the device and makes the user-device interface more efficient.
[0010] Suggesting prompt components based on previously presented content and / or device context allows the user to easily and efficiently select relevant prompt components to be used to influence the generation of an automatically-generated visual content, thereby reducing the number of inputs needed to generate automatically-generated visual media and reducing erroneous inputs to the electronic device.
[0011] Presenting options for personalizing the subject of automatically-generated visual media, including editing subjects and / or creating subjects provides efficient ways of customizing the content of automatically-generated visual media while reducing inputs and user errors.
[0012] Providing options to personalize template subjects allows a user to constrain certain appearance characteristics while allowing variability for other appearance characteristics therefore allowing for a broader range of representations of different types of subjects for the automatically-generated visual content.
[0013] The full descriptions of the embodiments are provided in the Drawings and the Detailed Description, and it is understood that the Summary provided above does not limit the scope of the disclosure in any way.
[0014] It is well understood that the use of personally identifiable information should follow privacy policies and practices that are generally recognized as meeting or exceeding industry or governmental requirements for maintaining the privacy of users. In particular, personally identifiable information data should be managed and handled so as to minimize risks of unintentional or unauthorized access or use, and the nature of authorized use should be clearly indicated to users. Brief Description of the Drawings
[0015] For a better understanding of the various described embodiments, reference should be made to the Detailed Description below, in conjunction with the following -3- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) drawings in which like reference numerals refer to corresponding parts throughout the figures.
[0016] Fig.1A is a block diagram illustrating a portable multifunction device with a touch-sensitive display in accordance with some embodiments.
[0017] Fig.1B is a block diagram illustrating exemplary components for event handling in accordance with some embodiments.
[0018] Fig.2 illustrates a portable multifunction device having a touch screen in accordance with some embodiments.
[0019] Figs.3A-3G is a block diagram of an exemplary multifunction device with a display and a touch-sensitive surface in accordance with some embodiments.
[0020] Fig.4A illustrates an exemplary user interface for a menu of applications on a portable multifunction device in accordance with some embodiments.
[0021] Fig.4B illustrates an exemplary user interface for a multifunction device with a touch-sensitive surface that is separate from the display in accordance with some embodiments.
[0022] Fig.5A illustrates a personal electronic device in accordance with some embodiments.
[0023] Fig.5B is a block diagram illustrating a personal electronic device in accordance with some embodiments.
[0024] Figs.5C-5D illustrate exemplary components of a personal electronic device having a touch-sensitive display and intensity sensors in accordance with some embodiments.
[0025] Figs.5E-5H illustrate exemplary components and user interfaces of a personal electronic device in accordance with some embodiments.
[0026] Figs.6A-6MM illustrate exemplary ways in which an electronic device displays recognized concepts and generates automatically-generated visual media in accordance with some embodiments of the disclosure.
[0027] Fig.7 illustrates a flow diagram illustrating a method in which an electronic device displays recognized concepts and generates automatically-generated visual media in accordance with some embodiments of the disclosure. -4- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0028] Fig.8 illustrates a flow diagram illustrating a method in which an electronic device edits one or more parameters associated with a previously generated automatically- generated visual content and regenerates a new automatically-generated visual content in accordance with some embodiments of the disclosure.
[0029] Figs.9A-9X illustrate exemplary ways in which an electronic device adds automatically-generated visual content to a first user interface of a first application according to some embodiments of the disclosure.
[0030] Fig.10 illustrates a flow diagram illustrating a method in which an electronic device adds automatically-generated visual content to a first user interface of a first application according to some embodiments of the disclosure.
[0031] Figs.11A-11J illustrate exemplary ways in which an electronic device selects a reference media item to be used to generate an automatically-generated visual content while displaying a user interface of a first application.
[0032] Fig.12 illustrates a flow diagram illustrating a method in which an electronic device selects a reference media item to be used to generate an automatically-generated visual content while displaying a user interface of a first application.
[0033] Figs.13A – 13HH illustrate exemplary visual effects applied to a user interfaces that include content related to an artificial intelligence model.
[0034] Fig.14 illustrates a flow diagram illustrating a method in which an electronic device displays visual information based upon an event that corresponds to an artificial intelligence model in accordance with some embodiments of the disclosure.
[0035] Fig.15 illustrates a flow diagram illustrating a method in which an electronic device displays an animation including high dynamic range luminance in accordance with some embodiments of the disclosure.
[0036] Figs.16A-16U illustrate exemplary ways in which an electronic device displays multiple variations of representations of automatically-generated visual content items in accordance with some embodiments of the disclosure.
[0037] Fig.17 illustrates a flow diagram illustrating a method in which an electronic device displays multiple variations of representations of automatically-generated visual content items in accordance with some embodiments of the disclosure. -5- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0038] Figs.18A-18BB illustrate exemplary ways in which an electronic device restricts automatically-generated visual content in accordance with some embodiments of the disclosure.
[0039] Fig.19 illustrates a flow diagram illustrating a method in which an electronic device restricts automatically-generated visual content in accordance with some embodiments of the disclosure.
[0040] Figs.20A-20AA illustrate exemplary ways in which an electronic device displays prompt component suggestions and generates an automatically-generated visual content in accordance with some embodiments of the disclosure.
[0041] Fig.21 illustrates a flow diagram illustrating a method in which an electronic device displays prompt component suggestions and generates an automatically-generated visual content in accordance with some embodiments of the disclosure.
[0042] Figs.22A through 22CCC illustrate exemplary ways in which an electronic device presents user interfaces for editing a subject of an automatically-generated visual content item in accordance with some embodiments of the disclosure.
[0043] Fig.23 illustrates a flow diagram illustrating a method in which an electronic device presents user interfaces for editing a subject of an automatically-generated visual content item in accordance with some embodiments of the disclosure.
[0044] Figs.24A-24E illustrate exemplary ways in which an electronic device detects an input to display automatically-generated visual content in a user interface of an application in accordance with some embodiments of the disclosure.
[0045] Figs.25A-25O illustrate exemplary ways in which an electronic device detects an input to display automatically-generated visual content in a user interface of an application in accordance with some embodiments of the disclosure and example visual effects and / or operations corresponding to visual effects and / or operations described in this disclosure.
[0046] Figs.26A-26P illustrate exemplary ways in which an electronic device presents user interfaces for selecting a template subject for generating an automatically- generated visual content item in accordance with some embodiments of the disclosure.
[0047] Fig.27 illustrates a flow diagram illustrating a method in which an electronic device presents user interfaces for selecting a template subject for generating an -6- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) automatically-generated visual content item in accordance with some embodiments of the disclosure. Detailed Description
[0048] In the following description of embodiments, reference is made to the accompanying drawings which form a part hereof, and in which it is shown by way of illustration specific embodiments that are optionally practiced. It is to be understood that other embodiments are optionally used, and structural changes are optionally made without departing from the scope of the disclosed embodiments.
[0049] Providing efficient ways of displaying representations of recognized concepts of a prompt used to generate an automatically-generated visual content allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors. Allowing a user to edit previously generated media allows the user to efficiently change recognized concepts of a previously generated media item, thereby reducing computing resources used by the electronic device, and also reduces erroneous inputs to the electronic device.
[0050] In some embodiments, an electronic device receives a prompt for generating an automatically-generated visual content. In some embodiments, the electronic device extracts recognized concepts from the prompt to be used to influence the generation of the automatically-generated visual content. In some embodiments, while displaying the user interface including the recognized concepts, the electronic device receives one or more inputs to modify the recognized concepts. In some embodiments, the electronic device generates (e.g., using an AI process or a generative AI process) an automatically-generated visual content using the one or more recognized concepts.
[0051] In some embodiments, the electronic device adds an automatically-generated visual content to a content entry field of a first application, different than the automatically- generated visual media application, without opening the automatically-generated visual media application. In some embodiments, the electronic device displays a user interface within the first application including one or more previously generated automatically- generated visual content to be added to the content entry field. In some embodiments, the electronic device edits previously generated automatically-generated visual content and -7- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) creates new automatically-generated visual content using reference media items while in the first application.
[0052] Displaying a user interface for inserting the automatically-generated visual content into the user interface of the first application allows a user to efficiently insert such visual media into the first application without opening a second application and / or navigating away from the first user interface, thereby reducing inputs needed to insert such visual media. Displaying a user interface for generating the automatically-generated visual content based on a reference media item in the user interface of the first application allows a user to generate an automatically-generated visual content for use with the first application without opening a second application and / or navigating away from the first application, thereby reducing inputs needed to generate such visual media.
[0053] Displaying a representation of second automatically-generated generative visual content based on a user selected prompt used to automatically generate first automatically-generated generative visual content in response to an input including a movement component provides an intuitive way to display representations of different automatically-generated generative visual content efficient interactions between the user and the electronic device.
[0054] Displaying an animation in which portions of a user interface are displayed with a degree of luminance that is greater than a standard dynamic range of luminance visually emphasizes that operations associated with of an electronic device are ongoing, and reduces input erroneously interrupting such operations.
[0055] Restricting automatically-generated visual content generated in response to user inputs enhances the operability of the device and makes the user-device interface more efficient.
[0056] Suggesting prompt components based on previously presented content and / or device context allows the user to easily and efficiently select relevant prompt components to be used to influence the generation of an automatically-generated visual content, thereby reducing the number of inputs needed to generate automatically-generated visual media and reducing erroneous inputs to the electronic device.
[0057] Presenting options for personalizing the subject of automatically-generated visual media, including editing subjects and / or creating subjects provides efficient ways of -8- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) customizing the content of automatically-generated visual media while reducing inputs and user errors.
[0058] Providing options to personalize template subjects allows a user to constrain certain appearance characteristics while allowing variability for other appearance characteristics therefore allowing for a broader range of representations of different types of subjects for the automatically-generated visual content.
[0059] Although the following description uses terms "first," "second," etc. to describe various elements, these elements should not be limited by the terms. These terms are only used to distinguish one element from another. For example, a first touch could be termed a second touch, and, similarly, a second touch could be termed a first touch, without departing from the scope of the various described embodiments. The first touch and the second touch are both touches, but they are not the same touch.
[0060] The terminology used in the description of the various described embodiments herein is for the purpose of describing particular embodiments only and is not intended to be limiting. As used in the description of the various described embodiments and the appended claims, the singular forms "a," "an," and "the" are intended to include the plural forms as well, unless the context clearly indicates otherwise. It will also be understood that the term "and / or" as used herein refers to and encompasses any and all possible combinations of one or more of the associated listed items. It will be further understood that the terms "includes," "including," "comprises," and / or "comprising," when used in this specification, specify the presence of stated features, integers, steps, operations, elements, and / or components, but do not preclude the presence or addition of one or more other features, integers, steps, operations, elements, components, and / or groups thereof.
[0061] The term "if" is, optionally, construed to mean "when" or "upon" or "in response to determining" or "in response to detecting," depending on the context. Similarly, the phrase "if it is determined" or "if [a stated condition or event] is detected" is, optionally, construed to mean "upon determining" or "in response to determining" or "upon detecting [the stated condition or event]" or "in response to detecting [the stated condition or event]," depending on the context. EXEMPLARY DEVICES
[0062] Embodiments of electronic devices, user interfaces for such devices, and associated processes for using such devices are described. In some embodiments, the device -9- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) is a portable communications device, such as a mobile telephone, that also contains other functions, such as PDA and / or music player functions. Exemplary embodiments of portable multifunction devices include, without limitation, the iPhone®, iPod Touch®, and iPad® devices from Apple Inc. of Cupertino, California. Other portable electronic devices, such as laptops or tablet computers with touch-sensitive surfaces (e.g., touch screen displays and / or touch pads), are, optionally, used. It should also be understood that, in some embodiments, the device is not a portable communications device, but is a desktop computer or a television with a touch-sensitive surface (e.g., a touch screen display and / or a touch pad). In some embodiments, the device does not have a touch screen display and / or a touch pad, but rather is capable of outputting display information (such as the user interfaces of the disclosure) for display on a separate display device, and capable of receiving input information from a separate input device having one or more input mechanisms (such as one or more buttons, a touch screen display and / or a touch pad). In some embodiments, the device has a display, but is capable of receiving input information from a separate input device having one or more input mechanisms (such as one or more buttons, a touch screen display and / or a touch pad). In some embodiments, the electronic device is a computer system that is in communication (e.g., via wireless communication, via wired communication) with a display generation component (e.g., a display device such as a head-mounted display (HMD), a display, a projector, a touch-sensitive display, or other device or component that presents visual content to a user, for example, on or in the display generation component itself or produced from the display generation component and visible elsewhere). The display generation component is configured to provide visual output, such as display via a CRT display, display via an LED display, or display via image projection. In some embodiments, the display generation component is integrated with the computer system. In some embodiments, the display generation component is separate from the computer system. As used herein, “displaying” content includes causing to display the content (e.g., video data rendered or decoded by display controller 156) by transmitting, via a wired or wireless connection, data (e.g., image data or video data) to an integrated or external display generation component to visually produce the content.
[0063] In the discussion that follows, an electronic device that includes a display and a touch-sensitive surface is described. It should be understood, however, that the electronic device optionally includes one or more other physical user-interface devices, such as a physical keyboard, a mouse and / or a joystick. Further, as described above, it should be -10- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) understood that the described electronic device, display and touch-sensitive surface are optionally distributed amongst two or more devices. Therefore, as used in this disclosure, information displayed on the electronic device or by the electronic device is optionally used to describe information outputted by the electronic device for display on a separate display device (touch-sensitive or not). Similarly, as used in this disclosure, input received on the electronic device (e.g., touch input received on a touch-sensitive surface of the electronic device) is optionally used to describe input received on a separate input device, from which the electronic device receives input information.
[0064] The device typically supports a variety of applications, such as one or more of the following: a drawing application, a presentation application, a word processing application, a website creation application, a disk authoring application, a spreadsheet application, a gaming application, a telephone application, a video conferencing application, an e-mail application, an instant messaging application, a workout support application, a photo management application, a digital camera application, a digital video camera application, a web browsing application, a digital music player application, a television channel browsing application, and / or a digital video player application.
[0065] The various applications that are executed on the device optionally use at least one common physical user-interface device, such as the touch-sensitive surface. One or more functions of the touch-sensitive surface as well as corresponding information displayed on the device are, optionally, adjusted and / or varied from one application to the next and / or within a respective application. In this way, a common physical architecture (such as the touch- sensitive surface) of the device optionally supports the variety of applications with user interfaces that are intuitive and transparent to the user.
[0066] Attention is now directed toward embodiments of portable or non-portable devices with touch-sensitive displays, though the devices need not include touch-sensitive displays or displays in general, as described above. Fig.1A is a block diagram illustrating portable or non-portable multifunction device 100 with touch-sensitive displays 112 in accordance with some embodiments. Touch-sensitive display 112 is sometimes called a "touch screen" for convenience, and is sometimes known as or called a touch-sensitive display system. Device 100 includes memory 102 (which optionally includes one or more computer readable storage mediums), memory controller 122, one or more processing units (CPU's) 120, peripherals interface 118, RF circuitry 108, audio circuitry 110, speaker 111, microphone 113, input / output (I / O) subsystem 106, other input or control devices 116, and -11- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) external port 124. Device 100 optionally includes one or more optical sensors 164. Device 100 optionally includes one or more contact intensity sensors 165 for detecting intensity of contacts on device 100 (e.g., a touch-sensitive surface such as touch-sensitive display system 112 of device 100). Device 100 optionally includes one or more tactile output generators 167 for generating tactile outputs on device 100 (e.g., generating tactile outputs on a touch- sensitive surface such as touch-sensitive display system 112 of device 100 or touchpad 355 of device 300). These components optionally communicate over one or more communication buses or signal lines 103.
[0067] As used in the specification and claims, the term "intensity" of a contact on a touch-sensitive surface refers to the force or pressure (force per unit area) of a contact (e.g., a finger contact) on the touch-sensitive surface, or to a substitute (proxy) for the force or pressure of a contact on the touch-sensitive surface. The intensity of a contact has a range of values that includes at least four distinct values and more typically includes hundreds of distinct values (e.g., at least 256). Intensity of a contact is, optionally, determined (or measured) using various approaches and various sensors or combinations of sensors. For example, one or more force sensors underneath or adjacent to the touch-sensitive surface are, optionally, used to measure force at various points on the touch-sensitive surface. In some implementations, force measurements from multiple force sensors are combined (e.g., a weighted average) to determine an estimated force of a contact. Similarly, a pressure- sensitive tip of a stylus is, optionally, used to determine a pressure of the stylus on the touch- sensitive surface. Alternatively, the size of the contact area detected on the touch-sensitive surface and / or changes thereto, the capacitance of the touch-sensitive surface proximate to the contact and / or changes thereto, and / or the resistance of the touch-sensitive surface proximate to the contact and / or changes thereto are, optionally, used as a substitute for the force or pressure of the contact on the touch-sensitive surface. In some implementations, the substitute measurements for contact force or pressure are used directly to determine whether an intensity threshold has been exceeded (e.g., the intensity threshold is described in units corresponding to the substitute measurements). In some implementations, the substitute measurements for contact force or pressure are converted to an estimated force or pressure and the estimated force or pressure is used to determine whether an intensity threshold has been exceeded (e.g., the intensity threshold is a pressure threshold measured in units of pressure). Using the intensity of a contact as an attribute of a user input allows for user access to additional device functionality that may otherwise not be accessible by the user on a -12- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) reduced-size device with limited real estate for displaying affordances (e.g., on a touch- sensitive display) and / or receiving user input (e.g., via a touch-sensitive display, a touch- sensitive surface, or a physical / mechanical control such as a knob or a button).
[0068] As used in the specification and claims, the term "tactile output" refers to physical displacement of a device relative to a previous position of the device, physical displacement of a component (e.g., a touch-sensitive surface) of a device relative to another component (e.g., housing) of the device, or displacement of the component relative to a center of mass of the device that will be detected by a user with the user's sense of touch. For example, in situations where the device or the component of the device is in contact with a surface of a user that is sensitive to touch (e.g., a finger, palm, or other part of a user's hand), the tactile output generated by the physical displacement will be interpreted by the user as a tactile sensation corresponding to a perceived change in physical characteristics of the device or the component of the device. For example, movement of a touch-sensitive surface (e.g., a touch-sensitive display or trackpad) is, optionally, interpreted by the user as a "down click" or "up click" of a physical actuator button. In some cases, a user will feel a tactile sensation such as a "down click" or "up click" even when there is no movement of a physical actuator button associated with the touch-sensitive surface that is physically pressed (e.g., displaced) by the user's movements. As another example, movement of the touch-sensitive surface is, optionally, interpreted or sensed by the user as "roughness" of the touch-sensitive surface, even when there is no change in smoothness of the touch-sensitive surface. While such interpretations of touch by a user will be subject to the individualized sensory perceptions of the user, there are many sensory perceptions of touch that are common to a large majority of users. Thus, when a tactile output is described as corresponding to a particular sensory perception of a user (e.g., an "up click," a "down click," "roughness"), unless otherwise stated, the generated tactile output corresponds to physical displacement of the device or a component thereof that will generate the described sensory perception for a typical (or average) user.
[0069] It should be appreciated that device 100 is only one example of a portable or non-portable multifunction device, and that device 100 optionally has more or fewer components than shown, optionally combines two or more components, or optionally has a different configuration or arrangement of the components. The various components shown in Fig.1A are implemented in hardware, software, or a combination of both hardware and software, including one or more signal processing and / or application specific integrated -13- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) circuits. Further, the various components shown in Fig.1A are optionally implemented across two or more devices; for example, a display and audio circuitry on a display device, a touch-sensitive surface on an input device, and remaining components on device 100. In such an embodiment, device 100 optionally communicates with the display device and / or the input device to facilitate operation of the system, as described in the disclosure, and the various components described herein that relate to display and / or input remain in device 100, or are optionally included in the display and / or input device, as appropriate.
[0070] Memory 102 optionally includes high-speed random access memory and optionally also includes non-volatile memory, such as one or more magnetic disk storage devices, flash memory devices, or other non-volatile solid-state memory devices. Memory controller 122 optionally controls access to memory 102 by other components of device 100.
[0071] Peripherals interface 118 can be used to couple input and output peripherals of the device to CPU 120 and memory 102. The one or more processors 120 run or execute various software programs and / or sets of instructions stored in memory 102 to perform various functions for device 100 and to process data.
[0072] In some embodiments, peripherals interface 118, CPU 120, and memory controller 122 are, optionally, implemented on a single chip, such as chip 104. In some other embodiments, they are, optionally, implemented on separate chips.
[0073] RF (radio frequency) circuitry 108 receives and sends RF signals, also called electromagnetic signals. RF circuitry 108 converts electrical signals to / from electromagnetic signals and communicates with communications networks and other communications devices via the electromagnetic signals. RF circuitry 108 optionally includes well-known circuitry for performing these functions, including but not limited to an antenna system, an RF transceiver, one or more amplifiers, a tuner, one or more oscillators, a digital signal processor, a CODEC chipset, a subscriber identity module (SIM) card, memory, and so forth. RF circuitry 108 optionally communicates with networks, such as the Internet, also referred to as the World Wide Web (WWW), an intranet and / or a wireless network, such as a cellular telephone network, a wireless local area network (LAN) and / or a metropolitan area network (MAN), and other devices by wireless communication. The RF circuitry 108 optionally includes well-known circuitry for detecting near field communication (NFC) fields, such as by a short-range communication radio. The wireless communication optionally uses any of a plurality of communications standards, protocols, and technologies, including but not limited -14- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) to Global System for Mobile Communications (GSM), Enhanced Data GSM Environment (EDGE), high-speed downlink packet access (HSDPA), high-speed uplink packet access (HSUPA), Evolution, Data-Only (EV-DO), HSPA, HSPA+, Dual-Cell HSPA (DC-HSPDA), long term evolution (LTE), near field communication (NFC), wideband code division multiple access (W-CDMA), code division multiple access (CDMA), time division multiple access (TDMA), Bluetooth, Bluetooth Low Energy (BTLE), Wireless Fidelity (Wi-Fi) (e.g., IEEE 802.11a, IEEE 802.11b, IEEE 802.11g, IEEE 802.11n, and / or IEEE 802.11ac), voice over Internet Protocol (VoIP), Wi-MAX, a protocol for e-mail (e.g., Internet message access protocol (IMAP) and / or post office protocol (POP)), instant messaging (e.g., extensible messaging and presence protocol (XMPP), Session Initiation Protocol for Instant Messaging and Presence Leveraging Extensions (SIMPLE), Instant Messaging and Presence Service (IMPS)), and / or Short Message Service (SMS), or any other suitable communication protocol, including communication protocols not yet developed as of the filing date of this document.
[0074] Audio circuitry 110, speaker 111, and microphone 113 provide an audio interface between a user and device 100. Audio circuitry 110 receives audio data from peripherals interface 118, converts the audio data to an electrical signal, and transmits the electrical signal to speaker 111. Speaker 111 converts the electrical signal to human-audible sound waves. Audio circuitry 110 also receives electrical signals converted by microphone 113 from sound waves. Audio circuitry 110 converts the electrical signal to audio data and transmits the audio data to peripherals interface 118 for processing. Audio data is, optionally, retrieved from and / or transmitted to memory 102 and / or RF circuitry 108 by peripherals interface 118. In some embodiments, audio circuitry 110 also includes a headset jack (e.g., 212, Fig.2). The headset jack provides an interface between audio circuitry 110 and removable audio input / output peripherals, such as output-only headphones or a headset with both output (e.g., a headphone for one or both ears) and input (e.g., a microphone).
[0075] I / O subsystem 106 couples input / output peripherals on device 100, such as touch screen 112 and other input control devices 116, to peripherals interface 118. I / O subsystem 106 optionally includes display controller 156, optical sensor controller 158, intensity sensor controller 159, haptic feedback controller 161 and one or more input controllers 160 for other input or control devices. The one or more input controllers 160 receive / send electrical signals from / to other input or control devices 116. The other input control devices 116 optionally include physical buttons (e.g., push buttons, rocker buttons, -15- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) etc.), dials, slider switches, joysticks, click wheels, and so forth. In some alternate embodiments, input controller(s) 160 are, optionally, coupled to any (or none) of the following: a keyboard, infrared port, USB port, and a pointer device such as a mouse. The one or more buttons (e.g., 208, Fig.2) optionally include an up / down button for volume control of speaker 111 and / or microphone 113. The one or more buttons optionally include a push button (e.g., 206, Fig.2).
[0076] A quick press of the push button optionally disengages a lock of touch screen 112 or optionally begins a process that uses gestures on the touch screen to unlock the device, as described in U.S. Patent Application 11 / 322,549, "Unlocking a Device by Performing Gestures on an Unlock Image," filed December 23, 2005, U.S. Pat. No.7,657,849, which is hereby incorporated by reference in its entirety. A longer press of the push button (e.g., 206) optionally turns power to device 100 on or off. The functionality of one or more of the buttons are, optionally, user-customizable. Touch screen 112 is used to implement virtual or soft buttons and one or more soft keyboards.
[0077] Touch-sensitive display 112 provides an input interface and an output interface between the device and a user. As described above, the touch-sensitive operation and the display operation of touch-sensitive display 112 are optionally separated from each other, such that a display device is used for display purposes and a touch-sensitive surface (whether display or not) is used for input detection purposes, and the described components and functions are modified accordingly. However, for simplicity, the following description is provided with reference to a touch-sensitive display. Display controller 156 receives and / or sends electrical signals from / to touch screen 112. Touch screen 112 displays visual output to the user. The visual output optionally includes graphics, text, icons, video, and any combination thereof (collectively termed "graphics"). In some embodiments, some or all of the visual output corresponds to user-interface objects.
[0078] Touch screen 112 has a touch-sensitive surface, sensor or set of sensors that accepts input from the user based on haptic and / or tactile contact. Touch screen 112 and display controller 156 (along with any associated modules and / or sets of instructions in memory 102) detect contact (and any movement or breaking of the contact) on touch screen 112 and convert the detected contact into interaction with user-interface objects (e.g., one or more soft keys, icons, web pages or images) that are displayed on touch screen 112. In an exemplary embodiment, a point of contact between touch screen 112 and the user corresponds to a finger of the user. -16- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0079] Touch screen 112 optionally uses LCD (liquid crystal display) technology, LPD (light emitting polymer display) technology, or LED (light emitting diode) technology, although other display technologies are used in other embodiments. Touch screen 112 and display controller 156 optionally detect contact and any movement or breaking thereof using any of a plurality of touch sensing technologies now known or later developed, including but not limited to capacitive, resistive, infrared, and surface acoustic wave technologies, as well as other proximity sensor arrays or other elements for determining one or more points of contact with touch screen 112. In an exemplary embodiment, projected mutual capacitance sensing technology is used, such as that found in the iPhone®, iPod Touch®, and iPad® from Apple Inc. of Cupertino, California.
[0080] A touch-sensitive display in some embodiments of touch screen 112 is, optionally, analogous to the multi-touch sensitive touchpads described in the following U.S. Patents: 6,323,846 (Westerman et al.), 6,570,557 (Westerman et al.), and / or 6,677,932 (Westerman), and / or U.S. Patent Publication 2002 / 0015024A1, each of which is hereby incorporated by reference in its entirety. However, touch screen 112 displays visual output from device 100, whereas touch-sensitive touchpads do not provide visual output.
[0081] A touch-sensitive display in some embodiments of touch screen 112 is described in the following applications: (1) U.S. Patent Application No.11 / 381,313, "Multipoint Touch Surface Controller," filed May 2, 2006; (2) U.S. Patent Application No. 10 / 840,862, "Multipoint Touchscreen," filed May 6, 2004; (3) U.S. Patent Application No. 10 / 903,964, "Gestures For Touch Sensitive Input Devices," filed July 30, 2004; (4) U.S. Patent Application No.11 / 48,264, "Gestures For Touch Sensitive Input Devices," filed January 31, 2005; (5) U.S. Patent Application No.11 / 38,590, "Mode-Based Graphical User Interfaces For Touch Sensitive Input Devices," filed January 18, 2005; (6) U.S. Patent Application No.11 / 228,758, "Virtual Input Device Placement On A Touch Screen User Interface," filed September 16, 2005; (7) U.S. Patent Application No.11 / 228,700, "Operation Of A Computer With A Touch Screen Interface," filed September 16, 2005; (8) U.S. Patent Application No.11 / 228,737, "Activating Virtual Keys Of A Touch-Screen Virtual Keyboard," filed September 16, 2005; and (9) U.S. Patent Application No.11 / 367,749, "Multi-Functional Hand-Held Device," filed March 3, 2006. All of these applications are incorporated by reference herein in their entirety.
[0082] Touch screen 112 optionally has a video resolution in excess of 100 dpi. In some embodiments, the touch screen has a video resolution of approximately 160 dpi. The -17- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) user optionally makes contact with touch screen 112 using any suitable object or appendage, such as a stylus, a finger, and so forth. In some embodiments, the user interface is designed to work primarily with finger-based contacts and gestures, which can be less precise than stylus-based input due to the larger area of contact of a finger on the touch screen. In some embodiments, the device translates the rough finger-based input into a precise pointer / cursor position or command for performing the actions desired by the user.
[0083] In some embodiments, in addition to the touch screen, device 100 optionally includes a touchpad (not shown) for activating or deactivating particular functions. In some embodiments, the touchpad is a touch-sensitive area of the device that, unlike the touch screen, does not display visual output. The touchpad is, optionally, a touch-sensitive surface that is separate from touch screen 112 or an extension of the touch-sensitive surface formed by the touch screen.
[0084] Device 100 also includes power system 162 for powering the various components. Power system 162 optionally includes a power management system, one or more power sources (e.g., battery, alternating current (AC)), a recharging system, a power failure detection circuit, a power converter or inverter, a power status indicator (e.g., a light- emitting diode (LED)) and any other components associated with the generation, management and distribution of power in portable or non-portable devices.
[0085] Device 100 optionally also includes one or more optical sensors 164. Fig.1A shows an optical sensor coupled to optical sensor controller 158 in I / O subsystem 106. Optical sensor 164 optionally includes charge-coupled device (CCD) or complementary metal-oxide semiconductor (CMOS) phototransistors. Optical sensor 164 receives light from the environment, projected through one or more lenses, and converts the light to data representing an image. In conjunction with imaging module 143 (also called a camera module), optical sensor 164 optionally captures still images or video. In some embodiments, an optical sensor is located on the back of device 100, opposite touch screen display 112 on the front of the device so that the touch screen display is enabled for use as a viewfinder for still and / or video image acquisition. In some embodiments, an optical sensor is located on the front of the device so that the user's image is, optionally, obtained for video conferencing while the user views the other video conference participants on the touch screen display. In some embodiments, the position of optical sensor 164 can be changed by the user (e.g., by rotating the lens and the sensor in the device housing) so that a single optical sensor 164 is -18- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) used along with the touch screen display for both video conferencing and still and / or video image acquisition.
[0086] Device 100 optionally also includes one or more contact intensity sensors 165. Fig.1A shows a contact intensity sensor coupled to intensity sensor controller 159 in I / O subsystem 106. Contact intensity sensor 165 optionally includes one or more piezoresistive strain gauges, capacitive force sensors, electric force sensors, piezoelectric force sensors, optical force sensors, capacitive touch-sensitive surfaces, or other intensity sensors (e.g., sensors used to measure the force (or pressure) of a contact on a touch-sensitive surface). Contact intensity sensor 165 receives contact intensity information (e.g., pressure information or a proxy for pressure information) from the environment. In some embodiments, at least one contact intensity sensor is collocated with, or proximate to, a touch-sensitive surface (e.g., touch-sensitive display system 112). In some embodiments, at least one contact intensity sensor is located on the back of device 100, opposite touch screen display 112 which is located on the front of device 100.
[0087] Device 100 optionally also includes one or more proximity sensors 166. Fig. 1A shows proximity sensor 166 coupled to peripherals interface 118. Alternately, proximity sensor 166 is, optionally, coupled to input controller 160 in I / O subsystem 106. Proximity sensor 166 optionally performs as described in U.S. Patent Application Nos.11 / 241,839, "Proximity Detector In Handheld Device"; 11 / 240,788, "Proximity Detector In Handheld Device"; 11 / 620,702, "Using Ambient Light Sensor To Augment Proximity Sensor Output"; 11 / 586,862, "Automated Response To And Sensing Of User Activity In Portable Devices"; and 11 / 638,251, "Methods And Systems For Automatic Configuration Of Peripherals," which are hereby incorporated by reference in their entirety. In some embodiments, the proximity sensor turns off and disables touch screen 112 when the multifunction device is placed near the user's ear (e.g., when the user is making a phone call).
[0088] Device 100 optionally also includes one or more tactile output generators 167. Fig.1A shows a tactile output generator coupled to haptic feedback controller 161 in I / O subsystem 106. Tactile output generator 167 optionally includes one or more electroacoustic devices such as speakers or other audio components and / or electromechanical devices that convert energy into linear motion such as a motor, solenoid, electroactive polymer, piezoelectric actuator, electrostatic actuator, or other tactile output generating component (e.g., a component that converts electrical signals into tactile outputs on the device). Contact intensity sensor 165 receives tactile feedback generation instructions from haptic feedback -19- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) module 133 and generates tactile outputs on device 100 that are capable of being sensed by a user of device 100. In some embodiments, at least one tactile output generator is collocated with, or proximate to, a touch-sensitive surface (e.g., touch-sensitive display system 112) and, optionally, generates a tactile output by moving the touch-sensitive surface vertically (e.g., in / out of a surface of device 100) or laterally (e.g., back and forth in the same plane as a surface of device 100). In some embodiments, at least one tactile output generator sensor is located on the back of device 100, opposite touch screen display 112 which is located on the front of device 100.
[0089] Device 100 optionally also includes one or more accelerometers 168. Fig.1A shows accelerometer 168 coupled to peripherals interface 118. Alternately, accelerometer 168 is, optionally, coupled to an input controller 160 in I / O subsystem 106. Accelerometer 168 optionally performs as described in U.S. Patent Publication No.20050190059, "Acceleration-based Theft Detection System for Portable Electronic Devices," and U.S. Patent Publication No.20060017692, "Methods And Apparatuses For Operating A Portable Device Based On An Accelerometer," both of which are incorporated by reference herein in their entirety. In some embodiments, information is displayed on the touch screen display in a portrait view or a landscape view based on an analysis of data received from the one or more accelerometers. Device 100 optionally includes, in addition to accelerometer(s) 168, a magnetometer (not shown) and a GPS (or GLONASS or other global navigation system) receiver (not shown) for obtaining information concerning the location and orientation (e.g., portrait or landscape) of device 100.
[0090] In some embodiments, the software components stored in memory 102 include operating system 126, communication module (or set of instructions) 128, contact / motion module (or set of instructions) 130, graphics module (or set of instructions) 132, text input module (or set of instructions) 134, Global Positioning System (GPS) module (or set of instructions) 135, and applications (or sets of instructions) 136. Furthermore, in some embodiments, memory 102 (Fig.1A) or 370 (Fig.3A) stores device / global internal state 157, as shown in Figs.1A and 3. Device / global internal state 157 includes one or more of: active application state, indicating which applications, if any, are currently active; display state, indicating what applications, views or other information occupy various regions of touch screen display 112; sensor state, including information obtained from the device's various sensors and input control devices 116; and location information concerning the device's location and / or attitude. -20- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0091] Operating system 126 (e.g., Darwin, RTXC, LINUX, UNIX, OS X, iOS, WINDOWS, or an embedded operating system such as VxWorks) includes various software components and / or drivers for controlling and managing general system tasks (e.g., memory management, storage device control, power management, etc.) and facilitates communication between various hardware and software components.
[0092] Communication module 128 facilitates communication with other devices over one or more external ports 124 and also includes various software components for handling data received by RF circuitry 108 and / or external port 124. External port 124 (e.g., Universal Serial Bus (USB), FIREWIRE, etc.) is adapted for coupling directly to other devices or indirectly over a network (e.g., the Internet, wireless LAN, etc.). In some embodiments, the external port is a multi-pin (e.g., 30-pin) connector that is the same as, or similar to and / or compatible with the 30-pin connector used on iPod (trademark of Apple Inc.) devices.
[0093] Contact / motion module 130 optionally detects contact with touch screen 112 (in conjunction with display controller 156) and other touch-sensitive devices (e.g., a touchpad or physical click wheel). Contact / motion module 130 includes various software components for performing various operations related to detection of contact, such as determining if contact has occurred (e.g., detecting a finger-down event), determining an intensity of the contact (e.g., the force or pressure of the contact or a substitute for the force or pressure of the contact) determining if there is movement of the contact and tracking the movement across the touch-sensitive surface (e.g., detecting one or more finger-dragging events), and determining if the contact has ceased (e.g., detecting a finger-up event or a break in contact). Contact / motion module 130 receives contact data from the touch-sensitive surface. Determining movement of the point of contact, which is represented by a series of contact data, optionally includes determining speed (magnitude), velocity (magnitude and direction), and / or an acceleration (a change in magnitude and / or direction) of the point of contact. These operations are, optionally, applied to single contacts (e.g., one finger contacts) or to multiple simultaneous contacts (e.g., "multitouch" / multiple finger contacts). In some embodiments, contact / motion module 130 and display controller 156 detect contact on a touchpad.
[0094] In some embodiments, contact / motion module 130 uses a set of one or more intensity thresholds to determine whether an operation has been performed by a user (e.g., to determine whether a user has "clicked" on an icon). In some embodiments at least a subset of -21- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) the intensity thresholds are determined in accordance with software parameters (e.g., the intensity thresholds are not determined by the activation thresholds of particular physical actuators and can be adjusted without changing the physical hardware of device 100). For example, a mouse "click" threshold of a trackpad or touch screen display can be set to any of a large range of predefined threshold values without changing the trackpad or touch screen display hardware. Additionally, in some implementations a user of the device is provided with software settings for adjusting one or more of the set of intensity thresholds (e.g., by adjusting individual intensity thresholds and / or by adjusting a plurality of intensity thresholds at once with a system-level click "intensity" parameter).
[0095] Contact / motion module 130 optionally detects a gesture input by a user. Different gestures on the touch-sensitive surface have different contact patterns (e.g., different motions, timings, and / or intensities of detected contacts). Thus, a gesture is, optionally, detected by detecting a particular contact pattern. For example, detecting a finger tap gesture includes detecting a finger-down event followed by detecting a finger-up (liftoff) event at the same position (or substantially the same position) as the finger-down event (e.g., at the position of an icon). As another example, detecting a finger swipe gesture on the touch-sensitive surface includes detecting a finger-down event followed by detecting one or more finger-dragging events, and subsequently followed by detecting a finger-up (liftoff) event.
[0096] Graphics module 132 includes various known software components for rendering and displaying graphics on touch screen 112 or other display, including components for changing the visual impact (e.g., brightness, transparency, saturation, contrast or other visual property) of graphics that are displayed. As used herein, the term "graphics" includes any object that can be displayed to a user, including without limitation text, web pages, icons (such as user-interface objects including soft keys), digital images, videos, animations and the like.
[0097] In some embodiments, graphics module 132 stores data representing graphics to be used. Each graphic is, optionally, assigned a corresponding code. Graphics module 132 receives, from applications etc., one or more codes specifying graphics to be displayed along with, if necessary, coordinate data and other graphic property data, and then generates screen image data to output to display controller 156. -22- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0098] Haptic feedback module 133 includes various software components for generating instructions used by tactile output generator(s) 167 to produce tactile outputs at one or more locations on device 100 in response to user interactions with device 100.
[0099] Text input module 134, which is, optionally, a component of graphics module 132, provides soft keyboards for entering text in various applications (e.g., contacts 137, e-mail 140, IM 141, browser 147, and any other application that needs text input).
[0100] GPS module 135 determines the location of the device and provides this information for use in various applications (e.g., to telephone 138 for use in location-based dialing, to camera 143 as picture / video metadata, and to applications that provide location- based services such as weather widgets, local yellow page widgets, and map / navigation widgets).
[0101] Applications 136 optionally include the following modules (or sets of instructions), or a subset or superset thereof: ^ contacts module 137 (sometimes called an address book or contact list); ^ telephone module 138; ^ video conferencing module 139; ^ e-mail client module 140; ^ instant messaging (IM) module 141; ^ workout support module 142; ^ camera module 143 for still and / or video images; ^ image management module 144; ^ video player module; ^ music player module; ^ browser module 147; ^ calendar module 148; ^ widget modules 149, which optionally include one or more of: weather widget 149-1, stocks widget 149-2, calculator widget 149-3, alarm clock widget 149-4, dictionary widget 149-5, and other widgets obtained by the user, as well as user-created widgets 149-6; -23- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) ^ widget creator module 150 for making user-created widgets 149-6; ^ search module 151; ^ video and music player module 152, which merges video player module and music player module; ^ notes module 153; ^ map module 154; and / or ^ online video module 155.
[0102] Examples of other applications 136 that are, optionally, stored in memory 102 include other word processing applications, other image editing applications, drawing applications, presentation applications, JAVA-enabled applications, encryption, digital rights management, voice recognition, and voice replication.
[0103] In conjunction with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, contacts module 137 are, optionally, used to manage an address book or contact list (e.g., stored in application internal state 192 of contacts module 137 in memory 102 or memory 370), including: adding name(s) to the address book; deleting name(s) from the address book; associating telephone number(s), e-mail address(es), physical address(es) or other information with a name; associating an image with a name; categorizing and sorting names; providing telephone numbers or e-mail addresses to initiate and / or facilitate communications by telephone 138, video conference module 139, e-mail 140, or IM 141; and so forth.
[0104] In conjunction with RF circuitry 108, audio circuitry 110, speaker 111, microphone 113, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, telephone module 138 are optionally, used to enter a sequence of characters corresponding to a telephone number, access one or more telephone numbers in contacts module 137, modify a telephone number that has been entered, dial a respective telephone number, conduct a conversation, and disconnect or hang up when the conversation is completed. As noted above, the wireless communication optionally uses any of a plurality of communications standards, protocols, and technologies.
[0105] In conjunction with RF circuitry 108, audio circuitry 110, speaker 111, microphone 113, touch screen 112, display controller 156, optical sensor 164, optical sensor controller 158, contact / motion module 130, graphics module 132, text input module 134, -24- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) contacts module 137, and telephone module 138, video conference module 139 includes executable instructions to initiate, conduct, and terminate a video conference between a user and one or more other participants in accordance with user instructions.
[0106] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, e-mail client module 140 includes executable instructions to create, send, receive, and manage e-mail in response to user instructions. In conjunction with image management module 144, e-mail client module 140 makes it very easy to create and send e-mails with still or video images taken with camera module 143.
[0107] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, the instant messaging module 141 includes executable instructions to enter a sequence of characters corresponding to an instant message, to modify previously entered characters, to transmit a respective instant message (for example, using a Short Message Service (SMS) or Multimedia Message Service (MMS) protocol for telephony-based instant messages or using XMPP, SIMPLE, or IMPS for Internet-based instant messages), to receive instant messages, and to view received instant messages. In some embodiments, transmitted and / or received instant messages optionally include graphics, photos, audio files, video files and / or other attachments as are supported in an MMS and / or an Enhanced Messaging Service (EMS). As used herein, "instant messaging" refers to both telephony-based messages (e.g., messages sent using SMS or MMS) and Internet-based messages (e.g., messages sent using XMPP, SIMPLE, or IMPS).
[0108] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, GPS module 135, map module 154, and music player module, workout support module 142 includes executable instructions to create workouts (e.g., with time, distance, and / or calorie burning goals); communicate with workout sensors (sports devices); receive workout sensor data; calibrate sensors used to monitor a workout; select and play music for a workout; and display, store, and transmit workout data.
[0109] In conjunction with touch screen 112, display controller 156, optical sensor(s) 164, optical sensor controller 158, contact / motion module 130, graphics module 132, and image management module 144, camera module 143 includes executable instructions to -25- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) capture still images or video (including a video stream) and store them into memory 102, modify characteristics of a still image or video, or delete a still image or video from memory 102.
[0110] In conjunction with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, and camera module 143, image management module 144 includes executable instructions to arrange, modify (e.g., edit), or otherwise manipulate, label, delete, present (e.g., in a digital slide show or album), and store still and / or video images.
[0111] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, browser module 147 includes executable instructions to browse the Internet in accordance with user instructions, including searching, linking to, receiving, and displaying web pages or portions thereof, as well as attachments and other files linked to web pages.
[0112] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, e-mail client module 140, and browser module 147, calendar module 148 includes executable instructions to create, display, modify, and store calendars and data associated with calendars (e.g., calendar entries, to -do lists, etc.) in accordance with user instructions.
[0113] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, and browser module 147, widget modules 149 are mini-applications that are, optionally, downloaded and used by a user (e.g., weather widget 149-1, stocks widget 149-2, calculator widget 149-3, alarm clock widget 149-4, and dictionary widget 149-5) or created by the user (e.g., user- created widget 149-6). In some embodiments, a widget includes an HTML (Hypertext Markup Language) file, a CSS (Cascading Style Sheets) file, and a JavaScript file. In some embodiments, a widget includes an XML (Extensible Markup Language) file and a JavaScript file (e.g., Yahoo! Widgets).
[0114] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, and browser module 147, the widget creator module 150 are, optionally, used by a user to create widgets (e.g., turning a user-specified portion of a web page into a widget). -26- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0115] In conjunction with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, search module 151 includes executable instructions to search for text, music, sound, image, video, and / or other files in memory 102 that match one or more search criteria (e.g., one or more user-specified search terms) in accordance with user instructions.
[0116] In conjunction with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, audio circuitry 110, speaker 111, RF circuitry 108, and browser module 147, video and music player module 152 includes executable instructions that allow the user to download and play back recorded music and other sound files stored in one or more file formats, such as MP3 or AAC files, and executable instructions to display, present, or otherwise play back videos (e.g., on touch screen 112 or on an external, connected display via external port 124). In some embodiments, device 100 optionally includes the functionality of an MP3 player, such as an iPod (trademark of Apple Inc.).
[0117] In conjunction with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, notes module 153 includes executable instructions to create and manage notes, to -do lists, and the like in accordance with user instructions.
[0118] In conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, GPS module 135, and browser module 147, map module 154 are, optionally, used to receive, display, modify, and store maps and data associated with maps (e.g., driving directions, data on stores and other points of interest at or near a particular location, and other location-based data) in accordance with user instructions.
[0119] In conjunction with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, audio circuitry 110, speaker 111, RF circuitry 108, text input module 134, e-mail client module 140, and browser module 147, online video module 155 includes instructions that allow the user to access, browse, receive (e.g., by streaming and / or download), play back (e.g., on the touch screen or on an external, connected display via external port 124), send an e-mail with a link to a particular online video, and otherwise manage online videos in one or more file formats, such as H.264. In some embodiments, instant messaging module 141, rather than e-mail client module 140, is used to send a link to a particular online video. Additional description of the online video application can be found -27- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in U.S. Provisional Patent Application No.60 / 936,562, "Portable Multifunction Device, Method, and Graphical User Interface for Playing Online Videos," filed June 20, 2007, and U.S. Patent Application No.11 / 968,67, "Portable Multifunction Device, Method, and Graphical User Interface for Playing Online Videos," filed December 31, 2007, the contents of which are hereby incorporated by reference in their entirety.
[0120] Each of the above -identified modules and applications corresponds to a set of executable instructions for performing one or more functions described above and the methods described in this application (e.g., the computer-implemented methods and other information processing methods described herein). These modules (e.g., sets of instructions) need not be implemented as separate software programs, procedures, or modules, and thus various subsets of these modules are, optionally, combined or otherwise rearranged in various embodiments. For example, video player module is, optionally, combined with music player module into a single module (e.g., video and music player module 152, Fig.1A). In some embodiments, memory 102 optionally stores a subset of the modules and data structures identified above. Furthermore, memory 102 optionally stores additional modules and data structures not described above.
[0121] In some embodiments, device 100 is a device where operation of a predefined set of functions on the device is performed exclusively through a touch screen and / or a touchpad. By using a touch screen and / or a touchpad as the primary input control device for operation of device 100, the number of physical input control devices (such as push buttons, dials, and the like) on device 100 is, optionally, reduced.
[0122] The predefined set of functions that are performed exclusively through a touch screen and / or a touchpad optionally include navigation between user interfaces. In some embodiments, the touchpad, when touched by the user, navigates device 100 to a main, home, or root menu from any user interface that is displayed on device 100. In such embodiments, a "menu button" is implemented using a touchpad. In some other embodiments, the menu button is a physical push button or other physical input control device instead of a touchpad.
[0123] Fig.1B is a block diagram illustrating exemplary components for event handling in accordance with some embodiments. In some embodiments, memory 102 (Fig. 1A) or 370 (Fig.3A) includes event sorter 170 (e.g., in operating system 126) and a respective application 136-1 (e.g., any of the aforementioned applications 137-151, 155, 380- 390). -28- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0124] Event sorter 170 receives event information and determines the application 136-1 and application view 191 of application 136-1 to which to deliver the event information. Event sorter 170 includes event monitor 171 and event dispatcher module 174. In some embodiments, application 136-1 includes application internal state 192, which indicates the current application view(s) displayed on touch-sensitive display 112 when the application is active or executing. In some embodiments, device / global internal state 157 is used by event sorter 170 to determine which application(s) is (are) currently active, and application internal state 192 is used by event sorter 170 to determine application views 191 to which to deliver event information.
[0125] In some embodiments, application internal state 192 includes additional information, such as one or more of: resume information to be used when application 136-1 resumes execution, user interface state information that indicates information being displayed or that is ready for display by application 136-1, a state queue for enabling the user to go back to a prior state or view of application 136-1, and a redo / undo queue of previous actions taken by the user.
[0126] Event monitor 171 receives event information from peripherals interface 118. Event information includes information about a sub-event (e.g., a user touch on touch- sensitive display 112, as part of a multi-touch gesture). Peripherals interface 118 transmits information it receives from I / O subsystem 106 or a sensor, such as proximity sensor 166, accelerometer(s) 168, and / or microphone 113 (through audio circuitry 110). Information that peripherals interface 118 receives from I / O subsystem 106 includes information from touch- sensitive display 112 or a touch-sensitive surface.
[0127] In some embodiments, event monitor 171 sends requests to the peripherals interface 118 at predetermined intervals. In response, peripherals interface 118 transmits event information. In other embodiments, peripherals interface 118 transmits event information only when there is a significant event (e.g., receiving an input above a predetermined noise threshold and / or for more than a predetermined duration).
[0128] In some embodiments, event sorter 170 also includes a hit view determination module 172 and / or an active event recognizer determination module 173.
[0129] Hit view determination module 172 provides software procedures for determining where a sub-event has taken place within one or more views when touch- -29- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) sensitive display 112 displays more than one view. Views are made up of controls and other elements that a user can see on the display.
[0130] Another aspect of the user interface associated with an application is a set of views, sometimes herein called application views or user interface windows, in which information is displayed and touch-based gestures occur. The application views (of a respective application) in which a touch is detected optionally correspond to programmatic levels within a programmatic or view hierarchy of the application. For example, the lowest level view in which a touch is detected is, optionally, called the hit view, and the set of events that are recognized as proper inputs are, optionally, determined based, at least in part, on the hit view of the initial touch that begins a touch-based gesture.
[0131] Hit view determination module 172 receives information related to sub-events of a touch-based gesture. When an application has multiple views organized in a hierarchy, hit view determination module 172 identifies a hit view as the lowest view in the hierarchy which should handle the sub-event. In most circumstances, the hit view is the lowest level view in which an initiating sub-event occurs (e.g., the first sub-event in the sequence of sub- events that form an event or potential event). Once the hit view is identified by the hit view determination module 172, the hit view typically receives all sub-events related to the same touch or input source for which it was identified as the hit view.
[0132] Active event recognizer determination module 173 determines which view or views within a view hierarchy should receive a particular sequence of sub-events. In some embodiments, active event recognizer determination module 173 determines that only the hit view should receive a particular sequence of sub-events. In other embodiments, active event recognizer determination module 173 determines that all views that include the physical location of a sub-event are actively involved views, and therefore determines that all actively involved views should receive a particular sequence of sub-events. In other embodiments, even if touch sub-events were entirely confined to the area associated with one particular view, views higher in the hierarchy would still remain as actively involved views.
[0133] Event dispatcher module 174 dispatches the event information to an event recognizer (e.g., event recognizer 180). In embodiments including active event recognizer determination module 173, event dispatcher module 174 delivers the event information to an event recognizer determined by active event recognizer determination module 173. In some -30- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) embodiments, event dispatcher module 174 stores in an event queue the event information, which is retrieved by a respective event receiver 182.
[0134] In some embodiments, operating system 126 includes event sorter 170. Alternatively, application 136-1 includes event sorter 170. In yet other embodiments, event sorter 170 is a stand-alone module, or a part of another module stored in memory 102, such as contact / motion module 130.
[0135] In some embodiments, application 136-1 includes a plurality of event handlers 190 and one or more application views 191, each of which includes instructions for handling touch events that occur within a respective view of the application's user interface. Each application view 191 of the application 136-1 includes one or more event recognizers 180. Typically, a respective application view 191 includes a plurality of event recognizers 180. In other embodiments, one or more of event recognizers 180 are part of a separate module, such as a user interface kit (not shown) or a higher level object from which application 136-1 inherits methods and other properties. In some embodiments, a respective event handler 190 includes one or more of: data updater 176, object updater 177, GUI updater 178, and / or event data 179 received from event sorter 170. Event handler 190 optionally utilizes or calls data updater 176, object updater 177, or GUI updater 178 to update the application internal state 192. Alternatively, one or more of the application views 191 include one or more respective event handlers 190. Also, in some embodiments, one or more of data updater 176, object updater 177, and GUI updater 178 are included in a respective application view 191.
[0136] A respective event recognizer 180 receives event information (e.g., event data 179) from event sorter 170 and identifies an event from the event information. Event recognizer 180 includes event receiver 182 and event comparator 184. In some embodiments, event recognizer 180 also includes at least a subset of: metadata 183, and event delivery instructions 188 (which optionally include sub-event delivery instructions).
[0137] Event receiver 182 receives event information from event sorter 170. The event information includes information about a sub-event, for example, a touch or a touch movement. Depending on the sub-event, the event information also includes additional information, such as location of the sub-event. When the sub-event concerns motion of a touch, the event information optionally also includes speed and direction of the sub-event. In some embodiments, events include rotation of the device from one orientation to another (e.g., from a portrait orientation to a landscape orientation, or vice versa), and the event -31- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) information includes corresponding information about the current orientation (also called device attitude) of the device.
[0138] Event comparator 184 compares the event information to predefined event or sub-event definitions and, based on the comparison, determines an event or sub-event, or determines or updates the state of an event or sub-event. In some embodiments, event comparator 184 includes event definitions 186. Event definitions 186 contain definitions of events (e.g., predefined sequences of sub-events), for example, event 1 (187-1), event 2 (187- 2), and others. In some embodiments, sub-events in an event (187) include, for example, touch begin, touch end, touch movement, touch cancellation, and multiple touching. In one example, the definition for event 1 (187-1) is a double tap on a displayed object. The double tap, for example, comprises a first touch (touch begin) on the displayed object for a predetermined phase, a first liftoff (touch end) for a predetermined phase, a second touch (touch begin) on the displayed object for a predetermined phase, and a second liftoff (touch end) for a predetermined phase. In another example, the definition for event 2 (187-2) is a dragging on a displayed object. The dragging, for example, comprises a touch (or contact) on the displayed object for a predetermined phase, a movement of the touch across touch- sensitive display 112, and liftoff of the touch (touch end). In some embodiments, the event also includes information for one or more associated event handlers 190.
[0139] In some embodiments, event definition 187 includes a definition of an event for a respective user-interface object. In some embodiments, event comparator 184 performs a hit test to determine which user-interface object is associated with a sub-event. For example, in an application view in which three user-interface objects are displayed on touch- sensitive display 112, when a touch is detected on touch-sensitive display 112, event comparator 184 performs a hit test to determine which of the three user-interface objects is associated with the touch (sub-event). If each displayed object is associated with a respective event handler 190, the event comparator uses the result of the hit test to determine which event handler 190 should be activated. For example, event comparator 184 selects an event handler associated with the sub-event and the object triggering the hit test.
[0140] In some embodiments, the definition for a respective event (187) also includes delayed actions that delay delivery of the event information until after it has been determined whether the sequence of sub-events does or does not correspond to the event recognizer's event type. -32- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0141] When a respective event recognizer 180 determines that the series of sub- events do not match any of the events in event definitions 186, the respective event recognizer 180 enters an event impossible, event failed, or event ended state, after which it disregards subsequent sub-events of the touch-based gesture. In this situation, other event recognizers, if any, that remain active for the hit view continue to track and process sub- events of an ongoing touch-based gesture.
[0142] In some embodiments, a respective event recognizer 180 includes metadata 183 with configurable properties, flags, and / or lists that indicate how the event delivery system should perform sub-event delivery to actively involved event recognizers. In some embodiments, metadata 183 includes configurable properties, flags, and / or lists that indicate how event recognizers interact, or are enabled to interact, with one another. In some embodiments, metadata 183 includes configurable properties, flags, and / or lists that indicate whether sub-events are delivered to varying levels in the view or programmatic hierarchy.
[0143] In some embodiments, a respective event recognizer 180 activates event handler 190 associated with an event when one or more particular sub-events of an event are recognized. In some embodiments, a respective event recognizer 180 delivers event information associated with the event to event handler 190. Activating an event handler 190 is distinct from sending (and deferred sending) sub-events to a respective hit view. In some embodiments, event recognizer 180 throws a flag associated with the recognized event, and event handler 190 associated with the flag catches the flag and performs a predefined process.
[0144] In some embodiments, event delivery instructions 188 include sub-event delivery instructions that deliver event information about a sub-event without activating an event handler. Instead, the sub-event delivery instructions deliver event information to event handlers associated with the series of sub-events or to actively involved views. Event handlers associated with the series of sub-events or with actively involved views receive the event information and perform a predetermined process.
[0145] In some embodiments, data updater 176 creates and updates data used in application 136-1. For example, data updater 176 updates the telephone number used in contacts module 137, or stores a video file used in video player module. In some embodiments, object updater 177 creates and updates objects used in application 136-1. For example, object updater 177 creates a new user-interface object or updates the position of a user-interface object. GUI updater 178 updates the GUI. For example, GUI updater 178 -33- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) prepares display information and sends it to graphics module 132 for display on a touch- sensitive display.
[0146] In some embodiments, event handler(s) 190 includes or has access to data updater 176, object updater 177, and GUI updater 178. In some embodiments, data updater 176, object updater 177, and GUI updater 178 are included in a single module of a respective application 136-1 or application view 191. In other embodiments, they are included in two or more software modules.
[0147] It shall be understood that the foregoing discussion regarding event handling of user touches on touch-sensitive displays also applies to other forms of user inputs to operate multifunction devices 100 with input devices, not all of which are initiated on touch screens. For example, mouse movement and mouse button presses, optionally coordinated with single or multiple keyboard presses or holds; contact movements such as taps, drags, scrolls, etc. on touchpads; pen stylus inputs; movement of the device; oral instructions; detected eye movements; biometric inputs; and / or any combination thereof are optionally utilized as inputs corresponding to sub-events which define an event to be recognized.
[0148] Fig.2 illustrates a portable or non-portable multifunction device 100 having a touch screen 112 in accordance with some embodiments. As stated above, multifunction device 100 is described as having the various illustrated structures (such as touch screen 112, speaker 111, accelerometer 168, microphone 113, etc.); however, it is understood that these structures optionally reside on separate devices. For example, display-related structures (e.g., display, speaker, etc.) and / or functions optionally reside on a separate display device, input- related structures (e.g., touch-sensitive surface, microphone, accelerometer, etc.) and / or functions optionally reside on a separate input device, and remaining structures and / or functions optionally reside on multifunction device 100.
[0149] The touch screen 112 optionally displays one or more graphics within user interface (UI) 200. In this embodiment, as well as others described below, a user is enabled to select one or more of the graphics by making a gesture on the graphics, for example, with one or more fingers 202 (not drawn to scale in the figure) or one or more styluses 203 (not drawn to scale in the figure). In some embodiments, selection of one or more graphics occurs when the user breaks contact with the one or more graphics. In some embodiments, the gesture optionally includes one or more taps, one or more swipes (from left to right, right to left, upward and / or downward) and / or a rolling of a finger (from right to left, left to right, -34- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) upward and / or downward) that has made contact with device 100. In some implementations or circumstances, inadvertent contact with a graphic does not select the graphic. For example, a swipe gesture that sweeps over an application icon optionally does not select the corresponding application when the gesture corresponding to selection is a tap.
[0150] Device 100 optionally also includes one or more physical buttons, such as "home" or menu button 204. As previously described, menu button 204 is, optionally, used to navigate to any application 136 in a set of applications that are, optionally executed on device 100. Alternatively, in some embodiments, the menu button is implemented as a soft key in a GUI displayed on touch screen 112.
[0151] In one embodiment, device 100 includes touch screen 112, menu button 204, push button 206 for powering the device on / off and locking the device, volume adjustment button(s) 208, Subscriber Identity Module (SIM) card slot 210, head set jack 212, and docking / charging external port 124. Push button 206 is, optionally, used to turn the power on / off on the device by depressing the button and holding the button in the depressed state for a predefined time interval; to lock the device by depressing the button and releasing the button before the predefined time interval has elapsed; and / or to unlock the device or initiate an unlock process. In an alternative embodiment, device 100 also accepts verbal input for activation or deactivation of some functions through microphone 113. Device 100 also, optionally, includes one or more contact intensity sensors 165 for detecting intensity of contacts on touch screen 112 and / or one or more tactile output generators 167 for generating tactile outputs for a user of device 100.
[0152] Fig.3A is a block diagram of an exemplary multifunction device with a display and a touch-sensitive surface in accordance with some embodiments. Device 300 need not include the display and the touch-sensitive surface, as described above, but rather, in some embodiments, optionally communicates with the display and the touch-sensitive surface on other devices. Additionally, device 300 need not be portable. In some embodiments, device 300 is a laptop computer, a desktop computer, a tablet computer, a multimedia player device (such as a television or a set-top box), a navigation device, an educational device (such as a child's learning toy), a gaming system, or a control device (e.g., a home or industrial controller). Device 300 typically includes one or more processing units (CPU's) 310, one or more network or other communications interfaces 360, memory 370, and one or more communication buses 320 for interconnecting these components. Communication buses 320 optionally include circuitry (sometimes called a chipset) that interconnects and -35- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) controls communications between system components. Device 300 includes input / output (I / O) interface 330 comprising display 340, which is typically a touch screen display. I / O interface 330 also optionally includes a keyboard and / or mouse (or other pointing device) 350 and touchpad 355, tactile output generator 357 for generating tactile outputs on device 300 (e.g., similar to tactile output generator(s) 167 described above with reference to Fig.1A), sensors 359 (e.g., optical, acceleration, proximity, touch-sensitive, and / or contact intensity sensors similar to contact intensity sensor(s) 165 described above with reference to Fig.1A). Memory 370 includes high-speed random access memory, such as DRAM, SRAM, DDR RAM or other random access solid state memory devices; and optionally includes non- volatile memory, such as one or more magnetic disk storage devices, optical disk storage devices, flash memory devices, or other non-volatile solid state storage devices. Memory 370 optionally includes one or more storage devices remotely located from CPU(s) 310. In some embodiments, memory 370 stores programs, modules, and data structures analogous to the programs, modules, and data structures stored in memory 102 of portable or non-portable multifunction device 100 (Fig.1A), or a subset thereof. Furthermore, memory 370 optionally stores additional programs, modules, and data structures not present in memory 102 of portable or non-portable multifunction device 100. For example, memory 370 of device 300 optionally stores drawing module 380, presentation module 382, word processing module 384, website creation module 386, disk authoring module 388, and / or spreadsheet module 390, while memory 102 of portable or non-portable multifunction device 100 (Fig.1A) optionally does not store these modules.
[0153] Each of the above identified elements in Fig.3A are, optionally, stored in one or more of the previously mentioned memory devices. Each of the above identified modules corresponds to a set of instructions for performing a function described above. The above identified modules or programs (e.g., sets of instructions) need not be implemented as separate software programs, procedures or modules, and thus various subsets of these modules are, optionally, combined or otherwise re-arranged in various embodiments. In some embodiments, memory 370 optionally stores a subset of the modules and data structures identified above. Furthermore, memory 370 optionally stores additional modules and data structures not described above.
[0154] Implementations within the scope of the present disclosure can be partially or entirely realized using a tangible computer-readable storage medium (or multiple tangible computer-readable storage media of one or more types) encoding one or more computer- -36- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) readable instructions. It should be recognized that computer-readable instructions can be organized in any format, including applications, widgets, processes, software, and / or components.
[0155] Implementations within the scope of the present disclosure include a computer-readable storage medium that encodes instructions organized as an application (e.g., application 3160) that, when executed by one or more processing units, control an electronic device (e.g., device 3150) to perform the method of FIG.3B, the method of FIG. 3C, and / or one or more other processes and / or methods described herein.
[0156] It should be recognized that application 3160 (shown in FIG.3D) can be any suitable type of application, including, for example, one or more of: a browser application, an application that functions as an execution environment for plug-ins, widgets or other applications, a fitness application, a health application, a digital payments application, a media application, a social network application, a messaging application, and / or a maps application. In some embodiments, application 3160 is an application that is pre-installed on device 3150 at purchase (e.g., a first-party application). In some embodiments, application 3160 is an application that is provided to device 3150 via an operating system update file (e.g., a first-party application or a second-party application). In some embodiments, application 3160 is an application that is provided via an application store. In some embodiments, the application store can be an application store that is pre-installed on device 3150 at purchase (e.g., a first-party application store). In some embodiments, the application store is a third-party application store (e.g., an application store that is provided by another application store, downloaded via a network, and / or read from a storage device).
[0157] Referring to FIG.3B and FIG.3F, application 3160 obtains information (e.g., 3010). In some embodiments, at 3010, information is obtained from at least one hardware component of device 3150. In some embodiments, at 3010, information is obtained from at least one software module of device 3150. In some embodiments, at 3010, information is obtained from at least one hardware component external to device 3150 (e.g., a peripheral device, an accessory device, and / or a server). In some embodiments, the information obtained at 3010 includes positional information, time information, notification information, user information, environment information, electronic device state information, weather information, media information, historical information, event information, hardware information, and / or motion information. In some embodiments, in response to and / or after -37- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) obtaining the information at 3010, application 3160 provides the information to a system (e.g., 3020).
[0158] In some embodiments, the system (e.g., 3110 shown in FIG.3E) is an operating system hosted on device 3150. In some embodiments, the system (e.g., 3110 shown in FIG.3E) is an external device (e.g., a server, a peripheral device, an accessory, and / or a personal computing device) that includes an operating system.
[0159] Referring to FIG.3C and FIG.3G, application 3160 obtains information (e.g., 3030). In some embodiments, the information obtained at 3030 includes positional information, time information, notification information, user information, environment information electronic device state information, weather information, media information, historical information, event information, hardware information, and / or motion information. In response to and / or after obtaining the information at 3030, application 3160 performs an operation with the information (e.g., 3040). In some embodiments, the operation performed at 3040 includes: providing a notification based on the information, sending a message based on the information, displaying the information, controlling a user interface of a fitness application based on the information, controlling a user interface of a health application based on the information, controlling a focus mode based on the information, setting a reminder based on the information, adding a calendar entry based on the information, and / or calling an API of system 3110 based on the information.
[0160] In some embodiments, one or more steps of the method of FIG.3B and / or the method of FIG.3C is performed in response to a trigger. In some embodiments, the trigger includes detection of an event, a notification received from system 3110, a user input, and / or a response to a call to an API provided by system 3110.
[0161] In some embodiments, the instructions of application 3160, when executed, control device 3150 to perform the method of FIG.3B and / or the method of FIG.3C by calling an application programming interface (API) (e.g., API 3190) provided by system 3110. In some embodiments, application 3160 performs at least a portion of the method of FIG.3B and / or the method of FIG.3C without calling API 3190.
[0162] In some embodiments, one or more steps of the method of FIG.3B and / or the method of FIG.3C includes calling an API (e.g., API 3190) using one or more parameters defined by the API. In some embodiments, the one or more parameters include a constant, a key, a data structure, an object, an object class, a variable, a data type, a pointer, an array, a -38- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) list or a pointer to a function or method, and / or another way to reference a data or other item to be passed via the API.
[0163] Referring to FIG.3D, device 3150 is illustrated. In some embodiments, device 3150 is a personal computing device, a smart phone, a smart watch, a fitness tracker, a head mounted display (HMD) device, a media device, a communal device, a speaker, a television, and / or a tablet. As illustrated in FIG.3D, device 3150 includes application 3160 and an operating system (e.g., system 3110 shown in FIG.3E). Application 3160 includes application implementation module 3170 and API-calling module 3180. System 3110 includes API 3190 and implementation module 3100. It should be recognized that device 3150, application 3160, and / or system 3110 can include more, fewer, and / or different components than illustrated in FIGS.3D and 3E.
[0164] In some embodiments, application implementation module 3170 includes a set of one or more instructions corresponding to one or more operations performed by application 3160. For example, when application 3160 is a messaging application, application implementation module 3170 can include operations to receive and send messages. In some embodiments, application implementation module 3170 communicates with API-calling module 3180 to communicate with system 3110 via API 3190 (shown in FIG.3E).
[0165] In some embodiments, API 3190 is a software module (e.g., a collection of computer-readable instructions) that provides an interface that allows a different module (e.g., API-calling module 3180) to access and / or use one or more functions, methods, procedures, data structures, classes, and / or other services provided by implementation module 3100 of system 3110. For example, API-calling module 3180 can access a feature of implementation module 3100 through one or more API calls or invocations (e.g., embodied by a function or a method call) exposed by API 3190 (e.g., a software and / or hardware module that can receive API calls, respond to API calls, and / or send API calls) and can pass data and / or control information using one or more parameters via the API calls or invocations. In some embodiments, API 3190 allows application 3160 to use a service provided by a Software Development Kit (SDK) library. In some embodiments, application 3160 incorporates a call to a function or method provided by the SDK library and provided by API 3190 or uses data types or objects defined in the SDK library and provided by API 3190. In some embodiments, API-calling module 3180 makes an API call via API 3190 to access and use a feature of implementation module 3100 that is specified by API 3190. In -39- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) such embodiments, implementation module 3100 can return a value via API 3190 to API- calling module 3180 in response to the API call. The value can report to application 3160 the capabilities or state of a hardware component of device 3150, including those related to aspects such as input capabilities and state, output capabilities and state, processing capability, power state, storage capacity and state, and / or communications capability. In some embodiments, API 3190 is implemented in part by firmware, microcode, or other low level logic that executes in part on the hardware component.
[0166] In some embodiments, API 3190 allows a developer of API-calling module 3180 (which can be a third-party developer) to leverage a feature provided by implementation module 3100. In such embodiments, there can be one or more API-calling modules (e.g., including API-calling module 3180) that communicate with implementation module 3100. In some embodiments, API 3190 allows multiple API-calling modules written in different programming languages to communicate with implementation module 3100 (e.g., API 3190 can include features for translating calls and returns between implementation module 3100 and API-calling module 3180) while API 3190 is implemented in terms of a specific programming language. In some embodiments, API-calling module 3180 calls APIs from different providers such as a set of APIs from an OS provider, another set of APIs from a plug-in provider, and / or another set of APIs from another provider (e.g., the provider of a software library) or creator of the another set of APIs.
[0167] Examples of API 3190 can include one or more of: a pairing API (e.g., for establishing secure connection, e.g., with an accessory), a device detection API (e.g., for locating nearby devices, e.g., media devices and / or smartphone), a payment API, a UIKit API (e.g., for generating user interfaces), a location detection API, a locator API, a maps API, a health sensor API, a sensor API, a messaging API, a push notification API, a streaming API, a collaboration API, a video conferencing API, an application store API, an advertising services API, a web browser API (e.g., WebKit API), a vehicle API, a networking API, a WiFi API, a Bluetooth API, an NFC API, a UWB API, a fitness API, a smart home API, contact transfer API, photos API, camera API, and / or image processing API. In some embodiments, the sensor API is an API for accessing data associated with a sensor of device 3150. For example, the sensor API can provide access to raw sensor data. For another example, the sensor API can provide data derived (and / or generated) from the raw sensor data. In some embodiments, the sensor data includes temperature data, image data, video data, audio data, heart rate data, IMU (inertial measurement unit) data, lidar data, location -40- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) data, GPS data, and / or camera data. In some embodiments, the sensor includes one or more of an accelerometer, temperature sensor, infrared sensor, optical sensor, heartrate sensor, barometer, gyroscope, proximity sensor, temperature sensor, and / or biometric sensor.
[0168] In some embodiments, implementation module 3100 is a system (e.g., operating system and / or server system) software module (e.g., a collection of computer- readable instructions) that is constructed to perform an operation in response to receiving an API call via API 3190. In some embodiments, implementation module 3100 is constructed to provide an API response (via API 3190) as a result of processing an API call. By way of example, implementation module 3100 and API-calling module 3180 can each be any one of an operating system, a library, a device driver, an API, an application program, or other module. It should be understood that implementation module 3100 and API-calling module 3180 can be the same or different type of module from each other. In some embodiments, implementation module 3100 is embodied at least in part in firmware, microcode, or hardware logic.
[0169] In some embodiments, implementation module 3100 returns a value through API 3190 in response to an API call from API-calling module 3180. While API 3190 defines the syntax and result of an API call (e.g., how to invoke the API call and what the API call does), API 3190 might not reveal how implementation module 3100 accomplishes the function specified by the API call. Various API calls are transferred via the one or more application programming interfaces between API-calling module 3180 and implementation module 3100. Transferring the API calls can include issuing, initiating, invoking, calling, receiving, returning, and / or responding to the function calls or messages. In other words, transferring can describe actions by either of API-calling module 3180 or implementation module 3100. In some embodiments, a function call or other invocation of API 3190 sends and / or receives one or more parameters through a parameter list or other structure.
[0170] In some embodiments, implementation module 3100 provides more than one API, each providing a different view of or with different aspects of functionality implemented by implementation module 3100. For example, one API of implementation module 3100 can provide a first set of functions and can be exposed to third-party developers, and another API of implementation module 3100 can be hidden (e.g., not exposed) and provide a subset of the first set of functions and also provide another set of functions, such as testing or debugging functions which are not in the first set of functions. In some embodiments, implementation module 3100 calls one or more other components via an underlying API and thus is both an -41- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) API-calling module and an implementation module. It should be recognized that implementation module 3100 can include additional functions, methods, classes, data structures, and / or other features that are not specified through API 3190 and are not available to API-calling module 3180. It should also be recognized that API-calling module 3180 can be on the same system as implementation module 3100 or can be located remotely and access implementation module 3100 using API 3190 over a network. In some embodiments, implementation module 3100, API 3190, and / or API-calling module 3180 is stored in a machine-readable medium, which includes any mechanism for storing information in a form readable by a machine (e.g., a computer or other data processing system). For example, a machine-readable medium can include magnetic disks, optical disks, random access memory; read only memory, and / or flash memory devices.
[0171] An application programming interface (API) is an interface between a first software process and a second software process that specifies a format for communication between the first software process and the second software process. Limited APIs (e.g., private APIs or partner APIs) are APIs that are accessible to a limited set of software processes (e.g., only software processes within an operating system or only software processes that are approved to access the limited APIs). Public APIs that are accessible to a wider set of software processes. Some APIs enable software processes to communicate about or set a state of one or more input devices (e.g., one or more touch sensors, proximity sensors, visual sensors, motion / orientation sensors, pressure sensors, intensity sensors, sound sensors, wireless proximity sensors, biometric sensors, buttons, switches, rotatable elements, and / or external controllers). Some APIs enable software processes to communicate about and / or set a state of one or more output generation components (e.g., one or more audio output generation components, one or more display generation components, and / or one or more tactile output generation components). Some APIs enable particular capabilities (e.g., scrolling, handwriting, text entry, image editing, and / or image creation) to be accessed, performed, and / or used by a software process (e.g., generating outputs for use by a software process based on input from the software process). Some APIs enable content from a software process to be inserted into a template and displayed in a user interface that has a layout and / or behaviors that are specified by the template.
[0172] Many software platforms include a set of frameworks that provides the core objects and core behaviors that a software developer needs to build software applications that can be used on the software platform. Software developers use these objects to display -42- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) content onscreen, to interact with that content, and to manage interactions with the software platform. Software applications rely on the set of frameworks for their basic behavior, and the set of frameworks provides many ways for the software developer to customize the behavior of the application to match the specific needs of the software application. Many of these core objects and core behaviors are accessed via an API. An API will typically specify a format for communication between software processes, including specifying and grouping available variables, functions, and protocols. An API call (sometimes referred to as an API request) will typically be sent from a sending software process to a receiving software process as a way to accomplish one or more of the following: the sending software process requesting information from the receiving software process (e.g., for the sending software process to take action on), the sending software process providing information to the receiving software process (e.g., for the receiving software process to take action on), the sending software process requesting action by the receiving software process, or the sending software process providing information to the receiving software process about action taken by the sending software process. Interaction with a device (e.g., using a user interface) will in some circumstances include the transfer and / or receipt of one or more API calls (e.g., multiple API calls) between multiple different software processes (e.g., different portions of an operating system, an application and an operating system, or different applications) via one or more APIs (e.g., via multiple different APIs). For example, when an input is detected the direct sensor data is frequently processed into one or more input events that are provided (e.g., via an API) to a receiving software process that makes some determination based on the input events, and then sends (e.g., via an API) information to a software process to perform an operation (e.g., change a device state and / or user interface) based on the determination. While a determination and an operation performed in response could be made by the same software process, alternatively the determination could be made in a first software process and relayed (e.g., via an API) to a second software process, that is different from the first software process, that causes the operation to be performed by the second software process. Alternatively, the second software process could relay instructions (e.g., via an API) to a third software process that is different from the first software process and / or the second software process to perform the operation. It should be understood that some or all user interactions with a computer system could involve one or more API calls within a step of interacting with the computer system (e.g., between different software components of the computer system or between a software component of the computer system and a software component of one or more remote computer systems). It should be understood that some or all user interactions -43- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) with a computer system could involve one or more API calls between steps of interacting with the computer system (e.g., between different software components of the computer system or between a software component of the computer system and a software component of one or more remote computer systems).
[0173] In some embodiments, the application can be any suitable type of application, including, for example, one or more of: a browser application, an application that functions as an execution environment for plug-ins, widgets or other applications, a fitness application, a health application, a digital payments application, a media application, a social network application, a messaging application, and / or a maps application.
[0174] In some embodiments, the application is an application that is pre-installed on the first computer system at purchase (e.g., a first-party application). In some embodiments, the application is an application that is provided to the first computer system via an operating system update file (e.g., a first-party application). In some embodiments, the application is an application that is provided via an application store. In some embodiments, the application store is pre-installed on the first computer system at purchase (e.g., a first-party application store) and allows download of one or more applications. In some embodiments, the application store is a third-party application store (e.g., an application store that is provided by another device, downloaded via a network, and / or read from a storage device). In some embodiments, the application is a third-party application (e.g., an app that is provided by an application store, downloaded via a network, and / or read from a storage device). In some embodiments, the application controls the first computer system to perform method 700, 800, 1000, 1200, 1400, 1500, 1700, 1900, 2100, 2300, and 2700 (FIG.7, 8, 10, 12, 14, 15, 17, 19, 21, 23, and 27) by calling an application programming interface (API) provided by the system process using one or more parameters.
[0175] In some embodiments, exemplary APIs provided by the system process include one or more of: a pairing API (e.g., for establishing secure connection, e.g., with an accessory), a device detection API (e.g., for locating nearby devices, e.g., media devices and / or smartphone), a payment API, a UIKit API (e.g., for generating user interfaces), a location detection API, a locator API, a maps API, a health sensor API, a sensor API, a messaging API, a push notification API, a streaming API, a collaboration API, a video conferencing API, an application store API, an advertising services API, a web browser API (e.g., WebKit API), a vehicle API, a networking API, a WiFi API, a Bluetooth API, an NFC -44- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) API, a UWB API, a fitness API, a smart home API, contact transfer API, a photos API, a camera API, and / or an image processing API.
[0176] In some embodiments, at least one API is a software module (e.g., a collection of computer-readable instructions) that provides an interface that allows a different module (e.g., API-calling module) to access and use one or more functions, methods, procedures, data structures, classes, and / or other services provided by an implementation module of the system process. The API can define one or more parameters that are passed between the API-calling module and the implementation module. In some embodiments, API 3190 defines a first API call that can be provided by API-calling module 3180. The implementation module is a system software module (e.g., a collection of computer-readable instructions) that is constructed to perform an operation in response to receiving an API call via the API. In some embodiments, the implementation module is constructed to provide an API response (via the API) as a result of processing an API call. In some embodiments, the implementation module is included in the device (e.g., 3150) that runs the application. In some embodiments, the implementation module is included in an electronic device that is separate from the device that runs the application.
[0177] Attention is now directed towards embodiments of user interfaces that are, optionally, implemented on, for example, portable multifunction device 100.
[0178] Fig.4A illustrates an exemplary user interface for a menu of applications on portable multifunction device 100 in accordance with some embodiments. Similar user interfaces are, optionally, implemented on device 300. In some embodiments, user interface 400 includes the following elements, or a subset or superset thereof: ^ Signal strength indicator(s) 402 for wireless communication(s), such as cellular and Wi-Fi signals; ^ Time 404; ^ Bluetooth indicator 405; ^ Battery status indicator 406; ^ Tray 408 with icons for frequently used applications, such as: ^ Icon 416 for telephone module 138, labeled “Phone,” which optionally includes an indicator 414 of the number of missed calls or voicemail messages; -45- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) ^ Icon 418 for e-mail client module 140, labeled “Mail,” which optionally includes an indicator 410 of the number of unread e-mails; ^ Icon 420 for browser module 147, labeled “Browser;” and ^ Icon 422 for video and music player module 152, also referred to as iPod (trademark of Apple Inc.) module 152, labeled “iPod;” and ^ Icons for other applications, such as: ^ Icon 424 for IM module 141, labeled “Messages;” ^ Icon 426 for calendar module 148, labeled “Calendar;” ^ Icon 428 for image management module 144, labeled “Photos;” ^ Icon 430 for camera module 143, labeled “Camera;” ^ Icon 432 for online video module 155, labeled “Online Video;” ^ Icon 434 for stocks widget 149-2, labeled “Stocks;” ^ Icon 436 for map module 154, labeled “Maps;” ^ Icon 438 for weather widget 149-1, labeled “Weather;” ^ Icon 440 for alarm clock widget 149-4, labeled “Clock;” ^ Icon 442 for workout support module 142, labeled “Workout Support;” ^ Icon 444 for notes module 153, labeled “Notes;” and ^ Icon 446 for a settings application or module, labeled “Settings,” which provides access to settings for device 100 and its various applications 136.
[0179] It should be noted that the icon labels illustrated in Fig.4A are merely exemplary. For example, icon 422 for video and music player module 152 is labeled “Music” or “Music Player.” Other labels are, optionally, used for various application icons. In some embodiments, a label for a respective application icon includes a name of an application corresponding to the respective application icon. In some embodiments, a label for a particular application icon is distinct from a name of an application corresponding to the particular application icon.
[0180] Fig.4B illustrates an exemplary user interface on a device (e.g., device 300, Fig.3) with a touch-sensitive surface 451 (e.g., a tablet or touchpad 355, Fig.3) that is -46- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) separate from the display 450 (e.g., touch screen display 112). Device 300 also, optionally, includes one or more contact intensity sensors (e.g., one or more of sensors 359) for detecting intensity of contacts on touch-sensitive surface 451 and / or one or more tactile output generators 357 for generating tactile outputs for a user of device 300.
[0181] Although some of the examples that follow will be given with reference to inputs on touch screen display 112 (where the touch-sensitive surface and the display are combined), in some embodiments, the device detects inputs on a touch-sensitive surface that is separate from the display, as shown in Fig.4B. In some embodiments, the touch-sensitive surface (e.g., 451 in Fig.4B) has a primary axis (e.g., 452 in Fig.4B) that corresponds to a primary axis (e.g., 453 in Fig.4B) on the display (e.g., 450). In accordance with these embodiments, the device detects contacts (e.g., 460 and 462 in Fig.4B) with the touch- sensitive surface 451 at locations that correspond to respective locations on the display (e.g., in Fig.4B, 460 corresponds to 468 and 462 corresponds to 470). In this way, user inputs (e.g., contacts 460 and 462, and movements thereof) detected by the device on the touch- sensitive surface (e.g., 451 in Fig.4B) are used by the device to manipulate the user interface on the display (e.g., 450 in Fig.4B) of the multifunction device when the touch-sensitive surface is separate from the display. It should be understood that similar methods are, optionally, used for other user interfaces described herein.
[0182] Additionally, while the following examples are given primarily with reference to finger inputs (e.g., finger contacts, finger tap gestures, finger swipe gestures), it should be understood that, in some embodiments, one or more of the finger inputs are replaced with input from another input device (e.g., a mouse-based input or stylus input). For example, a swipe gesture is, optionally, replaced with a mouse click (e.g., instead of a contact) followed by movement of the cursor along the path of the swipe (e.g., instead of movement of the contact). As another example, a tap gesture is, optionally, replaced with a mouse click while the cursor is located over the location of the tap gesture (e.g., instead of detection of the contact followed by ceasing to detect the contact). Similarly, when multiple user inputs are simultaneously detected, it should be understood that multiple computer mice are, optionally, used simultaneously, or a mouse and finger contacts are, optionally, used simultaneously.
[0183] Additionally, while the following examples are given primarily with reference to finger inputs (e.g., finger contacts, finger tap gestures, finger swipe gestures), it should be understood that, in some embodiments, one or more of the finger inputs are replaced with input from another input device (e.g., a mouse based input or stylus input). For example, a -47- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) swipe gesture is, optionally, replaced with a mouse click (e.g., instead of a contact) followed by movement of the cursor along the path of the swipe (e.g., instead of movement of the contact). As another example, a tap gesture is, optionally, replaced with a mouse click while the cursor is located over the location of the tap gesture (e.g., instead of detection of the contact followed by ceasing to detect the contact). Similarly, when multiple user inputs are simultaneously detected, it should be understood that multiple computer mice are, optionally, used simultaneously, or a mouse and finger contacts are, optionally, used simultaneously.
[0184] As used herein, the term "focus selector" refers to an input element that indicates a current part of a user interface with which a user is interacting. In some implementations that include a cursor or other location marker, the cursor acts as a "focus selector," so that when an input (e.g., a press input) is detected on a touch-sensitive surface (e.g., touchpad 355 in Fig.3A or touch-sensitive surface 451 in Fig.4B) while the cursor is over a particular user interface element (e.g., a button, window, slider or other user interface element), the particular user interface element is adjusted in accordance with the detected input. In some implementations that include a touch-screen display (e.g., touch-sensitive display system 112 in Fig.1A) that enables direct interaction with user interface elements on the touch-screen display, a detected contact on the touch-screen acts as a "focus selector," so that when an input (e.g., a press input by the contact) is detected on the touch-screen display at a location of a particular user interface element (e.g., a button, window, slider or other user interface element), the particular user interface element is adjusted in accordance with the detected input. In some implementations focus is moved from one region of a user interface to another region of the user interface without corresponding movement of a cursor or movement of a contact on a touch-screen display (e.g., by using a tab key or arrow keys to move focus from one button to another button); in these implementations, the focus selector moves in accordance with movement of focus between different regions of the user interface. Without regard to the specific form taken by the focus selector, the focus selector is generally the user interface element (or contact on a touch-screen display) that is controlled by the user so as to communicate the user's intended interaction with the user interface (e.g., by indicating, to the device, the element of the user interface with which the user is intending to interact). For example, the location of a focus selector (e.g., a cursor, a contact or a selection box) over a respective button while a press input is detected on the touch-sensitive surface (e.g., a touchpad or touch screen) will indicate that the user is intending to activate the -48- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) respective button (as opposed to other user interface elements shown on a display of the device).
[0185] As used in the specification and claims, the term "characteristic intensity" of a contact refers to a characteristic of the contact based on one or more intensities of the contact. In some embodiments, the characteristic intensity is based on multiple intensity samples. The characteristic intensity is, optionally, based on a predefined number of intensity samples, or a set of intensity samples collected during a predetermined time period (e.g., 0.05, 0.1, 0.2, 0.5, 1, 2, 5, 10 seconds) relative to a predefined event (e.g., after detecting the contact, prior to detecting liftoff of the contact, before or after detecting a start of movement of the contact, prior to detecting an end of the contact, before or after detecting an increase in intensity of the contact, and / or before or after detecting a decrease in intensity of the contact). A characteristic intensity of a contact is, optionally, based on one or more of: a maximum value of the intensities of the contact, a mean value of the intensities of the contact, an average value of the intensities of the contact, a top 10 percentile value of the intensities of the contact, a value at the half maximum of the intensities of the contact, a value at the 90 percent maximum of the intensities of the contact, or the like. In some embodiments, the duration of the contact is used in determining the characteristic intensity (e.g., when the characteristic intensity is an average of the intensity of the contact over time). In some embodiments, the characteristic intensity is compared to a set of one or more intensity thresholds to determine whether an operation has been performed by a user. For example, the set of one or more intensity thresholds optionally includes a first intensity threshold and a second intensity threshold. In this example, a contact with a characteristic intensity that does not exceed the first threshold results in a first operation, a contact with a characteristic intensity that exceeds the first intensity threshold and does not exceed the second intensity threshold results in a second operation, and a contact with a characteristic intensity that exceeds the second threshold results in a third operation. In some embodiments, a comparison between the characteristic intensity and one or more thresholds is used to determine whether or not to perform one or more operations (e.g., whether to perform a respective operation or forgo performing the respective operation), rather than being used to determine whether to perform a first operation or a second operation.
[0186] In some embodiments described herein, one or more operations are performed in response to detecting a gesture that includes a respective press input or in response to detecting the respective press input performed with a respective contact (or a plurality of -49- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) contacts), where the respective press input is detected based at least in part on detecting an increase in intensity of the contact (or plurality of contacts) above a press-input intensity threshold. In some embodiments, the respective operation is performed in response to detecting the increase in intensity of the respective contact above the press-input intensity threshold (e.g., a "down stroke" of the respective press input). In some embodiments, the press input includes an increase in intensity of the respective contact above the press-input intensity threshold and a subsequent decrease in intensity of the contact below the press-input intensity threshold, and the respective operation is performed in response to detecting the subsequent decrease in intensity of the respective contact below the press-input threshold (e.g., an "up stroke" of the respective press input).
[0187] In some embodiments, the device employs intensity hysteresis to avoid accidental inputs sometimes termed "jitter," where the device defines or selects a hysteresis intensity threshold with a predefined relationship to the press-input intensity threshold (e.g., the hysteresis intensity threshold is X intensity units lower than the press-input intensity threshold or the hysteresis intensity threshold is 75%, 90% or some reasonable proportion of the press-input intensity threshold). Thus, in some embodiments, the press input includes an increase in intensity of the respective contact above the press-input intensity threshold and a subsequent decrease in intensity of the contact below the hysteresis intensity threshold that corresponds to the press-input intensity threshold, and the respective operation is performed in response to detecting the subsequent decrease in intensity of the respective contact below the hysteresis intensity threshold (e.g., an "up stroke" of the respective press input). Similarly, in some embodiments, the press input is detected only when the device detects an increase in intensity of the contact from an intensity at or below the hysteresis intensity threshold to an intensity at or above the press-input intensity threshold and, optionally, a subsequent decrease in intensity of the contact to an intensity at or below the hysteresis intensity, and the respective operation is performed in response to detecting the press input (e.g., the increase in intensity of the contact or the decrease in intensity of the contact, depending on the circumstances).
[0188] For ease of explanation, the description of operations performed in response to a press input associated with a press-input intensity threshold or in response to a gesture including the press input are, optionally, triggered in response to detecting either: an increase in intensity of a contact above the press-input intensity threshold, an increase in intensity of a contact from an intensity below the hysteresis intensity threshold to an intensity above the -50- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) press-input intensity threshold, a decrease in intensity of the contact below the press-input intensity threshold, and / or a decrease in intensity of the contact below the hysteresis intensity threshold corresponding to the press-input intensity threshold. Additionally, in examples where an operation is described as being performed in response to detecting a decrease in intensity of a contact below the press-input intensity threshold, the operation is, optionally, performed in response to detecting a decrease in intensity of the contact below a hysteresis intensity threshold corresponding to, and lower than, the press-input intensity threshold.
[0189] Fig.5A illustrates a block diagram of an exemplary architecture for the device 500 according to some embodiments of the disclosure. In the embodiment of Fig.5A, media or other content is optionally received by device 500 via network interface 502, which is optionally a wireless or wired connection. The one or more processors 516 optionally execute any number of programs stored in memory 506 or storage, which optionally includes instructions to perform one or more of the methods and / or processes described herein (e.g., methods 700, 800, 1000, 1200, 1400, 1500, 1700, 1900, 2100, 2300, and / or 2700). A computer-readable storage medium can be any medium that can tangibly contain or store computer-executable instructions for use by or in connection with the instruction execution system, apparatus, or device. In some examples, the storage medium is a transitory computer-readable storage medium. In some examples, the storage medium is a non- transitory computer-readable storage medium. The non-transitory computer-readable storage medium can include, but is not limited to, magnetic, optical, and / or semiconductor storages. Examples of such storage include magnetic disks, optical discs based on CD, DVD, or Blu- ray technologies, as well as persistent solid-state memory such as flash, solid-state drives, and the like. Personal electronic device 500 is not limited to the components and configuration of Figs.5, but can include other or additional components in multiple configurations.
[0190] In addition, in methods described herein where one or more steps are contingent upon one or more conditions having been met, it should be understood that the described method can be repeated in multiple repetitions so that over the course of the repetitions all of the conditions upon which steps in the method are contingent have been met in different repetitions of the method. For example, if a method requires performing a first step if a condition is satisfied, and a second step if the condition is not satisfied, then a person of ordinary skill would appreciate that the claimed steps are repeated until the condition has been both satisfied and not satisfied, in no particular order. Thus, a method described with one or more steps that are contingent upon one or more conditions having been met could be -51- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) rewritten as a method that is repeated until each of the conditions described in the method has been met. This, however, is not required of system or computer readable medium claims where the system or computer readable medium contains instructions for performing the contingent operations based on the satisfaction of the corresponding one or more conditions and thus is capable of determining whether the contingency has or has not been satisfied without explicitly repeating steps of a method until all of the conditions upon which steps in the method are contingent have been met. A person having ordinary skill in the art would also understand that, similar to a method with contingent steps, a system or computer readable storage medium can repeat the steps of a method as many times as are needed to ensure that all of the contingent steps have been performed.
[0191] As used here, the term “affordance” refers to a user-interactive graphical user interface object that is, optionally, displayed on the display screen of devices 100, 300, and / or 500 (Figs.1A, 3, and 5A-5B). For example, an image (e.g., icon), a button, and text (e.g., hyperlink) each optionally constitute an affordance.
[0192] As used herein, the term “focus selector” refers to an input element that indicates a current part of a user interface with which a user is interacting. In some implementations that include a cursor or other location marker, the cursor acts as a “focus selector” so that when an input (e.g., a press input) is detected on a touch-sensitive surface (e.g., touchpad 355 in Fig.3A or touch-sensitive surface 451 in Fig.4B) while the cursor is over a particular user interface element (e.g., a button, window, slider, or other user interface element), the particular user interface element is adjusted in accordance with the detected input. In some implementations that include a touch screen display (e.g., touch-sensitive display system 112 in Fig.1A or touch screen 112 in Fig.4A) that enables direct interaction with user interface elements on the touch screen display, a detected contact on the touch screen acts as a “focus selector” so that when an input (e.g., a press input by the contact) is detected on the touch screen display at a location of a particular user interface element (e.g., a button, window, slider, or other user interface element), the particular user interface element is adjusted in accordance with the detected input. In some implementations, focus is moved from one region of a user interface to another region of the user interface without corresponding movement of a cursor or movement of a contact on a touch screen display (e.g., by using a tab key or arrow keys to move focus from one button to another button); in these implementations, the focus selector moves in accordance with movement of focus between different regions of the user interface. Without regard to the specific form taken by -52- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) the focus selector, the focus selector is generally the user interface element (or contact on a touch screen display) that is controlled by the user so as to communicate the user’s intended interaction with the user interface (e.g., by indicating, to the device, the element of the user interface with which the user is intending to interact). For example, the location of a focus selector (e.g., a cursor, a contact, or a selection box) over a respective button while a press input is detected on the touch-sensitive surface (e.g., a touchpad or touch screen) will indicate that the user is intending to activate the respective button (as opposed to other user interface elements shown on a display of the device).
[0193] As used in the specification and claims, the term “characteristic intensity” of a contact refers to a characteristic of the contact based on one or more intensities of the contact. In some embodiments, the characteristic intensity is based on multiple intensity samples. The characteristic intensity is, optionally, based on a predefined number of intensity samples, or a set of intensity samples collected during a predetermined time period (e.g., 0.05, 0.1, 0.2, 0.5, 1, 2, 5, 10 seconds) relative to a predefined event (e.g., after detecting the contact, prior to detecting liftoff of the contact, before or after detecting a start of movement of the contact, prior to detecting an end of the contact, before or after detecting an increase in intensity of the contact, and / or before or after detecting a decrease in intensity of the contact). A characteristic intensity of a contact is, optionally, based on one or more of: a maximum value of the intensities of the contact, a mean value of the intensities of the contact, an average value of the intensities of the contact, a top 10 percentile value of the intensities of the contact, a value at the half maximum of the intensities of the contact, a value at the 90 percent maximum of the intensities of the contact, or the like. In some embodiments, the duration of the contact is used in determining the characteristic intensity (e.g., when the characteristic intensity is an average of the intensity of the contact over time). In some embodiments, the characteristic intensity is compared to a set of one or more intensity thresholds to determine whether an operation has been performed by a user. For example, the set of one or more intensity thresholds optionally includes a first intensity threshold and a second intensity threshold. In this example, a contact with a characteristic intensity that does not exceed the first threshold results in a first operation, a contact with a characteristic intensity that exceeds the first intensity threshold and does not exceed the second intensity threshold results in a second operation, and a contact with a characteristic intensity that exceeds the second threshold results in a third operation. In some embodiments, a comparison between the characteristic intensity and one or more thresholds is used to determine whether or not to -53- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) perform one or more operations (e.g., whether to perform a respective operation or forgo performing the respective operation), rather than being used to determine whether to perform a first operation or a second operation.
[0194] Fig.5C illustrates detecting a plurality of contacts 552A-552E on touch- sensitive display screen 504 with a plurality of intensity sensors 524A-524D. Fig.5C additionally includes intensity diagrams that show the current intensity measurements of the intensity sensors 524A-524D relative to units of intensity. In this example, the intensity measurements of intensity sensors 524A and 524D are each 9 units of intensity, and the intensity measurements of intensity sensors 524B and 524C are each 7 units of intensity. In some implementations, an aggregate intensity is the sum of the intensity measurements of the plurality of intensity sensors 524A-524D, which in this example is 32 intensity units. In some embodiments, each contact is assigned a respective intensity that is a portion of the aggregate intensity. Fig.5D illustrates assigning the aggregate intensity to contacts 552A- 552E based on their distance from the center of force 554. In this example, each of contacts 552A, 552B, and 552E are assigned an intensity of contact of 8 intensity units of the aggregate intensity, and each of contacts 552C and 552D are assigned an intensity of contact of 4 intensity units of the aggregate intensity. More generally, in some implementations, each contact j is assigned a respective intensity Ij that is a portion of the aggregate intensity, A, in accordance with a predefined mathematical function, Ij = A^(Dj / ΣDi), where Dj is the distance of the respective contact j to the center of force, and ΣDi is the sum of the distances of all the respective contacts (e.g., i=1 to last) to the center of force. The operations described with reference to Figs.5C-5D can be performed using an electronic device similar or identical to device 100, 300, or 500. In some embodiments, a characteristic intensity of a contact is based on one or more intensities of the contact. In some embodiments, the intensity sensors are used to determine a single characteristic intensity (e.g., a single characteristic intensity of a single contact). It should be noted that the intensity diagrams are not part of a displayed user interface, but are included in Figs.5C-5D to aid the reader.
[0195] In some embodiments, a portion of a gesture is identified for purposes of determining a characteristic intensity. For example, a touch-sensitive surface optionally receives a continuous swipe contact transitioning from a start location and reaching an end location, at which point the intensity of the contact increases. In this example, the characteristic intensity of the contact at the end location is, optionally, based on only a portion of the continuous swipe contact, and not the entire swipe contact (e.g., only the -54- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) portion of the swipe contact at the end location). In some embodiments, a smoothing algorithm is, optionally, applied to the intensities of the swipe contact prior to determining the characteristic intensity of the contact. For example, the smoothing algorithm optionally includes one or more of: an unweighted sliding-average smoothing algorithm, a triangular smoothing algorithm, a median filter smoothing algorithm, and / or an exponential smoothing algorithm. In some circumstances, these smoothing algorithms eliminate narrow spikes or dips in the intensities of the swipe contact for purposes of determining a characteristic intensity.
[0196] The intensity of a contact on the touch-sensitive surface is, optionally, characterized relative to one or more intensity thresholds, such as a contact-detection intensity threshold, a light press intensity threshold, a deep press intensity threshold, and / or one or more other intensity thresholds. In some embodiments, the light press intensity threshold corresponds to an intensity at which the device will perform operations typically associated with clicking a button of a physical mouse or a trackpad. In some embodiments, the deep press intensity threshold corresponds to an intensity at which the device will perform operations that are different from operations typically associated with clicking a button of a physical mouse or a trackpad. In some embodiments, when a contact is detected with a characteristic intensity below the light press intensity threshold (e.g., and above a nominal contact-detection intensity threshold below which the contact is no longer detected), the device will move a focus selector in accordance with movement of the contact on the touch- sensitive surface without performing an operation associated with the light press intensity threshold or the deep press intensity threshold. Generally, unless otherwise stated, these intensity thresholds are consistent between different sets of user interface figures.
[0197] An increase of characteristic intensity of the contact from an intensity below the light press intensity threshold to an intensity between the light press intensity threshold and the deep press intensity threshold is sometimes referred to as a “light press” input. An increase of characteristic intensity of the contact from an intensity below the deep press intensity threshold to an intensity above the deep press intensity threshold is sometimes referred to as a “deep press” input. An increase of characteristic intensity of the contact from an intensity below the contact-detection intensity threshold to an intensity between the contact-detection intensity threshold and the light press intensity threshold is sometimes referred to as detecting the contact on the touch-surface. A decrease of characteristic intensity of the contact from an intensity above the contact-detection intensity threshold to an -55- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) intensity below the contact-detection intensity threshold is sometimes referred to as detecting liftoff of the contact from the touch-surface. In some embodiments, the contact-detection intensity threshold is zero. In some embodiments, the contact-detection intensity threshold is greater than zero.
[0198] In some embodiments described herein, one or more operations are performed in response to detecting a gesture that includes a respective press input or in response to detecting the respective press input performed with a respective contact (or a plurality of contacts), where the respective press input is detected based at least in part on detecting an increase in intensity of the contact (or plurality of contacts) above a press-input intensity threshold. In some embodiments, the respective operation is performed in response to detecting the increase in intensity of the respective contact above the press-input intensity threshold (e.g., a “down stroke” of the respective press input). In some embodiments, the press input includes an increase in intensity of the respective contact above the press-input intensity threshold and a subsequent decrease in intensity of the contact below the press-input intensity threshold, and the respective operation is performed in response to detecting the subsequent decrease in intensity of the respective contact below the press-input threshold (e.g., an “up stroke” of the respective press input).
[0199] Figs.5E-5H illustrate detection of a gesture that includes a press input that corresponds to an increase in intensity of a contact 562 from an intensity below a light press intensity threshold (e.g., “ITL”) in Fig.5E, to an intensity above a deep press intensity threshold (e.g., “ITD”) in Fig.5H. The gesture performed with contact 562 is detected on touch-sensitive surface 560 while cursor 576 is displayed over application icon 572B corresponding to App 2, on a displayed user interface 570 that includes application icons 572A-572D displayed in predefined region 574. In some embodiments, the gesture is detected on touch-sensitive display 504. The intensity sensors detect the intensity of contacts on touch-sensitive surface 560. The device determines that the intensity of contact 562 peaked above the deep press intensity threshold (e.g., “ITD”). Contact 562 is maintained on touch-sensitive surface 560. In response to the detection of the gesture, and in accordance with contact 562 having an intensity that goes above the deep press intensity threshold (e.g., “ITD”) during the gesture, reduced-scale representations 578A-578C (e.g., thumbnails) of recently opened documents for App 2 are displayed, as shown in Figs.5F-5H. In some embodiments, the intensity, which is compared to the one or more intensity thresholds, is the -56- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) characteristic intensity of a contact. It should be noted that the intensity diagram for contact 562 is not part of a displayed user interface, but is included in Figs.5E-5H to aid the reader.
[0200] In some embodiments, the display of representations 578A-578C includes an animation. For example, representation 578A is initially displayed in proximity of application icon 572B, as shown in Fig.5F. As the animation proceeds, representation 578A moves upward and representation 578B is displayed in proximity of application icon 572B, as shown in Fig.5G. Then, representations 578A moves upward, 578B moves upward toward representation 578A, and representation 578C is displayed in proximity of application icon 572B, as shown in Fig.5H. Representations 578A-578C form an array above icon 572B. In some embodiments, the animation progresses in accordance with an intensity of contact 562, as shown in Figs.5F-5G, where the representations 578A-578C appear and move upwards as the intensity of contact 562 increases toward the deep press intensity threshold (e.g., “ITD”). In some embodiments, the intensity, on which the progress of the animation is based, is the characteristic intensity of the contact. The operations described with reference to Figs.5E- 5H can be performed using an electronic device similar or identical to device 100, 300, or 500.
[0201] In some embodiments, the device employs intensity hysteresis to avoid accidental inputs sometimes termed “jitter,” where the device defines or selects a hysteresis intensity threshold with a predefined relationship to the press-input intensity threshold (e.g., the hysteresis intensity threshold is X intensity units lower than the press-input intensity threshold or the hysteresis intensity threshold is 75%, 90%, or some reasonable proportion of the press-input intensity threshold). Thus, in some embodiments, the press input includes an increase in intensity of the respective contact above the press-input intensity threshold and a subsequent decrease in intensity of the contact below the hysteresis intensity threshold that corresponds to the press-input intensity threshold, and the respective operation is performed in response to detecting the subsequent decrease in intensity of the respective contact below the hysteresis intensity threshold (e.g., an “up stroke” of the respective press input). Similarly, in some embodiments, the press input is detected only when the device detects an increase in intensity of the contact from an intensity at or below the hysteresis intensity threshold to an intensity at or above the press-input intensity threshold and, optionally, a subsequent decrease in intensity of the contact to an intensity at or below the hysteresis intensity, and the respective operation is performed in response to detecting the press input -57- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) (e.g., the increase in intensity of the contact or the decrease in intensity of the contact, depending on the circumstances).
[0202] For ease of explanation, the descriptions of operations performed in response to a press input associated with a press-input intensity threshold or in response to a gesture including the press input are, optionally, triggered in response to detecting either: an increase in intensity of a contact above the press-input intensity threshold, an increase in intensity of a contact from an intensity below the hysteresis intensity threshold to an intensity above the press-input intensity threshold, a decrease in intensity of the contact below the press-input intensity threshold, and / or a decrease in intensity of the contact below the hysteresis intensity threshold corresponding to the press-input intensity threshold. Additionally, in examples where an operation is described as being performed in response to detecting a decrease in intensity of a contact below the press-input intensity threshold, the operation is, optionally, performed in response to detecting a decrease in intensity of the contact below a hysteresis intensity threshold corresponding to, and lower than, the press-input intensity threshold.
[0203] As used herein, an “installed application” refers to a software application that has been downloaded onto an electronic device (e.g., devices 100, 300, and / or 500) and is ready to be launched (e.g., become opened) on the device. In some embodiments, a downloaded application becomes an installed application by way of an installation program that extracts program portions from a downloaded package and integrates the extracted portions with the operating system of the computer system.
[0204] As used herein, the terms “open application” or “executing application” refer to a software application with retained state information (e.g., as part of device / global internal state 157 and / or application internal state 192). An open or executing application is, optionally, any one of the following types of applications: ^ an active application, which is currently displayed on a display screen of the device that the application is being used on; ^ a background application (or background processes), which is not currently displayed, but one or more processes for the application are being processed by one or more processors; and ^ a suspended or hibernated application, which is not running, but has state information that is stored in memory (volatile and non-volatile, respectively) and that can be used to resume execution of the application. -58- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0205] As used herein, the term “closed application” refers to software applications without retained state information (e.g., state information for closed applications is not stored in a memory of the device). Accordingly, closing an application includes stopping and / or removing application processes for the application and removing state information for the application from the memory of the device. Generally, opening a second application while in a first application does not close the first application. When the second application is displayed and the first application ceases to be displayed, the first application becomes a background application.
[0206] Attention is now directed towards embodiments of user interfaces (“UI”) and associated processes that are implemented on an electronic device, such as device 100, device 300, or device 500. USER INTERFACES
[0207] As described herein, content is automatically generated by one or more computers in response to a request to generate the content. The automatically-generated content is optionally generated on-device (e.g., generated at least in part by a computer system at which a request to generate the content is received) and / or generated off-device (e.g., generated at least in part by one or more nearby computers that are available via a local network or one or more computers that are available via the internet). This automatically- generated content optionally includes visual content (e.g., images, graphics, and / or video), audio content, and / or text content.
[0208] In some embodiments, novel automatically-generated content that is generated via one or more artificial intelligence (AI) processes is referred to as generative content (e.g., generative images, generative graphics, generative video, generative audio, and / or generative text). Generative content is typically generated by an AI process based on a prompt that is provided to the AI process. An AI process typically uses one or more AI models to generate an output based on an input. An AI process optionally includes one or more pre-processing steps to adjust the input before it is used by the AI model to generate an output (e.g., adjustment to a user-provided prompt, creation of a system-generated prompt, and / or AI model selection). An AI process optionally includes one or more post-processing steps to adjust the output by the AI model (e.g., passing AI model output to a different AI model, upscaling, downscaling, cropping, formatting, and / or adding or removing metadata) before the output of the AI model used for other purposes such as being provided to a different -59- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) software process for further processing or being presented (e.g., visually or audibly) to a user. An AI process that generates generative content is sometimes referred to as a generative AI process.
[0209] A prompt for generating generative content can include one or more of: one or more words (e.g., a natural language prompt that is written or spoken), one or more images, one or more drawings, and / or one or more videos. AI processes can include machine learning models including neural networks. Neural networks can include transformer-based deep neural networks such as large language models (LLMs). Generative pre-trained transformer models are a type of LLM that can be effective at generating novel generative content based on a prompt. Some AI processes use a prompt that includes text to generate either different generative text, generative audio content, and / or generative visual content. Some AI processes use a prompt that includes visual content and / or an audio content to generate generative text (e.g., a transcription of audio and / or a description of the visual content). Some multi-modal AI processes use a prompt that includes multiple types of content (e.g., text, images, audio, video, and / or other sensor data) to generate generative content. A prompt sometimes also includes values for one or more parameters indicating an importance of various parts of the prompt. Some prompts include a structured set of instructions that can be understood by an AI process that include phrasing, a specified style, relevant context (e.g., starting point content and / or one or more examples), and / or a role for the AI process.
[0210] Generative content is generally based on the prompt but is not deterministically selected from pre-generated content and is, instead, generated using the prompt as a starting point. In some embodiments, pre-existing content (e.g., audio, text, and / or visual content) is used as part of the prompt for creating generative content (e.g., the pre-existing content is used as a starting point for creating the generative content). For example, a prompt could request that a block of text be summarized or rewritten in a different tone, and the output would be generative text that is summarized or written in the different tone. Similarly a prompt could request that visual content be modified to include or exclude content specified by a prompt (e.g., removing an identified feature in the visual content, adding a feature to the visual content that is described in a prompt, changing a visual style of the visual content, and / or creating additional visual elements outside of a spatial or temporal boundary of the visual content that are based on the visual content). In some embodiments, a random or pseudo-random seed is used as part of the prompt for creating generative content (e.g., the random or pseud-random seed content is used as a starting point for creating the -60- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) generative content). For example when generating an image from a diffusion model, a random noise pattern is iteratively denoised based on the prompt to generate an image that is based on the prompt. While specific types of AI processes have been described herein, it should be understood that a variety of different AI processes could be used to generate generative content based on a prompt.
[0211] Users interact with electronic devices in many different manners. In some embodiments, an electronic device is in communication with one or more input devices, a display generation component, and wireless circuitry. In some embodiments, the electronic device presents a user interface that receives a prompt to be used to influence generation of an automatically-generated visual media, such as described in more detail with reference to method 700. In some embodiments, automatically-generated visual media is media generated using autonomous processes. In some embodiments, generating the automatically-generated visual media includes using a real image as a starting point to build in additional concepts and / or features found in the prompt, as described in method 700. In some embodiments, a real image is an image captured of the real world such as an image captured by a camera or user created digital image or a modified version of a real-world image or user created digital image. In some embodiments, the real image is not a previously generated automatically- generated visual content. The embodiments described below provide ways in which the electronic device receives and displays concepts to be used to influence the generation of the automatically-generated visual media. The embodiments described below also provide ways in which the electronic device generates (e.g., using an AI process or a generative AI process) the automatically-generated visual media. Displaying representations of recognized concepts of a prompt allows a user to easily and efficiently see the concepts used to generate the automatically-generated visual media, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors. It is understood that people use devices. When a person uses a device, that person is optionally referred to as a user of the device.
[0212] In some embodiments, an input that is described as a contact is one or a tap input, a multi-tap input, or a long press input. In some embodiments, an input that is described as a selection input is a button press, air tap, tap input, multi-tap or a long press input, or other confirmation input that is detected while attention is directed to a corresponding user interface element. -61- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0213] Figs.6A-6MM illustrate exemplary ways in which an electronic device displays recognized concepts and generates automatically-generated visual media. The embodiments in these figures are used to illustrate the processes described below, including the processes described with reference to Fig.7 and Fig.8. Although Figs.6A-6MM illustrate various examples of ways an electronic device is able to perform the processes described below with respect to Fig.7 and Fig.8, it should be understood that these examples are not meant to be limiting, and the electronic device is able to perform one or more processes described below with reference to Fig.7 and Fig.8 in ways not expressly described with reference to Figs.6A-6MM.
[0214] In some embodiments, the user interface of Figs.6A-6MM corresponds to the user interface of Figs.16A-16U, 18A-18BB, 20D-20AA, 22A-22CCC, 24A-24E, and / or 26A-26P, and user interface elements with corresponding shapes, text, and / or glyphs have the same or similar functions, for example a “cancel” selectable user interface object illustrated in one figure optionally has some or all of the same functions as a “cancel” selectable user interface object in another figure. Similarly, a selectable user interface object with a “+” glyph has some or all of the same functions as a selectable user interface object with a “+” glyph in another figure.
[0215] Fig.6A-6F illustrates embodiments in which the electronic device generates (e.g., using an AI process or a generative AI process) an automatically-generated visual content using the automatically-generated visual media application. Fig.6A illustrates an electronic device 500 with a display generation component 504 (e.g., a touchscreen). In some embodiments, the electronic device 500 is a mobile device, such as a smartphone, tablet, or wearable device. In Fig.6A, the electronic device 500 displays user interface 600 that includes a gallery of automatically-generated visual content. The gallery of automatically- generated visual content includes automatically-generated visual content 602a through 602f. In some embodiments, the visual content 602a through 602f represent sample automatically- generated visual content. For example, the electronic device 500 provides examples of automatically-generated visual content in user interface 600 before the user has generated automatically-generated visual content. In Fig.6A, user interface 600 also includes a selectable option 610 that, when selected, causes the electronic device 500 to start the process to generate an automatically-generated visual media.
[0216] While displaying user interface 600 in Fig.6A, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 606 (e.g., a touch -62- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) input using a finger) directed towards selectable option 610. As a result, the electronic device 500 displays user interface 604, shown in Fig.6B. Fig.6B illustrates a user interface that receives a prompt, starting media, and / or recognized concepts for use in generating automatically-generated visual content. Examples of prompts, recognized concepts, and starting media are described in further detail in method 700. Although touch inputs including contacts are used in the embodiments described here as example inputs, other inputs are possible including voice, hardware input device inputs, and / or air gesture inputs.
[0217] In Fig.6B, the user interface 604 includes a visual indication 612 including text describing the generative visual content process. For example, to create an automatically- generated visual content, the user first needs to input a starting image and / or a prompt and / or recognized concept. In Fig.6B, user interface 604 also includes selectable options 608a through 608d, which, when selected, cause the electronic device 500 to add recognized concepts (e.g., including starting media) to an automatically-generated visual content. In some embodiments, selectable options 608a through 608b are suggested by the electronic device (e.g., at random or they are frequently used recognized concepts). In some embodiments, selectable options 608a through 608d include text describing the recognized concept and icons including visual representations of the recognized concept. For example, selectable option 608a includes a representation of “Jenna” and the text “Jenna”, selectable option 608b includes an icon of snowflakes and text “Snow”, selectable option 608c includes an icon of a sled and text “Sledding”, and selectable option 608d (e.g., representing a style recognized concept, as described in method 700) includes an icon of a person dancing and text “Animation”. Additionally, in Fig.6B, the user interface 604 includes a selectable option 616 that, when selected, causes the electronic device 500 to open a camera application or a photo library to select starting media, and a text input region 614 for receiving a prompt, both of which are described in further detail in method 700.
[0218] While displaying user interface 604 in Fig.6B, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 618 (e.g., a touch input using a finger or a gaze input using a user’s eyes) directed towards selectable option 608a, which is a starting media recognized concept of “Jenna”. In some embodiments, the electronic device has one or more images of Jenna (e.g., in a photos library or other content application), which were used to create the representation of Jenna shown in selectable option 608a. As a result of detecting the input, the electronic device displays “Jenna” as a -63- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) recognized concept in user interface 604, as shown in Fig.6C, which is now being used to influence the generation of the automatically-generated visual content.
[0219] Fig.6C illustrates user interface 604 after the user has initiated the process to generate an automatically-generated visual content (e.g., by selecting starting media). As described in method 700, the user optionally selects starting media using the suggestions (shown in Fig.6B) of recognized concepts and / or by choosing an image from a camera application and / or a photo library (e.g., by receiving an input directed towards selectable option 616). After initiating the process, the electronic device 500 displays the starting media as a recognized concept 620a, shown in Fig.6C. After initiating the process, the electronic device 500 also displays a representation 622a of the automatically-generated visual content. The representation of the automatically-generated visual content is described in further detail in method 700. The representation of the automatically-generated visual content is a preview of the automatically-generated visual content. The user interface 604 also includes selectable options 624a through 624b. In some embodiments, selectable option 624a, when selected, causes the electronic device 500 to end (e.g., cancel) the process to generate an automatically-generated visual content, and selectable option 624b, when selected, causes the electronic device 500 to generate the automatically-generated visual content, as described below. Selectable option 624a optionally corresponds to selectable option 1602a, 1802a, and / or 2202a in Figs.16A-16U, 18A-18BB, 20D-20AA, 22A-22CCC, and / or 26A-26P. Selectable option 624b optionally corresponds to selectable option 1602b, 1802b, and / or 2202b in Figs.16A-16U, 18A-18BB, 20D-20AA, 22A-22CCC, and / or 26A-26P.
[0220] While displaying the user interface 604 in Fig.6C, the electronic device 500 receives a swipe input including contact 626 (e.g., a drag motion using a finger in contact with the touch screen) directed towards the recognized concept suggestions. In response to receiving the input, the electronic device displays additional recognized concepts suggestions (e.g., selectable options 608e through 608f), shown in Fig.6D.
[0221] While displaying the additional selectable options 608e through 608f, the electronic device 500 receives a drag input using contacts 628a through 628b (e.g., a dragging motion using a finger remaining in contact with the touch screen) directed towards dragging selectable options 608d and 608e, respectively, to be used to influence generation of the automatically-generated visual content. As a result, the electronic device 500 displays selectable options 608d and 608e as recognized concepts 620b and 620c, respectively, in Fig. -64- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 6E. Alternatively, in some embodiments, the electronic device 500 receives a selection input towards selectable option 608d and 608e to add them as recognized concepts.
[0222] Fig.6E illustrates user interface 604 including recognized concepts 620a through 620c. In response to receiving the input in Fig.6D to add selectable options 608d and 608e as recognized concepts, the electronic device 500 updates the representation 622a of the automatically-generated visual content to include the additional recognized concepts, as described further in method 700. In some embodiments, in response to receiving the input selecting the option 608d, the electronic device 500 updates the user interface 604 including representation 622a to include the recognized concept associated with option 608d, and in response to receiving the input selecting the option 608e, the electronic device 500 updates the user interface 604 including representation 622a to include the recognized concept associated with option 608e. In some embodiments, the boundary of the representation 622a is animated, as described in method 700. As shown in Fig.6E, the boundary of the representation 622 is blob shape, and in some embodiments, the shape is animated. Additionally, in response to receiving the input in Fig.6D, the electronic device replaces the selectable options 608d and 608e with additional selectable options 608g and 608h corresponding to recognized concepts different from those corresponding to options 608d and 608e, as shown in Fig.6E.
[0223] While displaying the user interface 604 in Fig.6E, the electronic device 500 receives a selection input, (e.g., a tap or a long press) including contact 626, directed towards selectable option 624b, that when selected causes the electronic device to generate the automatically-generated visual content. In response to receiving the input in Fig.6E, the electronic device 500 ceases to display user interface 604 and displays user interface 629a, shown in Fig.6F. In some embodiments, user interface 629a includes the automatically- generated visual content that is optionally higher fidelity than the representation 622a, described in further detail in method 700.
[0224] Fig.6F is a user interface 629a including the automatically-generated visual content that is generated using autonomous processes, as described in method 700. In some embodiments, the automatically-generated visual content includes the recognized concepts 620b through 620c incorporated into the starting media (e.g., recognized concept 620a) to formulate a new image (e.g., the automatically-generated visual content). For example, the electronic device 500 uses the representation of Jenna as the starting image and incorporates the recognized concepts of “birthday” and “animation” in the automatically-generated visual -65- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) content in user interface 629a. As shown in user interface 629a, Jenna is having a birthday celebration with a banner and a birthday cake. In some embodiments, the automatically- generated visual content in Fig.6F is animated. For example, the candles are flickering, Jenna is shown blowing out the candles, and / or the “happy birthday” banner is moving and includes changing colors. As shown in Fig.6F, user interface 629a includes selectable options 630a through 630c. Options 630a through 630c optionally correspond to options 1602c and / or 2242a; options 1602d and / or 2242b; and options 1602e and 2242c, respectfully, in Figs.9A- 9X, 16A-16U, 18G-18BB, 20D-20AA, 22A-22CCC, and / or 26A-26P. In some embodiments, the electronic device 500 redisplays user interface 604 in Fig.6E in response to the electronic device receiving an input directed towards selectable option 630a. In some embodiments, the electronic device 500 saves the automatically-generated visual content (e.g., in the gallery shown by user interface 600 in Fig.6A) in response to the electronic device receiving an input directed towards selectable option 630b. Actions in response to the electronic device receiving an input directed towards selectable option 630b are described in further detail below. For example, the user optionally saves the image as a new image or overrides a previous image, as described in Fig.6Q. In some embodiments, the electronic device 500 transmits the automatically-generated visual content to a different application and / or a different device in response to the electronic device receiving an input directed towards selectable option 630c.
[0225] Fig.6G illustrates an embodiment wherein a user inputs a prompt to initiate a process to generate an automatically-generated visual content using a voice assistant (e.g., and not through user interface 600 shown in Fig.6A). As shown in Fig.6G, a user activates a voice assistant (e.g., using a voice command 634 and / or a button) while the electronic device displays user interface 632 (e.g., a lock screen user interface) and tells the voice assistant to generate an image. In some embodiments, the text illustrates the voice command 634 to the voice assistant.
[0226] In response to receiving the voice command 634, the electronic device 500 displays user interface 604, described in greater detail above, including the details from the voice command 634. Fig.6H-A illustrates a first embodiment of a user interface 604 including a recognized concept 620d representing “Jeremy” and a prompt 636 in text entry region 614. In some embodiments, the recognized concept 620d is a representation of “Jeremy” including one or more characteristics of the recognized concept 620a, as described above. Fig.6H-A also includes recognized concepts 620e and 620f, which are recognized -66- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) concepts extracted from prompt 636. In some embodiments, extracting recognized concepts from a prompt is described in further detail in method 700. Fig.6H-A also includes a representation 622b of the automatically-generated visual content, which as one or more characteristics of the representation 622a of another automatically-generated visual content, as described above.
[0227] Fig.6H-B illustrates a second embodiment of user interface 604. In some embodiments, Fig.6H-B includes one or more characteristics of Fig.6H-A. In response to receiving the voice command 634 (e.g., in some embodiments, the electronic device 500 receives a prompt and / or an additional prompt via text as described herein), the electronic device 500 displays recognized concept 620d representing “Jeremy” and a cluster of recognized concepts 690 representing the prompt. As described in method 700, in some embodiments, the electronic device 500 displays a cluster of icons and / or text in response to receiving a prompt with multiple recognized concepts.
[0228] While displaying the recognized concepts 620d through 620f in Fig.6H-A, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 640 directed towards a selectable option 638 in Fig.6H-A. In response to receiving the input, the electronic device 500 displays a menu 642 including a collection of recognized concepts that, when selected, influence the creation of an automatically-generated visual content, as shown in FIG.6I. In some embodiments, the collection of recognized concepts is described in further detail in the description of method 700 below.
[0229] FIG.6I illustrates the menu 642 of recognized concepts including recognized concepts 644a through 644f. In some embodiments, the menu 642 of recognized concepts are categorized by type of recognized concept. For example, recognized concepts 644a through 644f are activity recognized concepts (e.g., the recognized concepts 644a through 644f in this category are activities). In some embodiments, the menu 642 includes a plurality of categories of recognized concepts, as described in method 700. As shown in Fig.6I, the menu 642 includes indication 646, indicating the number of categories in the menu 642, and the location of the category currently displayed.
[0230] While displaying the menu 642 of recognized concepts including recognized concepts 644a through 644f in Fig.6I, the electronic device 500 receives a dragging input including movement of contact 648 directed towards dragging recognized concept 644b towards the region of user interface 604 including the recognized concepts being used to -67- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) influence the generation of the automatically-generated visual content. In response to receiving the input, the electronic device 500 updates the representation 622b of the automatically-generated visual content to include recognized concept 644b (e.g., “tennis”), shown in Fig.6J. For example, in Fig.6J, the electronic device 500 updates representation 622b to include a tennis racket. Additionally, the electronic device 500 updates the recognized concepts used to influence the generation of the automatically-generated visual content to include recognized concept 620g which corresponds to the selected recognized concept 644b (e.g., “tennis”).
[0231] In some embodiments, while displaying user interface 604 including the menu 642 in Fig.6I, the electronic device 500 receives a selection input directed towards option 624a, such as detecting contact 658a (e.g., a tap or long press input). In some embodiments, other selection inputs, such as an air gesture input, a voice input, and / or an input detected using a hardware input device, are possible. In some embodiments, the input is in place of or in addition to the input including contact 648 described above. In response to receiving the input directed towards option 624a, the electronic device 500 removes the recognized concepts 620d, 620f, and 620e from the prompt and ceases displaying the representations of recognized concepts 620d, 620f, and 620e. In response to removing the recognized concepts used to influence the generation of the automatically-generated visual media, the electronic device 500 displays user interface 604 with visual indication 612, as shown in Fig.6B. In some embodiments, the electronic device displays visual indication 612 because there are no longer any recognized concepts to be used to influence the generation of the automatically- generated visual content item.
[0232] In Fig.6J, the electronic device 500 receives a swipe input including contact 650 in the menu 642 region of the user interface 604. In response to receiving the input in Fig.6J, the electronic device 500 updates the menu 642 to include a display of a second category (e.g., “Styles”) of recognized concepts 644g through 644l, shown in Fig.6K. In some embodiments, the second category includes styles that the automatically-generated visual content is generated with. In some embodiments, in response to receiving the swipe input including movement of contact 650, the electronic device 500 also updates indication 646 to indicate the location of the second category (e.g., to the right of the activities category). While displaying the styles category, the electronic device 500 receives a dragging input including movement of contact 652 directed towards dragging recognized concept 644l (e.g., “sketch”) to the region of the user interface 604 including recognized concepts used to -68- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) influence the generation of the automatically-generated visual content. In response to receiving the input, the electronic device 500 updates the region of the user interface 604 to include recognized concept 620h, which corresponds to recognized concept 644i in the menu 642 of recognized concepts, in Fig.6L. Additionally, in response to receiving the dragging input including movement of contact 652, the electronic device 500 updates the representation 622b, in Fig.6L, to include a sketch style.
[0233] In Fig.6L, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 654 directed towards selectable option 624b. In response to receiving the selection input, the electronic device 500 ceases displaying user interface 604 and begins displaying user interface 629b, shown in Fig.6M. Fig.6M illustrates user interface 629b including the automatically-generated visual content generated using the recognized concepts 620d through 620h. For example, the automatically-generated visual content includes a sketch of Jeremy at the beach playing tennis with sunglasses on. In some embodiments, user interface 629b includes one or more characteristics of user interface 629a in Fig.6F. As described above and in method 700, the automatically-generated visual content is of higher fidelity than the representation 622b. For example, the automatically-generated visual content in Fig.6M includes more details (e.g., a tennis net, a tennis ball, an umbrella, and the sun) than the representation 622 illustrated in Fig.6L. As described with reference to method 700, the automatically-generated visual content is not visual media generated as a result of image manipulation, but instead the one or more recognized concepts (e.g., 620d through 620h) are used as seeds to modify a real-world image into an automatically- generated visual content.
[0234] In Fig.6M, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 656 directed towards selectable option 630a. In response to receiving the input, the electronic device 500 ceases displaying user interface 629b and redisplays user interface 604, including the recognized concepts (e.g., recognized concepts 620d through 620h) used to influence the generation of the automatically-generated visual content. Fig.6N illustrates the electronic device 500 displaying user interface 604 in response to the electronic device 500 receiving the selection input (e.g., a tap or long press input) including contact 656 in Fig.6M.
[0235] While displaying user interface 604 in Fig.6N, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 658 corresponding to a request to delete recognized concept 620e such that the recognized concept 620e no -69- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) longer influences the generation of the automatically-generated visual content. In some embodiments, the selection input is directed towards a selectable option 659 overlaid on the recognized concept 620e. In some embodiments, the recognized concepts used to influence the generation of the automatically-generated visual content include corresponding selectable options that, when selected, cause the electronic device 500 to delete the recognized concept corresponding to the selected option. In response to receiving the selection input in Fig.6N, the electronic device 500 ceases to display recognized concept 620e, as shown in Fig.6O. Additionally, in response to receiving the selection input, the electronic device 500 updates representation 622b, in Fig.6O, to reflect that the recognized concept 620e is no longer being used to influence the generation of the automatically-generated visual content. For example, in Fig.6N, the electronic device 500 receives an input corresponding to a request to delete the “sunglasses” recognized concept, and in response to receiving the input, the electronic device 500 updates representation 622b in Fig.6O to no longer include Jeremy wearing sunglasses.
[0236] In some embodiments, while displaying user interface 604 in Fig.6N, the electronic device 500 receives a selection input directed towards option 624a, such as detecting a tap input or long press input with contact 658b. In some embodiments, the input is in place of or in addition to the input including contact 658a described above. In some embodiments, other selection inputs, such as an air gesture input, a voice input, and / or an input detected using a hardware input device, are possible. In response to receiving the input directed towards option 624a, the electronic device 500 removes the recognized concepts 620d, 620f, 620h, 620e, and 620g. In response to removing the recognized concepts used to influence the generation of the automatically-generated visual media. the electronic device 500 displays user interface 604 with visual indication 612, as shown in Fig.6B, and no longer displays the recognized concepts 620d, 620f, 620h, 620e, and 620g. In some embodiments, the electronic device displays visual indication 612 because there are no longer any recognized concepts to be used to influence the generation of the automatically-generated visual content item.
[0237] In Fig.6O, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 660 directed towards selectable option 624b, to regenerate the automatically-generated visual content without recognized concept 620e. As a result, the electronic device 500 ceases to display user interface 604 and displays user interface 629c to include a regenerated version of the automatically-generated visual content not including recognized concept 620e (e.g., the automatically-generated visual content shown in Fig.6M). -70- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) Fig.6P illustrates the regenerated automatically-generated visual content when the process to generate automatically-generated visual content is deterministic. Specifically, Fig.6P is different from Fig.6O because the automatically-generated visual content in Fig.6P lacks sunglasses. However, in some embodiments, and as described in method 700, the generation of the automatically-generated visual content is non-deterministic. Specifically, one or more elements in addition to the removal of the sunglasses are changed in the resulting regenerated automatically-generated visual content. Figs.9D-9G describe regenerating a non- deterministic automatically-generated visual content. In some embodiments, user interface 629c has one or more characteristics of the user interfaces 629a through 629b described above. While displaying the regenerated automatically-generated visual content in Fig.6P, the electronic device 500 receives a selection input (e.g., a tap or a long press) using contact 662 directed towards selectable option 630c. In response, the electronic device 500 displays selectable options 630d through 630e, shown in Fig.6Q. In some embodiments, selectable option 630d, when selected, causes the electronic device 500 to save the regenerated automatically-generated visual content over the previously generated automatically-generated visual content (e.g., the automatically-generated visual content shown in Fig.6M), which is also described in greater detail in method 800. In some embodiments, selectable option 630e, when selected, causes the electronic device 500 to save the regenerated automatically- generated visual content as a new automatically-generated visual content, described in greater detail in method 800. In Fig.6Q, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 664 directed towards selectable option 630d. In response, the electronic device 500 saves the regenerated automatically-generated visual content, as shown in Fig.6R.
[0238] Fig.6R illustrates user interface 600 including automatically-generated visual content 602g through 602h. Automatically-generated visual content 602g corresponds to the automatically-generated visual content shown in Fig.6F, which was saved as a new image. Automatically-generated visual media 602h corresponds to the automatically-generated visual content shown in Fig.6Q, which was saved instead of the automatically-generated visual content shown in Fig.6M. In some embodiments, the automatically-generated visual content in user interface 600 are listed in the order that they are saved and / or generated (e.g., automatically-generated visual content 602g was saved and / or generated before automatically-generated visual content 602h). -71- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0239] In Fig.6R, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 666 directed towards automatically-generated visual content 602g. In response to receiving the input, the electronic device 500 redisplays the user interface 604 including the recognized concepts (e.g., recognized concepts 620a through 620c) used to generate automatically-generated visual content 602g, shown in Fig.6S. In some embodiments, the user interface 604 includes additional suggestions, different than the previous suggestions shown in Fig.6E, including selectable options 608h through 608i, shown in Fig.6S. In some embodiments, the text entry region 614 includes prompt 636 that is a textual description of the recognized concepts 620a through 620c. As described in method 700, the electronic device 500 updates prompt 636 as the user modifies adds, and / or removes recognized concepts to be used to influence the generation of the automatically-generated visual content.
[0240] In Fig.6S, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 668 directed towards selectable option 624b. In response to receiving the input, the electronic device 500 generates the automatically-generated visual content. After generating the automatically-generated visual content, the electronic device 500 ceases displaying user interface 604 and begins displaying user interface 629d, in Fig. 6T. In some embodiments, user interface 629d includes one or more characteristics of user interface 629a through 629c, described above.
[0241] Fig.6T illustrates a different automatically-generated visual content than Fig. 6F because, as described in method 800, the electronic device 500 uses a different random or pseudorandom seed value or seed content to regenerate the automatically-generated visual content. Although the automatically-generated visual content in Fig.6T and in Fig.6F use the same recognized concepts, they are different automatically-generated visual content. In Fig. 6T, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 670a directed towards selectable option 630a. In response to receiving the input, in Fig.6U, the electronic device ceases to display user interface 629d and redisplays user interface 604 including the recognized concepts (e.g., recognized concepts 620a through 620c) used to generate the automatically-generated visual content in Fig.6T.
[0242] In Fig.6U, the electronic device 500 receives a dragging input including movement of contact 672a directed towards recognized concept 644i in the style category of menu 642. In response to receiving the input shown in Fig.6U, in Fig.6V, the electronic device 500 ceases to display recognized concept 620b and displays recognized concept 620i, -72- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) corresponding to recognized concept 644l in menu 642. In some embodiments, as described in method 700, certain categories of recognized concepts (e.g., style category) are mutually exclusive; therefore, adding a recognized concept in the respective category results in the other recognized concept being removed from the recognized concepts used to generate the automatically-generated visual content. In response to receiving the input in Fig.6U, the electronic device 500 also updates representation 622a to include the sketch recognized concept (e.g., recognized concept 622i).
[0243] In Fig.6V, the electronic device receives a typing input directed towards the text entry region 614 to enter a new prompt 676. In some embodiments, the electronic device receives a voice command input to enter a new prompt 676. Fig.6V displays the new prompt 676. In Fig.6V, the electronic device receives a selection input 670b directed towards selectable option 674, which updates the prompt (“sketch image of Jenna’s birthday”) to the new prompt 676 (“Jerome on a shark volcano”), shown in Fig.6W.
[0244] Fig.6W illustrates the user interface 604 including the recognized concepts extracted from the prompt 676 (e.g., recognized concept 620j through 620k). Extracting the recognized concepts is described in greater detail in the description of method 700. Displaying user interface 604 including the new recognized concepts in response to receiving a new prompt 676 is also described in further detail in the description of method 800. In some embodiments, the recognized concepts used to generate the automatically-generated visual content are displayed as text (e.g., recognized concept 622k) without an image corresponding to the text, as described in greater detail in the description of method 700. In some embodiments, recognized concepts without images associated with the word are displayed with only text. For example, recognized concept 622k is displayed as text because “shark volcano” is an imaginary concept. Additionally, the user interface 604 includes additional suggestions such as selectable option 608j (e.g., “space”) and selectable option 608k (e.g., “Halloween”). Additionally, the user interface 604 includes a representation 622c of the automatically-generated visual content.
[0245] Fig.6X illustrates the electronic device 500 receiving a dragging input including movement of contact 672b directed to dragging selectable option 608j towards the region of the user interface 604 including the recognized concepts used to influence the generation of the automatically-generated visual content. In response to receiving the input, the electronic device 500 displays a recognized concept 620l corresponding to selectable option 608j, shown in Fig.6Y. Additionally, in Fig.6Y, the electronic device 500 updates -73- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) representation 622c to include the recognized concept 620l. While displaying the user interface 604 in Fig.6Y, the electronic device 500 receives a selection input directed towards selectable option 624b (e.g., a tap or a long press) including contact 670c. In response to receiving the input in Fig.6Y, the electronic device 500 ceases to display user interface 604, and begins displaying user interface 629e including the automatically-generated visual content, shown in Fig.6Z.
[0246] Fig.6Z illustrates user interface 629e which includes the automatically- generated visual content generated using the recognized concepts 620j through 620l in Fig. 6Y. User interface 629e has one or more characteristics of the user interface 629a through 629d described above. While displaying user interface 629e, the electronic device 500 receives a selection input directed towards 630b (e.g., a tap or a long press) including contact 670d. In response to receiving the input shown in Fig.6Z, the electronic device 500 displays additional selectable options 630d and 630e shown in Fig.6AA. In some embodiments, the selectable options 630d and 630e are described above with reference to Fig.6Q. However, in Fig.6AA, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 670e directed towards selectable option 630e to save the automatically- generated visual content in Fig.6AA as a new image instead of overriding the previously generated automatically-generated visual content (e.g., shown in Fig.6F). In response to receiving the input in Fig.6AA, the electronic device 500 displays the automatically- generated visual content in user interface 629e as a new image represented by automatically- generated visual content 602i in user interface 600, shown in Fig.6BB. For example, the automatically-generated visual content 602i does not replace the automatically-generated visual content 602g since the automatically-generated visual content 602i was generated by modifying the recognized concepts used to generate the automatically-generated visual content 602g and saved as a new automatically-generated visual content.
[0247] Fig.6Z also illustrates an input directed towards option 675a, such as detecting contact 680a (e.g., a tap or long press input). In some embodiments, the electronic device 500 receives the input separately (e.g., independently) from the input that the electronic device detects with contact 670d (e.g., a tap or a long press). Alternatively, in some embodiments, the electronic device 500 detects the input (e.g., with contact 680a (e.g., a tap or a long press)) concurrently with the input (e.g., with contact 670d (e.g., a tap or a long press)). In some embodiments, other selection inputs, such as an air gesture input, a voice input, and / or an input detected using a hardware input device, are possible. In response to -74- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) detecting the input (e.g., detecting contact 680a (e.g., a tap or long press input)), shown in Fig.6Z, the electronic device 500 regenerates the automatically-generated visual content item shown in Fig.6Z as a three-dimensional automatically-generated visual content item, shown in Fig.6AA’. Three-dimensional automatically-generated visual content items are described in greater detail in method 700. In some embodiments, the three-dimensional automatically- generated visual content item is a three-dimensional representation of the previously generated automatically-generated visual content item (e.g., shown in Fig.6Z). In some embodiments, the three-dimensional automatically-generated visual content item is regenerated using the recognized concepts and optionally has a different visual appearance than the automatically-generated visual content item shown in Fig.6Z.
[0248] In Fig.6AA’, the electronic device 500 detects an input directed towards option 630b, such as detecting contact 680b (e.g., a tap or long press input). In some embodiments, other selection inputs, such as an air gesture input, a voice input, and / or an input detected using a hardware input device, are possible. In response to detecting the input, the electronic device 500 displays additional selectable options 630d and 630e shown in Fig. 6AA”. In some embodiments, the selectable options 630d and 630e are described above with reference to Fig.6Q. In some embodiments, the electronic device 500 receives a selection input directed towards selectable option 630d, such as detecting contact 680c (e.g., a tap or long press input), to save the three-dimensional automatically-generated visual content item shown in Fig.6AA”, over the previously generated automatically-generated visual content item shown in Fig.6Z, In some embodiments, the electronic device 500 detects a selection input directed towards selectable option 630e, such as detecting contact 680d (e.g., a tap or long press input), to save the three-dimensional automatically-generated visual content item as a new image instead of overriding the previously generated automatically-generated visual content item (e.g., shown in Fig.6Z).
[0249] Fig.6BB illustrates the electronic device 500 displaying the automatically- generated visual content in user interface 629e as a new image represented by automatically- generated visual content 602i in user interface 600. While displaying the user interface 600 in Fig.6BB, the electronic device 500 receives an upward swipe input including contact 674a. As a result, the electronic device 500 displays additional automatically-generated visual content stored in the gallery of automatically-generated visual content, as shown in Fig.6CC. In some embodiments, the additional automatically-generated visual content (e.g., automatically-generated visual content 602j and 602k in Fig.6CC) are sample automatically- -75- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) generated visual content or previously generated visual content. For example, the automatically-generated visual content 602a through 602g and 602j through 602k are previously generated visual content generated using recognized concepts selected by the user similarly to the examples described herein instead of sample automatically-generated visual content (e.g., examples that use recognized concepts selected by the electronic device 500 or a server in communication with the electronic device 500). In some embodiments, the automatically-generated visual content displayed on user interface 600 include automatically- generated visual content of various people, as described in methods 700 and 800.
[0250] In Fig.6CC, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 670e directed towards automatically-generated visual content 602k. In response to receiving the input shown in Fig.6CC, the electronic device 500 ceases displaying user interface 600 and displays user interface 629f, which includes the image of the automatically-generated visual content 602k, shown in Fig.6DD. In some embodiments, user interface 629f includes one or more characteristics of the user interfaces 629a through 629e. While displaying the automatically-generated visual content in Fig.6DD, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 670f. In response, the electronic device 500 ceases displaying user interface 629f and begins displaying user interface 604 including the recognized concepts 620m through 620p used to generate the automatically-generated visual content 602k, shown in Fig.6EE. Fig. 6EE also includes a representation 622d of automatically-generated visual content 606k.
[0251] In Fig.6DD, the user interface 629f also includes a selectable option 675a, that when selected, causes the electronic device 500 to generate a three-dimensional (3D) version of automatically-generated visual media 602k. Alternatively, the electronic device 500 generates a new automatically-generated visual content item as a 3D automatically- generated visual content item using the one or more recognized concepts. In some embodiments, selectable option 627a is described in greater detail in Fig.6Z.
[0252] Fig.6FF illustrates an embodiment where user interface 604 includes recognized concepts used to generate a different automatically-generated visual content. In Fig.6FF, user interface 604 includes recognized concepts 620q through 620u and a representation 622e of the resulting automatically-generated visual content. In Fig.6FF, user interface 604 includes text recognized concepts (e.g., recognized concepts 620r, 620t, and 620u). As described above, some recognized concepts do not include a visual representation. -76- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0253] Fig.6GG illustrates the electronic device 500 detecting a user typing “gun” into the text entry field 614. The electronic device 500 also detects a selection input (e.g., a tap or a long press) using contact 670g to add “gun” as a recognized concept. Because “gun” is a high risk concept, as described in greater detail in method 700, the electronic device 500 does not add “gun” as a recognized concept, as shown in Fig.6HH. Alternatively, in some embodiments, the electronic device 500 displays a recognized concept corresponding to “gun” and also displays a visual indication indicating that the recognized concept is high risk. In some embodiments, “gun” is a high risk concept when combined with some words but not when combined other words. For example, “hot glue gun” is not a high risk concept while “shooting guns” is a high risk concept. Similarly, in some embodiments, there are words that are not high risk in isolation but are high risk when combined with other words. For example, “punch” is low risk on its own, but “punch” combined with a person (e.g., “punch my brother”) is a high risk word. In some embodiments, a high risk word is able to be turned into a low risk word when a term is added to lower the risk because it changes the interpretation of one or more words in the prompt. For example, “punch” is high risk but “my brother drinking punch” is not.
[0254] In Fig.6GG, the electronic device 500 detects a selection input (e.g., a tap or a long press) using contact 670h directed towards selectable option 624b. In response to receiving the input in Fig.6GG, the electronic device 500 generates the automatically- generated visual content without the high risk recognized concept, “gun”, including ceasing displays user interface 604 and displaying user interface 629g. As shown in Fig.6II, the user interface 629g does not include the high risk recognized concept, “gun”. Alternatively, in some embodiments and as described in method 700, the electronic device 500 forgoes generating an automatically-generated visual content all together. In some embodiments, the electronic device 500 displays an indication of an error (e.g., an error message) in response to detecting an input directed towards selectable option 624b.
[0255] In Fig.6II, the user interface 629g also includes a selectable option 675a, that when selected, causes the electronic device 500 to generate a three-dimensional (3D) version of the automatically-generated visual media shown in Fig.6II. Alternatively, the electronic device 500 generates a new automatically-generated visual content item as a 3D automatically-generated visual content item using the one or more recognized concepts. In some embodiments, selectable option 627a is described in greater detail in Fig.6Z. -77- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0256] Figs.6JJ-6MM illustrate an embodiment in which the electronic device 500 displays a plurality of representations of automatically-generated visual content that are generated using the same recognized concepts and / or prompt. The user interface shown in Fig.6JJ includes one or more characteristics of the user interface shown in Fig.6L. In Fig. 6JJ, the electronic device 500 displays representation 622b including a preview of the automatically-generated visual content to be generated using the recognized concepts 620d through 620h. In Fig.6JJ, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 670i directed towards the representation 622b. In response to receiving the input, the electronic device 500 displays a representation 622b-a corresponding to the representation 622b in Fig.6KK. In some embodiments, the electronic device 500 displays representation 622b-a corresponding to the representation 622b in Fig. 6KKwithout receiving the input directed towards the representation 622b. For example, the electronic device displays representation 622b-a corresponding to the representation 622b in Fig.6KK after a time threshold (e.g., 1 second, 5 seconds, 10 seconds, 30 seconds, 1 minute, or 5 minutes) where the electronic device 500 does not detect an input is reached.
[0257] Fig.6KK illustrates the electronic device 500 displaying user interface 604 including representation 622b-a. Representation 622b-a is a larger (in size) representation of representation 622b. In some embodiments, representation 622b-a includes one or more characteristics of 622b as described in method 700. While displaying the representation 622b- a, the electronic device detects a swipe (e.g., a motion) input including contact 674b directed towards the user interface 604 and / or directed towards representation 622b-a. In response to receiving the input, the electronic device 500 generates a representation 622b-b, shown in Fig.6LL, of a second automatically-generated visual content that is generated using the same prompt and / or recognized concepts (e.g., recognized concepts 620d through 620h) as the automatically-generated visual content and the corresponding representation 622b-a of the automatically-generated visual content. In some embodiments, the content of the representation 622b-b is different from the content of representation 622b-a even though they are based on the same recognized concepts. Generating a representation of a second automatically-generated visual content in response to the input is described in greater detail in method 700. In some embodiments, in Fig.6KK, if the electronic device 500 receives an input directed towards option 624b, then the electronic device 500 generates the automatically-generated visual content, which is based on representation 622b-a. -78- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0258] Fig.6LL illustrates the representation 622b-b that is generated using the same recognized concepts (e.g., recognized concepts 620d through 620h) as the representation 622b-a. In some embodiments, in Fig.6LL, if the electronic device 500 receives an input directed towards option 624b, then the electronic device 500 generates the second automatically-generated visual content, which is based on representation 622b-b. In some embodiments, if the electronic device receives right to left swipe (e.g., such as including contact 674b) in Fig.6LL, then the electronic device 500 generates a representation of a third automatically-generated visual media that is generated using recognized concepts 620d through 620h. In some embodiments, the content of the representation of the third automatically-generated visual media is different from the content of representation 622b-a and representation 622b-a even though they are based on the same recognized concepts. In some embodiments, if the electronic device receives a left to right swipe input in Fig.6LL, then the electronic device 500 redisplays representation 622b-a.
[0259] In Fig.6LL, the electronic device 500 receives a selection input (e.g., a tap or long press input) including contact 670j directed towards a location outside of the representation 622b-b. In response to receiving the input, the electronic device 500 ceases displaying the representation 622b-b and resumes the previous display of user interface 604 including the display of the recognized concepts 620d through 620h and the representation 622b, shown in Fig.6MM. In response to receiving the input, the electronic device 500 displays the representation 622b to correspond to the representation of the second automatically-generated visual content (e.g., instead of the first automatically-generated visual content, and different in content than the content of representation 622b corresponding to the first automatically-generated visual content, as in Fig.6JJ) while continuing to display the same recognized concepts 620d through 620h as before. In some embodiments, if the electronic device 500 receives an input directed towards option 624b, then the electronic device 500 initiates the generation of the second automatically-generated visual content, similar to as previously described.
[0260] Fig.7 illustrates a flow diagram illustrating a method in which an electronic device displays recognized concepts and generates automatically-generated visual media in accordance with some embodiments of the disclosure. The method 700 is optionally performed at first electronic device and / or electronic devices such as device 100, device 300, or device 500 as described above with reference to Figs.1A-1B, 2-3, 4A-4B and 5A-5H. -79- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) Some operations in method 700 are, optionally combined and / or order of some operations is, optionally, changed.
[0261] As described below, the method 700 provides ways in which an electronic device displays one or more recognized concepts and prompts to be used to influence generation of an automatically-generated visual content. Displaying representations of recognized concepts of a prompt allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0262] Method 700 is performed at an electronic device in communication with a display generation component and one or more input devices, such as electronic device 500 shown in Fig.6A. For example, a mobile device (e.g., a tablet, a smartphone, a media player, or a wearable device) including wireless communication circuitry, optionally in communication with one or more of a mouse (e.g., external), trackpad (optionally integrated or external), touchpad (optionally integrated or external), remote control device (e.g., external), another mobile device (e.g., separate from the electronic device), a handheld device (e.g., external), and / or a controller (e.g., external). In some embodiments, the display generation component is a display integrated with the electronic device (optionally a touch screen display), external display such as a monitor, projector, television, or a hardware component (optionally integrated or external) for projecting a user interface or causing a user interface to be visible to one or more users, etc. Examples of input devices include physical buttons, knobs, handles, and / or switches of a vehicle, a touch screen, mouse (e.g., external), trackpad (optionally integrated or external), touchpad (optionally integrated or external), microphone for capturing voice commands or other audio input, remote control device (e.g., external), another electronic device (e.g., mobile device that is separate from the electronic device), a handheld device (e.g., external), a controller (e.g., external), a camera, a depth sensor, an eye tracking device, and / or a motion sensor (e.g., a hand tracking device, a hand motion sensor).
[0263] In some embodiments, the electronic device receives (702a), via the one or more input devices, a prompt for use in creating automatically-generated visual media (e.g., generative visual media, such as images and / or videos, that is generated automatically and / or generated at least partially using one or more autonomous processes)) that is generated at least partially using one or more autonomous processes, such as the prompt received using voice command 634 shown in Fig.6G. In some embodiments, a user of the electronic device -80- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) inputs a prompt to generate visual media. In some embodiments, the user inputs a prompt using a keyboard (e.g., touch keyboard or physical keyboard), a voice command, and / or a touchpad. In some embodiments, the prompt is a text prompt describing concepts to be added to an image. In some embodiments, the prompt is and / or includes media such as still images, videos, avatars, emojis, and / or a combination of media. In some embodiments, the user chooses a prompt from a menu of concepts (e.g., activities, outfits, accessories, themes, and / or styles). In some embodiments, the menu presents concepts using images (e.g., icons). In some embodiments, the electronic device uses one or more autonomous processes such as artificial neural networks and / or machine learning to generate the automatically-generated visual media. For example, automatically-generated visual media are generated using prompts and are not images captured of the real world. In some embodiments, automatically- generated visual media corresponds to the media produced by a machine program that uses foundational models (FMs) (e.g., neural networks trained using data sourced from a variety of methods) to generate the media, optionally without user input providing visual content that ends up in the automatically-generated visual media. Optionally one or more or all visual portions of generative visual content are generated by the machine program, and not by a user providing such visual portions. For example, the electronic device generates (e.g., using an AI process or a generative AI process) automatically-generated visual media using a non- visual input (e.g., using text and / or audio) and not using a visual input such as visual media (e.g., images and / or photos). For example, the electronic device generates (e.g., using an AI process or a generative AI process) automatically-generated visual media using text prompts (e.g., a user telling the electronic device to generate an image of Max playing volleyball on the beach in space) rather than using visual inputs (e.g., visual inputs of Max playing volleyball on the beach and of space). In some embodiments, automatically-generated visual media is not visual media that is generated as a result of image manipulation (e.g., adding stickers or filters to a pre-existing image of the real world). In some embodiments, automatically-generated visual media is not visual media that is generated by a user providing visual content without semantic content. For example, automatically-generated visual media is not visual media that is generated using a drawing application (e.g., a user providing the visual content such as lines, shapes, contours, and / or colors for a drawing). In some embodiments, semantic content includes the semantics of the subjects in the image and / or the semantics of the style of the automatically-generated visual content (e.g., sketch, animation, cartoon, or pop-art). In some embodiments, the electronic device uses one or more items of pre-existing visual media (e.g., images captured of the real world such as images captured by -81- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) a camera or user created digital images or a modified version of a real-world image or user created digital image) as a seed for creating automatically-generated visual content. In some embodiments, the computer system uses a random seed such as visual noise or other as a starting point for generating the image or video. For example, automatically-generated visual content build additional concepts and / or features into a real image. In some embodiments, generating automatically-generated visual media includes using a real-world image as a starting point to create a new image using one or more recognized concepts. In some embodiments, automatically-generated visual media incorporates one or more recognized concepts by including the semantic characteristics of the recognized concepts without including an exact representation of the one or more recognized concepts. For example, the electronic device uses one or more autonomous processes (e.g., machine learning processes) to reimagine the starting image with the additional recognized concepts.
[0264] In some embodiments, in response to receiving the prompt, the electronic device displays (702b), via the display generation component, a user interface (e.g., a user interface of a content application) that includes prompt information (e.g., a prompt review user interface), such as user interface 604 shown in Fig.6H-A.
[0265] In some embodiments, displaying the user interface includes concurrently displaying a representation of the prompt (702c), such as prompt 636 displayed in text entry region 614 shown in Fig.6H-A. In some embodiments, the representation of the prompt includes text and / or images. In some embodiments, the representation of the prompt is text representing the prompt received via the one or more input devices. For example, a user inputs a prompt (e.g., via text or voice command) such as “birthday in space” and the representation of the prompt is text that includes “birthday in space”. In some embodiments, a user inputs a prompt by selecting one or more visual representations of recognized concepts in a collection of recognized concepts, as described below. For example, the electronic device displays a menu of recognized concepts including an icon representing “tennis” (e.g., a tennis racket), and an icon representing “beach” (e.g., a palm tree with sand). In some embodiments, a user selects the aforementioned recognized concepts by selecting the associated icons (e.g., a drag and drop input), which results in the electronic device combining the recognized concepts into a string (e.g., “tennis on the beach”) to be displayed as the representation of the prompt.
[0266] In some embodiments, displaying the user interface includes concurrently displaying a first visual indication that a first portion of the prompt (e.g., word, phrased, -82- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) portion of a word, visual input, or character string) has been identified as a first recognized concept that will influence generation of the automatically-generated visual media (702d), such as recognized concept 620f shown in Fig.6H-A representing a first portion of the prompt. In some embodiments, the recognized concepts include keywords, visual media, and / or style representations. In some embodiments, the electronic device and / or server processes the representation of the prompt using one or more language models, and optionally compares the prompt to a database of recognized concepts (e.g., keywords and / or media) to determine recognized concepts. For example, and as described below, the electronic device and / or server includes a database of recognized concepts. In some embodiments, the electronic device and / or server matches the recognized concepts to keywords and / or media in the prompt. In some embodiments, the recognized concepts in the prompt include the nouns in the prompt. In some embodiments, each recognized concept (e.g., concept tokens) includes a corresponding visual indication. In some embodiments, the visual indications includes text and / or images (e.g., an icon) describing the recognized concept. For example, the prompt is “birthday in space” and the recognized concepts include “birthday” and “space” and the first visual indication includes an image of a birthday cake and a second visual indication includes an image of a rocket. Displaying representations of recognized concepts of a prompt allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0267] In some embodiments, displaying the user interface that includes the prompt information further includes displaying a second visual indication that a second portion of the prompt has been identified as a second recognized concept that will influence the generation of the automatically-generated visual media, wherein the second visual indication is displayed concurrently with the first visual indication in the user interface, the second visual indication is different than the first visual indication and the second portion of the prompt is different than the first portion of the prompt, such as the recognized concept 620e shown in Fig.6H-A representing a second portion of the prompt. In some embodiments, the first visual indication is related to the first recognized concept and not the second recognized concept. In some embodiments, the second visual indication has one or more characteristics of the first visual indication as described above. For example, the second visual indication is an icon and / or text describing the second recognized concept and not the first recognized concept. For example, the prompt is “birthday in space” and the first portion of the prompt / the first -83- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) recognized concept is “birthday” and the second portion of the prompt / the second recognized concept is “space”. In some embodiments, the first visual indication is an illustration of a birthday cake and the second visual indication is an illustration of the solar system. Displaying more than one representation of recognized concepts corresponding to different recognized concepts of a prompt allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0268] In some embodiments, the prompt includes starting media for use in creating the automatically-generated visual media, and the starting media influences the generation of the automatically-generated visual media. For example, the prompt in Fig.6H-A includes “Jeremy”, which results in recognized concept 620d being used at the starting media. In some embodiments, starting media includes a photo, video, and / or other image, such as a photo, video, or an image from a photo library or a camera. In some embodiments, the prompt includes a photo and a string of text (which the electronic device will use to detect recognized concepts). In some embodiments, the starting media is a representation of a person generated by the electronic device. For example, the electronic device generates a representation of a person using a plurality of photos in a photo library that include the person. In some embodiments, the starting media is a person, a place, or other content captured by a photo. In some embodiments, the electronic device detects a user input (e.g., a selection input) that has characteristics of the inputs described below, for providing the starting media. In some embodiments, the starting media, and the first portion of prompt both concurrently influence the generation of the automatically-generated visual media. Including starting media allows the electronic device to generate a generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0269] In some embodiments, the first recognized concept includes a keyword or set of related keywords, such as recognized concepts 620f, 620e, and 690 shown in Figs.6H-A and 6H-B. In some embodiments, the first recognized concept is a keyword of a prompt. For example, nouns in the prompt are optionally keywords. In some embodiments, representations of keywords are displayed as text (e.g., there is not an image associated with the keyword). In some embodiments, the electronic device integrates concepts related to the keyword in to the automatically-generated visual media. For example, if the keyword is “Paris” then the electronic device integrates concepts related to “Paris” in the automatically- generated visual media (e.g., the Eiffel Tower, French pastries, French art, and / or other -84- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) French landmarks). Displaying keywords as recognized concepts allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0270] In some embodiments, the first recognized concept includes visual media, such as the picture of Jeremy represented by recognized concept 620d in Fig.6H-A. In some embodiments, the electronic device detects an input by the user to provide the user interface with visual media and then the visual media is a recognized concept. For example, the user selects an image (e.g., from a photo library or a camera application) to be used to generate the automatically-generated visual media, then the image is presented as the recognized concept. Alternatively, or additionally, in some embodiments, the electronic device detects an input identifying the recognized concept (e.g., via the prompt or an additional input), and the electronic device produces the visual media based on the input. For example, the electronic device displays an icon associated with a recognized concept. For example, if the portion of the prompt that has been identified is “Eiffel Tower”, then the electronic device displays a recognized concept that is an icon of the Eiffel Tower. In some embodiments, the visual media (e.g., the icon) is a photo, video, or a sketch. Displaying visual media as recognized concepts allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0271] In some embodiments, the first recognized concept includes a style prompt, such as recognized concept 620b shown in Fig.6E. In some embodiments, the electronic device detects an input selecting a style prompt from a predetermined set of style prompts, such as the style prompts associated with selectable options 644g through 644l in Fig.6N. Alternatively, or additionally, in some embodiments, the user inputs a style prompt using text (e.g., keywords or a set of keyword) and / or a voice input. In some embodiments, the style prompts dictate the style of the automatically-generated visual media. For example, style prompts include sketch, cartoon, realistic, black and white, animation, noir, or painting. In some embodiments, the recognized concepts only includes one style at a time, as described below. Displaying style prompts as recognized concepts allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors. -85- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0272] In some embodiments, while displaying the user interface that includes the prompt information, the electronic device detects one or more inputs corresponding to a request to modify one or more recognized concepts that will influence the generation of the automatically-generated visual media, such as input (e.g., a tap or a long press) including contact 647 in Fig.6I and input (e.g., a tap or a long press) including contact 658 in Fig.6N. In some embodiments, the one or more inputs includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments, the one or more inputs is directed towards a selectable option on or near the one or more visual indications associated with the one or more recognized concepts that is selectable to modify (e.g., change or delete) the one or more recognized concepts. For example, the one or more inputs are directed towards a selectable option on or near the first visual indication associated with the recognized concept.
[0273] In some embodiments, in response to detecting the one or more inputs, the electronic device modifies (e.g., adding, removing, and / or editing) the one or more recognized concepts that will influence the generation of the automatically-generated visual media in accordance with the one or more inputs, such as adding recognized concept 620g in Fig.6J in response to the input (e.g., a tap or a long press) including contact 647 and removing recognized concept 620e in Fig.6O in response to receiving the input (e.g., a tap or a long press) including contact 658. In some embodiments, a user modifies the first recognized concept by changing the recognized concept from a first concept to a second concept. For example, a user changes the recognized concept “birthday” to “graduation party” (e.g., by inputting via voice, text, or handwriting) which results in the electronic device changing the first visual indication representing “birthday” to a second visual indication representing “graduation party”. In some embodiments, a user modifies the first visual indication by changing the icon representing the recognized concept. For example, the first visual indication is a birthday cake representing the recognized concept, “birthday”. The user optionally changes the icon to an image of a birthday candle to represent the recognized concept, “birthday”. In some embodiments, modifying the first visual indication includes deleting the recognized concept such that the electronic device no longer displays the first visual indication. In some embodiments, modifying the recognized concept includes changing the recognized concepts that will influence the generation of the automatically-generated -86- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) visual media which also includes changing the representation of the automatically-generated visual media, described below, and ultimately, change the automatically-generated visual media. Modifying a recognized concept by selecting the visual indication associated with the recognized concept allows a user to easily and efficiently modify the recognized concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0274] In some embodiments, the one or more inputs corresponding to the request to modify the one or more recognized concepts that will influence the generation of the automatically-generated visual media includes an input to modify the prompt, such as inputting the new prompt 676 in text entry region 614 shown in Fig.6V. In some embodiments, the one or more inputs has one or more characteristics of the one or more inputs described above. In some embodiments, the one or more inputs includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments, the one or more inputs is directed towards the representation of the prompt. In some embodiments, the input is directed to a text entry region, described in detail below, and includes adding, deleting, or modifying text.
[0275] In some embodiments, in response to receiving the one or more inputs, the electronic device updates the representation of the prompt to a representation of a second prompt in accordance with modifications to the prompt indicated by the one or more inputs, such as shown by the new prompt in the text entry region 614 in Fig.6W. In some embodiments, the user changes a portion of the prompt or the entire prompt. In some embodiments, in response to receiving the second prompt, the electronic device updates the representation of the prompt to the representation of the second prompt. In some embodiments, modifications to the prompt includes changing the text of the prompt and / or changing an image associated with the prompt.
[0276] In some embodiments, in response to receiving the one or more inputs, the electronic device displays, in the user interface, one or more second visual indications corresponding to one or more portions of the second prompt that have been identified as recognized concepts, such as shown by recognized concepts 620j and 620k in Fig.6W. The one or more second visual indications have one or more characteristics of the first visual indication. In some embodiments, identifying the one or more portions of the second prompt -87- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) has one or more characteristics of identifying the one or more portions of the prompt. In some embodiments, changing the prompt includes deleting the portion of the prompt corresponding to the recognized concept. In some embodiments, if the portion of the prompt corresponding to the recognized concept is deleted, the electronic device ceases displaying the first visual indication corresponding to the recognized concept based on the portion of the prompt. For example, the electronic device ceases displaying recognized concepts 620a, 620c, and 620i in response to receiving the prompt 676 in Fig.6V and begins displaying recognized concepts 620j and 620k in Fig.6W. In some embodiments, if the user did not modify the portion of the prompt that includes the portion of the prompt associated with the first visual indication, then the electronic device continues to display the first visual indication corresponding to the recognized concept based on the portion of the prompt. In some embodiments, ceasing displaying the first visual indication means that the recognized concept associated with the first visual indication will no longer influence the generation of the automatically-generated visual media. In some embodiments, modifying the prompt includes changing the recognized concepts that will influence the generation of the automatically-generated visual media which also includes changing the representation of the automatically-generated visual media, described below, and ultimately, change the automatically-generated visual media Modifying recognized concepts by modifying the prompt allows a user to easily and efficiently modify the recognized concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0277] In some embodiments, the one or more inputs corresponding to the request to modify the one or more recognized concepts that will influence the generation of the automatically-generated visual media includes an input corresponding to a request to add a second recognized concept that will influence the generation of the automatically-generated visual media, such as shown by the input (e.g., a tap or a long press) including contact 672b to add selectable option 608j as a recognized concept in Fig.6X. In some embodiments, adding a second recognized concept includes changing the prompt to include the second recognize concept, such as described above. In some embodiments, adding the second recognized concept includes dragging (e.g., using a dragging input) to drag a second recognized concept or a selection input to select a second recognized concept from a plurality of recognized concepts (as described below) into the area where the other recognized concepts are displayed. Alternatively, or additionally, in some embodiments, adding the -88- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) second recognized concept includes selecting (e.g., via a tap) a recognized concept form a list of recognized concepts. In some embodiments, the first input has one or more characteristics of the selection inputs described above. In some embodiments, the first input is a dragging input including a tap and drag and drop motion using a contact or a gaze. In some embodiments, first input includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture).
[0278] In some embodiments, in response to receiving the one or more inputs, the electronic device displays a second visual indication corresponding to the second recognized concept concurrently with the first visual indication corresponding to the recognized concept, such as shown by recognized concept 620l in Fig.6Y. In some embodiments, the second visual indication is displayed near the first visual indication (e.g., next to or adjacent to). In some embodiments, displaying the second visual indication further includes displaying the second recognized concept in the representation of the prompt. In some embodiments, adding the second recognized concept also causes the automatically-generated visual media to be based on the second recognized concept and / or updates the representation of the automatically-generated visual media, described below. In some embodiments, adding new recognized concept includes changing the recognized concepts that will influence the generation of the automatically-generated visual media which also includes changing the representation of the automatically-generated visual media, described below, and ultimately, change the automatically-generated visual media. Modifying the recognized concept by adding recognized concepts allows a user to easily and efficiently change the recognized concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional inputs to correct such errors.
[0279] In some embodiments, the one or more inputs corresponding to the request to modify the one or more recognized concepts that will influence the generation of the automatically-generated visual media include an input corresponding to a request to delete the first recognized concept, such as input (e.g., a tap or a long press) including contact 658 to delete recognized concept 620e in Fig.6N. In some embodiments, the second input has one or more characteristics of the inputs described above. In some embodiments, the second input is a selection input directed towards a selectable option on or near the recognized concept that is selectable to delete the recognized concept. In some embodiments, the input includes a -89- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments, the second input is a drag input wherein a user taps on the first visual indication and, while continuing to make contact with the touch screen, drags the first visual indication off the user interface.
[0280] In some embodiments, in response to receiving the one or more inputs, the electronic device ceases display of the first visual indication, such as the electronic device 500 no longer displaying recognized concept 620e in Fig.6O. In some embodiments, the electronic device continues to display other visual indications of other recognized concepts (e.g., a second visual indication of a second recognized concept). In some embodiments, ceasing displaying the first visual indication includes updating the representation of the prompt to no longer include the portion of the prompt associated with the recognized concept. In some embodiments, deleting the recognized concept also causes the recognized concept to no longer influence the generation of the automatically-generated visual media and / or updates the representation of the automatically-generated visual media, described below, to no longer include the recognized concept. Allowing a user to easily and efficiently change the recognized concepts used to generate the generative image reduces errors in output of the electronic device, and avoids the need for additional inputs to correct such errors.
[0281] In some embodiments, while displaying the first visual indication corresponding to the first recognized concept (e.g., recognized concept 620b shown in Fig. 6U), the electronic device receives, via the one or more input devices, a second input corresponding to a request to add a second recognized concept, wherein the second recognized concept is in a same category as the first recognized concept, such as the input (e.g., a tap or a long press) including contact 672a directed towards option 644l in Fig.6U. In some embodiments, the second input has one or more characteristics of the inputs described above. In some embodiments, a user drags and drops the second recognized concept from a plurality of recognized concepts to the area occupied by the first recognized concept to add the second recognized concept. In some embodiments, the In some embodiments, a user selects the second recognized concept to add the second recognized concept. In some embodiments, adding a recognized concept includes adding a recognized concept to the recognized concepts that will influence the generation of the automatically-generated visual media. In some embodiments, recognized concepts in a category of recognized concepts are -90- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) related (e.g., same type of recognized concept that influences the same aspect of the generated media). For example, categories include styles (e.g., sketch, painting, realistic, or other styles), activities (e.g., types of sports), and locations. In some embodiments, some categories of recognized concepts are mutually exclusive such that if a recognized concept in a first category is added, then other recognized concepts in the same category that are being used to influence the generation of the automatically-generated visual media are removed. In some embodiments, the second input includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture).
[0282] In some embodiments, in response to receiving the input, the electronic device ceases display of the first visual indication corresponding to the first recognized concept (e.g., the electronic device does not display recognized concept 620b in Fig.6V). In some embodiments, ceasing displaying the first visual indication further includes no longer influencing the generation of the automatically-generated visual media using the first recognized concept. In some embodiments, the electronic device updates the representation of the prompt in response to removing a portion of the prompt (e.g., the first recognized concept). In some embodiments, the first visual indication and the first recognized concept associated with the first visual indication are removed without the electronic device receiving an input to remove them.
[0283] In some embodiments, in response to receiving the input, the electronic device displays a second visual indication corresponding to the second recognized concept that will influence the generation of the automatically-generated visual media, such as the electronic device 500 displaying recognized concept 620i in Fig.6V. In some embodiments, the second visual indication has one or more characteristics of the first visual indication and other visual indications as described herein. In some embodiments, displaying the second visual indication includes influencing the generation of the automatically-generated visual media using the second recognized concept. For example, recognized concepts in the style category are mutually exclusive such that the automatically-generated visual media only has one style. In some embodiments, recognized concepts in the activities category, clothing category, and / or food category are not mutually exclusive. For example, the automatically-generated visual media includes a plurality of activities. In some embodiments, as a result of the one or -91- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) more inputs, the representation of the automatically-generated visual media and the automatically-generated visual media (e.g., if generated) is based on the second recognized concept and not the first recognized concept. Allowing only one recognized concept in a specific category of recognized concepts by automatically removing the other recognized concept reduces the amount of erroneous inputs to select recognized concepts used to generate the automatically-generated visual media.
[0284] In some embodiments, displaying the user interface further includes displaying a first plurality of representations of concepts representing recognized concepts of a first category within a database of recognized concepts, such as menu 642 including recognized concepts 644a through 644f shown in Fig.6I. In some embodiments, the representation of concepts includes icons, images, and / or text. In some embodiments, the first category includes a first representations of concept (e.g., a first visual indication) representing a first recognized concept, and a second representations of concept (e.g., a second visual representation) representing a second recognized concept. In some embodiments, the database of recognized concepts includes a plurality of categories of recognized concepts (e.g., a first category, a second category, and a third category). In some embodiments, the categories include one or more characteristics of the categories as described above. For example, the electronic device 500 displays the category of recognized concepts associated with “activities” in Fig.6J and the category of recognized concepts associated with “Styles” in Fig.6K.In some embodiments, the database of recognized concepts is stored on the electronic device (e.g., in the content application) and / or on a storage device (e.g., cloud storage) in communication with the electronic device.
[0285] In some embodiments, while displaying the first plurality of representations of concepts, the electronic device receives, via the one or more input devices, an input corresponding to a request to display a second plurality of representations of concepts representing recognized concepts of a second category within the database, the second category different from the first category, such the swipe input including contact 650 shown in Fig.6J. In some embodiments, the input is a movement input (e.g., swipe input or a dragging input), and / or a selection input (e.g., tapping input or gaze input). In some embodiments, the input has one or more characteristics of the inputs as described above. In some embodiments, the input is directed towards a category selection control including a first visual indication corresponding to a first category and a second visual indication corresponding to a second category. In some embodiments, a user taps the visual indication to -92- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) navigate to the corresponding category. Alternatively, or additionally, in some embodiments, the electronic device displays a second category in response to receiving a swiping or dragging input directed towards representations of concepts in the first category (e.g., a right swipe or a left swipe to navigate to a different category).
[0286] In some embodiments, in response to receiving the input corresponding to the request to display the second plurality of representations of concepts, the electronic device displays, in the user interface, the second plurality of representations of concepts (e.g., and ceasing display of the first plurality of representations of concepts), such as recognized concepts 644g through 644l shown in Fig.6K. In some embodiments, the second plurality of representations of concepts is associated with a second category. In some embodiments, a user selects (e.g., via a tap input, a drag input, or a gaze input) an representations of concept in a second category (or a first category) of recognized concepts, and in response, the electronic device adds the respective recognized concept to the recognized concepts used to influence the generation of the automatically-generated visual media and displays a visual indication of the added recognized concept as described above with respect to visual indications of other recognized concepts. In some embodiments, recognized concepts in the first plurality of representations of concepts and recognized concepts in the second plurality of representations of concepts are selectable (e.g., using a selection or a movement input) to be added to the recognized concepts used to influence the generation of the automatically- generated visual media, such as using the swipe input including contact 652 directed towards recognized concept 644l in Fig.6K.. Allowing a user to navigate through a menu of recognized concepts grouped by category allows the user to easily identify and select recognized concepts to be used to generate the automatically-generated visual media, thereby reducing erroneous inputs to generate the automatically-generated visual media.
[0287] In some embodiments, while displaying the user interface that includes the prompt information, the electronic device displays a selectable option for removing a plurality of recognized concepts that will influence generation of the automatically-generated visual content, such as option 624a, shown in Fig.6N and Fig.6I. In some embodiments, the selectable option is a “remove all” or cancel option for removing the prompt and / or the one or more recognized concepts that comprise the prompt.
[0288] In some embodiments, while displaying the selectable option, the electronic device detects, via the one or more input devices, an input directed towards the selectable option, such as the electronic device detecting contact 658b (e.g., a tap or long press input) -93- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) directed towards option 624a, shown in Fig.6N, and the electronic device detecting contact 658a (e.g., a tap or long press input) directed towards option 624a, shown in Fig.6I. In some embodiments, the input is a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of the selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments, the input is a voice input corresponding to a request to select the selectable option.
[0289] In some embodiments, in response to detecting the input, the electronic device removes the plurality of recognized concepts from being used to influence generation of the automatically-generated visual content, including ceasing display of one or more visual indications corresponding to the plurality of the recognized concepts (e.g., the recognized concepts that were removed), and ceasing using the plurality of the recognized concepts (e.g., the recognized concepts that were removed) to influence the generation of the automatically- generated visual content. For example, in response to detecting the input (e.g., a tap or a long press) with contact 648b shown in Fig.6N, the electronic device ceases displaying the recognized concepts 620d through 620h, and begins displaying the user interface 604 shown in Fig.6B. In some embodiments, ceasing the display of the visual indications of the one or more recognized concepts includes removing the prompt. In some embodiments, ceasing the display of the visual indications of the one or more recognized concepts includes no longer using the one or more recognized concepts to influence the generation of the automatically- generated visual media. In some embodiments, the electronic device does not remove the subject and / or style to be used to generate the automatically-generated visual content. In some embodiments, in response to detecting the input, the electronic device displays the user interface that is displayed prior to the electronic device receiving the prompt. In some embodiments, the selectable option is selectable to remove all (e.g., or multiple of) the recognized concepts currently being used to influence the generation of the automatically- generated visual content. Displaying an option to remove recognized concepts allows a user to easily and efficiently remove a plurality of recognized concepts, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0290] In some embodiments, removing the plurality of recognized concepts includes removing all of the recognized concepts that will influence the generation of the -94- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) automatically-generated visual content, such as removing all the recognized concepts 620d through 620h, shown in Fig.6N. In some embodiments, as described above, removing all of the recognized concepts includes no longer using the recognized concepts to influence the generation of the automatically-generated visual media. In some embodiments, removing all of the recognized concepts includes ceasing the display of the respective visual indications corresponding to the recognized concepts in the user interface. In some embodiments, removing all of the recognized concepts does not include removing the style and / or subject. In some embodiments, the style and subject of the automatically-generated visual media is described in greater detail in methods 1700, 1900, 2100, 2300, and / or 2500. In some embodiments, removing all of the recognized concepts includes removing the style and / or subject. Displaying a button to remove all recognized concepts allows a user to easily and efficiently remove a plurality of recognized concepts, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0291] In some embodiments, while displaying the first visual indication corresponding to the first recognized concept (e.g., the first visual indication as described above), the electronic device receives, via the one or more input devices, a first input corresponding to a request to add a second recognized concept corresponding to a representation of a person (e.g., an image of a person) that will influence the generation of the automatically-generated visual media, such as input (e.g., a tap or a long press) including contact 618 shown in Fig.6B or if the electronic device had receives an input directed towards option 616 to add a representation of a person in Fig.6B. Adding a recognized concept corresponding to a representation of a person is described in greater detail in method 800. In some embodiments, the first input includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture).
[0292] In some embodiments, in response to receiving the first input, the electronic device displays, in the user interface, a second visual indication corresponding to the second recognized concept that will influence the generation of the automatically-generated visual media, such as displaying recognized concept 620a in Fig.6C in response to the electronic device receiving the input in Fig.6B. In some embodiments, the second visual indication includes a representation of the person. For example, the second visual indication is a picture -95- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) of the person. In some embodiments, the representation of a person is a mutually exclusive recognized concept. For example, adding a recognized concept of a representation of a second person includes removing the previously added recognized concept of a representation of a first person. In some embodiments, the person is the starting media for the generation of the automatically-generated visual media, as described above. In some embodiments, before receiving the first input, the automatically-generated visual media is influenced by the recognized concept and not by the second recognized concept. In some embodiments, after receiving the first input, the automatically-generated visual media is influenced by the recognized concept and the second recognized concept. In some embodiments, adding the representation of the person includes adding a representation of a person from a collection of representations of people identified (e.g., by the electronic device and / or a server) from a media library (e.g., photo and / or video library) of the user, as described in greater detail in method 800. In some embodiments, the user selects a representation of a person identified as being important (e.g., by being named in the collection of representations of people and / or by being marked as a favorite person in the collection of representations of people, such as in response to user input). Including starting media of a representation of a person allows the electronic device to generate a generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0293] In some embodiments, while displaying the first visual indication corresponding to the first recognized concept (e.g., the first visual indication as described above), the electronic device receives, via the one or more input devices, a first input corresponding to a request to add a second recognized concept corresponding to a representation of an animal (e.g., an image of a dog, cat, or another animal) that will influence the generation of the automatically-generated visual media, such as if input (e.g., a tap or a long press) including contact 618 was directed towards a selectable option of an animal in Fig.6B. Adding a recognized concept corresponding to a representation of an animal includes one or more characteristics of adding a recognized concept corresponding to a representation of a person as described in greater detail in method 800 and above. In some embodiments, the second recognized concept corresponds to an individual animal, such as a particular animal; for example, a pet. In some embodiments, the first input includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a -96- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments, the animals that are available to be used as recognized concepts are animals that have been identified in a personal media collection of a user, such as one or more pets of a user. Sometimes these pets will have been identified as being important to the user by being marked as “favorites” or by being assigned one or more names. In some embodiments, the animals that are available to be used as recognized concepts have been identified in multiple media items or more than a threshold number of different media items (e.g., 5, 10, 100, or 1,000 media items).
[0294] In some embodiments, in response to receiving the first input, the electronic device displays, in the user interface, a second visual indication corresponding to the second recognized concept that will influence the generation of the automatically-generated visual media, such as if recognized concept 620a includes a representation of an animal in Fig.6C. In some embodiments, the second visual indication includes a representation of the animal. For example, the second visual indication is a picture of the animal. In some embodiments, the representation of the animal is a mutually exclusive recognized concept. For example, adding a recognized concept of a representation of a second animal includes removing the previously added recognized concept of a representation of a first animal. In some embodiments, adding the representation of the animal includes removing the previously added recognized concept of a person or other object that is being used as the starting media. In some embodiments, adding the representation of the animal includes adding a representation of an animal from a collection of representations of animals identified (e.g., by the electronic device and / or by a server) from a media library (e.g., photo and / or video library) of the user, as described in greater detail in method 800. In some embodiments, the user selects a representation of an animal identified as being important (e.g., by being named in the collection of representations of animals and / or by being marked as a favorite animal in the collection of representations of animal, such as in response to a user input). Including starting media of a representation of an animal allows the electronic device to generate a generative image using starting media that is not a person, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0295] In some embodiments, displaying the user interface further comprises displaying a text entry region, such as text entry region 614, shown in Fig.6G, configured to receive text corresponding to additional prompt information corresponding to one or more concepts that will influence the generation of the automatically-generated visual media. In -97- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) some embodiments, a user adds and / or modifies the prompt to the second prompt (e.g., as described above) using the text entry region. In some embodiments, the electronic device receives text using a physical or virtual keyboard, with a voice command, or with handwriting (e.g., using a stylus). In some embodiments, in response to receiving text, the electronic device updates the representation of the prompt with the additional prompt information. In some embodiments, in response to receiving text, the electronic device displays representations of the additional recognized concepts corresponding to portions of the text that have been identified as recognized concepts. In some embodiments, the generation of the automatically-generated visual media is influenced by the additional recognized concepts. In some embodiments, the electronic device updates the representation of the automatically-generated visual media to include the representations of the additional recognized concepts, described below. In some embodiments, the prompt received by the electronic device in step(s) 702 that corresponds to the recognized concept is a text string, and the electronic device receives the prompt via input directed to the text entry region. Including a text entry region that receives text corresponding to additional prompt information on the user interface allows the user to easily input additional prompt information to be used to generate the automatically-generated visual media, thereby reducing erroneous inputs to generate the automatically-generated visual media.
[0296] In some embodiments, while displaying the user interface including the text entry region, the electronic device receives, via the one or more input devices, a first input that includes text directed towards the text entry region, such as shown by the voice command 634 in Fig.6G resulting in prompt 636 being added to the text entry region 614. In some embodiments, the first input includes typing using a physical or virtual keyboard, with a voice command, or with handwriting using a stylus.
[0297] In some embodiments, in response to receiving the first input, in accordance with a determination that the first input corresponds to a plurality of recognized concepts that will influence the generation of the automatically-generated visual media, the electronic device displays, in the user interface, a first visual representation corresponding to the first input that indicates the plurality of recognized concepts of the text, such as shown by recognized concepts 629 as a cluster in Fig.6H-B. In some embodiments, the electronic device identifies the one or more recognized concepts in the text using methods as described above. For example, the electronic device identifies keywords in the text. In some embodiments, the first visual representation has one or more characteristics of the first visual -98- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) indication as described above. In some embodiments, the visual representation is a cluster of visual indications of recognized concepts. For example, when text includes and / or corresponds to a plurality of recognized concepts, the electronic device displays the visual indications as a first visual representation (e.g., a cluster of visual indications). For example, if the text is “on the beach wearing sunglasses”, then the first visual representation includes a visual indication of “beach” and a visual indication of “sunglasses”. In some embodiments, the visual representation for an individual recognized concept in the cluster includes icons / images, text, and / or a mix of images and text. For example, if there is no image or icon associated with a recognized concept, then the electronic device displays text associated with the recognized concept. In some embodiments, in accordance with a determination that the first input does not correspond to a plurality of recognized concepts that will influence the generation of the automatically-generated visual media, the electronic device forgoes displaying the first visual representation that indicates the plurality of recognized concepts of the text. In some embodiments, the first input includes text associated with one recognized concept. In some embodiments, the electronic device displays a visual indication corresponding to the recognized concept but the electronic device does not display a first visual representation including a cluster of visual indications (e.g., because there is only one recognized concept). Displaying more than one representation of recognized concepts corresponding to different recognized concepts of a prompt as a cluster allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0298] In some embodiments, while displaying the first visual representation corresponding to the first input that indicates the plurality of recognized concepts of the text, the electronic device receives, via the one or more input devices, an input directed towards the visual representation, such as if the electronic device receives an input directed towards recognized concept 620c and the recognized concept 620c does not include the text “birthday” in Fig.6S. In some embodiments, the input has one or more characteristics of the inputs as described above. In some embodiments, the input is a selection input (e.g., a tap input, an air gesture, or a gaze input).
[0299] In some embodiments, in response to receiving the input, the electronic device displays a second visual representation corresponding to the text wherein the second visual representation is different from the first visual representation corresponding to the first input, -99- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) such as if recognized concept 620c only includes the text “birthday” after receiving the input in Fig.6S. In some embodiments, the second visual representation is displayed concurrently with the first visual representation. In some embodiments, the second visual representation is displayed in place of the first visual representation (e.g., the electronic device ceases displaying the visual representation). In some embodiments, the second visual representation is the previously inputted text describing the recognized concepts while the first visual representation is a cluster of icons, text, or a combination thereof describing the text. In some embodiments, the second visual representation is displayed overlaid over (e.g., on top of) the visual representation. In some embodiments, the second visual representation is displayed to the right, left, above, or below the first visual representation. In some embodiments, the second visual representation is displayed in the text entry region that originally received the prompt and / or text. Redisplaying the text phrase corresponding to text after receiving an input directed towards the visual representation (e.g., a cluster of images) allows a user to easily and efficiently see the text used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0300] In some embodiments, the text entry region includes the representation of the prompt (e.g., as described above), such as text entry region 614 including prompt 636 in Fig. 6S.
[0301] In some embodiments, while displaying the user interface with the text entry region including the representation of the prompt, the electronic device receives, via the one or more input devices, an input corresponding to a request to change one or more recognized concepts that will influence the generation of the automatically-generated visual media, such as if the electronic device 500 receives a typing or voice command input in Fig.6S. In some embodiments, the input has one or more characteristics of the inputs as described above. In some embodiments, the input includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments, changing the one or more recognized concepts includes changing the prompt from the first prompt to the second prompt, as described above. In some embodiments, changing the one or more recognized concepts includes modifying, adding, or deleting recognized concepts, as described above. -100- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0302] In some embodiments, in response to receiving the input, the electronic device displays, via the display generation component, an updated representation of the prompt in accordance with the request to change the one or more recognized concepts (e.g., by modifying or replacing the prior representation of the prompt), such as shown by the new prompt in the text entry field 614 in Fig.6W. In some embodiments, updating the representation of the prompt includes updating the text to include the modifications, additions, or deletions of recognized concepts. For example, if the representation of the prompt includes “watching a movie in space” and the user adds “monkey” to the recognized concepts, then the representation of the prompt is updated to include “monkey” (e.g., “watching a movie in space with a monkey”). In some embodiments, updating the representation of the prompt includes changing the representation of the prompt from the first prompt to the second prompt, as described above. In some embodiments, the representation of the prompt changes randomly (e.g., the user randomly generates a prompt). Updating the representation of the prompt as the prompt changes (e.g., by modifying the recognized concepts or changing the prompt) allows a user to easily and efficiently see the concepts used to generate the generative image, thereby reducing errors in output of the electronic device, and avoiding the need for additional input to correct such errors.
[0303] In some embodiments, in response to receiving the prompt (e.g., the first prompt or the second prompt, as described above), the electronic device displays, via the display generation component, a representation of the automatically-generated visual media in the user interface that is influenced by the prompt information (e.g., the recognized concepts identified from one or more portions of the prompt), such as shown by representation 622c in Fig.6W. In some embodiments, the representation of the automatically-generated visual media is a preview of the automatically-generated visual media that is based on the recognized concepts. In some embodiments, the representation of the automatically-generated visual media (e.g., the preview) is a lower resolution version of the automatically-generated visual media. In some embodiments, the representation of the automatically-generated visual media does not include all the characteristics of the automatically-generated visual media. Displaying a representation of the automatically- generated visual media allows a user to easily and efficiently see the representation of the automatically-generated visual media while also seeing the recognized concepts, thereby reducing the need for additional inputs to change the concepts used to generate the automatically-generated visual media. -101- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1)
[0304] In some embodiments, while displaying the first visual indication corresponding to the first recognized concept and while displaying the representation of the automatically-generated visual media, the electronic device receives, via the one or more input devices, a first input corresponding to a request to modify one or more recognized concepts that will influence the generation of the automatically-generated visual media, such as the input (e.g., a tap or a long press) including contact 672b in Fig.6X. In some embodiments, the first input has one or more characteristics of the inputs as described above. In some embodiments, the first input includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments modifying the one or more recognized concepts includes modifying the prompt, such as described above. In some embodiments, modifying the one or more recognized concepts includes adding a recognized concept from a database of recognized concepts, as described above. In some embodiments, modifying the one or more recognized concepts includes deleting / removing one or more recognized concepts that are currently being used to influence the generation of the automatically-generated visual media.
[0305] In some embodiments, in response to receiving the first input, the electronic device displays, via the display generation component, an updated representation of the automatically-generated visual media in accordance with the modifications of the one or more recognized concepts, such as shown by the representation 622c in Fig.6Y including stars in response to the addition of the recognized concept 620l. In some embodiments, the modification is an addition of a second recognized concept and updating the representation of the automatically-generated visual media includes integrating the second recognized concept with the recognized concept in the representation of the automatically-generated visual media. In some embodiments, updating the representation of the automatically-generated visual media includes regenerating the representation of the automatically-generated visual media to include the second recognized concept. Additionally, in some embodiments, removing a recognized concept causes the electronic device to update the representation of the automatically-generated visual media to no longer include the recognized concept. Additionally, in some embodiments, modifying a recognized concept causes the electronic device to update the representation of the automatically-generated visual media to be influenced by the modification of the recognized concept Updating the representation of the -102- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) automatically-generated visual media in response to modifying the recognized concepts allows a user to visualize the automatically-generated visual media thereby the need for additional input to regenerate the representation of the automatically-generated visual media after modifying the recognized concepts.
[0306] In some embodiments, while displaying the user interface including the representation of the automatically-generated visual media, the electronic device receives, via the one or more input devices, a first input corresponding to a request to generate the automatically-generated visual media based on the prompt corresponding to the one or more recognized concepts, such as input (e.g., a tap or a long press) including contact 670c in Fig. 6Y. In some embodiments, the first input includes one or more characteristics of the inputs described above. In some embodiments, the first input is a selection input directed towards a selectable option, that when selected, generates the automatically-generated visual media. In some embodiments, the first input includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option, or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture).
[0307] In some embodiments, in response to receiving the first input, the electronic device ceases display of the representation of the automatically-generated visual media, such as no longer displaying representation 622c in Fig.6Z. In some embodiments, the electronic device also ceases displaying the user interface including the recognized concepts used to influence the generation of the automatically-generated visual media. In some embodiments, the representation of the automatically-generated visual media is a low fidelity representation of the automatically-generated visual media. In some embodiments, the representation of the automatically-generated visual media does not include all the features / characteristics that are included in the automatically-generated visual media. For example, the representation of the automatically-generated visual media does not fully encapsulate all the recognized concepts. For example, if the recognized concept is “birthday in space”, the representation of automatically-generated visual media optionally includes stars and a birthday cake, wherein the automatically-generated visual media include stars, a birthday cake, birthday banners, all while set inside a space shuttle capsule.
[0308] In some embodiments, in response to receiving the first input, the electronic device displays, via the display generation component, a second representation of the automatically-generated visual media, wherein the second representation of the -103- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) automatically-generated visual media is of higher fidelity than the representation of the automatically-generated visual media, such as the automatically-generated visual content shown in user interface 629e in Fig.6Z. In some embodiments, and as described above, the automatically-generated visual media includes additional features and characteristics that are not included in the representation of the automatically-generated visual media. For example, the automatically-generated visual media is of higher resolution than the representation of the automatically-generated visual media. For example, the automatically-generated visual media includes more color depth, higher dynamic range, higher resolution, more detailed, larger scale, different colors, additional objects, and / or additional representations of recognized concepts than the representation of the automatically-generated visual media. In some embodiments, displaying the automatically-generated visual media includes displaying the automatically-generated visual media in a second user interface while ceasing display of the user interface including the recognized concepts. In some embodiments, the electronic device displays the automatically-generated visual media at a second size greater than a first size that the representation of the automatically-generated visual media is displayed with. Displaying a high fidelity representation of the automatically-generated visual media in response to receiving an input allows the user to quickly and efficiently see the automatically-generated visual media thereby reducing erroneous inputs to the electronic device.
[0309] In some embodiments, displaying the representation of the automatically- generated visual media includes animating the representation of the automatically-generated visual media over time, such as the boundary of the representation 622c changing over time in Fig.6Y. In some embodiments, animating the representation of the automatically- generated visual media includes changing the boundary location between the representation of the automatically-generated visual media and the remainder of the user interface, which optionally changes the separation and / or distance between one or more visual indications of the one or more recognized concepts and the representation of the automatically-generated visual media. In some embodiments, animating the representation of the automatically- generated visual media includes changing the color, pattern, and / or texture of the boundary between the representation of the automatically-generated visual media and the remainder of the user interface. In some embodiments, the electronic device animates the representation of the automatically-generated visual media in response to detecting an input including a prompt and / or a recognized concept. Animating the representation of the automatically-generated -104- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) visual media allows the user to quickly identify the representation of the automatically- generated visual media thereby reducing erroneous inputs to the electronic device.
[0310] In some embodiments, while displaying a second representation of the automatically-generated visual media (e.g., the second representation of the automatically- generated visual media is displayed at a larger size but at a same resolution and / or fidelity as the representation of the automatically-generated visual media. In some embodiments, the second representation of the automatically-generated visual media is the same as the representation of the automatically-generated visual media), the electronic device receives, via the one or more input devices, a sequence of one or more inputs corresponding to a request to display different automatically-generated visual media based on the prompt, such as the input (e.g., a tap or a long press) including contact 670i in Fig. JJ and the input (e.g., a tap or a long press) including contact 674b in Fig.6KK. In some embodiments, the second representation of the automatically-generated visual media is displayed as a result of the sequence of one or more inputs (e.g., a selection input, such as a tap or a click or an air gesture) directed towards the representation of the automatically-generated visual media. In some embodiments, the input has one or more characteristics of the inputs as described above. In some embodiments, the representation of the automatically-generated visual media displayed in response to receiving the prompt is the same as the second representation of the automatically-generated visual media. In some embodiments, the sequence of one or more inputs directed to the second representation includes a selection input, such as a tap or long press with a contact (e.g., a finger, or stylus), selection with an indirect input device (e.g., mouse, remote control, or trackpad) that is directed to a location of a selectable option (e.g., the location of the second representation), or focus indicator such as a cursor or selection ring) and / or a gaze input (optionally as part of an air gesture). In some embodiments, the sequence of one or more inputs is a swipe input directed towards the second representation of the automatically-generated visual media. For example, the swipe input is a swipe (e.g., a tap and drag) with a contact, a swipe with an indirect input device and / or a gaze input. In some embodiments, the representation of the second automatically-generated visual media is a different image than the second representation of the automatically-generated visual media. In some embodiments, the representation of the second automatically-generated visual media is generated using the same prompt as the second representation of the automatically- generated visual media. In some embodiments, since the process to generate automatically- generated visual media is non-deterministic, described below, the second representation of -105- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) the automatically-generated visual media is a different image (e.g., includes one or more different features) than the representation of the second automatically-generated visual media. In some embodiments, the input corresponding to the request to replace the second representation of the automatically-generated visual media with the representation of second automatically-generated visual media different from the automatically-generated visual media and based on the prompt does not alter the prompt and / or one or more or any of the recognized concepts used to generate the representation of the automatically-generated visual media and / or the second representation of the automatically-generated visual media. For example, the recognized concepts in Fig.6JJ and Fig.6MM are the same.
[0311] In some embodiments, in response to receiving the sequence of one or more inputs corresponding to the request to display different automatically-generated visual media based on the prompt, the electronic device displays, via the display generation component, a representation of the second automatically-generated visual media, such as representation 622b-b in Fig.6LL. In some embodiments, the second automatically-generated visual media is different from the automatically-generated visual media (e.g., representation 622b-b of the second automatically-generated visual content in Fig.6LL is different than representation 622b-a of the automatically-generated visual content in Fig.6KK), the second automatically- generated visual media was generated based on the prompt (e.g., the same prompt that was used to generate the second representation of the automatically-generated visual media) (e.g., the representation 622b-b of the second automatically-generated visual content is generated using the same recognized concepts (e.g., recognized concepts 620d through 620h) as the representation 622b-a in Fig.6KK), and the second automatically-generated visual media is displayed at a location that was previously occupied by the second representation of the automatically-generated visual content (e.g., representation 622b-b in Fig.6LL is displayed at the same place as representation 622b-a in Fig.6KK). In some embodiments, the second representation of the automatically-generated visual media ceases to be displayed. In some embodiments, a visual prominence of the second representation of the automatically- generated visual media is reduce (e.g., by reducing a size, opacity, brightness, saturation, and / or other visual property of the second representation of the automatically-generated visual content). In some embodiments, the second representation of the automatically- generated visual media is moved to a different location in the user interface that includes the second automatically-generated visual media. In some em...
Claims
Attorney Docket No.106842222440 (P65924WO1) CLAIMS What is claimed is:
1. A method comprising: at an electronic device in communication with a display generation component and one or more input devices: receiving, via the one or more input devices, a prompt for use in creating automatically-generated visual media that is generated at least partially using one or more autonomous processes; and in response to receiving the prompt, displaying, via the display generation component, a user interface that includes prompt information, wherein displaying the user interface includes concurrently displaying: a representation of the prompt; and a first visual indication that a first portion of the prompt has been identified as a first recognized concept that will influence generation of the automatically-generated visual media.
2. The method of claim 1, wherein displaying the user interface that includes the prompt information further includes displaying a second visual indication that a second portion of the prompt has been identified as a second recognized concept that will influence the generation of the automatically-generated visual media, wherein the second visual indication is displayed concurrently with the first visual indication in the user interface, the second visual indication is different than the first visual indication and the second portion of the prompt is different than the first portion of the prompt.
3. The method of any of claims 1-2, wherein the prompt includes starting media for use in creating the automatically-generated visual media, and the starting media influences the generation of the automatically-generated visual media.
4. The method of any of claims 1-3, wherein the first recognized concept includes a keyword or set of related keywords.
5. The method of any of claims 1-4, wherein the first recognized concept includes visual media. -574- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 6. The method of any of claim 1-5, wherein the first recognized concept includes a style prompt.
7. The method of any of claims 1-6, further comprising: while displaying the user interface that includes the prompt information: detecting one or more inputs corresponding to a request to modify one or more recognized concepts that will influence the generation of the automatically-generated visual media; and in response to detecting the one or more inputs: modifying the one or more recognized concepts that will influence the generation of the automatically-generated visual media in accordance with the one or more inputs.
8. The method of claim 7, wherein the one or more inputs corresponding to the request to modify the one or more recognized concepts that will influence the generation of the automatically-generated visual media include an input to modify the prompt; and in response to receiving the one or more inputs: updating the representation of the prompt to a representation of a second prompt in accordance with modifications to the prompt indicated by the one or more inputs; and displaying, in the user interface, one or more second visual indications corresponding to one or more portions of the second prompt that have been identified as recognized concepts.
9. The method of any claims 7-8, wherein: the one or more inputs corresponding to the request to modify the one or more recognized concepts that will influence the generation of the automatically-generated visual media include an input corresponding to a request to add a second recognized concept that will influence the generation of the automatically-generated visual media; and in response to receiving the one or more inputs: displaying a second visual indication corresponding to the second recognized concept concurrently with the first visual indication corresponding to the recognized concept.
10. The method of any of claims 7-9, wherein: -575- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) the one or more inputs corresponding to the request to modify the one or more recognized concepts that will influence the generation of the automatically-generated visual media include an input corresponding to a request to delete the first recognized concept; and in response to receiving the one or more inputs: ceasing display of the first visual indication.
11. The method of any of claims 1-10, further comprising: while displaying the first visual indication corresponding to the first recognized concept, receiving, via the one or more input devices, a second input corresponding to a request to add a second recognized concept, wherein the second recognized concept is in a same category as the first recognized concept; and in response to receiving the second input: ceasing display of the first visual indication corresponding to the first recognized concept; and displaying a second visual indication corresponding to the second recognized concept that will influence the generation of the automatically-generated visual media.
12. The method of any of claims 1-11, wherein displaying the user interface further includes: displaying a first plurality of representations of concepts representing recognized concepts of a first category within a database of recognized concepts; while displaying the first plurality of representations of concepts: receiving, via the one or more input devices, an input corresponding to a request to display a second plurality of representations of concepts representing recognized concepts of a second category within the database, the second category different from the first category; and in response to receiving the input corresponding to the request to display the second plurality of representations of concepts, displaying, in the user interface, the second plurality of representations of concepts.
13. The method of any of claims 1-12, further comprising: while displaying the user interface that includes the prompt information, displaying a selectable option for removing a plurality of recognized concepts that will influence generation of the automatically-generated visual content; -576- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) while displaying the selectable option, detecting, via the one or more input devices, an input directed towards the selectable option; and in response to detecting the input, removing the plurality of recognized concepts from being used to influence generation of the automatically-generated visual content: ceasing display of one or more visual indications corresponding to the plurality of the recognized concepts; and ceasing using the plurality of the recognized concepts to influence the generation of the automatically-generated visual content.
14. The method of claim 13, wherein: removing the plurality of recognized concepts includes removing all of the recognized concepts that will influence the generation of the automatically-generated visual content.
15. The method of any of claims 1-14, further comprising: while displaying the first visual indication corresponding to the first recognized concept, receiving, via the one or more input devices, a first input corresponding to a request to add a second recognized concept corresponding to a representation of a person that will influence the generation of the automatically-generated visual media; and in response to receiving the first input: displaying, in the user interface, a second visual indication corresponding to the second recognized concept that will influence the generation of the automatically- generated visual media.
16. The method of any of claims 1-15, further comprising: while displaying the first visual indication corresponding to the first recognized concept, receiving, via the one or more input devices, a first input corresponding to a request to add a second recognized concept corresponding to a representation of an animal that will influence the generation of the automatically-generated visual media; and in response to receiving the first input: displaying, in the user interface, a second visual indication corresponding to the second recognized concept that will influence the generation of the automatically- generated visual media. -577- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 17. The method of any of claims 1-16, wherein displaying the user interface further comprises displaying a text entry region configured to receive text corresponding to additional prompt information corresponding to one or more concepts that will influence the generation of the automatically-generated visual media.
18. The method of claim 17, further comprising: while displaying the user interface including the text entry region: receiving, via the one or more input devices, a first input that includes text directed towards the text entry region; and in response to receiving the first input: in accordance with a determination that the first input corresponds to a plurality of recognized concepts that will influence the generation of the automatically- generated visual media, displaying, in the user interface, a first visual representation corresponding to the first input that indicates the plurality of recognized concepts of the text.
19. The method of claim 18, further comprising: while displaying the first visual representation corresponding to the first input that indicates the plurality of recognized concepts of the text: receiving, via the one or more input devices, an input directed towards the visual representation; in response to receiving the input: displaying a second visual representation corresponding to the text wherein the second visual representation is different from the first visual representation corresponding to the first input.
20. The method of any of claims 17-18, wherein the text entry region includes the representation of the prompt; and the method further comprises: while displaying the user interface with the text entry region including the representation of the prompt: receiving, via the one or more input devices, an input corresponding to a request to change one or more recognized concepts that will influence the generation of the automatically-generated visual media; and in response to receiving the input: -578- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) displaying, via the display generation component, an updated representation of the prompt in accordance with the request to change the one or more recognized concepts.
21. The method of any of claims 1-20, further comprising, in response to receiving the prompt, displaying, via the display generation component, a representation of the automatically-generated visual media in the user interface that is influenced by the prompt information.
22. The method of claim 21, further comprising: while displaying the first visual indication corresponding to the first recognized concept and while displaying the representation of the automatically-generated visual media, receiving, via the one or more input devices, a first input corresponding to a request to modify one or more recognized concepts that will influence the generation of the automatically-generated visual media; and in response to receiving the first input: displaying, via the display generation component, an updated representation of the automatically-generated visual media in accordance with the modifications of the one or more recognized concepts.
23. The method of any of claims 21-22, further comprising: while displaying the user interface including the representation of the automatically- generated visual media: receiving, via the one or more input devices, a first input corresponding to a request to generate the automatically-generated visual media based on the prompt corresponding to the one or more recognized concepts; and in response to receiving the first input: ceasing display of the representation of the automatically-generated visual media; and displaying, via the display generation component, a second representation of the automatically-generated visual media, wherein the second representation of the automatically-generated visual media is of higher fidelity than the representation of the automatically-generated visual media. -579- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 24. The method of any of claims 21-23, wherein displaying the representation of the automatically-generated visual media includes animating the representation of the automatically-generated visual media over time.
25. The method of any of claims 21-24, further comprising: while displaying a second representation of the automatically-generated visual media, receiving, via the one or more input devices, a sequence of one or more inputs corresponding to a request to display different automatically-generated visual media based on the prompt; in response to receiving the sequence of one or more inputs corresponding to the request to display different automatically-generated visual media based on the prompt: displaying, via the display generation component, a representation of the second automatically-generated visual media wherein: the second automatically-generated visual media is different from the automatically-generated visual media; the second automatically-generated visual media was generated based on the prompt; and the second automatically-generated visual media is displayed at a location that was previously occupied by the second representation of the automatically- generated visual content.
26. The method of claim 25, wherein the input is a movement input directed towards the second representation of the automatically-generated visual media.
27. The method of any of claims 25-26, wherein prior to receiving the sequence of one or more inputs corresponding to the request to display different automatically-generated visual media based on the prompt, the representation of the second automatically-generated visual media had not been generated, the method further comprising in response to receiving the sequence of one or more inputs corresponding to the request to display different automatically-generated visual media based on the prompt, generating the representation of the second automatically-generated visual media based on the prompt.
28. The method of any of claims 21-27, wherein displaying the representation of the automatically-generated visual content in the user interface includes displaying a selectable -580- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) option for generating a three-dimensional representation of the automatically-generated visual content, and the method further comprises: detecting, via the one or more input device, an input directed towards the selectable option; and in response to detecting the input, generating a three-dimensional representation of the automatically-generated visual content.
29. The method of any of claims 1-28, further comprising: while displaying the user interface including the representation of the prompt and the first visual indication: receiving, via the one or more input devices, an input corresponding to a request to generate the automatically-generated visual media; and in response to receiving the input: initiating a process to generate the automatically-generated visual media.
30. The method of any of claims 1-29, further comprising: while the automatically-generated visual media is influenced by one or more concepts of the prompt, receiving, via the one or more input devices, a first input corresponding to a request to regenerate the automatically-generated visual media; and in response to receiving the first input: displaying, via the display generation component, a second automatically-generated visual media based on the prompt corresponding to the one or more recognized concepts, wherein the second automatically-generated visual media is different from the automatically-generated visual media.
31. The method of any of claims 1-30, further comprising: displaying a gallery of automatically-generated visual content; while displaying the gallery of automatically-generated visual content, receiving, via the one or more input devices, a first input directed towards a selectable option to display the user interface that includes the prompt information; and in response to receiving the first input: -581- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) ceasing display of the gallery of automatically-generated visual content; and displaying, via the display generation component, the user interface that includes the prompt information.
32. The method of any of claims 1-31, further comprising: receiving, via the one or more input devices, a first input corresponding to a request to generate second automatically-generated visual media based on a second prompt, wherein the second prompt corresponds to a second recognized concept; and in response to receiving the first input: in accordance with a determination that the second recognized concept satisfies one or more criteria, initiating a process to generate the second visual generative visual content based on the second recognized concept; and in accordance with a determination that the second recognized concept does not satisfy the one or more criteria, forgoing initiating the process to generate the second automatically-generated visual media based on the second recognized concept.
33. The method of claim 32, wherein forgoing initiating the process to generate the second automatically-generated visual media based on the second recognized concept in response to receiving the first input further comprises: initiating a process to generate third automatically-generated visual media based on a third recognized concept corresponding to the second prompt that satisfies the one or more criteria and not based on the second recognized concept that does not satisfy the one or more criteria.
34. The method of any of claims 32-33, wherein forgoing initiating the process to generate the second automatically-generated visual media based on the second recognized concept in response to receiving the first input further comprises: forgoing generating the second automatically-generated visual media.
35. The method of any of claims 32-34, wherein forgoing initiating the process to generate the second visual generative visual content based on the second recognized concept further comprises: -582- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) displaying, via the display generation component, a visual indication indicating a portion of the second prompt that does not satisfy the one or more criteria in the user interface including a representation of the second prompt.
36. An electronic device that is in communication with a display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: receiving, via the one or more input devices, a prompt for use in creating automatically-generated visual media that is generated at least partially using one or more autonomous processes; and in response to receiving the prompt, displaying, via the display generation component, a user interface that includes prompt information, wherein displaying the user interface includes concurrently displaying: a representation of the prompt; and a first visual indication that a first portion of the prompt has been identified as a first recognized concept that will influence generation of the automatically- generated visual media.
37. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: receiving, via one or more input devices, a prompt for use in creating automatically- generated visual media that is generated at least partially using one or more autonomous processes; and in response to receiving the prompt, displaying, via a display generation component, a user interface that includes prompt information, wherein displaying the user interface includes concurrently displaying: a representation of the prompt; and -583- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) a first visual indication that a first portion of the prompt has been identified as a first recognized concept that will influence generation of the automatically-generated visual media.
38. An electronic device, comprising: one or more processors; memory; means for, receiving, via one or more input devices, a prompt for use in creating automatically-generated visual media that is generated at least partially using one or more autonomous processes; and means for, in response to receiving the prompt, displaying, via a display generation component, a user interface that includes prompt information, wherein displaying the user interface includes concurrently displaying: a representation of the prompt; and a first visual indication that a first portion of the prompt has been identified as a first recognized concept that will influence generation of the automatically-generated visual media.
39. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for, receiving, via one or more input devices, a prompt for use in creating automatically-generated visual media that is generated at least partially using one or more autonomous processes; and means for, in response to receiving the prompt, displaying, via a display generation component, a user interface that includes prompt information, wherein displaying the user interface includes concurrently displaying: a representation of the prompt; and a first visual indication that a first portion of the prompt has been identified as a first recognized concept that will influence generation of the automatically-generated visual media.
40. An electronic device, comprising: one or more processors; memory; and -584- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 1-35.
41. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 1-35.
42. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 1-35.
43. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 1-35.
44. A method comprising: at an electronic device in communication with a display generation component and one or more input devices: displaying, via the display generation component, a generative visual content user interface including one or more representations of one or more previously generated automatically-generated visual content that were generated at least partially using one or more autonomous processes; while displaying the generative visual content user interface, receiving, via the one or more input devices, a first input directed to a first representation of a first previously generated automatically-generated visual content of the one or more previously generated automatically-generated visual content; in response to receiving the input, displaying, via the display generation component, an editing user interface for editing the first previously generated automatically-generated visual content including one or more selectable options to edit one or more parameters used to generate the first previously generated automatically-generated visual content, wherein the -585- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) editing user interface indicates one or more parameters that were used to generate the first previously automatically-generated visual content; and while displaying the editing user interface, detecting a set of one or more inputs including a second input directed to a respective selectable option of the one or more selectable options; and in response to detecting the set of one or more inputs including the second input directed to the respective selectable option of the one or more selectable options, changing the one or more parameters.
45. The method of claim 44, wherein the editing user interface includes a first visual indication of a first parameter used to generate the first previously generated automatically- generated visual content.
46. The method of any of claims 44 – 45, further comprising: while displaying the editing user interface: changing the one or more parameters used to generate the first previously generated automatically-generated visual content; after changing the one or more parameters, receiving, via the one or more input devices, a second input corresponding to a request to generate a second automatically- generated visual content based on the changed one or more parameters; and in response to receiving the second input: initiating a process to generate the second automatically-generated visual content, different than the first previously automatically-generated visual content, based on the changed one or more parameters; and after generating the second automatically-generated visual content, displaying, via the display generation component, the second automatically-generated visual content.
47. The method of claim 46, further comprising: after generating the second automatically-generated visual content: receiving, via the one or more input devices, a third input corresponding to a request to save the second automatically-generated visual content; and in response to receiving the third input: -586- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with a determination that the third input is directed towards a first option, replacing the first previously generated visual content in the generative visual content user interface with the second automatically-generated visual content; and in accordance with a determination that the third input is directed towards a second option, saving the second automatically-generated visual content as a new automatically-generated visual content to be displayed in the generative visual content user interface.
48. The method of any of claims 44-47, further comprising: while displaying the editing user interface for editing the first previously generated automatically-generated visual content, receiving a second input, via the one or more input devices, directed towards an option to generate a second automatically-generated visual content using the one or more parameters used to generate the first previously generated automatically-generated visual content; in response to receiving the second input, initiating a process to generate a second automatically-generated visual content, different than the first previously automatically-generated visual content, based on the one or more parameters; and after generating the second automatically-generated visual content, displaying, via the display generation component, the second automatically-generated visual content.
49. The method of any of claims 44-48, wherein displaying the editing user interface further includes concurrently displaying at least a portion of the generative visual content user interface.
50. The method of claim 44-49, wherein the one or more parameters includes a person represented by the first previously generated automatically-generated visual content, and changing the one or more parameters includes receiving an input, via the one or more input devices, corresponding to a request to change the person represented by the first previously generated visual content from a first person to a second person different from the first person.
51. The method of claim 50, further comprising: -587- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) while the first previously automatically-generated visual content is associated with the first person, receiving an input corresponding to selection of an image of a second person; and in response to receiving the input corresponding to selection of the image of the second person, using the image of the second person as a recognized concept that will influence generation of the automatically-generated visual media.
52. The method of claim 51, wherein receiving the input corresponding to the selection of the image of the second person comprises: receiving a first input corresponding to a request to open a view of a media library; in response to receiving the first input, displaying the view of the media library; and while displaying the view of the media library, receiving a second input corresponding to the selection of the image of the second person in the media library.
53. The method of any of claims 51-52, wherein receiving the input corresponding to the selection of the image of the second person comprises: receiving a first input corresponding to a request to open a view of a camera application including a live preview of a physical environment; in response to receiving the first input, displaying the view of the camera application; and while displaying the view of the camera application, receiving a second input capturing the image of the second person in the camera application of the electronic device.
54. The method of any of claims 50-53, wherein changing the person represented by the first previously generated visual content from the first person to the second person further comprises: receiving, via the one or more input devices, an input corresponding to a request to select the second person from a collection of identified people that are identified from a media library associated with the electronic device; and in response to receiving the input corresponding to the request to select the second person from the collection of identified people, updating the person represented by the first previously generated visual content from the first person to the second person based on one or more media items identified as including the second person. -588- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 55. The method of claim 44-54, wherein the one or more parameters includes a person represented by the first previously generated automatically-generated visual content, and changing the one or more parameters includes receiving an input, via the one or more input devices, corresponding to a request to change the person represented by the first previously generated visual content from a person to an animal different from the person.
56. The method of any of claims 44-55, wherein displaying the generative visual content user interface includes concurrently displaying, via the display generation component, in the generative visual content user interface: a representation of the first previously generated visual content associated with a first person; and a representation of a second previously generated visual content associated with a second person different from the first person.
57. An electronic device that is in communication with a display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: displaying, via the display generation component, a generative visual content user interface including one or more representations of one or more previously generated automatically-generated visual content that were generated at least partially using one or more autonomous processes; while displaying the generative visual content user interface, receiving, via the one or more input devices, a first input directed to a first representation of a first previously generated automatically-generated visual content of the one or more previously generated automatically-generated visual content; in response to receiving the input, displaying, via the display generation component, an editing user interface for editing the first previously generated automatically- generated visual content including one or more selectable options to edit one or more parameters used to generate the first previously generated automatically-generated visual -589- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) content, wherein the editing user interface indicates one or more parameters that were used to generate the first previously automatically-generated visual content; and while displaying the editing user interface, detecting a set of one or more inputs including a second input directed to a respective selectable option of the one or more selectable options; and in response to detecting the set of one or more inputs including the second input directed to the respective selectable option of the one or more selectable options, changing the one or more parameters.
58. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: displaying, via a display generation component, a generative visual content user interface including one or more representations of one or more previously generated automatically-generated visual content that were generated at least partially using one or more autonomous processes; while displaying the generative visual content user interface, receiving, via an one or more input devices, a first input directed to a first representation of a first previously generated automatically-generated visual content of the one or more previously generated automatically-generated visual content; in response to receiving the input, displaying, via the display generation component, an editing user interface for editing the first previously generated automatically-generated visual content including one or more selectable options to edit one or more parameters used to generate the first previously generated automatically- generated visual content, wherein the editing user interface indicates one or more parameters that were used to generate the first previously automatically-generated visual content; and while displaying the editing user interface, detecting a set of one or more inputs including a second input directed to a respective selectable option of the one or more selectable options; and in response to detecting the set of one or more inputs including the second input directed to the respective selectable option of the one or more selectable options, changing the one or more parameters.
59. An electronic device, comprising: -590- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) one or more processors; memory; means for, displaying, via a display generation component, a generative visual content user interface including one or more representations of one or more previously generated automatically-generated visual content that were generated at least partially using one or more autonomous processes; means for, while displaying the generative visual content user interface, receiving, via an one or more input devices, a first input directed to a first representation of a first previously generated automatically-generated visual content of the one or more previously generated automatically-generated visual content; in response to receiving the input, displaying, via the display generation component, an editing user interface for editing the first previously generated automatically-generated visual content including one or more selectable options to edit one or more parameters used to generate the first previously generated automatically-generated visual content, wherein the editing user interface indicates one or more parameters that were used to generate the first previously automatically-generated visual content; and means for, while displaying the editing user interface, detecting a set of one or more inputs including a second input directed to a respective selectable option of the one or more selectable options; and means for, in response to detecting the set of one or more inputs including the second input directed to the respective selectable option of the one or more selectable options, changing the one or more parameters.
60. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for, displaying, via a display generation component, a generative visual content user interface including one or more representations of one or more previously generated automatically-generated visual content that were generated at least partially using one or more autonomous processes; means for, while displaying the generative visual content user interface, receiving, via an one or more input devices, a first input directed to a first representation of a first previously generated automatically-generated visual content of the one or more previously generated automatically-generated visual content; -591- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to receiving the input, displaying, via the display generation component, an editing user interface for editing the first previously generated automatically-generated visual content including one or more selectable options to edit one or more parameters used to generate the first previously generated automatically-generated visual content, wherein the editing user interface indicates one or more parameters that were used to generate the first previously automatically-generated visual content; and means for, while displaying the editing user interface, detecting a set of one or more inputs including a second input directed to a respective selectable option of the one or more selectable options; and means for, in response to detecting the set of one or more inputs including the second input directed to the respective selectable option of the one or more selectable options, changing the one or more parameters.
61. An electronic device, comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 44-56.
62. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 44-56.
63. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 44-56.
64. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 44-56. -592- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 65. A method comprising: at an electronic device in communication with a display generation component and one or more input devices: displaying, via the display generation component, a first user interface of a first application; while displaying the first user interface of the first application, receiving, via the one or more input devices, a first input corresponding to a request to insert a first automatically- generated visual content into the first user interface; in response to receiving first input, displaying a second user interface, wherein the second user interface is a user interface for inserting one or more automatically-generated visual content into the first user interface of the first application; while displaying the second user interface, detecting a second set of one or more inputs directed to the second user interface; and in response to detecting the second set of one or more inputs, adding a representation of the automatically-generated visual content that was generated at least partially using one or more autonomous processes to the first user interface.
66. The method of claim 65, wherein displaying the second user interface further comprises concurrently displaying a plurality of representations of recommended automatically-generated visual content including the automatically-generated visual content.
67. The method of claim 66, wherein the recommended automatically-generated visual content are previously generated automatically-generated visual content.
68. The method of any of claims 66-67, wherein adding the automatically-generated visual content to the first user interface further comprises adding a representation of the automatically-generated visual content to a content entry field in the first user interface of the first application.
69. The method of claim 68, wherein the representation of the automatically-generated visual content in the first user interface includes a visual indication that the automatically- generated visual content is automatically-generated visual media, and the method further comprises: -593- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) displaying a third user interface for inserting non-automatically-generated visual media into the first user interface of the first application; while displaying the third user interface for inserting the non-automatically-generated visual media into the first user interface of the first application: detecting, via the one or more input devices, a respective set of one or more inputs directed to the third user interface; and in response to detecting the respective set of one or more inputs directed to the third user interface, displaying, via the display generation component, a representation of a non-automatically-generated visual content in the first user interface without displaying a visual indication that the non-automatically-generated visual content is automatically- generated visual media.
70. The method of any of claims 68-69, wherein the first application is a messages application.
71. The method of any of claims 68-70, further comprising: while displaying the representation of the automatically-generated visual content in the content entry field in the first user interface of the first application, receiving, via the one or more input devices, a respective set of one or more inputs corresponding to a request to send a message to a messaging conversion; and in response to receiving the respective set of one or more inputs corresponding to the request to send the message to the messaging conversion: sending a message to the messaging conversation that includes the automatically-generated visual content; and displaying, via the display generation component, a representation of the message that includes the automatically-generated visual content in the first user interface that includes a representation of the automatically-generated visual content.
72. The method of any of claims 68-71, further comprising: while displaying the representation of the automatically-generated visual content in the content entry field, receiving, via the one or more input devices, a respective set of one or more inputs corresponding to a request to enter text in the content entry field; and in response to receiving the respective set of one or more inputs corresponding to the request to enter text in the content entry field, displaying, via the display generation -594- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) component, the text in the content entry field concurrently with the representation of the automatically-generated visual content.
73. The method of any of claims 68-72, further comprising: while displaying, via the display generation component, the representation of the automatically-generated visual content, receiving, via the one or more input devices, a respective set of one or more inputs corresponding to a request to edit the automatically- generated visual content; and in response to receiving the respective set of one or more inputs corresponding to the request to edit the automatically-generated visual content, displaying an editing user interface including one or more controls for modifying one or more parameters used to influence the generation of the automatically-generated visual content.
74. The method of any of claims 65-73, further comprising: while displaying the second user interface, receiving, via the one or more input devices, a respective set of one or more inputs directed towards a selectable option that is selectable to create a new automatically-generated visual content; and in response to receiving the respective set of one or more inputs directed towards the selectable option that is selectable to create a new automatically-generated visual content, displaying, via the display generation component, an automatically-generated visual media creation user interface for generating the new automatically-generated visual content.
75. The method of claim 74, further comprising: while displaying the automatically-generated visual media creation user interface, receiving, via the one or more input devices, one or more parameters that will influence the generation of the new automatically-generated visual media that is generated at least partially using one or more autonomous processes.
76. The method of claim 75, wherein the one or more parameters includes text corresponding to one or more recognized concepts.
77. The method of any of claims 74-76, further comprising: after generating the new automatically-generated visual content and while displaying the automatically-generated visual media creation user interface, receiving, via the one or -595- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) more input devices, a respective set of one or more inputs corresponding to a request to save the new automatically-generated visual content; and in response to receiving the respective set of one or more inputs corresponding to the request to save the new automatically-generated visual content: in accordance with a determination that the respective set of one or more inputs corresponding to the request to save the new automatically-generated visual content is directed towards a first option, replacing the automatically-generated visual content with the new automatically-generated visual content; and in accordance with a determination that the respective set of one or more inputs corresponding to the request to save the new automatically-generated visual content is directed towards a second option, saving the new automatically-generated visual content as a second automatically-generated visual content to be displayed in the second user interface.
78. The method of claim 77, further comprising: in response to detecting the respective set of one or more inputs corresponding to the request to save the new automatically-generated visual content, adding the new automatically-generated visual content that was generated at least partially using the one or more autonomous processes into the first user interface of the first application.
79. An electronic device that is in communication with a display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: displaying, via the display generation component, a first user interface of a first application; while displaying the first user interface of the first application, receiving, via the one or more input devices, a first input corresponding to a request to insert a first automatically-generated visual content into the first user interface; in response to receiving first input, displaying a second user interface, wherein the second user interface is a user interface for inserting one or more automatically-generated visual content into the first user interface of the first application; -596- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) while displaying the second user interface, detecting a second set of one or more inputs directed to the second user interface; and in response to detecting the second set of one or more inputs, adding a representation of the automatically-generated visual content that was generated at least partially using one or more autonomous processes to the first user interface.
80. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: displaying, via a display generation component, a first user interface of a first application; while displaying the first user interface of the first application, receiving, via a one or more input devices, a first input corresponding to a request to insert a first automatically- generated visual content into the first user interface; in response to receiving first input, displaying a second user interface, wherein the second user interface is a user interface for inserting one or more automatically-generated visual content into the first user interface of the first application; while displaying the second user interface, detecting a second set of one or more inputs directed to the second user interface; and in response to detecting the second set of one or more inputs, adding a representation of the automatically-generated visual content that was generated at least partially using one or more autonomous processes to the first user interface.
81. An electronic device, comprising: one or more processors; memory; means for, displaying, via a display generation component, a first user interface of a first application; means for, while displaying the first user interface of the first application, receiving, via a one or more input devices, a first input corresponding to a request to insert a first automatically-generated visual content into the first user interface; -597- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) means for, in response to receiving first input, displaying a second user interface, wherein the second user interface is a user interface for inserting one or more automatically- generated visual content into the first user interface of the first application; means for, while displaying the second user interface, detecting a second set of one or more inputs directed to the second user interface; and means for, in response to detecting the second set of one or more inputs, adding a representation of the automatically-generated visual content that was generated at least partially using one or more autonomous processes to the first user interface.
82. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for, displaying, via a display generation component, a first user interface of a first application; means for, while displaying the first user interface of the first application, receiving, via a one or more input devices, a first input corresponding to a request to insert a first automatically-generated visual content into the first user interface; means for, in response to receiving first input, displaying a second user interface, wherein the second user interface is a user interface for inserting one or more automatically- generated visual content into the first user interface of the first application; means for, while displaying the second user interface, detecting a second set of one or more inputs directed to the second user interface; and means for, in response to detecting the second set of one or more inputs, adding a representation of the automatically-generated visual content that was generated at least partially using one or more autonomous processes to the first user interface.
83. An electronic device, comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 65-78.
84. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more -598- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) processors of an electronic device, cause the electronic device to perform any of the methods of claims 65-75.
85. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 65-78.
86. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 65-78.
87. A method comprising: at an electronic device in communication with a display generation component and one or more input devices: displaying, via the display generation component, a first user interface of a first application, wherein the first application is not an automatically-generated visual media application; while displaying the first user interface of the first application, receiving, via the one or more input devices, an input corresponding to a request to select a reference media item as a basis for generating an automatically-generated visual content that will be generated at least partially using one or more autonomous processes using the reference media item as input; in response to receiving the input, displaying an automatically-generated visual media creation user interface for generating an automatically-generated visual content that is based on the reference media item, wherein the automatically-generated visual media creation user interface includes one or more controls for selecting options to be used as additional inputs for generating an automatically-generated visual content based on the reference media item.
88. The method of claim 87, wherein the first application is a messaging application.
89. The method of claim 88, further comprising: while displaying the automatically-generated visual media creation user interface, receiving, via the one or more input devices, a respective set of one or more inputs corresponding to a request to generate the automatically-generated visual content; -599- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to receiving the respective set of one or more inputs corresponding to the request to generate the automatically-generated visual content: generating the automatically-generated visual content; and adding a representation of the automatically-generated visual content to a content entry field in the first user interface of the messaging application.
90. The method of any of claims 87-89, wherein receiving the input corresponding to the request to select the reference media item as the basis of generating the automatically- generated visual content further comprises receiving the input directed towards a first visual indication displayed, via the display generation component, in conjunction with a representation of the reference media item.
91. The method of claim 90, further comprising: while displaying the automatically-generated visual media creating user interface, receiving a respective set of one or more inputs corresponding to a request to generate the automatically-generated visual content; in response to receiving the respective set of one or more inputs corresponding to a request to generate the automatically-generated visual content, generating the automatically- generated visual content that is based on the reference media item; and after generating the automatically-generated visual content, displaying, in the first user interface: a representation of the automatically-generated visual content; and a second visual indication displayed in conjunction with the representation of the automatically-generated visual content, wherein the first visual indication has a first value for a visual characteristic and the second visual indication has a second value for the visual characteristics that is different from the first value.
92. The method of claim 91, further comprising: while displaying the representation of the automatically-generated visual content in the first user interface of the first application with the second visual indication displayed in conjunction with the representation of the automatically-generated visual content, receiving, via the one or more input devices, a respective set of one or more inputs corresponding to a request to send a message including the automatically-generated visual content to a messaging conversation; and -600- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to receiving the respective set of one or more inputs corresponding to the request to send a message including the automatically-generated visual content to the messaging conversation: sending the message including the automatically-generated visual content to the messaging conversation; and displaying, via the display generation component, in the first user interface, a representation of the message that includes the representation of the automatically-generated visual media including a third visual indication displayed in conjunction with the representation of the automatically-generated visual content with a third value for the visual characteristic that is different from the second value.
93. The method of claim 91, further comprising: while displaying the representation of the automatically-generated visual content in the first user interface of the first application with the second visual indication displayed in conjunction with the representation of the automatically-generated visual content, receiving, via the one or more input devices, a respective set of one or more inputs corresponding to a request to send a message including the automatically-generated visual content to a messaging conversation; and in response to receiving the respective set of one or more inputs corresponding to the request to send a message including the automatically-generated visual content to a messaging conversation: sending the message including the automatically-generated visual content to the messaging conversation; and displaying, in the first user interface, a representation of the message that includes the representation of the automatically-generated visual media without displaying the second visual indication.
94. The method of any of claims 91-93, further comprising: while displaying the first user interface of the first application, receiving an indication of a first message including a representation of a second automatically-generated visual content; and in response to receiving the indication of the first message, displaying, in the first user interface, the representation of the message that includes the representation of the second -601- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) automatically-generated visual content including the third visual indication displayed in conjunction with the representation of the second automatically-generated visual content.
95. The method of claim 94, wherein the first user interface of the first application concurrently includes the representation of the automatically-generated visual content including the third visual indication and the representation of the second automatically- generated visual content including the third visual indication.
96. The method of any of claims 94-95, further comprising: while displaying the first user interface of the first application, receiving an indication of a second message that includes a representation of a non-automatically-generated visual content; and in response to receiving the indication of the second message, displaying a representation of the second message including the representation of the non-generative automatically-generated visual content without displaying the third visual indication displayed in conjunction with the representation of the non-automatically-generated visual content.
97. The method of any of claims 87-96, wherein the input corresponding to the request to select the reference media item is included in a sequence of one or more inputs that further comprises: receiving a respective set of one or more inputs corresponding to a request to display a media library; in response to receiving the respective set of one or more inputs corresponding to the request to display the media library, displaying, via the display generation component the media library; while displaying the media library, receiving a respective set of one or more inputs corresponding to a request to select a respective media item from a plurality of media items in the media library as the reference media item; and in response to receiving the respective set of one or more inputs corresponding to the request to select the reference media item, using the respective media item as the reference media item. -602- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 98. The method of any of claims 87-97, wherein the input corresponding to the request to select the reference media item is included in a sequence of one or more inputs that further comprises: receiving a respective set of one or more inputs corresponding to a request to display a user interface of a camera application including a live preview of a physical environment; in response to receiving the respective set of one or more inputs corresponding to the request to display a user interface of a camera, displaying, on the first user interface, the user interface of the camera application; while displaying the user interface of the camera application, receiving a respective set of one or more inputs corresponding to a request to capture a new media item using the camera application for use as the reference media item; and in response to receiving the respective set of one or more inputs corresponding to the request to capture a new media item using the camera application, capturing a new media item using the camera application and using the new media item as the reference media item.
99. The method of any of claims 87-98, wherein displaying the automatically-generated visual media creation user interface in response to receiving the input corresponding to the request to the select the reference media item includes displaying a representation of the reference media item in the automatically-generated visual media creation user interface.
100. The method of any of claims 87-99, further comprising: while displaying the automatically-generated visual media creation user interface, receiving, via the one or more input devices, one or more parameters to be used as the options to be used as additional inputs for generating the automatically-generated visual content.
101. The method of claim 100, wherein receiving, via the one or more input devices, the one or more parameters includes receiving a language prompt to be used to influence the generation of the automatically-generated visual content.
102. The method of any of claims 87-101, further comprising: receiving, via the one or more input devices, a respective set of one or more inputs providing one or more parameters for generating the automatically-generated visual content and requesting to generate the automatically-generated visual content based on the reference media item and the one or more parameters; and -603- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to receiving the respective set of one or more inputs providing one or more parameters for generating the automatically-generated visual content: displaying, via the display generation component, a representation of the automatically-generated visual content in the first user interface of the first application generated based on the one or more parameters, wherein displaying the representation of the generative visual content includes: in accordance with a determination that the one or more parameters are one or more first parameters, displaying the representation of the automatically-generated visual content with a first appearance; and in accordance with a determination that the one or more parameters are one or more second parameters different from the one or more first parameters, displaying the representation of the automatically-generated visual content with a second appearance that is different from the first appearance.
103. The method of claim 102, further comprising: after generating the automatically-generated visual content, receiving, via the one or more input devices, a respective set of one or more inputs corresponding to a request to save the automatically-generated visual content; and in response to receiving the respective set of one or more inputs corresponding to the request to save the automatically-generated visual content: saving the automatically-generated visual content to a gallery of automatically-generated visual content in an automatically-generated visual media application.
104. An electronic device that is in communication with a display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: displaying, via the display generation component, a first user interface of a first application, wherein the first application is not an automatically-generated visual media application; -604- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) while displaying the first user interface of the first application, receiving, via the one or more input devices, an input corresponding to a request to select a reference media item as a basis for generating an automatically-generated visual content that will be generated at least partially using one or more autonomous processes using the reference media item as input; in response to receiving the input, displaying an automatically-generated visual media creation user interface for generating an automatically-generated visual content that is based on the reference media item, wherein the automatically-generated visual media creation user interface includes one or more controls for selecting options to be used as additional inputs for generating an automatically-generated visual content based on the reference media item.
105. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: displaying, via a display generation component, a first user interface of a first application, wherein the first application is not an automatically-generated visual media application; while displaying the first user interface of the first application, receiving, via a one or more input devices, an input corresponding to a request to select a reference media item as a basis for generating an automatically-generated visual content that will be generated at least partially using one or more autonomous processes using the reference media item as input; in response to receiving the input, displaying an automatically-generated visual media creation user interface for generating an automatically-generated visual content that is based on the reference media item, wherein the automatically-generated visual media creation user interface includes one or more controls for selecting options to be used as additional inputs for generating an automatically-generated visual content based on the reference media item.
106. An electronic device, comprising: one or more processors; memory; -605- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) means for, displaying, via a display generation component, a first user interface of a first application, wherein the first application is not an automatically-generated visual media application; means for, while displaying the first user interface of the first application, receiving, via a one or more input devices, an input corresponding to a request to select a reference media item as a basis for generating an automatically-generated visual content that will be generated at least partially using one or more autonomous processes using the reference media item as input; means for, in response to receiving the input, displaying an automatically-generated visual media creation user interface for generating an automatically-generated visual content that is based on the reference media item, wherein the automatically-generated visual media creation user interface includes one or more controls for selecting options to be used as additional inputs for generating an automatically-generated visual content based on the reference media item.
107. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for, displaying, via a display generation component, a first user interface of a first application, wherein the first application is not an automatically-generated visual media application; means for, while displaying the first user interface of the first application, receiving, via a one or more input devices, an input corresponding to a request to select a reference media item as a basis for generating an automatically-generated visual content that will be generated at least partially using one or more autonomous processes using the reference media item as input; means for, in response to receiving the input, displaying an automatically-generated visual media creation user interface for generating an automatically-generated visual content that is based on the reference media item, wherein the automatically-generated visual media creation user interface includes one or more controls for selecting options to be used as additional inputs for generating an automatically-generated visual content based on the reference media item.
108. An electronic device, comprising: one or more processors; -606- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 87-103.
109. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 87-103.
110. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 87-103.
111. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 87-103.
112. A method comprising: at an electronic device in communication with a display generation component, and one or more input devices: while displaying a first user interface, detecting a first event; and in response to detecting the first event: in accordance with a determination that the first event corresponds to a first functionality that outputs content generated using a first artificial intelligence (AI) model, displaying first visual information corresponding to the first event in the first user interface with a first visual effect, wherein the first visual effect includes a first visual characteristic; and in accordance with a determination that the first event corresponds to a second functionality, different from the first functionality, that outputs content generated using a second AI model, displaying second visual information corresponding to the first event in the first user interface with a second visual effect, different from the first visual effect, wherein the second visual effect includes the first visual characteristic. -607- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 113. The method of claim 112, further comprising: while displaying the first user interface, and in response to detecting the first event: in accordance with a determination that the first event corresponds to a third functionality, different from the first functionality and the second functionality, that outputs content generated using a third AI model displaying third visual information corresponding to the first event in the first user interface with a third visual effect, different from the first visual effect and the second visual effect, wherein the third visual effect includes the first visual characteristic.
114. The method of claim 113, further comprising: while displaying the first user interface, and in response to detecting the first event: in accordance with a determination that the first event corresponds to a fourth functionality, different from the first functionality, the second functionality, and the third functionality, that outputs content generated using a fourth AI model, displaying fourth visual information corresponding to the first event in the first user interface with a fourth visual effect, different from the first visual effect, the second visual effect, and the third visual effect, wherein the fourth visual effect includes the first visual characteristic.
115. The method of any one of claims 112-114, further comprising: while displaying the first user interface, and in response to detecting the first event: in accordance with a determination that the event corresponds to a respective functionality, different from the first functionality and the second functionality, that does not output content generated using an AI model, displaying respective visual information corresponding to the first event in the first user interface, wherein the respective visual information does not include the first visual characteristic.
116. The method of any one of claims 112-115, further comprising: while displaying the first user interface, detecting a second event, different from the first event; and in response to detecting the second event; -608- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with a determination that the second event corresponds to the first functionality, displaying a first animation sequence corresponding to the first functionality, wherein the first animation sequence corresponds to a respective animation sequence that was displayed as part of displaying the first visual information corresponding to the first functionality in response to the first event.
117. The method of any one of claims 112-116, wherein displaying the visual effects with the first visual characteristic includes displaying the visual effects with a graphical element.
118. The method of any one of claims 112-117, wherein displaying the first visual effect includes displaying the first visual effect with a first animated visual effect, and wherein displaying the second visual effect includes displaying the second visual effect with a second animated visual effect.
119. The method of claim 118, wherein displaying the visual effects with the first visual characteristic includes displaying the visual effects with an animated change in color.
120. The method of claim 119, wherein displaying the animated change in color includes displaying the animated change in color with an animated change through a first range of colors.
121. The method of claim 120, further comprising: while displaying the first user interface, and in response to detecting the first event: in accordance with the determination that the first event corresponds to the first functionality that outputs content using the first AI model, animating the animated change through the first range of colors in a first manner; and in accordance with the determination that the first event corresponds to the second functionality that outputs content using the second AI model, animating the animated changed through the first range of colors in a second manner, different from the first manner.
122. The method of any one of claims 112-121, further comprising: while displaying the first user interface, and in response to detecting the first event: -609- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with the determination that the first event corresponds to the first functionality that outputs content using the first AI model, animating the animated change through a first portion of the first range of colors; and in accordance with the determination that the first event corresponds to the second functionality that outputs content using the second AI model, animating the animated change through a second portion of the first range of colors, different than the first portion of the first range of colors.
123. The method of any one of claims 118-122, wherein the displaying the visual effects with the first visual characteristic includes displaying the visual effects with an animation of a spatially varying color pattern.
124. The method of any one of claims 118-123, wherein displaying visual effects with the first visual characteristic includes displaying the visual effects with a glowing effect that extends from a boundary associated with the first user interface, and gradually fades according to a non-linear manner as a distance from the boundary associated with the first user interface increases.
125. The method of any one of claims 118-124, wherein the method further comprises: while displaying the first user interface, and in response to detecting the first event corresponding to the first functionality: displaying the first visual effect with a first spatial parameter at a first time; and displaying the first visual effect with a second spatial parameter, different from the first spatial parameter, at a second time, different from the first time.
126. The method of any one of claims 118-125, wherein the method further comprises: while displaying the first user interface, and in response to detecting the first event: displaying the first visual effect with a first visual intensity at a first time; and -610- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) displaying the first visual effect with a second visual intensity, different from the first visual intensity, at a second time, different from the first time.
127. The method of claim 126, wherein the first visual effect at least temporarily includes displaying the first visual content with High Dynamic Range (HDR)luminance that is greater than a standard range of luminance that is available to display content in the first user interface.
128. The method of any one of claims 118-127, wherein the first user interface includes a plurality of user interface elements, and wherein the method further comprises: in response to the first event, and in accordance with the determination that the first event corresponds to the first functionality, displaying the first visual effect comprises displaying a first user interface element of the first user interface with the first visual effect, wherein the first visual effect includes displaying the first user interface element with a common animated layer, and wherein the common animated layer varies spatially over time; and in response to detecting a second event, different from the first event, and in accordance with a determination that the second event corresponds to a second functionally that outputs content generated using a third AI model, displaying a second user interface element of the first user interface with a third visual effect, wherein the third visual effect includes displaying the second user interface element with the common animated layer, and wherein the common animated layer varies spatially over time.
129. The method of any one of claims 118-128, wherein displaying first visual information corresponding to the first event in the first user interface with the first visual effect comprises: while displaying the first user interface, and while displaying the first visual effect in response to detecting the first event: receiving a first user input at the first user interface; and in response to receiving the first user input, modifying one or more aspects of the first visual effect. the first effect is applied to content generated based on the first functionality, and the second effect is applied to content generated based on the second functionality. -611- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 130. The method of any one of claims 112-129, displaying first visual information includes applying the first visual effect to content generated based on the first functionality, and wherein displaying second visual information includes applying the second visual effect to content generated based on the second functionality.
131. The method of any one of claims 112-130, wherein the first functionality includes a status indicator associated with a virtual assistant.
132. The method of any one of claims 112-131, wherein the first functionality includes a response generated by a virtual assistant.
133. The method of any one of claims 112-132, wherein the first functionality includes displaying an input region on the first user interface for accepting one or more inputs to specify parameters for generation of content using an AI model.
134. The method of any one of claims 112-133, wherein the first functionality includes displaying a block of text that was impacted by the first AI model.
135. The method of claim 134, wherein the block of text includes a first portion and a second portion, and wherein displaying first visual information corresponding to the first event in the first user interface with the first visual effect includes displaying the first portion of the block of text with the first visual effect, and displaying the second portion of the block of text without the first visual effect.
136. The method of any one of claims 134-135, wherein the block of text is included in a notification, the block of text includes summarized text, and the summarized text is generated by the first AI model.
137. The method of any one of claims 112-136, wherein the first functionality includes displaying visual media content generated by the first AI model.
138. The method of any one of claims 112-137, wherein the first functionality includes displaying an element of a visual media content that is removeable. -612- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 139. The method of any one of claims 112-138, wherein the first functionality includes displaying one or more animated media items as part of a visual media content.
140. The method of any one of claims 112-139, wherein the first functionality includes displaying an indication of one or more search results, wherein the one or more search results are generated, using the first AI model, in response to a search query.
141. The method of any one of claims 112-140, wherein the first AI model is a generative AI model.
142. The method of any one of claims 112-141, wherein the first event is detected while the electronic device displays a first user interface corresponding to a first software application, the method further comprising: while displaying a second user interface, different from the first user interface, corresponding to a second software application that is different from the first software application, detecting a second event, different from the first event; and in response to detecting the second event: in accordance with a determination that the second event corresponds to the first functionality that outputs content generated using the first artificial intelligence (AI) model, displaying the first visual information corresponding to the second event in the first user interface with the first visual effect, wherein the first visual effect includes the first visual characteristic; and in accordance with a determination that the second event corresponds to the second functionality, different from the first functionality, that outputs content generated using the second AI model, displaying the second visual information corresponding to the first event in the first user interface with the second visual effect, different from the first visual effect, wherein the second visual effect includes the first visual characteristic.
143. An electronic device that is in communication with a display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and -613- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: while displaying a first user interface, detecting a first event; and in response to detecting the first event: in accordance with a determination that the first event corresponds to a first functionality that outputs content generated using a first artificial intelligence (AI) model, displaying first visual information corresponding to the first event in the first user interface with a first visual effect, wherein the first visual effect includes a first visual characteristic; and in accordance with a determination that the first event corresponds to a second functionality, different from the first functionality, that outputs content generated using a second AI model, displaying second visual information corresponding to the first event in the first user interface with a second visual effect, different from the first visual effect, wherein the second visual effect includes the first visual characteristic.
144. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: while displaying a first user interface, detecting a first event; and in response to detecting the first event: in accordance with a determination that the first event corresponds to a first functionality that outputs content generated using a first artificial intelligence (AI) model, displaying first visual information corresponding to the first event in the first user interface with a first visual effect, wherein the first visual effect includes a first visual characteristic; and in accordance with a determination that the first event corresponds to a second functionality, different from the first functionality, that outputs content generated using a second AI model, displaying second visual information corresponding to the first event in the first user interface with a second visual effect, different from the first visual effect, wherein the second visual effect includes the first visual characteristic.
145. An electronic device, comprising: -614- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) one or more processors; memory; means for while displaying a first user interface, detecting a first event; and means for in response to detecting the first event: in accordance with a determination that the first event corresponds to a first functionality that outputs content generated using a first artificial intelligence (AI) model, displaying first visual information corresponding to the first event in the first user interface with a first visual effect, wherein the first visual effect includes a first visual characteristic; and in accordance with a determination that the first event corresponds to a second functionality, different from the first functionality, that outputs content generated using a second AI model, displaying second visual information corresponding to the first event in the first user interface with a second visual effect, different from the first visual effect, wherein the second visual effect includes the first visual characteristic.
146. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for while displaying a first user interface, detecting a first event; and means for in response to detecting the first event: in accordance with a determination that the first event corresponds to a first functionality that outputs content generated using a first artificial intelligence (AI) model, displaying first visual information corresponding to the first event in the first user interface with a first visual effect, wherein the first visual effect includes a first visual characteristic; and in accordance with a determination that the first event corresponds to a second functionality, different from the first functionality, that outputs content generated using a second AI model, displaying second visual information corresponding to the first event in the first user interface with a second visual effect, different from the first visual effect, wherein the second visual effect includes the first visual characteristic.
147. An electronic device, comprising: one or more processors; memory; and -615- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 112-142.
148. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 112-142.
149. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 112-142.
150. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 112-142.
151. A method comprising: at an electronic device in communication with one or more input devices and a display generation component: detecting, via the one or more input devices, an event; in response to detecting the event: in accordance with a determination that the event is a first type of event, displaying, via the display generation component, an animation indicative of the event, wherein the displaying of the animation includes a first portion of the animation that includes displaying a portion of a user interface with a respective luminance that is higher than a standard dynamic range of luminance for the user interface, wherein the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface is associated with content that is presented after the event; and after displaying the first portion of the animation and when the animation concludes, displaying the portion of the user interface with luminance that is within the standard dynamic range for the user interface. -616- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 152. The method of claim 151, wherein displaying the animation includes: in accordance with a determination that a result of an operation performed in response to the event corresponds to a first region in the user interface, the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface corresponds to the first region, and in accordance with a determination that the result of the operation corresponds to a second region in the user interface, different from the first region in the user interface, the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface corresponds to the second region in the user interface.
153. The method of claim 152, wherein the first region of the user interface includes text while displaying the animation indicative of the event.
154. The method of any of claims 152-153, wherein the first region of the user interface includes at least a portion of an image while displaying the animation indicative of the event.
155. The method of any of claims 151-154, further comprising: in response to detecting the event, and in accordance with a determination that the event is a second type of event, different from the first type of event, displaying, via the display generation component, an animation indicative of the second type of event, wherein displaying the animation indicative of the second type of event includes displaying a portion of the user interface with respective second luminance, different from the respective luminance, that is within the standard dynamic range of luminance for the user interface.
156. The method of any of claims 151-155, wherein the first type of event includes an operation associated with an artificial intelligence model.
157. The method of any of claims 151-156, wherein the first type of event includes an operation associated with generating the content.
158. The method of any of claims 151-157, wherein the first type of event includes an operation associated with modifying a content item. -617- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 159. The method of any of claims 151-158, wherein the displayed animation indicative of the event corresponds to one or more content items related to an operation associated with the first type of event.
160. The method of claim 159, wherein the one or more content items includes a first content item that is used to generate a result of the operation associated with the first type of event, and the displayed animation is displayed over the first content item.
161. The method of claim 159, wherein the one or more content items includes a first content item that is a result of the operation associated with the first type of event, and the displayed animation is displayed over the first content item.
162. The method of any of claims 159-161, wherein displaying the first portion of the animation includes changing one or more visual characteristics of the one or more content items.
163. The method of any of claims 159-162, wherein displaying the first portion of the animation includes changing one or more colors of the one or more content items.
164. The method of any of claims 159-163, wherein displaying the first portion of the animation comprises moving the animation over different portions of the one or more content items over time.
165. The method of any of claims 159-164, wherein displaying the first portion of the animation comprises: displaying, via the display generation component, a respective first portion of the user interface with one or more first colors during the first portion of the animation; and displaying, via the display generation component, the respective first portion of the user interface with one or more second colors during a second portion of the animation, different from the first portion of the animation, wherein the one or more first colors and the one or more second colors are colors along a color gradient. -618- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 166. The method of claim 165, wherein the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface corresponds to the respective first portion of the user interface during the first portion of the animation, and the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface corresponds to a respective second portion of the user interface, different from the respective first portion of the user interface, during the second portion of the animation.
167. The method of any of claims 165-166, wherein: in accordance with a determination that a first portion of the content presented after the event is associated with first depth information, the color gradient displayed at the first portion of the content includes one or more third colors, and in accordance with a determination that the first portion of the content presented after the event is associated with second depth information, different from the first depth information, the color gradient displayed at the first portion of the content includes one or more fourth colors, different from the one or more third colors.
168. The method of any of claims 151-167, wherein the displaying the first portion of the animation includes displaying the first portion of the animation on a first spatial area of the user interface, and wherein the portion of the user interface displayed with the respective luminance is a second spatial area of the user interface, smaller than the first spatial area of the user interface.
169. The method of any of claims 151-168, wherein displaying the first portion of the animation includes changing an intensity of the animation over time.
170. The method of claim 169, wherein the intensity of the animation includes a luminance of the animation.
171. The method of claim 170, wherein the intensity of the animation corresponds to an amount of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface. -619- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 172. The method of any of claims 170-171, wherein the displaying the portion of the user interface with the respective luminance that is higher that a standard dynamic range of luminance for the user interface occurs a period of time after the animation indicative of the event begins, wherein during the period of time the animation does not include displaying the portion of the user interface with the respective luminance that is higher than the standard dynamic range of luminance for the user interface.
173. The method of any of claims 170-172, wherein displaying the portion of the user interface with respective luminance that is higher than the standard dynamic range of luminance for the user interface comprises: displaying the portion of the user interface with respective luminance that is higher than the standard dynamic range of luminance for the user interface during a first time period; ceasing display of the portion of the user interface with respective luminance that is higher than a standard dynamic range of luminance for the user interface during a second time period that is after the first time period but before a third time period; and displaying the portion of the user interface with respective luminance that is higher than the standard dynamic range of luminance for the user interface during the third time period.
174. The method of any of claims 170-173, wherein displaying the animation indicative of the event includes changing a saturation of the portion of the user interface over time.
175. The method of any of claims 170-173, wherein displaying the animation indicative of the event includes changing a degree of blurring associated with the animation over time.
176. The method of any of claims 170-175, wherein displaying the animation indicative of the event includes changing a degree of spatial distortion associated with the animation over time.
177. The method of any of claims 170-176, wherein displaying the animation indicative of the event includes changing a spatial distribution of the animation over the user interface over time. -620- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 178. The method of any of claims 151-177, wherein the user interface that is displayed when the event is detected is a first user interface that corresponds to a first software application, the method further comprising: while displaying a second user interface, different from the first user interface, corresponding to a second software application, different from the first software application: detecting, via the one or more input devices, a respective event; in response to detecting the respective event: in accordance with a determination that the event is the first type of event, displaying, via the display generation component, a respective animation indicative of the event, wherein the displaying of the respective animation includes a first portion of the respective animation that includes displaying a portion of the second user interface with the respective luminance that is higher than the standard dynamic range of luminance for the second user interface, wherein the portion of the second user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the second user interface is associated with content that is presented after the respective event; and after displaying the first portion of the respective animation and when the respective animation concludes, displaying the portion of the second user interface with luminance that is within the standard dynamic range for the second user interface.
179. An electronic device that is in communication with a display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: detecting, via the one or more input devices, an event; in response to detecting the event: in accordance with a determination that the event is a first type of event, displaying, via the display generation component, an animation indicative of the event, wherein the displaying of the animation includes a first portion of the animation that includes displaying a portion of a user interface with a respective luminance that is higher than a standard dynamic range of luminance for the user interface, wherein the portion of the user -621- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface is associated with content that is presented after the event; and after displaying the first portion of the animation and when the animation concludes, displaying the portion of the user interface with luminance that is within the standard dynamic range for the user interface.
180. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: detecting, via the one or more input devices, an event; in response to detecting the event: in accordance with a determination that the event is a first type of event, displaying, via the display generation component, an animation indicative of the event, wherein the displaying of the animation includes a first portion of the animation that includes displaying a portion of a user interface with a respective luminance that is higher than a standard dynamic range of luminance for the user interface, wherein the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface is associated with content that is presented after the event; and after displaying the first portion of the animation and when the animation concludes, displaying the portion of the user interface with luminance that is within the standard dynamic range for the user interface.
181. An electronic device, comprising: one or more processors; memory; means for detecting, via the one or more input devices, an event; means for in response to detecting the event: in accordance with a determination that the event is a first type of event, displaying, via the display generation component, an animation indicative of the event, wherein the displaying of the animation includes a first portion of the animation that includes displaying a portion of a user interface with a respective luminance that is higher than a -622- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) standard dynamic range of luminance for the user interface, wherein the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface is associated with content that is presented after the event; and after displaying the first portion of the animation and when the animation concludes, displaying the portion of the user interface with luminance that is within the standard dynamic range for the user interface.
182. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for detecting, via the one or more input devices, an event; means for in response to detecting the event: in accordance with a determination that the event is a first type of event, displaying, via the display generation component, an animation indicative of the event, wherein the displaying of the animation includes a first portion of the animation that includes displaying a portion of a user interface with a respective luminance that is higher than a standard dynamic range of luminance for the user interface, wherein the portion of the user interface displayed with the respective luminance that is higher than the standard dynamic range of luminance for the user interface is associated with content that is presented after the event; and after displaying the first portion of the animation and when the animation concludes, displaying the portion of the user interface with luminance that is within the standard dynamic range for the user interface.
183. An electronic device, comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 151-177.
184. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more -623- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) processors of an electronic device, cause the electronic device to perform any of the methods of claims 151-177.
185. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 151-177.
186. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 151-177.
187. A method comprising: at an electronic device in communication with one or more display generation components and one or more input devices: while displaying, via the one or more display generation components, a representation of first generative visual content based on a first prompt, detecting, via the one or more input devices, a navigation input including a movement component; and in response to detecting the navigation input including the movement component, displaying, via the one or more display generation components, a representation of second generative visual content based on the first prompt, wherein the second generative visual content is different from the first generative visual content.
188. The method of claim 187, wherein detecting the navigation input includes detecting, via the one or more input devices, a swipe gesture.
189. The method of any of claims 187-188, wherein detecting the navigation input includes detecting, via the one or more input devices, a dragging input. -624- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 190. The method of any of claims 187-189, wherein displaying the representation of the second generative visual content includes displaying, via the one or more display generation components, a representation of a generative visual content item.
191. The method of any of claims 187-190, wherein displaying the representation of the second generative visual content includes displaying, via the one or more display generation components, text-based content.
192. The method of any of claims 187-191, wherein displaying the representation of the first generative visual content includes displaying the first generative visual content at a first location in a user interface, and the method further comprises: in response to detecting the navigation input, displaying, via the one or more display generation components, the representation of the second generative visual content at the first location in the user interface.
193. The method of any of claims 187-192, wherein displaying the representation of the first generative visual content includes displaying the first generative visual content at a first location in the user interface, and the method further comprises: in response to detecting the navigation input, partially displaying the representation of the second generative visual content at the first location.
194. The method of any of claims 187-193, further comprising: while displaying the representation of the first generative visual content based on the first prompt, concurrently displaying, via the one or more display generation components, one or more indicators corresponding to a number of generative visual content available based on the first prompt.
195. The method of claim 194, wherein the one or more indicators include a plurality of indicators having a first visual appearance and corresponding to a plurality of generative visual content available based on the first prompt, and the method further comprises: while displaying the one or more indicators, concurrently displaying, via the one or more display generation components, an indication of a third generative visual content being generated in response to detecting the navigation input, wherein the indication of the third -625- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) generative visual content has a second visual appearance, different from the first visual appearance.
196. The method of claim 195, wherein displaying the indication of the third generative visual content includes: while the third generative visual content is being generated, displaying, via the one or more display generation components, the indication of the third generative visual content having the second visual appearance; and in accordance with a determination that the third generative visual content has been generated, displaying, via the one or more generation components, the indication of the third generative visual content having the first visual appearance, different from the second visual appearance.
197. The method of any of claims 187-196, further comprising: prior to displaying the representation of the first generative visual content, detecting, via the one or more input devices, one or more inputs providing the first prompt; in response to detecting the one or more inputs providing the first prompt and before detecting the navigation input, generating a first set of generative visual content including the first generative visual content and the second generative visual content without generating one or more additional generative visual content; and in response to detecting the navigation input, initiating a process to generate the one or more additional generative visual content based on the first prompt, wherein the one or more additional generative visual content are different from the first set of generative visual content.
198. The method of any of claims 187-197, further comprising: while displaying the representation of the second generative visual content, detecting, via the one or more input devices, an input directed to a user interface element; and in response to detecting the input, displaying, via the one or more display generation components, a second representation of the second generative visual content, wherein the second representation of the second generative visual content is of higher fidelity than the representation of the second generative visual content. -626- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 199. The method of any of claims 187-198, further comprising: while displaying the representation of the first generative visual content, concurrently displaying, via the one or more display generation components, one or more selectable options to edit the first prompt; while concurrently displaying the representation of the first generative visual content and the one or more selectable options, detecting, via the one or more input devices, an input directed to a first selection option of the one or more selectable options; and in response to detecting the input, initiating an event that will influence generation of the first generative visual content.
200. The method of any of claims 187-199, further comprising: while displaying a respective representation of a respective generative visual content in a generative visual content variants user interface, wherein the respective generative visual content is based on a respective prompt, detecting, via the one or more input devices, a prompt editing input; and in response to detecting the prompt editing input: in accordance with a determination that the prompt editing input satisfies one or more respective criteria: ceasing displaying, via the one or more display generation components, the generative visual content variants user interface; and displaying, via the one or more display generation components, an editing user interface for editing the respective prompt; and in accordance with a determination that the prompt editing input does not satisfy the one or more respective criteria: forgoing displaying the editing user interface for editing the respective generative visual content; and maintaining display of the generative visual content variants user interface. -627- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 201. The method of claim 200, wherein the one or more respective criteria include a criterion that is satisfied when the prompt editing input corresponds to a request to add a representation of a prompt component to the first prompt.
202. The method of any of claims 200-201, wherein the one or more respective criteria include a criterion that is satisfied when the prompt editing input is an input directed to a location outside a region of the respective representation of the respective generative visual content.
203. The method of any of claims 200-202, wherein displaying the respective representation of the respective generative visual content in the generative visual content variants user interface includes displaying the respective representation with a first amount of visual emphasis, and wherein displaying the editing user interface includes displaying the respective representation with a second amount of visual emphasis, less than the first amount of visual emphasis.
204. The method of any of claims 200-203, wherein: displaying the respective representation of the respective generative visual content in the generative visual content variants user interface includes displaying an edge of the respective representation with a first visual appearance; and displaying the editing user interface includes displaying the edge of the respective representation with a second visual appearance, different than the first visual appearance.
205. The method of any of claims 200-204, wherein displaying the editing user interface includes displaying, via the one or more display generation components, one or more visual indications of one or more prompt components associated with the first prompt.
206. The method of claim 205, wherein: displaying the one or more visual indications of the one or more prompt components in the generative visual content variants user interface includes displaying the one or more visual indications of the one or more prompt components with a first amount of visual emphasis; and -628- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) displaying the one or more visual indications of the one or more prompt components in the editing user interface includes displaying the one or more visual indications of the one or more prompt components with a second amount of visual emphasis, greater than the first amount of visual emphasis.
207. The method of any of claims 200-206, wherein: displaying the respective representation of the respective generative visual content based on a respective prompt in the generative visual content variants user interface includes displaying, via the one or more display generation components, one or more visual indications of one or more second respective generative visual content associated with the respective prompt; and displaying the editing user interface for editing the respective generative visual content includes forgoing displaying the one or more visual indications of the one or more second respective generative visual content associated with the respective prompt.
208. The method of any of claims 187-207, further comprising: prior to displaying the representation of the first generative visual content, displaying, via the one or more display generation components, an editing user interface for detecting one or more prompt components for use in generating generative visual content.
209. The method of claim 208, further comprising: while displaying the editing user interface, detecting, via the one or more input devices, a second prompt for use in generating generative visual content; and in response to detecting the second prompt component, concurrently displaying, via the one or more display generation components, a representation of third generative visual content that was generated based on the second prompt.
210. The method of claim 209, further comprising: while displaying the representation of the third generative visual content, detecting, via the one or more input devices, an event; and in response to detecting the event: in accordance with a determination that the event satisfies one or more respective criteria, displaying, via the one or more display generation components: a second representation of the third generative visual content based on the second prompt; and -629- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) a selectable option to navigate to a second representation of fourth generative visual content based on the second prompt; and in accordance with a determination that the event does not satisfy the one or more respective criteria, forgoing displaying the second representation of the third generative visual content and the selectable option.
211. The method of claim 210, wherein the one or more respective criteria includes a criterion that is satisfied when the event is a selection input directed to the representation of the third generative visual content.
212. The method of any of claims 210-211, wherein the one or more respective criteria includes a criterion that is satisfied when the electronic device does not detect one or more inputs editing the second prompt for a threshold period of time.
213. The method of any of claims 210-212, wherein the editing user interface includes one or more indications of the one or more prompt components displayed, via the one or more display generation components, with a first amount of visual prominence, and the method further comprises: while displaying the editing user interface, detecting, via the one or more input devices, an event; and in response to detecting the event: displaying, via the one or more display generation components, the representation of the first generative visual content; and reducing the visual prominence of the one or more indications of the one or more prompt components.
214. The method of claim 213, wherein reducing the visual prominence of the one or more indications of the one or more prompt components includes ceasing display of the one or more indications of the one or more prompt components.
215. The method of any of claims 213-214, wherein reducing the visual prominence of the one or more indications of the one or more prompt components includes displaying one or more second indications corresponding to the one or more prompt components with a second amount of visual prominence that is less than the first amount of visual prominence of the one or more indications of the one or more prompt components. -630- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 216. The method of any of claims 187-215, further comprising: while displaying a respective representation of a respective generative visual content based on the first prompt, detecting, via the one or more input devices, one or more inputs corresponding to creation of a second prompt, different from the first prompt; and in response to detecting one or more inputs corresponding to creation of the second prompt, displaying, via the one or more display generation components, a representation of third generative visual content based on the second prompt, wherein the third generative visual content is different from the respective generative visual content; while displaying the representation of the third generative visual content based on the second prompt, detecting, via the one or more input devices, a second navigation input including a second movement component; and in response to detecting the second navigation input including the second movement component, displaying, via the one or more display generation components, a representation of fourth generative visual content based on the second prompt, wherein the fourth generative visual content is different from the third generative visual content.
217. The method of any of claims 187-216, wherein displaying the representation of the first generative visual content includes displaying the first generative visual content in a first user interface of a first application, and the method further comprises: detecting, via the one or more input devices, a sequence of one or more inputs corresponding to a request to display generative visual content in a second user interface of a second application, different from the first application; in response to detecting the sequence of one or more inputs, displaying, via the one or more display generation components, the second user interface including a representation of third generative visual content based on a second prompt; while displaying the representation of third generative visual content, detecting, via the one or more input devices, a second navigation input including a movement component; and in response to detecting the second navigation input including the movement component, displaying, via the one or more display generation components, a representation of fourth generative visual content based on the second prompt, wherein the fourth generative visual content is different from the third generative visual content.
218. The method of any of claims 189-217, further comprising: -631- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) while displaying, via the one or more display generation components, the representation of the first generative visual content based on the first prompt displaying, via the one or more display generation components, a first selectable option that, when selected, causes the electronic device to initiate a process to preserve the first generative visual content; and receiving, via the one or more input devices, an input selecting the first selectable option; and in response to receiving the input selecting the first selectable option, initiating the process to preserve the first generative visual content.
219. The method of claim 218, wherein initiating the process to preserve the first generative visual content includes displaying, via the one or more display generation components, a second selectable option that, when selected, causes the electronic device to perform a first action with respect to the first generative visual content, and a third selectable option that, when selected, causes the electronic device to perform a second action with respect to the first generative visual content different from the first action.
220. The method of claim 219, further comprising: while displaying, via the one or more display generation components, the second selectable option and the third selectable option, receiving, via the one or more input devices a second input selecting the second selectable option or the third selectable option; and in response to receiving the second input: in accordance with a determination that the second input includes selection of the second selectable option, performing the first action with respect to the first generative visual content including preserving the first generative visual content in a first manner; and in accordance with a determination that the second input includes selection of the third selectable option, performing the second action with respect to the first generative visual content including preserving the first generative visual content in a second manner different from the first manner.
221. The method of any of claims 218-220, further comprising: while displaying, via the one or more display generation components, the representation of the second generative visual content based on the first prompt: -632- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) displaying, via the one or more display generation components, a second selectable option that, when selected, causes the electronic device to initiate a process to preserve the second generative visual content; and receiving, via the one or more input devices, an input selecting the second selectable option; and in response to receiving the input selecting the second selectable option, initiating the process to preserve the second generative visual content.
222. An electronic device that is in communication with one or more display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: while displaying, via the one or more display generation components, a representation of first generative visual content based on a first prompt, detecting, via the one or more input devices, a navigation input including a movement component; and in response to detecting the navigation input including the movement component, displaying, via the one or more display generation components, a representation of second generative visual content based on the first prompt, wherein the second generative visual content is different from the first generative visual content.
223. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: while displaying, via the one or more display generation components, a representation of first generative visual content based on a first prompt, detecting, via the one or more input devices, a navigation input including a movement component; and in response to detecting the navigation input including the movement component, displaying, via the one or more display generation components, a representation of second generative visual content based on the first prompt, wherein the second generative visual content is different from the first generative visual content. -633- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 224. An electronic device, comprising: one or more processors; memory; means for while displaying, via the one or more display generation components, a representation of first generative visual content based on a first prompt, detecting, via the one or more input devices, a navigation input including a movement component; and means for in response to detecting the navigation input including the movement component, displaying, via the one or more display generation components, a representation of second generative visual content based on the first prompt, wherein the second generative visual content is different from the first generative visual content.
225. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for while displaying, via the one or more display generation components, a representation of first generative visual content based on a first prompt, detecting, via the one or more input devices, a navigation input including a movement component; and means for in response to detecting the navigation input including the movement component, displaying, via the one or more display generation components, a representation of second generative visual content based on the first prompt, wherein the second generative visual content is different from the first generative visual content.
226. An electronic device, comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 187-221.
227. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 189-221. -634- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 228. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 189-221.
229. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 189-221.
230. A method comprising: at an electronic device in communication with one or more display generation components and one or more input devices: detecting, via the one or more input devices, an event for generating generative visual content based on a user selected prompt that will influence the generation of the generative visual content; and in response to detecting the event: in accordance with a determination that the user selected prompt satisfies one or more respective criteria, initiating a process to generate a representation of the generative visual content based on the event; and in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria, forgoing initiating the process to generate the representation of the generative visual content based on the event.
231. The method of claim 230, wherein satisfaction of the one or more respective criteria occurs automatically.
232. The method of claim 231, wherein automatically initiating the process to generate the representation of the generative visual content based on the event includes automatically initiating the process to generate the representation of the generative visual content based on the event in accordance with a determination that the electronic device does not detect one or more inputs editing the user selected prompt for a threshold period of time.
233. The method of any of claims 231-232, wherein the event is a prompt editing event. -635- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 234. The method of any of claims 230-233, wherein detecting the event for generating generative visual content based on the user selected prompt that will influence the generation of the generative visual content includes: detecting, via the one or more input devices, a user input corresponding to a request to generate generative visual content based on the user selected prompt.
235. The method of any of claims 230-234, wherein: the one or more respective criteria are not satisfied when the user selected prompt includes at least one term that has a high-risk meaning; and the one or more respective criteria include a requirement that none of the terms in the user selected prompt have a high-risk meaning in order for the one or more respective criteria to be satisfied.
236. The method of any of claims 230-235, wherein: the one or more respective criteria are not satisfied when the user selected prompt includes a combination of two or more terms that collectively have a high-risk meaning; and the one or more respective criteria include a requirement that none of the combinations of two or more terms in the user selected prompt have a high-risk meaning in order for the one or more respective criteria to be satisfied.
237. The method of any of claims 230-236, further comprising: in response to detecting the event: in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria, displaying, via the one or more display generation components, a user interface element that, when selected, causes the electronic device to perform an undo operation directed to the user selected prompt.
238. The method of claim 237, further comprising: while displaying the user interface element for performing the undo operation, detecting, via the one or more input devices, an input directed to the user interface element; in response to detecting the input directed to the user interface element for performing the undo operation, performing the undo operation directed to the user selected prompt; and after performing the undo operation directed to the user selected prompt: -636- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with a determination that performing the undo operation causes the user selected prompt to satisfy the one or more respective criteria, initiating the process to generate a second representation of the generative visual content based on the user selected prompt; and in accordance with a determination that performing the undo operation causes the user selected prompt to not satisfy the one or more respective criteria, forgoing initiating the process to generate the second representation of the generative visual content based on the user selected prompt.
239. The method of any of claims 237-238, wherein performing the undo operation directed to the user selected prompt includes removing a last-added prompt component from the user selected prompt.
240. The method of any of claims 237-239, wherein performing the undo operation directed to the user selected prompt includes returning the user selected prompt to a previous state of the user selected prompt that satisfied the one or more respective criteria.
241. The method of any of claims 237-240, further comprising: while displaying the user interface element, detecting, via the one or more input devices, an input corresponding to a request to edit the user selected prompt; in response to detecting the input directed to the user selected prompt, editing the user selected prompt in accordance with the input, and initiating a second event for generating generative visual content based on the user selected prompt that will influence generation of the generative visual content; after initiating the second event: in accordance with a determination that editing the user selected prompt in accordance with the input causes the user selected prompt to satisfy the one or more respective criteria, cease displaying, via the one or more display generation components, the user interface element for performing the undo operation; and in accordance with a determination that editing the user selected prompt in accordance with the input causes the user selected prompt to not satisfy the one or more respective criteria, maintaining display of the user interface element for performing the undo operation. -637- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 242. The method of any of claims 237-241, further comprising: while the user interface element for performing the undo operation is not displayed via the one or more display generation components, detecting, via the one or more input devices, an input corresponding to a request to perform the undo operation directed to the user selected prompt; and in response to detecting the input, performing the undo operation directed to the user selected prompt.
243. The method of any of claims 230-232, further comprising: while forgoing initiating the process to generate the representation of the generative visual content based on the event, detecting, via the one or more input devices, an input corresponding to a request to remove a prompt component from the user selected prompt; and in response to detecting the input: removing the prompt component from the user selected prompt; and in accordance with a determination that the user selected prompt with the prompt component removed satisfies the one or more respective criteria, initiating the process to generate the representation of the generative visual content.
244. The method of any of claims 230-241, further comprising: while the process to generate the representation of the generative visual content based on the event is not occurring, detecting, via the one or more input devices, an input corresponding to a request to add a prompt component to the user selected prompt; and in response to detecting the input: adding the prompt component to the user selected prompt; and in accordance with a determination that the user selected prompt with the prompt component added satisfies the one or more respective criteria, initiating the process to generate the representation of the generative visual content.
245. The method of any of claims 230-244, further comprising: in response to detecting the event: in accordance with a determination that a prompt component of the user selected prompt requires that the user selected prompt includes a subject in order for the one or more respective criteria to be satisfied, and that the user selected prompt does not include a -638- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) subject, displaying, via the one or more display generation components, a user interface element for selecting the subject; and in accordance with a determination that the user selected prompt does not include a prompt component that requires that the user selected prompt include a subject for the one or more respective criteria to be satisfied, or that the user selected prompt includes a subject, forgoing displaying, via the one or more display generation components, the user interface element for selecting the subject.
246. The method of claim 245, wherein displaying the user interface element for selecting the subject includes automatically selecting a particular character or person as the subject.
247. The method of claim 246, wherein automatically selecting the subject includes: in accordance with a determination that context information of the electronic device corresponds to a first person, selecting a representation of the first person as the subject; and in accordance with a determination that context information of the electronic device corresponds to a second person, selecting a representation of the second person as the subject based on a person who was recently used as a subject.
248. The method of any of claims 246-247, wherein automatically selecting the subject includes: in accordance with a determination that a first individual was selected as the subject within a time threshold of a current time, selecting the first individual as the subject; and in accordance with a determination that a second individual was selected as the subject within the time threshold of the current time, selecting the second individual as the subject.
249. The method of any of claims 246-247, wherein displaying the user interface element for selecting the subject includes: in accordance with a determination that the user selected prompt does not include a prompt component that requires that the user selected prompt include a subject for the one or more respective criteria to be satisfied, or that the user selected prompt includes a subject, displaying the user interface element for selecting the subject includes displaying, via the one or more display generation components, the user interface element for selecting the subject with a first amount of visual emphasis; and -639- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with a determination that a prompt component of the user selected prompt requires that the user selected prompt includes a subject in order for the one or more respective criteria to be satisfied, and that the user selected prompt does not include a subject, displaying the user interface element for selecting the subject with a second amount of visual emphasis, different from the first amount of visual emphasis.
250. The method of claim 249, further comprising: while displaying the user interface element for selecting the subject, detecting, via the one or more input devices, an input directed to the user interface element; and in response to detecting the input directed to the user interface element, displaying, via the one or more display generation components, a second user interface element for changing the subject.
251. The method of any of claims 245-250, wherein displaying the user interface element for selecting the subject includes a prompt for a user of the electronic device to select the subject.
252. The method of any of claims 245-251, further comprising: while displaying the user interface element for selecting the subject and while the prompt component is included in the user selected prompt, detecting, via the one or more input devices, an input corresponding to a request to select the subject; in response to detecting the input, adding the subject as a second prompt component to the user selected prompt; after adding the subject as the second prompt component to the user selected prompt, detecting, via the one or more input devices, a second input that corresponds to a request to remove the prompt component from the user selected prompt; and in response to detecting the second input, removing the prompt component while maintaining the subject as part of the user selected prompt.
253. The method of any of claims 230-251, further comprising: in response to detecting the event: in accordance with a determination that a prompt component of the user selected prompt requires that the user selected prompt includes a subject in order for the one or more respective criteria to be satisfied, and that the user selected prompt does not include a -640- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) subject, displaying, via the one or more display generation components, a user interface element for selecting the subject among a plurality of subjects identified in a media library associated with a user of the electronic device; and in accordance with a determination that the user selected prompt does not include a prompt component that requires that the user selected prompt include a subject for the one or more respective criteria to be satisfied, or that the user selected prompt includes a subject, forgoing displaying, via the one or more display generation components, the user interface element for selecting the subject.
254. The method of any of claims 230-252, further comprising: in response to detecting the event: in accordance with a determination that a prompt component of the user selected prompt requires that the user selected prompt includes a subject in order for the one or more respective criteria to be satisfied, and that the user selected prompt does not include a subject, displaying, via the one or more display generation components, a user interface element for selecting the subject among a plurality of subjects including one or more computer-generated avatars; and in accordance with a determination that the user selected prompt does not include a prompt component that requires that the user selected prompt include a subject for the one or more respective criteria to be satisfied, or that the user selected prompt includes a subject, forgoing displaying, via the one or more display generation components, the user interface element for selecting the subject.
255. The method of any of claims 230-253, wherein initiating the process to generate the representation of the generative visual content based on the event includes displaying, via the one or more display generation components, a preview of the generative visual content, and the method further comprises: while displaying the preview of the generative visual content based on the user selected prompt, detecting, via the one or more input devices, an input corresponding to a request to edit the user selected prompt; and in response to detecting the input, editing the user selected prompt in accordance with the input, and automatically changing the appearance of the preview of the generative visual content based on the edited user selected prompt. -641- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 256. The method of any of claims 230-254, including, in response to detecting the event: in accordance with a determination that the user selected prompt satisfies one or more respective criteria displaying, via the one or more display generation components, a first generation status indication; and in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria displaying, via the one or more display generation components, a second generation status indication, different from the first generation status indication.
257. The method of claim 256, wherein: displaying the first generation status indication includes displaying the first generation status indication with a first animation effect; and displaying the second generation status indication includes displaying the second generation status indication with a second animation effect, different from the first animation effect.
258. The method of any of claims 230-256, wherein detecting the event for generating generative visual content based on the user selected prompt that will influence the generation of the generative visual content includes detecting the event in a first user interface of a first application, and the method further comprises: detecting, via the one or more input devices, a second event for generating generative visual content based on a second user selected prompt that will influence the generation of the generative visual content in a second user interface of a second application, different from the first application; and in response to detecting the second event: in accordance with a determination that the second user selected prompt satisfies the one or more respective criteria, initiating a process to generate a representation of the generative visual content based on the second event; and in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria, forgoing initiating the process to generate the representation of the generative visual content based on the second event.
259. An electronic device that is in communication with one or more display generation component and one or more input devices, the electronic device comprising: -642- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: detecting, via the one or more input devices, an event for generating generative visual content based on a user selected prompt that will influence the generation of the generative visual content; and in response to detecting the event: in accordance with a determination that the user selected prompt satisfies one or more respective criteria, initiating a process to generate a representation of the generative visual content based on the event; and in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria, forgoing initiating the process to generate the representation of the generative visual content based on the event.
260. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: detecting, via the one or more input devices, an event for generating generative visual content based on a user selected prompt that will influence the generation of the generative visual content; and in response to detecting the event: in accordance with a determination that the user selected prompt satisfies one or more respective criteria, initiating a process to generate a representation of the generative visual content based on the event; and in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria, forgoing initiating the process to generate the representation of the generative visual content based on the event.
261. An electronic device, comprising: one or more processors; memory; -643- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) means for detecting, via the one or more input devices, an event for generating generative visual content based on a user selected prompt that will influence the generation of the generative visual content; and means for in response to detecting the event: in accordance with a determination that the user selected prompt satisfies one or more respective criteria, initiating a process to generate a representation of the generative visual content based on the event; and in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria, forgoing initiating the process to generate the representation of the generative visual content based on the event.
262. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for detecting, via the one or more input devices, an event for generating generative visual content based on a user selected prompt that will influence the generation of the generative visual content; and means for in response to detecting the event: in accordance with a determination that the user selected prompt satisfies one or more respective criteria, initiating a process to generate a representation of the generative visual content based on the event; and in accordance with a determination that the user selected prompt does not satisfy the one or more respective criteria, forgoing initiating the process to generate the representation of the generative visual content based on the event.
263. An electronic device, comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 230-258.
264. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more -644- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) processors of an electronic device, cause the electronic device to perform any of the methods of claims 230-258.
265. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 230-258.
266. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 230-258.
267. A method comprising: at an electronic device in communication with one or more display generation components and one or more input devices: detecting, via the one or more input devices, a first input corresponding to a request to display a user interface for detecting a prompt for use in generating automatically-generated visual media; and in response to detecting the input to display the user interface for detecting the prompt: in accordance with a determination that the electronic device is in a first context, displaying, via the one or more display generation components, one or more visual indications of a first set of one or more prompt component suggestions; and in accordance with a determination that the electronic device is in a second context, different from the first context, displaying, via the one or more display generation components, one or more visual indications of a second set of one or more prompt component suggestions, different from the first set of one or more prompt component suggestions.
268. The method of claim 267, wherein displaying the user interface for detecting the prompt for use in creating automatically-generated visual media includes displaying, via the one or more display generation components, a keyboard user interface element including a keyboard and a content entry field in the user interface for detecting the prompt for use in creating automatically-generated visual media. -645- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 269. The method of claim 268, wherein: displaying, via the display generation component, the one or more visual indications of the first set of the one or more prompt component suggestions includes displaying the first set of the one or more prompt components suggestions in the keyboard user interface element; and displaying, via the display generation component, the one or more visual indications of the second set of one or more prompt component suggestions includes displaying the second set of the one or more prompt components suggestions in the keyboard user interface element.
270. The method of any of claims 267-269, wherein: displaying, via the one or more display generation components, the one or more visual indications of the first set of the one or more prompt component suggestions includes displaying the first set of the one or more prompt components in a prompt component suggestion region near the content entry field; and displaying, via the display generation component, the one or more visual indications of the second set of one or more prompt component suggestions includes displaying the second set of the one or more prompt components in the prompt component suggestion region near the content entry field.
271. The method of any of claims 267-270, further comprising: while displaying one or more visual indications of a respective set of one or more prompt component suggestions and while the electronic device is in a respective context, detecting, via the one or more input devices, one or more second inputs corresponding to a request to add text to the content entry field; and in response to detecting the one or more second inputs: adding the text to the content entry field in accordance with the one or more second inputs; ceasing displaying one or more visual indications of a respective set of one or more prompt component suggestions; and displaying, via the one or more display generation components, one or more visual indications of a third set of one or more prompt component suggestions corresponding to the text in the content entry field.
272. The method of claim 271, further comprising: in response to detecting the one or more inputs: -646- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with a determination that the text corresponds to recognized concept, replacing the text with a representation of the recognized concept in the content entry field.
273. The method of any of claims 271-272, further comprising: in response to detecting the one or more inputs: in accordance with a determination that the text corresponds to a recognized concept, displaying, via the one or more display generation components, a graphical representation of the recognized concept in the content entry field.
274. The method of any of clams 264-266, further comprising: after adding the text to the content entry field, detecting, via the one or more input devices, a third input corresponding to a request to add the text as a prompt component to be used to influence generation of an automatically-generated visual content item; in response to detecting the third input: in accordance with a determination that the text corresponds to a recognized concept, displaying, via the display generation component, an indication of the recognized concept in an automatically-generated visual media user interface and using the recognized concept to influence the generation of the automatically-generated visual content item; and in accordance with a determination that the text does not correspond to a recognized concept, displaying, via the display generation component, an indication of the text as an indication of a prompt component in the automatically-generated visual media user interface and using the prompt component to influence the generation of the automatically- generated visual content item.
275. The method of any of claims 267-274, wherein displaying the user interface for detecting the prompt for use in creating automatically-generated visual media includes displaying a set of prompt editing options, the method further comprising: while the prompt is a respective prompt, detecting, via the one or more input devices, a sequence of one or more inputs including an input directed towards a first prompt editing option of the set of prompt editing options; and in response to detecting the sequence of one or more inputs, updating the respective prompt in accordance with the first prompt editing option. -647- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 276. The method of claim 275, wherein the set of prompt editing options includes a keyboard invocation option, and the method further comprises: detecting, via the one or more input devices, a second input selecting the keyboard invocation option; and in response to detecting the second input, displaying, via the one or more display generation components, a keyboard user interface element on the user interface for detecting the prompt.
277. The method of any of claims 275-276, wherein the set of prompt editing options includes a subject selection option, and the method further comprises: detecting, via the one or more input devices, a second input selecting the subject selection option; and in response to detecting the second input, displaying, via the one or more display generation components, a subject selection user interface for detecting a candidate subject for use in creating the automatically-generated visual media.
278. The method of claim 277, wherein displaying the subject selection user interface includes: in accordance with a determination that the electronic device is in the first context, displaying, via the one or more display generation components, one or more selectable options of a first set of candidate subjects in the subject selection user interface; and in accordance with a determination that the electronic device is in the second context, displaying, via the one or more display generation components, one or more selectable options of a second set of candidate subjects in the subject selection user interface, wherein the second set of candidate subjects is different from the first set of candidate subjects.
279. The method of any of claims 277-278, wherein displaying the subject selection user interface includes displaying one or more selectable options of a set of previously used candidate subjects.
280. The method of any of claims 277-279, wherein displaying the subject selection user interface includes displaying an option to capture a media item of a new subject, and the method further comprises: detecting, via the one or more input devices, a second input selecting the option; and -648- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to detecting the second input, displaying, via the one or more display generation component, on the subject selection user interface, a user interface of a camera application including a live preview of a media item of a physical environment of the electronic device; while displaying the user interface of the camera application, detecting, via the one or more input devices, one or more inputs corresponding to a request to capture a media item using the camera application for use as a candidate subject; and in response to detecting the one or more inputs corresponding to the request to capture the media item using the camera application, capturing the media item using the camera application and using the media item as the candidate subject in generating the automatically- generated visual media.
281. The method of any of claims 277-280, wherein displaying the subject selection user interface includes displaying an option to display a media library to select a candidate subject, and the method further comprises: detecting, via the one or more input devices, a second input selecting the option; and in response to detecting the second input, displaying, via the one or more display generation components, in the subject selection user interface, a representation of the media library; while displaying the media library, detecting, via the one or more input devices, one or more inputs corresponding to a request to select a respective media item from a plurality of media items in the media library as the candidate subject; and in response to detecting the one or more inputs corresponding to the request to select the respective media item, using the respective media item as the candidate subject in generating the automatically-generated visual media.
282. The method of any of claims 275-281, wherein the set of prompt editing options includes a style selection option; and the method further comprises: detecting, via the one or more input devices, a second input selecting the style selection option; and in response to detecting the second input, displaying, via the one or more display generation components, a style selection user interface including one or more selectable options corresponding to different styles. -649- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 283. The method of claim 282, wherein prior to detecting a third input corresponding to a request to select a respective style, automatically selecting a first style to be used to influence the generation of an automatically-generated visual content item.
284. The method of claim 283, wherein the first style was a last-used style to influence the generation of an automatically-generated visual content item, the method further comprising: while displaying the style selection user interface, detecting, via the one or more input devices, the third input corresponding to the request to select a respective style; and in response to detecting the third input corresponding to a request to select a respective style: ceasing using the first style to influence the generation of the automatically- generated visual content item; and selecting the respective style to be used to influence the generation of the automatically-generated visual content item.
285. The method of any of claims 283-284, wherein the one or more selectable options corresponding to different styles includes a selectable option corresponding to a three- dimensional style, and the method further comprises: while displaying the style selection user interface, detecting, via the one or more input devices, a third input corresponding to the request to select the selectable option corresponding to the three-dimensional style; and in response to detecting the third input corresponding to the request to select the selectable option corresponding to the three-dimensional style, selecting the three- dimensional style to be used to influence the generation of the automatically-generated visual content item, including generating the automatically-generated visual content item as a three- dimensional object.
286. The method of any of claims 275-285, wherein displaying the user interface for detecting the prompt for use in creating automatically-generated visual media includes displaying a selectable option for displaying the set of prompt editing options, the method further comprising: while not displaying the set of prompt editing options, detecting, via the one or more input devices, a second input directed towards the selectable option for displaying the set of prompt editing options; and -650- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to detecting the second input, displaying, via the one or more display generation components, the set of prompt editing options including one or more style selection options or one or more subject selection options.
287. The method of any of claims 267-286, wherein the electronic device is in the first context based on a first prior device activity and the electronic device is in the second context based on a second prior device activity.
288. The method of claim 287, wherein the first prior device activity is based on a first previously used application and the second prior device activity is based on a second previously used application.
289. The method of any of claims 287-288, wherein the first prior device activity is based on a first previous communication and the second prior device activity is based on a second previous communication.
290. The method of any of claims 287-289, wherein the first prior device activity is based on one or more topics of content from a previously used application and the second prior device activity is based on one or more topics of content from a previously used application.
291. The method of any of claims 267-290, further comprising: while displaying the user interface for detecting the prompt for use in generating the automatically-generated visual media: while the prompt is a first prompt, displaying, via the one or more display generation components, a representation of the automatically-generated visual content item based on the first prompt in a preview region of a first region of the user interface for detecting the prompt, wherein the representation of the automatically-generated visual content item is influenced by the prompt; while the prompt is the first prompt and while displaying the representation of the automatically-generated visual content item based on the first prompt, detecting, via the one or more input devices, one or more second inputs corresponding to a request to modify one or more prompt components of the first prompt; and in response to detecting the one or more second inputs, updating the first prompt in accordance with the one or more second inputs, and displaying, via the one or more display -651- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) generation components, an updated representation of the automatically-generated visual content item in accordance with the modifications of one or more prompt components.
292. The method of any of claims 267-291, further comprising: while displaying the user interface for detecting the prompt for use in generating the automatically-generated visual media and while displaying one or more visual indications of a respective set of one or more prompt component suggestions: displaying, via the one or more display generation components, a representation of the automatically-generated visual content item based on a first prompt in a preview region of the first region of the user interface for detecting the prompt, wherein the representation of the automatically-generated visual content item is influenced by the prompt, and a first visual indication corresponding to a first prompt component of the first prompt; and detecting, via the one or more input devices, a second input corresponding to a request to add a first prompt component suggestion as a second prompt component for use in generating the automatically-generated visual content item; in response to detecting the second input: updating, via the one or more display generation components, the display of the representation of the automatically-generated visual content item to a second representation of the automatically-generated visual content item; and displaying, via the one or more display generation components, the first visual indication corresponding to the first prompt component and a second visual indication corresponding to the second prompt component in a prompt region of the first region of the user interface.
293. The method of claim 292, wherein: while displaying the first visual indication corresponding to the first prompt component without displaying the second visual indication corresponding to the second prompt component, displaying, via the one or more display generation components, a first visual effect with a first value for a first visual characteristic between the first visual indication and the representation of the automatically-generated visual content item; in response to detecting the second input, displaying, via the one or more display generation components, the first visual effect with the first value between the first visual indication, and the second representation of the automatically-generated visual content item, -652- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) and a second visual effect with a second value for the first visual characteristic between the second visual indication and the second representation of the automatically-generated visual content item; while displaying the first visual effect with the first value between the first visual indication, and the second representation of the automatically-generated visual content item, and the second visual effect with the second value for the first visual characteristic between the second visual indication and the second representation of the automatically- generated visual content item, detecting, via the one or more input devices, a third input corresponding to a request to move the first visual indication from a first location to a second location, different than the first location in the first region of the user interface; in response to detecting the third input: displaying, via the one or more display generation components, the first visual indication at the second location in the first region of the user interface; and displaying, via the one or more display generation components, the second visual effect with a third value, different than the first value and the second value, for the first visual characteristic between the first visual indication and the second representation of the automatically-generated visual content item.
294. The method of any of claims 291-293, further comprising: displaying, via the one or more display generation components, one or more indications of one or more first prompt components used to influence the generation of the automatically-generated visual content item in the prompt region of the first region adjacent to the preview region of the first region of the user interface.
295. The method of any of claims 291-294, further comprising: displaying, via the one or more display generation components, one or more indications of one or more second prompt components that indicate variations for predefined characteristics of the automatically-generated visual content item, different than the one or more first prompt components, used to influence the generation of the automatically- generated visual content item in a second region of the user interface, different from a first region of the user interface.
296. The method of claim 295, wherein the one or more second prompt components include a subject prompt component. -653- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 297. The method of any of claims 295-296, wherein the one or more second prompt components include a style prompt component.
298. The method of any of claims 292-297, further comprising: while the first prompt including the first prompt component and not including the second prompt component corresponding to the first prompt component suggestion influences generation of the automatically-generated visual content item, displaying the first visual indication corresponding to the first prompt component at a first location in the prompt region of the first region of the user interface and displaying a third visual indication correspond to the first prompt component suggestion in a second location in a second region of the user interface ; and in response to detecting the second input: displaying the first visual indication corresponding to the first prompt component at a third location in the prompt region of the first region of the user interface, different than the first location; and displaying the second visual indication corresponding to the second prompt component at a fourth location in the prompt region of the first region of the user interface, different than the first location and the second location.
299. The method of any of claims 267-298, wherein displaying the user interface for detecting the prompt for use in generating automatically-generated visual media includes displaying the user interface for detecting the prompt in a first user interface of a first application, and the method further comprises: detecting, via the one or more input devices, a sequence of one or more inputs corresponding to a request to display the user interface for detecting the prompt in a second application, different from the first application; and in response to detecting the sequence of one or more inputs: in accordance with a determination that the electronic device is in the first context, displaying, via the one or more display generation components, the one or more visual indications of the first set of one or more prompt component suggestions while running the second application; and in accordance with a determination that the electronic device is in the second context, different from the first context, displaying, via the display generation component, the -654- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) one or more visual indications of the second set of one or more prompt component suggestions, different from the first set of one or more prompt component suggestions.
300. An electronic device that is in communication with one or more display generation component and one or more input devices, the electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: detecting, via the one or more input devices, a first input corresponding to a request to display a user interface for detecting a prompt for use in generating automatically- generated visual media; and in response to detecting the input to display the user interface for detecting the prompt: in accordance with a determination that the electronic device is in a first context, displaying, via the one or more display generation components, one or more visual indications of a first set of one or more prompt component suggestions; and in accordance with a determination that the electronic device is in a second context, different from the first context, displaying, via the one or more display generation components, one or more visual indications of a second set of one or more prompt component suggestions, different from the first set of one or more prompt component suggestions.
301. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: detecting, via the one or more input devices, a first input corresponding to a request to display a user interface for detecting a prompt for use in generating automatically-generated visual media; and in response to detecting the input to display the user interface for detecting the prompt: -655- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with a determination that the electronic device is in a first context, displaying, via one or more display generation components, one or more visual indications of a first set of one or more prompt component suggestions; and in accordance with a determination that the electronic device is in a second context, different from the first context, displaying, via the display generation component, one or more visual indications of a second set of one or more prompt component suggestions, different from the first set of one or more prompt component suggestions.
302. An electronic device, comprising: one or more processors; memory; means for, detecting, via the one or more input devices, a first input corresponding to a request to display a user interface for detecting a prompt for use in generating automatically-generated visual media; and means for, in response to detecting the input to display the user interface for detecting the prompt: in accordance with a determination that the electronic device is in a first context, displaying, via the one or more display generation components, one or more visual indications of a first set of one or more prompt component suggestions; and in accordance with a determination that the electronic device is in a second context, different from the first context, displaying, via the one or more display generation components, one or more visual indications of a second set of one or more prompt component suggestions, different from the first set of one or more prompt component suggestions.
303. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for, detecting, via the one or more input devices, a first input corresponding to a request to display a user interface for detecting a prompt for use in generating automatically-generated visual media; and means for, in response to detecting the input to display the user interface for detecting the prompt: in accordance with a determination that the electronic device is in a first context, displaying, via the one or more display generation components, one or more visual indications of a first set of one or more prompt component suggestions; and -656- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in accordance with a determination that the electronic device is in a second context, different from the first context, displaying, via the one or more display generation components, one or more visual indications of a second set of one or more prompt component suggestions, different from the first set of one or more prompt component suggestions.
304. An electronic device, comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 267-299.
305. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 267-299.
306. An electronic device, comprising: one or more processors; memory; and means for performing any of the methods of claims 267-299.
307. An information processing apparatus for use in an electronic device, the information processing apparatus comprising: means for performing any of the methods of claims 267-299.
308. A method comprising: at an electronic device in communication with one or more display generation components and one or more input devices: while composing a prompt for generation of automatically-generated visual content, detecting, via the one or more input devices, a first set of one or more inputs corresponding to a request to customize an appearance of a subject of the prompt, wherein the subject of the prompt is an anthropomorphic subject; and -657- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to detecting the first set of one or more inputs, displaying, via the one or more display generation components, an editing user interface for customizing the appearance of the subject of the prompt, including one or more selectable options for customizing an appearance of the subject in an automatically-generated visual content generated based on the prompt; while displaying the editing user interface, detecting, via the one or more input devices, a second set of one or more inputs corresponding to selection of a respective option of the one or more selectable options for customizing the appearance of the subject; and in response to detecting the second set of one or more inputs corresponding to selection of the respective option, initiating a process for customizing the appearance of the subject based on the respective option.
309. The method of claim 308, wherein: the editing user interface includes a first representation of a first candidate subject for the prompt, and a second representation of a second candidate subject for the prompt, the second set of one or more inputs includes selection of the first representation of the first candidate subject, and initiating the process for customizing the appearance of the subject includes customizing the appearance of the subject based on the first candidate subject.
310. The method of claim 309, wherein the editing user interface includes a preview representation of the automatically-generated visual content generated based on the prompt, and initiating the process for customizing the appearance of the subject based on the first candidate subject includes updating the preview representation of the automatically-generated visual content to be based on the first candidate subject.
311. The method of any of claims 309-310, wherein the editing user interface includes a first selectable option for initiating a process to capture a photo using a camera associated with the electronic device, the method further comprising: while displaying, via the one or more display generation components, the editing user interface, detecting, via the one or more input devices, a third set of one or more inputs including selection of the first selectable option; -658- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to detecting the third set of one or more inputs, displaying, via the one or more display generation components, a camera user interface that includes a representation of a physical environment of the electronic device that is captured using the camera; while displaying the camera user interface, detecting, via the one or more input devices, a fourth set of one or more inputs corresponding to a request to use a captured subject in the representation of the physical environment of the electronic device that is captured using the camera as the subject of the prompt; and in response to detecting the fourth set of one or more inputs, initiating a process to use the captured subject in the representation of the physical environment of the electronic device that is captured using the camera as the subject of the prompt.
312. The method of any of claims 309-311, wherein the editing user interface includes a first selectable option for initiating a process to generate a new candidate subject for the prompt, the method further comprising: while displaying, via the one or more display generation components, the editing user interface, detecting, via the one or more input devices, a third set of one or more inputs including selection of the first selectable option; in response to detecting the third set of one or more inputs, displaying, via the one or more display generation components, a subject generation user interface that includes one or more options for customizing a generated candidate subject; while displaying the subject generation user interface, detecting, via the one or more input devices, a fourth set of one or more inputs corresponding to a request to use the generated candidate subject as the subject of the prompt; and in response to detecting the fourth set of one or more inputs, initiating a process to use the generated candidate subject as the subject of the prompt.
313. The method of any of claims 309-312, wherein the one or more selectable options include one or more options to select a body style for the subject of the prompt.
314. The method of claim 313, wherein the body style that can be selected for the subject of the prompt is based on gender. -659- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 315. The method of any of claims 308-314, wherein the one or more options to select the body style for the subject of the prompt include a first option associated with a first body style, the method further comprising: displaying, via the one or more display generation components, a representation of the first body style including a plurality of representations of template subjects having the first body style.
316. The method of claim 315, wherein displaying the representation of the first body style includes changing an appearance of the plurality of representations of template subjects over time, including: at a first time, displaying, via the one or more display generation components, a first plurality of representations of template subjects having the first body style, and at a second time different from the first time, displaying, via the one or more display generation components, a second plurality of representations of template subjects having the first body style different from the first plurality of representations of template subjects having the first body style.
317. The method of any of claims 315-316, wherein the one or more selectable options include one or more options to select a skin tone for the subject of the prompt, and displaying the representation of the first body style including the plurality of representations of template subjects having the first body style includes: in accordance with a determination that the one or more options to select the skin tone for the subject of the prompt are not selected, displaying a first plurality of representations of template subjects having the first body style, where the first plurality of representations include template representations with different skin tones; and in accordance with a determination that a respective skin tone option of the one or more options to select the skin tone for the subject of the prompt is selected, displaying a second plurality of representations of template subjects having the first body style different from the first plurality of representations of template subjects having the first body style, where the second plurality of representations have the respective skin tone.
318. The method of any of claims 315-317, wherein displaying the representation of the first body style includes, prior to receiving a second set of one or more inputs, displaying, via -660- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) the one or more display generation components, a first plurality of representations of template subjects having the first body style, and the method further comprises: receiving, via the one or more input devices, the second set of one or more inputs; and in response to receiving the second set of one or more inputs, updating the representation of the first body style to include a second plurality of representations of template subjects having the first body style different from the first plurality of representations of template subjects having the first body style in accordance with the second set of one or more inputs.
319. The method of claim 318, wherein: displaying the first plurality of representations of template subjects having the first body style includes displaying, via the one or more display generation components, the first plurality of representations of template subjects having the first body style with a first amount of visual emphasis, and displaying the second plurality of representations of template subjects having the first body style includes displaying, via the one or more display generation components, the second plurality of representations of template subjects having the first body style with a second amount of visual emphasis greater than the first amount of visual emphasis.
320. The method of any of claims 308-319, wherein the one or more selectable options for customizing the appearance of the subject include a plurality of options for customizing different aspects of the appearance of the subject.
321. The method of any of claims 308-320, wherein the editing user interface includes a preview representation of the automatically-generated visual content generated based on the prompt, including the subject having the customized appearance.
322. The method of claim 321, further comprising: in response to detecting the second set of one or more inputs, updating the preview representation of the automatically-generated visual content to be based on the appearance of the subject as customized by the selection of the respective option.
323. The method of any of claims 321-322, further comprising: -661- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) after detecting the second set of one or more inputs, detecting, via the one or more input devices, a third set of one or more inputs corresponding to confirmation of customization of the appearance of the subject based on the second set of one or more inputs; and in response to detecting the third set of one or more inputs, updating the preview representation of the automatically-generated visual content to be based on the appearance of the subject as customized by the second set of one or more inputs.
324. The method of claim 323, wherein prior to detecting the third set of one or more inputs, the prompt includes one or more components, and the updated preview representation of the automatically-generated visual content is based on the appearance of the subject as customized by the second set of one or more inputs and the one or more components of the prompt.
325. The method of any of claims 308-324, wherein the anthropomorphic subject is a human subject.
326. The method of any of claims 308-325, wherein the anthropomorphic subject is a human subject identified from a media collection of a user of the electronic device.
327. The method of claim 326, wherein the one or more selectable options in the editing user interface include: a first selectable option that is selectable to select a first appearance for the human subject for use in generating the automatically-generated visual content based on the prompt, wherein the first appearance corresponds to a first set of one or more media items in the media collection of the user of the electronic device; and a second selectable option that is selectable to select a second appearance, different from the first appearance, for the human subject for use in generating the automatically- generated visual content based on the prompt, wherein the second appearance corresponds to a second set of one or more media items in the media collection of the user of the electronic device.
328. The method of claim 327, wherein: the first set of one or more media items is from a first time period, and -662- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) the second set of one or more media items is from a second time period, different from the first time period.
329. The method of claim 328, wherein: in accordance with a determination that the human subject has a current age that is a first age, the first time period has a first length and the second time period has the first length, and in accordance with a determination that the human subject has the current age that is a second age, different from the first age, the first time period has a second length and the second time period has the second length, wherein the second length is different from the first length.
330. The method of any of claims 327-329, wherein: the first set of one or more media items depict the human subject with a first visual appearance, and the second set of one or more media items depict the human subject with a second visual appearance, different from the first visual appearance.
331. The method of any of claims 326-330, wherein the one or more selectable options in the editing user interface include a first selectable option that is selectable to display selectable options for selecting any media item included in the media collection of the user of the electronic device as an appearance for the human subject, including a selectable option for a media item that does not include the human subject.
332. The method of any of claims 326-331, wherein: the one or more selectable options in the editing user interface include a first selectable option that is selectable to display selectable options for selecting any media item included in the media collection of the user of the electronic device that has been identified as including the human subject as an appearance for the human subject, and the one or more selectable options in the editing user interface do not include a selectable option for a media item that does not include the human subject.
333. The method of any of claims 326-332, wherein the one or more selectable options in the editing user interface include a first selectable option that is selectable to cause the -663- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) electronic device to automatically select a set of one or more media items as an appearance for the human subject, the method further comprising: while displaying, via the one or more display generation components, the editing user interface, detecting, via the one or more input devices, a third set of one or more inputs corresponding to selection of the first selectable option; and in response to detecting the third set of one or more inputs, automatically selecting the set of one or more media items as the appearance for the human subject.
334. The method of any of claims 326-333, further comprising: before the anthropomorphic subject is the human subject identified from a media collection of a user of the electronic device, detecting, via the one or more input devices, one or more respective inputs selecting the human subject identified from the media collection; and in response to detecting the one or more respective inputs, displaying, via the one or more display generation components, a representation of a candidate subject corresponding to the human subject, wherein the representation of the candidate subject is automatically- generated based on one or more media items from the media collection featuring the human subject.
335. The method of claim 334, wherein: the prompt includes one or more prompt components other than the subject of the prompt, and generating the representation of the candidate subject is based on the one or more media items without being based on the one or more prompt components other than the subject of the prompt.
336. The method of any of claims 334-335, wherein displaying the editing user interface includes: while displaying the representation of the candidate subject, displaying, via the one or more display generation components, a representation of a first media item of the one or more media items from the media collection featuring the human subject. -664- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 337. The method of any of claims 334-336, wherein the prompt includes a visual style for the automatically-generated visual content, and displaying the representation of the candidate subject includes: in accordance with a determination that the prompt includes a first visual style, displaying, via the one or more display generation components, the representation of the candidate subject in the first visual style; and in accordance with a determination that the prompt includes a second visual style different from the first visual style, displaying, via the one or more display generation components, the representations of the candidate subject in the second visual style.
338. The method of any of claims 334-337, further comprising: while displaying the representation of the candidate subject, detecting, via the one or more input devices, a respective input corresponding to a request to display a representation of a second candidate subject corresponding to the human subject; and in response to detecting the respective input corresponding to the request to display the representation of the second candidate subject, displaying, via the one or more display generation components, the representation of the second candidate subject different from the representation of the candidate subject, wherein the representation of the second candidate subject is automatically-generated based on one or more second media items from the media collection featuring the human subject, the one or more second media items different from the one or more media items.
339. The method of claim 338, further comprising: while a respective candidate subject corresponding to the human subject is selected for use in generating the automatically-generated visual media, the respective candidate subject being one of the candidate subject or the second candidate subject: detecting, via the one or more input devices, a third set of one or more inputs corresponding to a request to generate the automatically-generated visual media using the respective candidate subject; in response to detecting the third set of one or more inputs, generating first automatically-generated visual media with the respective candidate subject; after generating the first automatically-generated visual media: -665- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) receiving, via the one or more input devices, a fourth set of one or more inputs corresponding to a request to generate second automatically-generated visual media based on a second prompt that includes the human subject as the subject of the second prompt; in response to receiving the fourth set of one or more inputs, generating the second automatically-generated visual media with the respective candidate subject, wherein: in accordance with a determination that the third set of one or more inputs selected a first image or set of images as a basis for generating the representation of the human subject when generating the first automatically-generated visual media , the first image or set of images is used as a basis for generating the representation of the human subject when generating the second automatically-generated visual media; and in accordance with a determination that the third set of one or more inputs selected a second image or set of images as a basis for generating the representation of the human subject when generating the first automatically-generated visual media, the second image or set of images is used as a basis for generating the representation of the human subject when generating the second automatically-generated visual media.
340. The method of any of claims 308-339, wherein the anthropomorphic subject is an avatar that is generated by the electronic device, wherein the avatar has one or more user- specified characteristics.
341. The method of claim 340, wherein the one or more selectable options include a respective set of one or more selectable options for customizing a skin tone of the avatar.
342. The method of claim 341, wherein: in accordance with a determination that the automatically-generated visual content is an automatically-generated emoji, the respective set of one or more selectable options corresponds to a first set of available skin tones for the avatar, and in accordance with a determination that the automatically-generated visual content is not an emoji, the respective set of one or more selectable options corresponds to a second set of available skin tones for the avatar, different from the first set of available skin tones for the avatar. -666- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 343. The method of any of claims 340-342, wherein the one or more selectable options include a content entry field for entering content for customizing the avatar, the method further comprising: while displaying, via the one or more display generation components, the editing user interface including the content entry field, detecting, via the one or more input devices, entry of respective content into the content entry field; and in response to detecting the entry of the respective content into the content entry field: in accordance with a determination that the respective content is first content, modifying the avatar in a first manner corresponding to the first content; and in accordance with a determination that the respective content is second content, different from the first content, modifying the avatar in a second manner corresponding to the second content, wherein the second manner is different from the first manner.
344. The method of any of claims 340-343, further comprising: while displaying, via the one or more display generation components, the editing user interface, detecting, via the one or more input devices, a third set of one or more inputs corresponding to a request to add the avatar that has the one or more user-specified characteristics to a respective set of candidate subjects available for use in prompts for generating an automatically-generated visual content; and in response to detecting the third set of one or more inputs, adding the avatar that has the one or more user-specified characteristics to the respective set of candidate subjects available for use in prompts for generating an automatically-generated visual content.
345. The method of claim 344, wherein: in accordance with a determination that the automatically-generated visual content is a first type of automatically-generated visual content, the respective set of candidate subjects is a first set of candidate subjects available for use in prompts for generating the automatically-generated visual content, and in accordance with a determination that the automatically-generated visual content is a second type of automatically-generated visual content, different from the first type of automatically-generated visual content, the respective set of candidate subjects is a second set of candidate subjects available for use in prompts for generating the automatically-generated -667- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) visual content, different from the first set of candidate subjects available for use in prompts for generating the automatically-generated visual content.
346. The method of any of claims 308-345, wherein the second set of one or more inputs includes input directed to a plurality of the one or more selectable options in the editing user interface, the method further comprising: while displaying, via the one or more display generation components, the editing user interface and after detecting the second set of one or more inputs, detecting, via the one or more input devices, a third set of one or more inputs corresponding to a request to complete customization of the appearance of the subject of the prompt; and in response to detecting the third set of one or more inputs, adding a plurality of different components to the prompt for generation of the automatically-generated visual content based on the second set of one or more inputs.
347. The method of any of claims 308-346, further comprising: while displaying, via the one or more display generation components, the editing user interface, detecting, via the one or more input devices, a third set of one or more inputs corresponding to a request to complete customization of the appearance of the subject of the prompt; and in response to detecting the third set of one or more inputs, displaying, via the one or more display generation components, a automatically-generated visual content generation user interface, including displaying, in the automatically-generated visual content generation user interface, a visual indication of the subject having the customized appearance as a single recognized concept for the prompt.
348. The method of any of claims 308-347, further comprising: while displaying, via the one or more display generation components, the editing user interface, detecting, via the one or more input devices, a third set of one or more inputs corresponding to a request to complete customization of the appearance of the subject of the prompt and a request to generate the automatically-generated visual content; and in response to detecting the third set of one or more inputs, generating the automatically-generated visual content based on the customized appearance of the subject, including adding the automatically-generated visual content to a collection of automatically- generated visual content. -668- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 349. The method of any of claims 308-348, further comprising: while displaying, via the one or more display generation components, the editing user interface, detecting, via the one or more input devices, a third set of one or more inputs corresponding to a request to complete customization of the appearance of the subject of the prompt and a request to generate the automatically-generated visual content; and in response to detecting the third set of one or more inputs, generating the automatically-generated visual content based on the customized appearance of the subject, including adding the automatically-generated visual content to a collection of automatically- generated visual content that are available to be used in a plurality of different applications.
350. The method of claim 349, wherein the editing user interface is a user interface of a first application, the method further comprising: after generating the automatically-generated visual content based on the customized appearance of the subject, displaying, via the one or more display generation components, a user interface of a second application, different from the first application, wherein the user interface of the second application includes a keyboard user interface; while displaying the keyboard user interface, detecting, via the one or more input devices, a fourth set of one or more inputs directed to the keyboard user interface; and in response to detecting the fourth set of one or more inputs, displaying, in the user interface of the second application, a plurality of representations of a plurality of automatically-generated visual content that are available for use in the second application, including a representation of the automatically-generated visual content that was generated based on the customized appearance of the subject.
351. An electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: while composing a prompt for generation of automatically-generated visual content, detecting, via the one or more input devices, a first set of one or more inputs corresponding to -669- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) a request to customize an appearance of a subject of the prompt, wherein the subject of the prompt is an anthropomorphic subject; and in response to detecting the first set of one or more inputs, displaying, via the one or more display generation components, an editing user interface for customizing the appearance of the subject of the prompt, including one or more selectable options for customizing an appearance of the subject in a automatically-generated visual content generated based on the prompt; while displaying the editing user interface, detecting, via the one or more input devices, a second set of one or more inputs corresponding to selection of a respective option of the one or more selectable options for customizing the appearance of the subject; and in response to detecting the second set of one or more inputs corresponding to selection of the respective option, initiating a process for customizing the appearance of the subject based on the respective option.
352. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: while composing a prompt for generation of automatically-generated visual content, detecting, via the one or more input devices, a first set of one or more inputs corresponding to a request to customize an appearance of a subject of the prompt, wherein the subject of the prompt is an anthropomorphic subject; and in response to detecting the first set of one or more inputs, displaying, via the one or more display generation components, an editing user interface for customizing the appearance of the subject of the prompt, including one or more selectable options for customizing an appearance of the subject in a automatically-generated visual content generated based on the prompt; while displaying the editing user interface, detecting, via the one or more input devices, a second set of one or more inputs corresponding to selection of a respective option of the one or more selectable options for customizing the appearance of the subject; and in response to detecting the second set of one or more inputs corresponding to selection of the respective option, initiating a process for customizing the appearance of the subject based on the respective option. -670- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 353. An electronic device comprising: one or more processors; memory; means for while composing a prompt for generation of automatically-generated visual content, detecting, via the one or more input devices, a first set of one or more inputs corresponding to a request to customize an appearance of a subject of the prompt, wherein the subject of the prompt is an anthropomorphic subject; and means for in response to detecting the first set of one or more inputs, displaying, via the one or more display generation components, an editing user interface for customizing the appearance of the subject of the prompt, including one or more selectable options for customizing an appearance of the subject in a automatically-generated visual content generated based on the prompt; while displaying the editing user interface, detecting, via the one or more input devices, a second set of one or more inputs corresponding to selection of a respective option of the one or more selectable options for customizing the appearance of the subject; and means for in response to detecting the second set of one or more inputs corresponding to selection of the respective option, initiating a process for customizing the appearance of the subject based on the respective option.
354. An electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 308-350.
355. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 308-350.
356. An electronic device comprising: one or more processors; memory; and -671- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) means for performing any of the methods of claims 308-350.
357. A method comprising: at an electronic device in communication with one or more display generation components and one or more input devices: detecting, via the one or more input devices, one or more first inputs corresponding to a request to generate a visual media content based on a first prompt; and in response to detecting the one or more first inputs: in accordance with a determination that the one or more first inputs satisfy one or more first criteria including a criterion that is satisfied when the one or more first inputs correspond to a request to generate a visual media content that includes an anthropomorphic subject, displaying, via the one or more display generation components, an automatically- generated visual content using a randomly or pseudorandomly selected representation of a template subject from a set of template subjects that includes template subjects that have different appearances for at least one appearance characteristic; and in accordance with a determination that the one or more first inputs do not satisfy the one or more first criteria, generating the automatically-generated visual content without using a representation of a template subject from the set of template subjects.
358. The method of claim 357, further comprising: in response to detecting the one or more first inputs: in accordance with a determination that the one or more first inputs satisfy one or more second criteria including a criterion that is satisfied when the one or more first inputs correspond to a request to generate a visual media item based on a subject in a media item, displaying, via the one or more display generation components, the automatically-generated visual content including a candidate subject based on an appearance of the subject in the media item.
359. The method of any of claims 357-358, further comprising: after generating the automatically-generated visual content using the randomly or pseudorandomly selected representation of the template subject, detecting, via the one or more input devices, one or more second inputs corresponding to a request to regenerate the automatically-generated visual media content that includes the anthropomorphic subject; and -672- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to detecting the one or more second inputs, regenerating the automatically-generated visual media content using a second randomly or pseudorandomly selected representation of a second template subject from the set of template subjects.
360. The method of claim 359, wherein the one or more second inputs include an input corresponding to a request to change the first prompt to a second prompt to be used to generate the visual media content.
361. The method of claim 359-360, wherein the one or more second inputs include an input corresponding to a request to generate a second visual media content based on the first prompt, different from the visual media content.
362. The method of any of claims 357-361, wherein the set of template subjects share a first appearance characteristic selected in response to detecting, via the one or more input devices, one or more user inputs.
363. The method of claim 362, wherein the one or more user inputs includes an input corresponding to a request to select a body type characteristic as the first appearance characteristic; and in response to detecting the input, selecting the body type characteristic as the first appearance characteristic.
364. The method of claim 363, further comprising: in response to receiving the input, constraining the set of template subjects with the body type characteristic prior to generating the automatically-generated visual content.
365. The method of claim 364, wherein displaying the set of template subjects includes displaying an animation of the set of template subjects having a set of appearance characteristics that change over time, including: displaying, at a first time, a first set of template subjects with a first set of appearance characteristics including the body type characteristic; and displaying, at a second time, different than the first time, a second set of template subjects with a second set of appearance characteristics including the body type characteristic. -673- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 366. The method of claim 365, wherein displaying the animation includes: in accordance with a determination that the one or more user inputs include an input corresponding to a request to constrain the set of template subjects based on a first set of appearance characteristic selected by a user, displaying, via the one or more display generation components, the animation of a first set of template subjects sharing the first set of appearance characteristics; and in accordance with a determination that the one or more user inputs include an input corresponding to a request to constrain the set of template subjects based on a second set of appearance characteristic selected by a user, displaying, via the one or more display generation components, the animation of a second set of template subjects sharing the second set of appearance characteristics.
367. The method of any of claims 364-366, further comprising, in response to detecting the one or more first inputs, displaying a representation of the first appearance characteristic, including: in accordance with a determination that the request to generate the visual media content is a first type of request, displaying, via the one or more display generation components, a first representation of the first appearance characteristic with a first style associated with the first type of request; and in accordance with a determination that the request to generate the visual media content is a second type of request, different from the first type of request, displaying, via the one or more display generation components, a second representation of the first appearance characteristic with a second style associated with the second type of request, wherein the second style is different from the first style.
368. The method of any of claims 363-367, wherein the one or more user inputs includes an input corresponding to a request to select a skin tone characteristic as the first appearance characteristic; and in response to detecting the input, selecting the skin tone characteristic as the first appearance characteristic.
369. The method of claim 368, further comprising: -674- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) in response to receiving the input, constraining the set of template subjects with skin tone characteristic prior to generating the automatically-generated visual content.
370. The method of any of claims 362-369, wherein constraining the set of template subjects based on the first appearance characteristic includes forgoing constraining the set of template subjects in one or more second appearance characteristics.
371. The method of any of claims 357-370, wherein the at least one appearance characteristic includes a hair type characteristic.
372. The method of any of claims 357-371, wherein the at least one appearance characteristic includes an age characteristic.
373. An electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for: detecting, via one or more input devices, one or more first inputs corresponding to a request to generate a visual media content based on a first prompt; and in response to detecting the one or more first inputs: in accordance with a determination that the one or more first inputs satisfy one or more first criteria including a criterion that is satisfied when the one or more first inputs correspond to a request to generate a visual media content that includes an anthropomorphic subject, displaying, via one or more display generation components, an automatically- generated visual content using a randomly or pseudorandomly selected representation of a template subject from a set of template subjects that includes template subjects that have different appearances for at least one appearance characteristic; and in accordance with a determination that the one or more first inputs do not satisfy the one or more first criteria, generating the automatically-generated visual content without using a representation of a template subject from the set of template subjects. -675- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 374. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform a method comprising: detecting, via one or more input devices, one or more first inputs corresponding to a request to generate a visual media content based on a first prompt; and in response to detecting the one or more first inputs: in accordance with a determination that the one or more first inputs satisfy one or more first criteria including a criterion that is satisfied when the one or more first inputs correspond to a request to generate a visual media content that includes an anthropomorphic subject, displaying, via one or more display generation components, an automatically- generated visual content using a randomly or pseudorandomly selected representation of a template subject from a set of template subjects that includes template subjects that have different appearances for at least one appearance characteristic; and in accordance with a determination that the one or more first inputs do not satisfy the one or more first criteria, generating the automatically-generated visual content without using a representation of a template subject from the set of template subjects.
375. An electronic device comprising: one or more processors; memory; means for detecting, via one or more input devices, one or more first inputs corresponding to a request to generate a visual media content based on a first prompt; and means for, in response to detecting the one or more first inputs: in accordance with a determination that the one or more first inputs satisfy one or more first criteria including a criterion that is satisfied when the one or more first inputs correspond to a request to generate a visual media content that includes an anthropomorphic subject, displaying, via one or more display generation components, an automatically- generated visual content using a randomly or pseudorandomly selected representation of a template subject from a set of template subjects that includes template subjects that have different appearances for at least one appearance characteristic; and in accordance with a determination that the one or more first inputs do not satisfy the one or more first criteria, generating the automatically-generated visual content without using a representation of a template subject from the set of template subjects. -676- 4923-0100-0750, v.1Attorney Docket No.106842222440 (P65924WO1) 376. An electronic device comprising: one or more processors; memory; and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for performing any of the methods of claims 357-372.
377. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the electronic device to perform any of the methods of claims 357-372.
378. An electronic device comprising: one or more processors; memory; and means for performing any of the methods of claims 357-372. -677- 4923-0100-0750, v.1
Citation Information
Patent Citations
Method and apparatus for integrating manual input
US20020015024A1
Acceleration-based theft detection system for portable electronic devices
US20050190059A1
Methods and apparatuses for operating a portable device based on an accelerometer
US20060017692A1
Method and system to pre-fetch data in a network
US20130346490A1
Si / C COMPOSITE MATERIAL, METHOD FOR MANUFACTURING THE SAME, AND ELECTRODE
US20140234722A1