Aggregated content item user interface

By enabling the playback of visual and audio content within an aggregated media item interface, along with user input detection, the inefficiencies of existing media management systems are addressed, resulting in a more efficient and user-friendly media navigation experience.

JP2025087741AActive Publication Date: 2025-06-10APPLE INC
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
JP2025028123
Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
Priority Date
2021-12-06
Filing Date
2025-02-25
Publication Date
2025-06-10
Estimated Expiration
2042-05-23

AI Technical Summary

Technical Problem

Existing media management interfaces are cumbersome and inefficient, requiring users to navigate through multiple directories and interfaces to find relevant media content, leading to wasted time and resources.

Method used

The implementation of a method that plays visual content of an aggregated content item, which is an ordered sequence of media items selected based on specific criteria, while allowing for the separate playback of audio content and user input detection, enabling seamless modification of the audio content without interrupting the visual content.

Benefits of technology

This approach facilitates faster and more efficient navigation, browsing, and editing of media items, reducing cognitive burden and conserving power in battery-operated devices, while providing a contextually relevant and user-friendly interface.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2025087741000001_ABST
    Figure 2025087741000001_ABST
Patent Text Reader

Abstract

To provide user interfaces for navigating, viewing, and editing content items, including aggregated content items.SOLUTION: A method comprises, at a computer system that is in communication with a display generation component and one or more input devices: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items that are selected from a set of content items based on a first set of selection criteria; while playing the visual content of the first aggregated content item, playing audio content that is separate from the content items; while playing the visual content of the first aggregated content item and the audio content, detecting, via the one or more input devices, a user input; and in response to detecting the user input, modifying audio content that is playing while continuing to play visual content of the first aggregated content item.SELECTED DRAWING: Figure 7
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] (Cross - Reference to Related Applications) This application claims priority to U.S. Patent Application No. 17 / 542,947, entitled "AGGREGATED CONTENT ITEM USER INTERFACES," filed on December 6, 2021, and U.S. Provisional Patent Application No. 63 / 195,645, entitled "AGGREGATED CONTENT ITEM USER INTERFACES," filed on June 1, 2021, the entire contents of each of which are incorporated herein by reference.

[0002] This disclosure generally relates to computer user interfaces, and more specifically to techniques for navigating, browsing, and editing a collection of media items that include aggregated content items.

Background Art

[0003] As the storage capacity and processing power of devices continue to increase, in conjunction with the rising of media sharing that requires no effort between interconnected devices, the size of a user's library of media items (e.g., photos and videos) continues to increase.

Summary of the Invention

[0004] However, as the library of media items continues to grow and an archive of the user's life and experiences is created, the library can become cumbersome to navigate. For example, many libraries arrange media items in a substantially inflexible manner by default. A user's browsing of media may wish to view media relevant to the current context over different periods. However, some interfaces require the user to navigate to an excessive number of different media directories or interfaces to search for the content they want. This is inefficient and a waste of the user's time and resources. Therefore, it is desirable to facilitate the presentation of media items in a contextually relevant manner, thereby providing an improved interface for engaging with media content.

[0005] Furthermore, some techniques for navigating, browsing, and / or editing a collection of media items using an electronic device are generally cumbersome and inefficient. For example, some existing techniques use complex and time-consuming user interfaces that may involve multiple key presses or keystrokes. Existing techniques take more time than necessary and waste the user's time and the device's energy. This latter consideration is particularly important in battery-operated devices.

[0006] Accordingly, the present technology provides an electronic device with faster and more efficient methods and interfaces for navigating, browsing, and editing a collection of media items that includes aggregated content items (e.g., aggregated media items). Such methods and interfaces optionally complement or replace other methods for navigating, browsing, and editing a collection of media items. Such methods and interfaces reduce the user's cognitive burden and create a more efficient human-machine interface. In the case of battery-operated computing devices, such methods and interfaces conserve power and extend the battery charging intervals.

[0007] According to some embodiments, a method is described. The method includes, in a computer system communicating with a display generation component and one or more input devices, playing, via the display generation component, visual content of a first aggregated content item that includes an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria; playing, while playing the visual content of the first aggregated content item, audio content separate from the content items; detecting, while playing the visual content and the audio content of the first aggregated content item, user input via the one or more input devices; and modifying the playing audio content while continuing to play the visual content of the first aggregated content item in response to detecting the user input.

[0008] According to some embodiments, a non-transitory computer-readable storage medium is described. The non-transitory computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system communicating with a display generation component and one or more input devices, the one or more programs including instructions to play, via the display generation component, visual content of a first aggregated content item that includes an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria; play audio content separate from the content items while playing the visual content of the first aggregated content item; detect user input via the one or more input devices while playing the visual content and the audio content of the first aggregated content item; and modify the playing audio content while continuing to play the visual content of the first aggregated content item in response to detecting the user input.

[0009] According to some embodiments, a non-transitory computer-readable storage medium is described. The non-transitory computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices. The one or more programs, via the display generation component, play visual content of a first aggregated content item, the first aggregated content item including an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria. While playing the visual content of the first aggregated content item, play audio content separate from the content items. While playing the visual content and the audio content of the first aggregated content item, detect user input via the one or more input devices, and in response to detecting the user input, modify the audio content being played while continuing to play the visual content of the first aggregated content item.

[0010] According to some embodiments, a computer system is described. The computer system is configured to communicate with a display generation component and one or more input devices, and includes one or more processors and a memory storing one or more programs configured to be executed by the one or more processors. The one or more programs cause, via the display generation component, visual content of a first aggregated content item, which includes an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria, to be played. While playing the visual content of the first aggregated content item, audio content separate from the content items is played. While playing the visual content and the audio content of the first aggregated content item, user input is detected via the one or more input devices, and in response to detecting the user input, the audio content being played is modified while continuing to play the visual content of the first aggregated content item.

[0011] According to some embodiments, a method is described. The method is in a computer system communicating with a display generation component and one or more input devices, and via the display generation component, a first aggregated content item, which is a first plurality of content items selected from a media library including photos and / or videos taken by a user of the computer system, and is ordered according to a first set of selection criteria, playing visual content of the first aggregated content item; playing audio content while playing the visual content of the first aggregated content item; detecting that playing of the visual content of the first aggregated content item meets one or more end criteria after playing at least a portion of the visual content of the first aggregated content item; and after detecting that playing of the visual content of the first aggregated content item meets one or more end criteria, playing visual content of a second aggregated content item different from the first aggregated content item, which is a second plurality of content items different from the first plurality of content items, and is further selected from a media library including photos and / or videos taken by a user of the computer system and is ordered according to a second set of selection criteria, according to a determination that a first set of playback conditions among one or more playback conditions is satisfied.

[0012] According to some embodiments, a non-transitory computer-readable storage medium is described. The non-transitory computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system communicating with a display generation component and one or more input devices. The one or more programs, via the display generation component, play visual content of a first aggregated content item, which is a sequence of ordered first plural content items selected from a media library including photos and / or videos taken by a user of the computer system, based on a first set of selection criteria, play audio content while playing the visual content of the first aggregated content item, detect that playback of the visual content of the first aggregated content item meets one or more end criteria after playing at least a portion of the visual content of the first aggregated content item, and, after detecting that playback of the visual content of the first aggregated content item meets one or more end criteria, play visual content of a second aggregated content item different from the first aggregated content item, which is a sequence of ordered second plural content items different from the first plural content items, further selected from a media library including photos and / or videos taken by a user of the computer system and selected based on a second set of selection criteria, according to a determination that a first set of playback conditions among one or more playback conditions is satisfied, including instructions.

[0013] According to some embodiments, a non-transitory computer-readable storage medium is described. The non-transitory computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices. The one or more programs, via the display generation component, play visual content of a first aggregated content item, which is a sequence of ordered first plural content items selected from a media library including photos and / or videos taken by a user of the computer system, and selected based on a first set of selection criteria. While playing the visual content of the first aggregated content item, play audio content. After playing at least a portion of the visual content of the first aggregated content item, detect that the playing of the visual content of the first aggregated content item satisfies one or more end criteria. After detecting that the playing of the visual content of the first aggregated content item satisfies one or more end criteria, according to a determination that a first set of playback conditions among one or more playback conditions is satisfied, play visual content of a second aggregated content item different from the first aggregated content item, which is a sequence of ordered second plural content items different from the first plural content items, and further selected from a media library including photos and / or videos taken by a user of the computer system and selected based on a second set of selection criteria.

[0014] According to some embodiments, a computer system is described. The computer system is configured to communicate with a display generation component and one or more input devices, and includes one or more processors and a memory storing one or more programs configured to be executed by the one or more processors. The one or more programs cause, via the display generation component, a first aggregated content item, which is a sequential sequence of a first plurality of content items selected from a media library including photographs and / or videos taken by a user of the computer system, to be selected based on a first set of selection criteria, and play visual content of the first aggregated content item; play audio content while playing the visual content of the first aggregated content item; detect that playback of the visual content of the first aggregated content item meets one or more end criteria after playing at least a portion of the visual content of the first aggregated content item; and, after detecting that playback of the visual content of the first aggregated content item meets one or more end criteria, play visual content of a second aggregated content item different from the first aggregated content item, which is a sequential sequence of a second plurality of content items different from the first plurality of content items, and which is selected from a media library including photographs and / or videos taken by a user of the computer system and selected based on a second set of selection criteria, according to a determination that a first set of playback conditions among one or more playback conditions is satisfied.

[0015] According to some embodiments, a method is described. The method is in a computer system communicating with a display generation component and one or more input devices, and via the display generation component, plays visual content of a first aggregated content item, which includes an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria; while playing the visual content of the first aggregated content item, detects user input via the one or more input devices; in response to detecting the user input, pauses the playback of the visual content of the first aggregated content item; and via the display generation component, displays a user interface including a first representation of a first content item among the first plurality of content items and a second representation of a second content item among the first plurality of content items, simultaneously displaying a plurality of representations of content items within the first plurality of content items.

[0016] According to some embodiments, a non-transitory computer-readable storage medium is described. The non-transitory computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices. The one or more programs, via the display generation component, play visual content of a first aggregated content item, which includes an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria, and while playing the visual content of the first aggregated content item, detect user input via the one or more input devices, in response to detecting the user input, temporarily pause the playback of the visual content of the first aggregated content item, and via the display generation component, display a user interface including simultaneously displaying a first representation of a first content item among the first plurality of content items and a second representation of a second content item among the first plurality of content items.

[0017] According to some embodiments, a non-transitory computer-readable storage medium is described. The non-transitory computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs, via the display generation component, reproduce visual content of a first aggregated content item that includes an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria, detect user input via the one or more input devices while reproducing the visual content of the first aggregated content item, in response to detecting the user input, temporarily pause the reproduction of the visual content of the first aggregated content item, and display a user interface including simultaneously displaying a first representation of a first content item among the first plurality of content items and a second representation of a second content item among the first plurality of content items.

[0018] According to some embodiments, a computer system is described. The computer system is configured to communicate with a display generation component and one or more input devices, and includes one or more processors and a memory storing one or more programs configured to be executed by the one or more processors. The one or more programs, via the display generation component, reproduce visual content of a first aggregated content item, which is an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria, and while reproducing the visual content of the first aggregated content item, detect user input via the one or more input devices, and in response to detecting the user input, temporarily pause the reproduction of the visual content of the first aggregated content item, and via the display generation component, display a user interface including a first representation of a first content item among the first plurality of content items and a second representation of a second content item among the first plurality of content items, the plurality of representations of content items within the first plurality of content items being displayed simultaneously.

[0019] The executable instructions for performing these functions are optionally included in a non-transitory computer-readable storage medium or other computer program product configured to be executed by one or more processors. The executable instructions for performing these functions are optionally included in a transitory computer-readable storage medium or other computer program product configured to be executed by one or more processors.

[0020] Thus, a faster and more efficient method and interface for navigating, browsing, and editing media items are provided to a device, thereby increasing the effectiveness, efficiency, and user satisfaction of such a device. Such a method and interface can complement or replace other methods for navigating, browsing, and editing media items.

Brief Description of the Drawings

[0021] To better understand the various embodiments described, the following "Modes for Carrying Out the Invention" should be referred to in conjunction with the following drawings, and like reference numerals refer to corresponding parts throughout the following figures.

[0022]

Figure 1A

[0023]

Figure 1B

[0024]

Figure 2

[0025]

Figure 3

[0026]

Figure 4A

[0027]

Figure 4B

[0028]

Figure 5A

[0029]

Figure 5B

[0030]

Figure 6A

Figure 6B

Figure 6C

Figure 6D

Figure 6E

Figure 6F

Figure 6G

Figure 6H

Figure 6I

Figure 6J

Figure 6K

Figure 6L

Figure 6M

Figure 6N

Figure 6O

Figure 6P

Figure 6Q

Figure 6R

Figure 6S

Figure 6T

Figure 6U

Figure 6V

Figure 6W

Figure 6X

Figure 6Y

Figure 6Z

Figure 6AA

Figure 6AB

Figure 6AC

Figure 6AD

Figure 6AE

Figure 6AF

Figure 6AG

[0031]

Figure 7

[0032]

Figure 8A

Figure 8B

Figure 8C

Figure 8D

Figure 8E

Figure 8F

Figure 8G

Figure 8H

Figure 8I

Figure 8J

Figure 8K

Figure 8L

[0033]

Figure 9

[0034]

Figure 10A

Figure 10B

Figure 10C

Figure 10D

Figure 10E

Figure 10F

Figure 10G

Figure 10H

Figure 10I

Figure 10J

Figure 10K

Figure 10L

Figure 10M

Figure 10N

Figure 10O

Figure 10P

Figure 10Q

Figure 10R

Figure 10S

[0035]

Figure 11

[0036]

Figure 12A

Figure 12B

Figure 12C

Figure 12D

Figure 12E

Figure 12F

Figure 12G

Figure 12H

Figure 12I

Figure 12J

Figure 12K

Figure 12L

Figure 12M

Figure 12N

Figure 12O

Figure 12P

Figure 12Q

Figure 12R

Figure 12S

Figure 12T

Figure 12U

Figure 12V

Figure 12W

DETAILED DESCRIPTION

[0037] The following description sets forth exemplary methods, parameters, and the like. However, it should be recognized that such description is not intended as a limitation on the scope of the present disclosure, but rather as a description of exemplary embodiments.

[0038] There is a need for an electronic device that provides an efficient method and interface for navigating, viewing, and editing content items (e.g., media items such as photos and / or videos). For example, there is a need for technology that eliminates the significant manual effort by the user to retrieve media content relevant to the current context, and / or technology that eliminates the significant manual effort by the user to modify content items such as aggregated content items. Such technology can reduce the cognitive burden on the user who navigates, views, and / or edits content items, thereby improving productivity. Further, such technology can reduce the power of the processor and battery that would otherwise be wasted on redundant user input.

[0039] The following FIGS. 1A-1B, 2, 3, 4A-4B, and 5A-5B provide an illustration of an exemplary device that performs techniques for viewing, navigating, and editing content items. FIGS. 6A-6AG illustrate an exemplary user interface for viewing and modifying content items while visual content is playing continuously. FIG. 7 is a flowchart showing a method for modifying content items while visual content is playing continuously, according to some embodiments. The user interfaces of FIGS. 6A-6AG are used to illustrate the processes described below, including the process of FIG. 7. FIGS. 8A-8L illustrate an exemplary user interface for managing the playback of content after a content item has been played. FIG. 9 is a flowchart showing a method for managing the playback of content after a content item has been played, according to some embodiments. The user interfaces of FIGS. 8A-8L are used to illustrate the processes described below, including the process of FIG. 9. FIGS. 10A-10S illustrate an exemplary user interface for viewing the presentation of a content item. FIG. 11 is a flowchart showing a method for viewing the presentation of a content item, according to some embodiments. The user interfaces of FIGS. 10A-10S are used to illustrate the processes described below, including the process of FIG. 11. FIGS. 12A-12W illustrate an exemplary user interface for viewing, navigating, and editing content items. The user interfaces of FIGS. 12A-12W are used to illustrate the processes described below, including the processes of FIGS. 7, 9, and 11.

[0040] The processes described below improve the device's operability and make the user interface with the device more efficient (e.g., by helping to provide appropriate input and reducing user errors when the user operates / interacts with the device) by various techniques, including providing improved visual feedback to the user, reducing the number of inputs required to perform an operation, providing additional control options without cluttering the user interface with additional controls being displayed, performing an operation without requiring further user input when a set of conditions is met, and / or other techniques. These techniques also reduce power usage and improve the device's battery life by enabling the user to use the device more quickly and efficiently.

[0041] Furthermore, in the method described herein where one or more steps are conditional upon one or more conditions being met, it should be understood that the method can be repeated in multiple iterations such that all of the conditions upon which the steps of the method are conditional are met in different iterations of the method over the course of the repetition. For example, if a method requires performing a first step when a condition is met and a second step when the condition is not met, one of ordinary skill in the art would understand that the steps recited in the claims are repeated in no particular order until the condition is met and then ceases to be met. Thus, a method described in terms of one or more steps that are conditional upon one or more conditions being met can be rewritten as a method that is repeated until each condition described in the method is met. However, this is not required in claims for a system or computer-readable medium that includes instructions to perform conditional operations based on the fulfillment of the corresponding one or more conditions, and thus can determine whether the contingency is met without explicitly repeating the steps of the method until all of the conditions upon which the steps of the method are conditional are met. One of ordinary skill in the art will also understand that a system or computer-readable storage medium can repeat the steps of the method as many times as necessary to ensure that all of the conditional steps are executed, similar to a method with conditional steps.

[0042] In the following description, terms such as "first", "second", etc. are used to describe various elements, but these elements should not be limited by these terms. These terms are only used to distinguish one element from another. For example, without departing from the scope of the various embodiments described, a first touch could be referred to as a second touch, and similarly a second touch could be referred to as a first touch. Both the first touch and the second touch are touches, but they are not the same touch.

[0043] The terms used in the description of the various embodiments described herein are for the purpose of describing particular embodiments only and are not intended to be limiting. When used in the description of the various embodiments described and the appended claims, the singular forms "a", "an", and "the" are intended to include the plural as well, unless the context clearly dictates otherwise. Also, as used herein, the term "and / or" refers to and includes any and all possible combinations of one or more of the associated listed items. It should be understood that the terms "includes", "including", "comprises", and / or "comprising", when used herein, specify the presence of the stated features, integers, steps, operations, elements, and / or components, but do not preclude the presence or addition of one or more other features, integers, steps, operations, elements, components, and / or groups thereof.

[0044] The term "if" optionally, depending on context, is interpreted to mean "when" or "upon", or "in response to determining" or "in response to detecting". Similarly, the phrases "if it is determined" or "if [a stated condition or event] is detected" optionally, depending on context, are interpreted to mean "upon determining" or "in response to determining", or "upon detecting [the stated condition or event]" or "in response to detecting [the stated condition or event]".

[0045] Embodiments of electronic devices, user interfaces for such devices, and related processes for using such devices are described. In some embodiments, the device is a portable communication device such as a cellular phone that also includes other functions such as PDA functionality and / or music player functionality. Exemplary embodiments of portable multifunctional devices include, but are not limited to, the iPhone®, iPod Touch®, and iPad® devices from Apple Inc. of Cupertino, California. Optionally, other portable electronic devices such as laptop computers or tablet computers having a touch-sensitive surface (e.g., a touch screen display and / or a touch pad) are also used. Also, in some embodiments, it should be understood that the device is not a portable communication device but a desktop computer having a touch-sensitive surface (e.g., a touch screen display and / or a touch pad). In some embodiments, the electronic device is a computer system that communicates (e.g., via wireless communication, via wired communication) with a display generation component. The display generation component is configured to provide a visual output such as a display via a CRT display, a display via an LED display, or a display via image projection. In some embodiments, the display generation component is integrated with the computer system. In some embodiments, the display generation component is separate from the computer system. As used herein, "displaying" content includes causing content (e.g., video data rendered or decoded by a display controller 156) to be displayed by sending data (e.g., image data or video data) via a wired or wireless connection to an integrated or external display generation component for visually generating the content.

[0046] In the following discussion, an electronic device including a display and a touch sensing surface will be described. However, it should be understood that the electronic device optionally includes one or more other physical user interface devices such as a physical keyboard, a mouse, and / or a joystick.

[0047] The device typically supports various applications such as one or more of a drawing application, a presentation application, a word processing application, a website creation application, a disk authoring application, a spreadsheet application, a game application, a phone application, a video conferencing application, an email application, an instant messaging application, a training support application, a photo management application, a digital camera application, a digital video camera application, a web browsing application, a digital music player application, and / or a digital video player application.

[0048] The various applications executed on the device optionally use at least one common physical user interface device such as a touch sensing surface. One or more functions of the touch sensing surface, as well as the corresponding information displayed on the device, are optionally adjusted and / or changed for each application and / or within an individual application. Thus, the common physical architecture of the device (such as the touch sensing surface) optionally supports various applications with a user interface that is intuitive and transparent to the user.

[0049] Attention is now directed to an embodiment of a portable device equipped with a touch-sensing display. FIG. 1A is a block diagram showing a portable multifunctional device 100 having a touch-sensing display system 112 according to some embodiments. The touch-sensing display 112 may be referred to as a "touch screen" for convenience and may be known or referred to as a "touch-sensing display system". The device 100 includes a memory 102 (optionally including one or more computer-readable storage media), a memory controller 122, one or more processing units (CPUs) 120, a peripheral device interface 118, an RF circuit 108, an audio circuit 110, a speaker 111, a microphone 113, an input / output (I / O) subsystem 106, other input control devices 116, and an external port 124. The device 100 optionally includes one or more optical sensors 164. The device 100 optionally includes one or more contact intensity sensors 165 (e.g., a touch-sensing surface such as the touch-sensing display system 112 of the device 100) for detecting the intensity of contact on the device 100. The device 100 optionally includes one or more haptic output generators 167 for generating haptic output on the device 100 (e.g., generating haptic output on a touch-sensing surface such as the touch-sensing display system 112 of the device 100 or the touch pad 355 of the device 300). These components communicate, optionally, via one or more communication buses or signal lines 103.

[0050] As used in this specification and the claims, the term "intensity" of a contact on a touch-sensing surface refers to the force or pressure (force per unit area) of a contact (e.g., a finger contact) on the touch-sensing surface, or an alternative (proxy) for the force or pressure of a contact on the touch-sensing surface. The intensity of a contact has a range of values that includes at least four distinct values, and more typically, hundreds (e.g., at least 256) of distinct values. The intensity of a contact is optionally determined (or measured) using a variety of techniques and a variety of sensors or combinations of sensors. For example, one or more force sensors under or adjacent to the touch-sensing surface are optionally used to measure the force at various points on the touch-sensing surface. In some implementations, force measurements from multiple force sensors are combined (e.g., weighted averaged) to determine the estimated force of a contact. Similarly, a pressure-sensitive tip of a stylus is optionally used to determine the pressure of the stylus on the touch-sensing surface. Alternatively, the size and / or change thereof of the contact area detected on the touch-sensing surface, the capacitance and / or change thereof of the touch-sensing surface proximate to the contact, and / or the resistance and / or change thereof of the touch-sensing surface proximate to the contact are optionally used as an alternative for the force or pressure of a contact on the touch-sensing surface. In some implementations, an alternative measurement of the force or pressure of a contact is directly used to determine whether it exceeds an intensity threshold (e.g., the intensity threshold is described in units corresponding to the alternative measurement). In some implementations, an alternative measurement of the force or pressure of a contact is converted to an estimated value of force or pressure, and the estimated value of force or pressure is used to determine whether it exceeds an intensity threshold (e.g., the intensity threshold is a pressure threshold measured in units of pressure). By using the intensity of a contact as an attribute of user input, a user can access additional device functions that may otherwise be inaccessible to the user on a reduced-size device with a limited implementation area for displaying affordances (e.g., on a touch-sensing display), and / or receive user input (e.g., via a touch-sensing display, a touch-sensing surface, or a physical / mechanical control such as a knob or button).

[0051] As used in this specification and the claims, the term "haptic output" refers to the physical displacement of the device relative to its previous position, the physical displacement of a component of the device (e.g., a touch-sensitive surface) relative to another component of the device (e.g., the housing), or the displacement of a component relative to the center of mass of the device, which will be detected by the user's sense of touch. For example, in a situation where the device or a component of the device is in contact with a touch-sensitive surface of the user (e.g., the finger, palm, or other part of the user's hand), the haptic output generated by the physical displacement will be interpreted by the user as a tactile sensation corresponding to a perceived change in the physical characteristics of the device or the component of the device. For example, the movement of a touch-sensitive surface (e.g., a touch-sensitive display or a trackpad) may optionally be interpreted by the user as a "down click" or "up click" of a physical actuator button. In some cases, even when there is no movement of the physical actuator button associated with the touch-sensitive surface physically pressed (e.g., displaced) by the user's action, the user may feel a tactile sensation such as a "down click" or "up click". As another example, the movement of the touch-sensitive surface may optionally be interpreted or perceived by the user as the "roughness" of the touch-sensitive surface even when there is no change in the smoothness of the touch-sensitive surface. Such an interpretation of touch by the user depends on the user's individual sensory perception, but there are many sensory perceptions of touch that are common to the majority of users. Therefore, when the haptic output is described as corresponding to a particular sensory perception of the user (e.g., "up click", "down click", "roughness"), unless otherwise stated, the generated haptic output corresponds to the physical displacement of the device or a component of the device that generates the described sensory perception of a typical (or average) user.

[0052] Device 100 is merely an example of a portable multifunctional device, and it should be understood that Device 100 may optionally have more or fewer components than those shown, may optionally combine two or more components, or may optionally have different configurations or arrangements of those components. The various components shown in FIG. 1A are implemented in a combination of hardware, software, or both hardware and software, including one or more signal processing circuits and / or application specific integrated circuits.

[0053] Memory 102 optionally includes high-speed random access memory and also optionally includes non-volatile memory such as one or more magnetic disk storage devices, flash memory devices, or other non-volatile solid state memory devices. Memory controller 122 optionally controls access to memory 102 by other components of device 100.

[0054] Peripheral interface 118 can be used to couple input and output peripheral devices of the device to CPU 120 and memory 102. One or more processors 120 operate or execute various software programs (such as computer programs including instructions) and / or instruction sets stored in memory 102 to perform various functions for device 100 and process data. In some embodiments, peripheral interface 118, CPU 120, and memory controller 122 are optionally implemented on a single chip such as chip 104. In some other embodiments, they are optionally implemented on separate chips.

[0055] The RF (radio frequency) circuit 108 transmits and receives RF signals, also called electromagnetic signals. The RF circuit 108 converts electrical signals into electromagnetic signals or vice versa and communicates with a communication network and other communication devices via the electromagnetic signals. The RF circuit 108 optionally includes well-known circuits for performing these functions, such as, but not limited to, an antenna system, an RF transceiver, one or more amplifiers, a tuner, one or more oscillators, a digital signal processor, a CODEC chipset, a subscriber identity module (SIM) card, a memory, and the like. The RF circuit 108 optionally communicates wirelessly with a network such as the Internet, also called the World Wide Web (WWW), an intranet, and / or a wireless network such as a cellular telephone network, a wireless local area network (LAN), and / or a metropolitan area network (MAN), as well as with other devices. The RF circuit 108 optionally includes well-known circuits for detecting a near field communication (NFC) field, such as by a short-range communication radio. Wireless communication optionally includes, but is not limited to, Global System for Mobile Communications (GSM) for mobile communication, Enhanced Data GSM Environment (EDGE), high-speed downlink packet access (HSDPA), high-speed uplink packet access (HSUPA), Evolution, Data-Only (EV-DO), HSPA, HSPA+, Dual-Cell HSPA (DC-HSPDA), Long Termevolution, LTE), Near Field Communication (NFC), Wideband Code Division Multiple Access (W-CDMA), Code Division Multiple Access (CDMA), Time Division Multiple Access (TDMA), Bluetooth, Bluetooth Low Energy (BTLE), Wireless Fidelity (Wi-Fi) (e.g., IEEE802.11a, IEEE802.11b, IEEE802.11g, IEEE802.11n, and / or IEEE802.11ac), Voice over Internet Protocol (VoIP), Wi-MAX, protocols for email (e.g., Internet Message Access Protocol (IMAP) and / or Post Office Protocol (POP)), instant messaging (e.g., Extensible Messaging and Presence Protocol (XMPP), Session Initiation Protocol for Instant Messaging and Presence Leveraging Extensions (SIMPLE), Instant Messaging and Presence Service (IMPS)), and / or Short Message Service (SMS), or any other suitable communication protocol including communication protocols not yet developed as of the filing date of this specification. Any one of a plurality of communication standards, protocols, and technologies is used.

[0056] The audio circuit 110, speaker 111, and microphone 113 provide an audio interface between the user and the device 100. The audio circuit 110 receives audio data from the peripheral device interface 118, converts this audio data into an electrical signal, and transmits this electrical signal to the speaker 111. The speaker 111 converts the electrical signal into human audible sound waves. Also, the audio circuit 110 receives the electrical signal converted from sound waves by the microphone 113. The audio circuit 110 converts the electrical signal into audio data and transmits this audio data to the peripheral device interface 118 for processing. The audio data is optionally retrieved from and / or transmitted to the memory 102 and / or the RF circuit 108 by the peripheral device interface 118. In some embodiments, the audio circuit 110 also includes a headset jack (e.g., 212 of FIG. 2). The headset jack provides an interface between the audio circuit 110 and a removable audio input / output peripheral device such as an output-only headset or a headset having both an output (e.g., mono or stereo headphones) and an input (e.g., microphone).

[0057] The I / O subsystem 106 couples input / output peripheral devices on the device 100, such as the touch screen 112 and other input control devices 116, to the peripheral device interface 118. The I / O subsystem 106 optionally includes a display controller 156, an optical sensor controller 158, a depth camera controller 169, an intensity sensor controller 159, a haptic feedback controller 161, and one or more input controllers 160 for other input devices or control devices. The one or more input controllers 160 receive electrical signals from and transmit electrical signals to the other input control devices 116. The other input control devices 116 optionally include physical buttons (e.g., push buttons, rocker buttons, etc.), dials, slider switches, joysticks, click wheels, etc. In some embodiments, the input controller(s) 160 are optionally coupled to (or not coupled to any of) a pointer device such as a keyboard, an infrared port, a USB port, and a mouse. One or more buttons (e.g., 208 in FIG. 2) optionally include up / down buttons for volume control of the speaker 111 and / or the microphone 113. One or more buttons optionally include push buttons (e.g., 206 in FIG. 2). In some embodiments, the electronic device is a computer system that communicates with one or more input devices (e.g., via wireless communication, via wired communication). In some embodiments, the one or more input devices include a touch sensing surface (e.g., a trackpad as part of a touch sensing display). In some embodiments, the one or more input devices include one or more camera sensors (e.g., one or more optical sensors 164 and / or one or more depth camera sensors 175) for tracking a user's gesture (e.g., a hand gesture) as an input. In some embodiments, the one or more input devices are integrated with the computer system. In some embodiments, the one or more input devices are separate from the computer system.

[0058] As described in U.S. Patent Application No. 11 / 322,549, filed December 23, 2005, "Unlocking a Device by Performing Gestures on an Unlock Image," and U.S. Patent No. 7,657,849, which are hereby incorporated by reference in their entirety, a quick press of a push button optionally unlocks the touch screen 112 or, optionally, initiates a process of unlocking the device using gestures on the touch screen. A longer press of a push button (e.g., 206) optionally turns the power to the device 100 on or off. The function of one or more of the buttons can optionally be customized by the user. The touch screen 112 is used to implement virtual or soft buttons and one or more soft keyboards.

[0059] The touch-sensitive display 112 provides an input interface and an output interface between the device and the user. The display controller 156 receives electrical signals from the touch screen 112 and / or transmits electrical signals to the touch screen 112. The touch screen 112 displays visual output to the user. This visual output optionally includes graphics, text, icons, video, and any combination thereof (collectively referred to as "graphics"). In some embodiments, some or all of the visual output optionally corresponds to user interface objects.

[0060] The touch screen 112 has a touch sensing surface, sensor, or set of sensors that accepts input from a user based on tactile and / or haptic contact. The touch screen 112 and the display controller 156 detect contact (and any movement or interruption of the contact) on the touch screen 112 (along with any associated modules and / or instruction sets within the memory 102), and convert the detected contact into an interaction with user interface objects (e.g., one or more soft keys, icons, web pages, or images) displayed on the touch screen 112. In an exemplary embodiment, the point of contact between the touch screen 112 and the user corresponds to the user's finger.

[0061] The touch screen 112 optionally uses LCD (liquid crystal display) technology, LPD (light emitting polymer display) technology, or LED (light emitting diode) technology, although in other embodiments other display technologies are used. The touch screen 112 and the display controller 156 optionally detect contact and any movement or interruption thereof using any of a plurality of touch sensing technologies currently known or later developed, including but not limited to capacitive, resistive, infrared, and surface acoustic wave technologies, as well as other proximity sensor arrays or other elements for determining one or more points of contact with the touch screen 112. In an exemplary embodiment, projected mutual capacitance sensing technology such as that found in the iPhone (registered trademark) and iPod Touch (registered trademark) from Apple Inc. of Cupertino, California is used.

[0062] The touch-sensing display in some embodiments of touch screen 112 is optionally similar to a multi-touch sensing touch pad described in U.S. Patent No. 6,323,846 (Westerman et al.), No. 6,570,557 (Westerman et al.), and / or No. 6,677,932 (Westerman), and / or U.S. Patent Application Publication No. 2002 / 0015024 (A1), each of which is hereby incorporated by reference in its entirety. However, while touch screen 112 displays visual output from device 100, the touch-sensing touch pad does not provide visual output.

[0063] Touch sensing displays in some embodiments of the touch screen 112 are described in the following applications: (1) U.S. Patent Application No. 11 / 381,313, filed May 2, 2006, "Multipoint Touch Surface Controller"; (2) U.S. Patent Application No. 10 / 840,862, filed May 6, 2004, "Multipoint Touchscreen"; (3) U.S. Patent Application No. 10 / 903,964, filed Jul. 30, 2004, "Gestures For Touch Sensitive Input Devices"; (4) U.S. Patent Application No. 11 / 048,264, filed Jan. 31, 2005, "Gestures For Touch Sensitive Input Devices"; (5) U.S. Patent Application No. 11 / 038,590, filed Jan. 18, 2005, "Mode-Based Graphical User Interfaces For Touch Sensitive Input Devices"; (6) U.S. Patent Application No. 11 / 228,758, filed Sep. 16, 2005, "Virtual Input Device Placement On A Touch Screen User Interface"; (7) U.S. Patent Application No. 11 / 228,700, filed Sep. 16, 2005, "Operation Of A Computer With A Touch Screen Interface"; (8) U.S. Patent Application No. 11 / 228,737, filed Sep. 16, 2005, "Activating Virtual Keys Of A Touch-Screen Virtual Keyboard"; and (9) U.S. Patent Application No. 11 / 367,749, filed Mar. 3, 2006, "Multi-Functional Hand-Held Device". All of these applications are hereby incorporated by reference in their entirety.

[0064] The touch screen 112 optionally has a video resolution greater than 100 dpi. In some embodiments, the touch screen has a video resolution of about 160 dpi. The user optionally touches the touch screen 112 using any suitable object or appendage such as a stylus, finger, etc. In some embodiments, the user interface is designed to operate primarily using finger-based contact and gestures, although this may be less accurate than stylus-based input because the contact area of the finger on the touch screen is larger. In some embodiments, the device converts the finger-based rough input into an accurate pointer / cursor position or command for performing the action desired by the user.

[0065] In some embodiments, in addition to the touch screen, the device 100 optionally includes a touch pad for activating or deactivating certain functions. In some embodiments, the touch pad, unlike the touch screen, is a touch-sensitive area of the device that does not display a visual output. The touch pad is optionally a separate touch-sensitive surface from the touch screen 112 or an extension of the touch-sensitive surface formed by the touch screen.

[0066] The device 100 also includes a power system 162 that powers various components. The power system 162 optionally includes a power management system, one or more power sources (e.g., battery, alternating current (AC)), a recharge system, a power outage detection circuit, a power converter or inverter, a power status indicator (e.g., light-emitting diode (LED)), and any other components associated with the generation, management, and distribution of power within a portable device.

[0067] In addition, device 100 optionally includes one or more optical sensors 164. FIG. 1A shows an optical sensor coupled to an optical sensor controller 158 within I / O subsystem 106. Optical sensor 164 optionally includes a charge-coupled device (CCD) or a complementary metal-oxide semiconductor (CMOS) phototransistor. Optical sensor 164 receives light from the environment projected through one or more lenses and converts that light into data representing an image. Optical sensor 164 cooperates with imaging module 143 (also referred to as a camera module) to optionally capture a still image or video. In some embodiments, the optical sensor is located on the back surface of device 100 opposite touch screen display 112 on the front of the device, and thus the touch screen display can be used as a viewfinder for acquiring still and / or video images. In some embodiments, the optical sensor is disposed on the front of the device such that an image of the user is optionally acquired for a video conference while the user is viewing other video conference participants on the touch screen display. In some embodiments, the position of optical sensor 164 can be changed by the user (e.g., by rotating the lens and sensor within the device housing), and thus a single optical sensor 164 is used with the touch screen display for both video conferencing and for acquiring still and / or video images.

[0068] Device 100 optionally also includes one or more depth camera sensors 175. FIG. 1A shows a depth camera sensor coupled to a depth camera controller 169 within I / O subsystem 106. The depth camera sensor 175 receives data from the environment and creates a three-dimensional model of an object (e.g., a face) within the scene from the perspective (e.g., the depth camera sensor). In some embodiments, in conjunction with imaging module 143 (also referred to as the camera module), the depth camera sensor 175 is optionally used to determine depth maps of different portions of an image captured by imaging module 143. In some embodiments, while a user is viewing other video conferencing participants on a touch screen display, a depth camera sensor is disposed on the front face of device 100 to optionally acquire an image of the user with depth information for video conferencing and also to capture a self-portrait image with depth map data. In some embodiments, the depth camera sensor 175 is disposed on the back of the device, or on both the back and front faces of device 100. In some embodiments, the position of the depth camera sensor 175 can be changed by the user (e.g., by rotating the lens and sensor within the device housing), such that the depth camera sensor 175 is used for both video conferencing and for acquiring still and / or video images with the touch screen display.

[0069] Device 100 also optionally includes one or more contact intensity sensors 165. FIG. 1A shows a contact intensity sensor coupled to an intensity sensor controller 159 within I / O subsystem 106. The contact intensity sensor 165 optionally includes one or more piezoresistive strain gauges, capacitive force sensors, electro-force sensors, piezoelectric force sensors, optical force sensors, capacitive touch sensing surfaces, or other intensity sensors (e.g., sensors used to measure the force (or pressure) of contact on a touch sensing surface). The contact intensity sensor 165 receives contact intensity information (e.g., pressure information, or a proxy for pressure information) from the environment. In some embodiments, at least one contact intensity sensor is juxtaposed with or proximate to a touch sensing surface (e.g., touch sensing display system 112). In some embodiments, at least one contact intensity sensor is disposed on the back of device 100, opposite a touch screen display 112 disposed on the front of device 100.

[0070] Device 100 also optionally includes one or more proximity sensors 166. FIG. 1A shows a proximity sensor 166 coupled to the peripheral device interface 118. Alternatively, the proximity sensor 166 is optionally coupled to an input controller 160 within the I / O subsystem 106. The proximity sensor 166 functions optionally as described in U.S. Patent Application Nos. 11 / 241,839, "Proximity Detector In Handheld Device", 11 / 240,788, "Proximity Detector In Handheld Device", 11 / 620,702, "Using Ambient Light Sensor To Augment Proximity Sensor Output", 11 / 586,862, "Automated Response To And Sensing Of User Activity In Portable Devices", and 11 / 638,251, "Methods And Systems For Automatic Configuration Of Peripherals", which are hereby incorporated by reference in their entirety. In some embodiments, when a multifunctional device is placed near the user's ear (e.g., when the user is making a phone call), the proximity sensor turns off and disables the touch screen 112.

[0071] Device 100 also optionally includes one or more haptic output generators 167. FIG. 1A shows a haptic output generator coupled to a haptic feedback controller 161 within I / O subsystem 106. The haptic output generator 167 optionally includes one or more electroacoustic devices, such as speakers or other audio components, and / or electromechanical devices that convert energy, such as motors, solenoids, electroactive polymers, piezoelectric actuators, electrostatic actuators, or other haptic output generating components (e.g., components that convert an electrical signal into a haptic output on the device), into linear movement. The contact intensity sensor 165 receives haptic feedback generation instructions from the haptic feedback module 133 and generates a haptic output on device 100 that can be sensed by a user of device 100. In some embodiments, at least one haptic output generator is juxtaposed with or proximate to a touch sensing surface (e.g., touch sensing display system 112) and optionally generates a haptic output by moving the touch sensing surface in a vertical direction (e.g., in / out of the surface of device 100) or in a horizontal direction (e.g., back and forth within the same plane as the surface of device 100). In some embodiments, at least one haptic output generator sensor is disposed on the back of device 100, which is opposite the touch screen display 112 disposed on the front of device 100.

[0072] Device 100 also optionally includes one or more accelerometers 168. FIG. 1A shows an accelerometer 168 coupled to the peripheral device interface 118. Alternatively, the accelerometer 168 is optionally coupled to the input controller 160 within the I / O subsystem 106. The accelerometer 168 functions optionally as described in both U.S. Patent Application Publication No. 20050190059, "Acceleration-based Theft Detection System for Portable Electronic Devices", and U.S. Patent Application Publication No. 20060017692, "Methods And Apparatuses For Operating A Portable Device Based On An Accelerometer", which are hereby incorporated by reference in their entirety. In some embodiments, information is displayed on the touch screen display in a portrait or landscape display based on analysis of data received from one or more accelerometers. In addition to the accelerometer(s) 168, device 100 optionally includes a magnetometer and a GPS (or GLONASS or other global navigation system) receiver for obtaining information regarding the position and orientation (e.g., portrait or landscape orientation) of device 100.

[0073] In some embodiments, the software components stored in the memory 102 include an operating system 126, a communication module (or instruction set) 128, a touch / motion module (or instruction set) 130, a graphics module (or instruction set) 132, a text input module (or instruction set) 134, a Global Positioning System (GPS) module (or instruction set) 135, and an application (or instruction set) 136. Further, in some embodiments, the memory 102 (FIG. 1A) or 370 (FIG. 3) stores a device / global internal state 157 as shown in FIGS. 1A and 3. The device / global internal state 157 includes an active application state indicating which application is active if there is a currently active application, a display state indicating which application, view, or other information occupies various regions of the touch screen display 112, a sensor state including information obtained from various sensors and input control devices 116 of the device, and one or more of position information regarding the position and / or orientation of the device.

[0074] The operating system 126 (e.g., an embedded operating system such as Darwin, RTXC, LINUX, UNIX, OS X, iOS, WINDOWS, or VxWorks) includes various software components and / or drivers that control and manage general system tasks (e.g., memory management, storage device control, power management, etc.) and facilitate communication between various hardware components and software components.

[0075] The communication module 128 facilitates communication with other devices via one or more external ports 124 and also includes various software components for processing data received by the RF circuit 108 and / or the external port 124. The external ports 124 (e.g., Universal Serial Bus (USB), FIREWIRE, etc.) are adapted to couple to other devices either directly or indirectly via a network (e.g., the Internet, a wireless LAN, etc.). In some embodiments, the external port is a multi-pin (e.g., 30-pin) connector that is the same as or similar to and / or compatible with the 30-pin connector used on iPod (registered trademark) devices (a trademark of Apple Inc.).

[0076] The contact / motion module 130 optionally detects contact with the touch screen 112 and other touch-sensing devices (e.g., a touch pad or a physical click wheel) (in cooperation with the display controller 156). The contact / motion module 130 includes various software components for performing various operations related to the detection of contact, such as determining whether contact has occurred (e.g., detecting a finger-down event), determining the intensity of the contact (e.g., the force or pressure of the contact, or an alternative to the force or pressure of the contact), determining whether there is movement of the contact, tracking movement across the touch-sensing surface (e.g., detecting one or more finger-drag events), and determining whether the contact has ceased (e.g., detecting a finger-up event or an interruption of the contact). The contact / motion module 130 receives contact data from the touch-sensing surface. Determining the movement of the contact point, represented by a series of contact data, optionally includes determining the speed (magnitude), velocity (magnitude and direction), and / or acceleration (change in magnitude and / or direction) of the contact point. These operations are optionally applicable to a single contact (e.g., contact with one finger) or multiple simultaneous contacts (e.g., "multi-touch" / contact with multiple fingers). In some embodiments, the contact / motion module 130 and the display controller 156 detect contact on the touch pad.

[0077] In some embodiments, the contact / motion module 130 uses a set of one or more intensity thresholds to determine whether an action has been performed by the user (e.g., to determine whether the user has "clicked" on an icon). In some embodiments, at least a subset of the intensity thresholds are determined according to software parameters (e.g., the intensity thresholds can be adjusted without changing the physical hardware of the device 100, rather than being determined by the activation threshold of a particular physical actuator). For example, the mouse "click" threshold for a trackpad or touch screen display can be set to any of a wide range of predefined thresholds without changing the trackpad or touch screen display hardware. Additionally, in some implementations, the user of the device is provided with software settings to adjust one or more of the set of intensity thresholds (e.g., by adjusting individual intensity thresholds and / or adjusting multiple intensity thresholds at once according to a system-level click "intensity" parameter).

[0078] The contact / motion module 130 optionally detects gesture inputs by the user. Different gestures on the touch-sensitive surface have different contact patterns (e.g., different detected contact motions, timings, and / or intensities). Thus, gestures are optionally detected by detecting a particular contact pattern. For example, detecting a finger tap gesture includes detecting a finger down event followed by detecting a finger up (lift off) event at the same position (or substantially the same position) (e.g., the position of an icon) as the finger down event. As another example, detecting a finger swipe gesture on the touch-sensitive surface includes detecting a finger down event followed by detecting one or more finger drag events, followed by detecting a finger up (lift off) event.

[0079] The graphic module 132 includes various known software components for rendering and displaying graphics on the touch screen 112 or other display, including components for changing the visual effects of the displayed graphics (e.g., brightness, transparency, saturation, contrast, or other visual properties). As used herein, the term "graphic" includes, but is not limited to, any object that can be displayed to the user, including text, web pages, icons (such as user interface objects including soft keys), digital images, videos, animations, and the like.

[0080] In some embodiments, the graphic module 132 stores data representing the graphics that will be used. Each graphic is optionally assigned a corresponding code. The graphic module 132 receives from an application, as needed, one or more codes specifying the graphics to be displayed, along with coordinate data and other graphic property data, and then generates the image data of the screen to be output to the display controller 156.

[0081] The tactile feedback module 133 includes various software components for generating the instructions used by the tactile output generator(s) 167 to generate tactile output at one or more locations on the device 100 in response to the user's interaction with the device 100.

[0082] The text input module 134 is optionally a component of the graphic module 132 and provides a soft keyboard for entering text in various applications (e.g., contacts 137, email 140, IM 141, browser 147, and any other application that requires text input).

[0083] The GPS module 135 determines the location of the device and provides this information for use within various applications (e.g., to the phone 138 for use in location-based dialing, to the camera 143 as photo / video metadata, and to applications that provide location-based services such as weather widgets, local yellow page widgets, and map / navigation widgets).

[0084] The application 136 optionally includes the following modules (or sets of instructions) or subsets or supersets thereof. ● Contact module 137 (also sometimes referred to as an address book or contact list), ● Phone module 138, ● Video conferencing module 139, ● Email client module 140, ● Instant messaging (IM) module 141, ● Training support module 142, ● Camera module 143 for still and / or video images, ● Image management module 144, ● Video player module, ● Music player module, ● Browser module 147, ● Calendar module 148, ● Optionally, a widget module 149 that includes one or more of weather widget 149-1, stock price widget 149-2, calculator widget 149-3, alarm clock widget 149-4, dictionary widget 149-5, and other widgets acquired by the user, as well as user-created widget 149-6, ● Widget creator module 150 for creating user-created widget 149-6, ● Search module 151, ● A video and music player module 152 that integrates a video player module and a music player module. ● A memo module 153. ● A map module 154, and / or ● An online video module 155.

[0085] Examples of other applications 136 that are optionally stored in the memory 102 include other word processing applications, other image editing applications, drawing applications, presentation applications, Java-compatible applications, encryption, digital rights management, voice recognition, and voice replication.

[0086] The contact module 137 cooperates with the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, and the text input module 134 to optionally manage an address book or contact list (e.g., stored in the application internal state 192 of the contact module 137 in the memory 102 or the memory 370), including adding a name(s) to the address book, deleting a name(s) from the address book, associating a phone number(s), an email address(es), an address(es), or other information with a name, associating an image with a name, classifying and sorting names, providing a phone number or email address to initiate and / or facilitate communication by the phone 138, the video conferencing module 139, the email 140, or the IM 141, etc.

[0087] The telephone module 138 works in conjunction with the RF circuit 108, the audio circuit 110, the speaker 111, the microphone 113, the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, and the text input module 134, and is optionally used for the input of a series of characters corresponding to a telephone number, access to one or more telephone numbers in the contact module 137, modification of the input telephone number, dialing of an individual telephone number, execution of a call, and disconnection or call hold at the end of a call. As described above, the wireless communication optionally uses any of a plurality of communication standards, protocols, and technologies.

[0088] The videoconference module 139 works in conjunction with the RF circuit 108, the audio circuit 110, the speaker 111, the microphone 113, the touch screen 112, the display controller 156, the optical sensor 164, the optical sensor controller 158, the contact / motion module 130, the graphic module 132, the text input module 134, the contact module 137, and the telephone module 138, and includes executable instructions for starting, executing, and ending a videoconference between the user and one or more other participants according to the user's instructions.

[0089] The email client module 140 works in conjunction with the RF circuit 108, the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, and the text input module 134, and includes executable instructions for creating, sending, receiving, and managing emails in response to the user's instructions. The email client module 140 works in conjunction with the image management module 144 to facilitate the creation and sending of emails with still or video images captured by the camera module 143.

[0090] The instant messaging module 141, in cooperation with the RF circuit 108, touch screen 112, display controller 156, contact / motion module 130, graphic module 132, and text input module 134, includes executable instructions for inputting a series of characters corresponding to an instant message, modifying previously input characters, and transmitting (e.g., using the Short Message Service (SMS) or Multimedia Message Service (MMS) protocol for phone communication-based instant messages, or XMPP, SIMPLE, or IMPS for Internet-based instant messages), receiving, and viewing received instant messages. In some embodiments, the instant messages transmitted and / or received optionally include graphics, photos, audio files, video files, and / or other attachment files supported by MMS and / or Enhanced Messaging Service (EMS). As used herein, "instant messaging" refers to both phone communication-based messages (e.g., messages transmitted using SMS or MMS) and Internet-based messages (e.g., messages transmitted using XMPP, SIMPLE, or IMPS).

[0091] The training support module 142, in cooperation with the RF circuit 108, touch screen 112, display controller 156, contact / motion module 130, graphic module 132, text input module 134, GPS module 135, map module 154, and music player module, creates training (e.g., having time, distance, and / or calorie burn goals), communicates with a training sensor (sports device), receives training sensor data, calibrates sensors used to monitor the training, selects and plays music for the training, and includes executable instructions to display, store, and transmit training data.

[0092] The camera module 143, in cooperation with the touch screen 112, display controller 156, optical sensor(s) 164, optical sensor controller 158, contact / motion module 130, graphic module 132, and image management module 144, captures still images or videos (including video streams), stores them in the memory 102, modifies the characteristics of the still images or videos, or deletes the still images or videos from the memory 102, and includes executable instructions.

[0093] The image management module 144, in cooperation with the touch screen 112, display controller 156, contact / motion module 130, graphic module 132, text input module 134, and camera module 143, arranges, modifies (e.g., edits), or otherwise operates on, labels, deletes, presents (e.g., in a digital slide show or album), and stores still images and / or video images, and includes executable instructions.

[0094] The browser module 147 includes executable instructions for browsing the Internet according to user instructions, including searching for, linking to, receiving, and displaying a web page or a part thereof, as well as attached files and other files linked to the web page, in cooperation with the RF circuit 108, the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, and the text input module 134.

[0095] The calendar module 148 includes executable instructions for creating, displaying, modifying, and storing a calendar and data associated with the calendar (e.g., calendar items, to-do lists, etc.) according to user instructions, in cooperation with the RF circuit 108, the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, the text input module 134, the email client module 140, and the browser module 147.

[0096] The widget module 149 cooperates with the RF circuit 108, the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, the text input module 134, and the browser module 147, and optionally, mini-applications (e.g., weather widget 149-1, stock price widget 149-2, calculator widget 149-3, alarm clock widget 149-4, and dictionary widget 149-5) that are downloaded and used by the user, or mini-applications created by the user (e.g., user-created widget 149-6). In some embodiments, the widget includes an HTML (Hypertext Markup Language) file, a CSS (Cascading Style Sheets) file, and a JavaScript file. In some embodiments, the widget includes an XML (Extensible Markup Language) file and a JavaScript file (e.g., Yahoo! widget).

[0097] The widget creator module 150 cooperates with the RF circuit 108, the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, the text input module 134, and the browser module 147, and is used by the user to optionally create a widget (e.g., turn a user-specified portion of a web page into a widget).

[0098] The search module 151 cooperates with the touch screen 112, the display controller 156, the contact / motion module 130, the graphic module 132, and the text input module 134, and includes executable instructions to search for text, music, sound, images, video, and / or other files in the memory 102 that match one or more search criteria (e.g., one or more user-specified search terms) according to the user's instructions.

[0099] The video and music player module 152, in cooperation with the touch screen 112, the display controller 156, the touch / motion module 130, the graphic module 132, the audio circuit 110, the speaker 111, the RF circuit 108, and the browser module 147, includes executable instructions that enable a user to download and play recorded music and other sound files stored in one or more file formats such as MP3 or AAC files, and executable instructions to display, present, or otherwise play videos (e.g., on the touch screen 112 or on an external display connected via the external port 124). In some embodiments, the device 100 optionally includes the functionality of an MP3 player such as an iPod (a trademark of Apple Inc.).

[0100] The memo module 153, in cooperation with the touch screen 112, the display controller 156, the touch / motion module 130, the graphic module 132, and the text input module 134, includes executable instructions to create and manage memos, to-do lists, etc. according to a user's instructions.

[0101] The map module 154, in cooperation with the RF circuit 108, the touch screen 112, the display controller 156, the touch / motion module 130, the graphic module 132, the text input module 134, the GPS module 135, and the browser module 147, is optionally used to receive, display, modify, and store maps and data associated with the maps (e.g., driving routes, data regarding stores and other target locations at or near a specific position, and other location-based data) according to a user's instructions.

[0102] The online video module 155, in cooperation with the touch screen 112, display controller 156, contact / motion module 130, graphic module 132, audio circuit 110, speaker 111, RF circuit 108, text input module 134, email client module 140, and browser module 147, includes instructions that enable a user to access a particular online video, browse a particular online video, receive it (e.g., by streaming and / or downloading), play it (e.g., on the touch screen or on an external display connected via the external port 124), send an email having a link to a particular online video, and perform other management of online videos in one or more file formats such as H.264. In some embodiments, instead of the email client module 140, the instant messaging module 141 is used to send a link to a particular online video. Additional explanation of the online video application can be found in U.S. Provisional Patent Application No. 60 / 936,562, filed Jun. 20, 2007, "Portable Multifunction Device, Method, and Graphical User Interface for Playing Online Videos," and U.S. Patent Application No. 11 / 968,067, filed Dec. 31, 2007, "Portable Multifunction Device, Method, and Graphical User Interface for Playing Online Videos," the entire contents of which are hereby incorporated by reference.

[0103] The modules and applications identified above each correspond to a set of executable instructions that perform one or more of the functions described above and the methods described in this application (e.g., the computer-executed methods and other information processing methods described herein). These modules (e.g., sets of instructions) need not be implemented as separate software programs (e.g., computer programs including instructions), procedures, or modules, and thus, in various embodiments, various subsets of these modules are optionally combined or otherwise reconfigured. For example, the video player module may optionally be combined with the music player module to form a single module (e.g., the video and music player module 152 of FIG. 1A). In some embodiments, the memory 102 optionally stores a subset of the modules and data structures identified above. Further, the memory 102 optionally stores additional modules and data structures not described above.

[0104] In some embodiments, the device 100 is a device in which the operation of a set of default functions in the device is performed only via the touch screen and / or touch pad. By using the touch screen and / or touch pad as the main input control device for the device 100 to operate, the number of physical input control devices (push buttons, dials, etc.) on the device 100 is optionally reduced.

[0105] The set of default functions that are executed only via a touch screen and / or a touch pad optionally includes navigation between user interfaces. In some embodiments, the touch pad, when touched by a user, navigates the device 100 from any user interface displayed on the device 100 to a main menu, a home menu, or a root menu. In such embodiments, the "menu button" is implemented using the touch pad. In some other embodiments, the menu button is a physical push button or other physical input control device rather than a touch pad.

[0106] FIG. 1B is a block diagram showing exemplary components for event processing according to some embodiments. In some embodiments, the memory 102 (FIG. 1A) or 370 (FIG. 3) includes an event sorting unit 170 (e.g., within the operating system 126) and an individual application 136-1 (e.g., any of the aforementioned applications 137-151, 155, 380-390).

[0107] The event sorting unit 170 receives event information and determines the application 136-1 to which the event information is to be delivered and the application view 191 of the application 136-1. The event sorting unit 170 includes an event monitor 171 and an event dispatcher module 174. In some embodiments, the application 136-1 includes an application internal state 192 indicating the current application view(s) displayed on the touch-sensitive display 112 when the application is active or running. In some embodiments, the device / global internal state 157 is used by the event sorting unit 170 to determine which application(s) is / are currently active, and the application internal state 192 is used by the event sorting unit 170 to determine the application view 191 to which the event information is to be delivered.

[0108] In some embodiments, the application internal state 192 includes additional information such as resume information to be used when application 136-1 resumes execution, user interface state information indicating or ready to display the information being displayed by application 136-1, a state queue that enables the user to return to the previous state or view of application 136-1, and a redo / undo queue of previous actions performed by the user.

[0109] The event monitor 171 receives event information from the peripheral device interface 118. The event information includes information regarding sub-events (e.g., a user touch as part of a multi-touch gesture on the touch-sensitive display 112). The peripheral device interface 118 transmits information received from the I / O subsystem 106, or sensors such as the proximity sensor 166, the accelerometer(s) 168, and / or the microphone 113 (via the audio circuit 110). The information that the peripheral device interface 118 receives from the I / O subsystem 106 includes information from the touch-sensitive display 112 or a touch-sensitive surface.

[0110] In some embodiments, the event monitor 171 transmits requests to the peripheral device interface 118 at predetermined intervals. In response, the peripheral device interface 118 transmits event information. In other embodiments, the peripheral device interface 118 transmits event information only when there is an important event (e.g., receipt of an input that exceeds a predetermined noise threshold and / or exceeds a predetermined duration).

[0111] In some embodiments, the event sorter 170 also includes a hit view determination module 172 and / or an active event recognition unit determination module 173.

[0112] The hit view determination module 172 provides software procedures for determining where in one or more views a sub-event occurs when the touch-sensitive display 112 is displaying two or more views. A view is composed of controls and other elements that a user can see on the display.

[0113] Another aspect of the user interface associated with an application is a set of views, sometimes referred to herein as application views or user interface windows, in which information is displayed and touch-based gestures occur. The application view (of an individual application) in which a touch is detected optionally corresponds to a program level within the program hierarchy or view hierarchy of the application. For example, the lowest level view in which a touch is detected is optionally referred to as the hit view, and the set of events recognized as appropriate input is optionally determined based at least in part on the hit view of the initial touch that initiates a touch-based gesture.

[0114] The hit view determination module 172 receives information related to sub-events of touch-based gestures. When an application has a plurality of hierarchically structured views, the hit view determination module 172 identifies the hit view as the lowest level view within the hierarchy in which the sub-event is to be processed. In most situations, the hit view is the lowest level view in which a start sub-event (e.g., the first sub-event in a series of sub-events that form an event or potential event) occurs. Once the hit view is identified by the hit view determination module 172, the hit view typically receives all sub-events related to the same touch or input source as the touch or input source identified as the hit view.

[0115] The active event recognition unit determination module 173 determines which view(s) within the view hierarchy should receive a particular series of sub-events. In some embodiments, the active event recognition unit determination module 173 determines that only the hit view should receive a particular series of sub-events. In other embodiments, the active event recognition unit determination module 173 determines that all views including the physical location of the sub-events are views that are actively involved, and thus, determines that all views that are actively involved should receive a particular series of sub-events. In other embodiments, even if a touch sub-event is completely limited to an area associated with one particular view, the upper-level views within the hierarchy remain views that are still actively involved.

[0116] The event dispatcher module 174 dispatches event information to the event recognition unit (e.g., event recognition unit 180). In embodiments including the active event recognition unit determination module 173, the event dispatcher module 174 distributes event information to the event recognition unit determined by the active event recognition unit determination module 173. In some embodiments, the event dispatcher module 174 stores the event information retrieved by the individual event receiver 182 in the event queue.

[0117] In some embodiments, the operating system 126 includes the event sorter 170. Alternatively, the application 136-1 includes the event sorter 170. In still other embodiments, the event sorter 170 is a stand-alone module or part of another module stored in the memory 102 such as the touch / motion module 130.

[0118] In some embodiments, application 136-1 includes a plurality of event processing units 190 and one or more application views 191, each including instructions for processing touch events that occur within an individual view of the application's user interface. Each application view 191 of application 136-1 includes one or more event recognition units 180. Typically, an individual application view 191 includes a plurality of event recognition units 180. In other embodiments, one or more of the event recognition units 180 are part of a separate module, such as a user interface kit or a higher-level object from which application 136-1 inherits methods and other properties. In some embodiments, an individual event processing unit 190 includes one or more of event data 179 received from data update unit 176, object update unit 177, GUI update unit 178, and / or event sorting unit 170. The event processing unit 190 optionally utilizes or invokes the data update unit 176, object update unit 177, or GUI update unit 178 to update the internal state 192 of the application. Alternatively, one or more of the application views 191 include one or more respective event processing units 190. Also, in some embodiments, one or more of the data update unit 176, object update unit 177, and GUI update unit 178 are included within an individual application view 191.

[0119] An individual event recognition unit 180 receives event information (e.g., event data 179) from the event sorting unit 170 and identifies an event from the event information. The event recognition unit 180 includes an event receiving unit 182 and an event comparing unit 184. In some embodiments, the event recognition unit 180 also includes at least a subset of metadata 183 and event distribution instructions 188 (optionally including sub-event distribution instructions).

[0120] The event receiving unit 182 receives event information from the event sorting unit 170. The event information includes sub-events, for example, information about a touch or a movement of a touch. Depending on the sub-event, the event information also includes additional information such as the position of the sub-event. When the sub-event is related to the movement of a touch, the event information also optionally includes the speed and direction of the sub-event. In some embodiments, the event includes a rotation of the device from one orientation to another (e.g., from portrait to landscape or vice versa), and the event information includes corresponding information about the current orientation of the device (also referred to as the posture of the device).

[0121] The event comparison unit 184 compares the event information with the definition of a predefined event or sub-event, and based on the comparison, determines an event or sub-event, or determines or updates the state of an event or sub-event. In some embodiments, the event comparison unit 184 includes an event definition 186. The event definition 186 includes definitions of events (e.g., a predefined series of sub-events) such as event 1 (187-1) and event 2 (187-2). In some embodiments, the sub-events within an event (187) include, for example, a touch start, a touch end, a touch movement, a touch cancellation, and multiple touches. In one example, the definition of event 1 (187-1) is a double-tap on a displayed object. The double-tap includes, for example, a first touch (touch start) on the displayed object for a predetermined stage, a first lift-off (touch end) for the predetermined stage, a second touch (touch start) on the displayed object for the predetermined stage, and a second lift-off (touch end) for the predetermined stage. In another example, the definition of event 2 (187-2) is a drag on a displayed object. The drag includes, for example, a touch (or contact) on the displayed object for a predetermined stage, a movement of the touch across the touch-sensitive display 112, and a lift-off of the touch (touch end). In some embodiments, the event also includes information about one or more associated event processing units 190.

[0122] In some embodiments, the event definition 187 includes the definition of events for individual user interface objects. In some embodiments, the event comparison unit 184 performs a hit test to determine which user interface object is associated with the sub - event. For example, within an application view where three user interface objects are displayed on the touch - sensitive display 112, when a touch is detected on the touch - sensitive display 112, the event comparison unit 184 performs a hit test to determine which of the three user interface objects is associated with the touch (sub - event). If each of the displayed objects is associated with an individual event processing unit 190, the event comparison unit determines which event processing unit 190 should be activated using the result of the hit test. For example, the event comparison unit 184 selects the event processing unit associated with the sub - event and object that triggered the hit test.

[0123] In some embodiments, the definition of an individual event (187) also includes a delay action that delays the delivery of event information until it is determined whether a series of sub - events corresponds to the event type of the event recognition unit.

[0124] If the individual event recognition unit 180 determines that a series of sub - events does not match any of the events of the event definition 186, the individual event recognition unit 180 enters a state of event impossible, event failure, or event end, and then ignores subsequent sub - events of the touch - based gesture. In this situation, if there are other event recognition units that remain active for the hit view, those event recognition units continue to track and process the sub - events of the ongoing touch - based gesture.

[0125] In some embodiments, the individual event recognition unit 180 includes metadata 183 having configurable properties, flags, and / or lists indicating how the event delivery system should actively participate in the event recognition unit that should perform sub-event delivery. In some embodiments, the metadata 183 includes configurable properties, flags, and / or lists indicating how the event recognition units interact with each other or how they can interact with each other. In some embodiments, the metadata 183 includes configurable properties, flags, and / or lists indicating whether sub-events are distributed at various levels in the view hierarchy or program hierarchy.

[0126] In some embodiments, the individual event recognition unit 180 activates the event processing unit 190 associated with the event when one or more specific sub-events of the event are recognized. In some embodiments, the individual event recognition unit 180 distributes the event information associated with the event to the event processing unit 190. Activating the event processing unit 190 is separate from sending (and deferring sending) sub-events to individual hit views. In some embodiments, the event recognition unit 180 sets a flag associated with the recognized event, and the event processing unit 190 associated with the flag catches the flag and executes a predefined process.

[0127] In some embodiments, the event delivery command 188 includes a sub-event delivery command that distributes event information about sub-events without activating the event processing unit. Instead, the sub-event delivery command distributes the event information to the event processing unit associated with a series of sub-events or to the view actively participating in the event. The event processing unit associated with a series of sub-events or the view actively participating in the event receives the event information and executes a predetermined process.

[0128] In some embodiments, the data update unit 176 creates and updates data used in the application 136-1. For example, the data update unit 176 updates the phone numbers used in the contact module 137 or stores video files used in the video player module. In some embodiments, the object update unit 177 creates and updates objects used in the application 136-1. For example, the object update unit 177 creates a new user interface object or updates the position of a user interface object. The GUI update unit 178 updates the GUI. For example, the GUI update unit 178 prepares display information and sends the display information to the graphic module 132 for display on the touch-sensitive display.

[0129] In some embodiments, the event processing unit(s) 190 includes or has access to the data update unit 176, the object update unit 177, and the GUI update unit 178. In some embodiments, the data update unit 176, the object update unit 177, and the GUI update unit 178 are included in a single module of the individual application 136-1 or the application view 191. In other embodiments, they are included in two or more software modules.

[0130] The foregoing description regarding event processing of a user's touch on the touch-sensitive display also applies to other forms of user input for operating the multifunctional device 100 using an input device, but it should be understood that not all of them are initiated on the touch screen. For example, the movement of a mouse and the pressing of a mouse button, the movement of a contact such as a tap, drag, or scroll on a touch pad, a pen stylus input, the movement of the device, a spoken command, a detected eye movement, a biometric input, and / or any combination thereof, optionally in association with a single or multiple presses or holds of a keyboard, are used as inputs corresponding to sub-events that define events to be optionally recognized.

[0131] FIG. 2 shows a portable multifunctional device 100 having a touch screen 112, according to some embodiments. The touch screen optionally displays one or more graphics within a user interface (UI) 200. In this embodiment, as well as in other embodiments described below, the user can select one or more of those graphics by performing gestures on the graphics using, for example, one or more fingers 202 (not drawn to scale in the figure) or one or more styli 203 (not drawn to scale in the figure). In some embodiments, the selection of one or more graphics is performed when the user interrupts contact with the one or more graphics. In some embodiments, the gesture optionally includes one or more taps, one or more swipes (from left to right, from right to left, upward and / or downward), and / or rolling (from right to left, from left to right, upward and / or downward) of a finger in contact with the device 100. In some implementations or situations, an unexpected contact with a graphic does not select the graphic. For example, if the gesture corresponding to the selection is a tap, a swipe gesture that sweeps over an application icon does not optionally select the corresponding application.

[0132] The device 100 also optionally includes one or more physical buttons, such as a "home" button or a menu button 204. As described above, the menu button 204 is optionally used to navigate to any application 136 within a set of applications optionally executed on the device 100. Alternatively, in some embodiments, the menu button is implemented as a soft key within a GUI displayed on the touch screen 112.

[0133] In some embodiments, device 100 includes a touch screen 112, a menu button 204, a push button 206 for turning the device on / off and locking the device, volume adjustment button(s) 208, a subscriber identity module (SIM) card slot 210, a headset jack 212, and a docking / charging external port 124. The push button 206 is optionally used to turn the device on / off by pressing the button and holding it pressed for a predetermined period, to lock the device by pressing the button and releasing it before a predetermined time has elapsed, and / or to unlock the device or initiate an unlock process. In an alternative embodiment, device 100 also accepts verbal input via a microphone 113 to activate or deactivate some functions. Device 100 optionally also includes one or more contact intensity sensors 165 for detecting the intensity of contact on the touch screen 112 and / or one or more haptic output generators 167 for generating haptic output to the user of device 100.

[0134] FIG. 3 is a block diagram of an exemplary multifunctional device having a display and a touch sensing surface, according to some embodiments. Device 300 need not be portable. In some embodiments, device 300 is a laptop computer, desktop computer, tablet computer, multimedia player device, navigation device, educational device (such as a child's learning toy), game system, or control device (e.g., a home or business controller). Device 300 typically includes one or more processing units (CPUs) 310, one or more networks or other communication interfaces 360, memory 370, and one or more communication buses 320 that interconnect these components. Communication bus 320 optionally includes circuitry (sometimes called a chipset) that interconnects and controls communication between system components. Device 300 includes an input / output (I / O) interface 330 that includes a display 340, which is typically a touch screen display. I / O interface 330 also optionally includes a keyboard and / or mouse (or other pointing device) 350, a touch pad 355, a haptic output generator 357 that generates haptic output on device 300 (e.g., similar to the haptic output generator(s) 167 described above with reference to FIG. 1A), and a sensor 359 (e.g., light, acceleration, proximity, touch sensing, and / or a contact intensity sensor similar to the contact intensity sensor(s) 165 described above with reference to FIG. 1A). Memory 370 includes high-speed random access memory such as DRAM, SRAM, DDR RAM, or other random access solid state memory devices, and optionally includes non-volatile memory such as one or more magnetic disk storage devices, optical disk storage devices, flash memory devices, or other non-volatile solid state storage devices. Memory 370 optionally includes one or more storage devices located remotely from the CPU(s) 310.In some embodiments, the memory 370 stores programs, modules, and data structures similar to, or a subset of, the programs, modules, and data structures stored in the memory 102 of the portable multifunctional device 100 (FIG. 1A). Further, the memory 370 optionally stores additional programs, modules, and data structures that are not present in the memory 102 of the portable multifunctional device 100. For example, the memory 370 of the device 300 optionally stores a drawing module 380, a presentation module 382, a word processing module 384, a website creation module 386, a disk authoring module 388, and / or a spreadsheet module 390, whereas the memory 102 of the portable multifunctional device 100 (FIG. 1A) optionally does not store these modules.

[0135] Each of the elements identified above in FIG. 3 is optionally stored in one or more of the memory devices described above. Each of the modules identified above corresponds to a set of instructions for performing the functions described above. The modules or computer programs identified above (e.g., sets of instructions or instructions) need not be implemented as separate software programs (e.g., computer programs including instructions), procedures, or modules, and thus in various embodiments, various subsets of these modules are optionally combined or otherwise reconfigured. In some embodiments, the memory 370 optionally stores a subset of the modules and data structures identified above. Further, the memory 370 optionally stores additional modules and data structures not described above.

[0136] Next, optionally direct attention to an embodiment of a user interface implemented, for example, on the portable multifunctional device 100.

[0137] Figure 4A shows an exemplary user interface of a menu of an application on a portable multifunctional device 100 according to some embodiments. A similar user interface is optionally implemented on device 300. In some embodiments, the user interface 400 includes the following elements, or a subset or superset thereof. ● Signal strength indicator(s) 402 for wireless communication(s) such as cellular signal and Wi-Fi signal, ● Time 404, ● Bluetooth indicator 405, ● Battery status indicator 406, ● A tray 408 having icons of frequently used applications such as ○ An icon 416 of the phone module 138 labeled "Phone", optionally including an indicator 414 of the number of missed calls or voicemail messages, ○ An icon 418 of the email client module 140 labeled "Mail", optionally including an indicator 410 of the number of unread emails, ○ An icon 420 of the browser module 147 labeled "Browser", and ○ An icon 422 of the video and music player module 152, also referred to as the iPod (trademark of Apple Inc.) module 152, labeled "iPod", and ● Icons of other applications such as ○ An icon 424 of the IM module 141 labeled "Message", ○ An icon 426 of the calendar module 148 labeled "Calendar", ○ An icon 428 of the image management module 144 labeled "Photos", ○ An icon 430 of the camera module 143 labeled "Camera", ○ An icon 432 of the online video module 155 labeled "Online Video", ○ The icon 434 of the stock price widget 149-2, labeled "Stock Price", ○ The icon 436 of the map module 154, labeled "Map", ○ The icon 438 of the weather widget 149-1, labeled "Weather", ○ The icon 440 of the alarm clock widget 149-4, labeled "Clock", ○ The icon 442 of the training support module 142, labeled "Training Support", ○ The icon 444 of the memo module 153, labeled "Memo", and ○ The icon 446 of the settings application or module, labeled "Settings", which provides access to the settings of the device 100 and its various applications 136.

[0138] Note that the icon labels shown in FIG. 4A are merely exemplary. For example, the icon 422 of the video and music player module 152 is labeled "Music" or "Music Player". Other labels may optionally be used for the various application icons. In some embodiments, the label for an individual application icon includes the name of the application corresponding to the individual application icon. In some embodiments, the label for a particular application icon is different from the name of the application corresponding to that particular application icon.

[0139] FIG. 4B shows an exemplary user interface on a device (e.g., device 300 of FIG. 3) having a touch sensing surface 451 (e.g., the tablet or touch pad 355 of FIG. 3) separate from the display 450 (e.g., touch screen display 112). The device 300 also optionally includes one or more contact intensity sensors (e.g., one or more of sensors 359) that detect the intensity of contact on the touch sensing surface 451, and / or one or more haptic output generators 357 that generate haptic output for a user of the device 300.

[0140] Some of the following examples are provided with reference to inputs on a touch screen display 112 (where a touch sensing surface and a display are combined), but in some embodiments, the device detects inputs on a touch sensing surface separate from the display, as shown in FIG. 4B. In some embodiments, the touch sensing surface (e.g., 451 in FIG. 4B) has a primary axis (e.g., 452 in FIG. 4B) corresponding to a primary axis (e.g., 453 in FIG. 4B) on the display (e.g., 450). According to these embodiments, the device detects contact with the touch sensing surface 451 (e.g., 460 and 462 in FIG. 4B) at positions corresponding to respective positions on the display (e.g., in FIG. 4B, 460 corresponds to 468 and 462 corresponds to 470). In this way, user inputs (e.g., contacts 460 and 462 and their movements) detected by the device on the touch sensing surface (e.g., 451 in FIG. 4B) are used by the device to operate the user interface on the display (e.g., 450 in FIG. 4B) of the multifunctional device when the touch sensing surface is separate from the display. It should be understood that a similar method is optionally used for other user interfaces described herein.

[0141] In addition, while the following examples are given primarily with reference to finger inputs (e.g., finger contact, finger tap gesture, finger swipe gesture), it should be understood that in some embodiments, one or more of the finger inputs may be replaced by inputs from another input device (e.g., mouse-based input or stylus input). For example, a swipe gesture may optionally be a mouse click (e.g., instead of a contact) followed by a mouse click with movement of the cursor along the path of the swipe (e.g., instead of movement of the contact). As another example, a tap gesture may optionally be a mouse click while the cursor is positioned over the location of the tap gesture (e.g., instead of detecting a contact and then ceasing to detect the contact). Similarly, it should be understood that when multiple user inputs are detected simultaneously, multiple computer mice may optionally be used simultaneously, or mouse and finger contacts may optionally be used simultaneously.

[0142] FIG. 5A shows an exemplary personal electronic device 500. The device 500 includes a body 502. In some embodiments, the device 500 can include some or all of the functions described with respect to devices 100 and 300 (e.g., FIGS. 1A - 4B). In some embodiments, the device 500 has a touch-sensitive display screen 504, hereinafter the touch screen 504. Alternatively, or in addition to the touch screen 504, the device 500 has a display and a touch-sensitive surface. Similar to devices 100 and 300, in some embodiments, the touch screen 504 (or touch-sensitive surface) optionally includes one or more intensity sensors that detect the intensity of an applied contact (e.g., a touch). One or more intensity sensors of the touch screen 504 (or touch-sensitive surface) can provide output data representative of the intensity of the touch. The user interface of the device 500 can respond to the touch(es) based on its intensity, which means that touches of different intensities can invoke different user interface operations on the device 500.

[0143] Exemplary techniques for detecting and processing touch intensity are described, for example, in International Patent Application No. PCT / US2013 / 040061, filed May 8, 2013, published as International Publication No. WO / 2013 / 169849, "Device, Method, and Graphical User Interface for Displaying User Interface Objects Corresponding to an Application", and International Patent Application No. PCT / US2013 / 069483, filed November 11, 2013, published as International Publication No. WO / 2014 / 105276, "Device, Method, and Graphical User Interface for Transitioning Between Touch Input to Display Output Relationships", each of which is hereby incorporated by reference in its entirety.

[0144] In some embodiments, device 500 has one or more input mechanisms 506 and 508. Input mechanisms 506 and 508, if included, can be physical. Examples of physical input mechanisms include push buttons and rotatable mechanisms. In some embodiments, device 500 has one or more attachment mechanisms. Such attachment mechanisms, if included, can enable device 500 to be attached, for example, to hats, glasses, earrings, necklaces, shirts, jackets, bracelets, watch bands, chains, pants, belts, shoes, wallets, backpacks, and the like. These attachment mechanisms enable the user to wear device 500.

[0145] FIG. 5B shows an exemplary personal electronic device 500. In some embodiments, device 500 can include some or all of the components described with respect to FIGS. 1A, 1B, and 3. Device 500 has a bus 512 that operably couples an I / O section 514 to one or more computer processors 516 and a memory 518. The I / O section 514 can be connected to a display 504, and the display 504 can have a touch sensing component 522 and optionally an intensity sensor 524 (e.g., a contact intensity sensor). Additionally, the I / O section 514 can be connected to a communication unit 530 that receives application and operating system data using Wi-Fi, Bluetooth, near field communication (NFC), cellular, and / or other wireless communication technologies. The device 500 can include an input mechanism 506 and / or 508. The input mechanism 506 can optionally be, for example, a rotatable input device or a depressible and rotatable input device. In some examples, the input mechanism 508 can optionally be a button.

[0146] In some examples, the input mechanism 508 can optionally be a microphone. The personal electronic device 500 can optionally include various sensors such as a GPS sensor 532, an accelerometer 534, a direction sensor 540 (e.g., a compass), a gyroscope 536, a motion sensor 538, and / or combinations thereof, all of which can be operably connected to the I / O section 514.

[0147] The memory 518 of the personal electronic device 500 can include one or more non-transitory computer-readable storage media for storing computer-executable instructions, which, when executed by one or more computer processors 516, can cause the computer processor to execute, for example, the techniques described below, including processes 700 to 1100 (FIGS. 7, 9, and 11). A computer-readable storage media can be any media that can tangibly contain or store computer-executable instructions used by or related to an instruction execution system, apparatus, or device. In some embodiments, the storage media is a transitory computer-readable storage media. In some embodiments, the storage media is a non-transitory computer-readable storage media. Non-transitory computer-readable storage media can include, but are not limited to, magnetic storage devices, optical storage devices, and / or semiconductor storage devices. Examples of such storage devices include magnetic disks, CDs, DVDs, or optical disks based on Blu-ray technology, and persistent solid-state memories such as flash, solid-state drives, etc. The personal electronic device 500 is not limited to the components and configurations of FIG. 5B and can include other or additional components in multiple configurations.

[0148] As used herein, the term "affordance" refers to user interaction graphical user interface objects that are optionally displayed on the display screens of devices 100, 300, and / or 500 (FIGS. 1A, 3, and 5A-5B). For example, images (e.g., icons), buttons, and text (e.g., hyperlinks) each optionally constitute an affordance.

[0149] As used herein, the term "focus selector" refers to an input element that indicates the current part of the user interface with which the user is interacting. In some implementations that include a cursor or other position marker, the cursor acts as the "focus selector," and thus while the cursor is positioned over a particular user interface element (e.g., a button, window, slider, or other user interface element), when an input (e.g., a press input) is detected on a touch-sensitive surface (e.g., the touchpad 355 of FIG. 3 or the touch-sensitive surface 451 of FIG. 4B), the particular user interface element is adjusted according to the detected input. In some implementations that include a touch screen display (e.g., the touch-sensitive display system 112 of FIG. 1A or the touch screen 112 of FIG. 4A) that enables direct interaction with user interface elements on the touch screen display, the detected contact on the touch screen acts as the "focus selector," and thus when an input (e.g., a press input by contact) is detected at the location of a particular user interface element (e.g., a button, window, slider, or other user interface element) on the touch screen display, the particular user interface element is adjusted according to the detected input. In some implementations, the focus is moved from one area of the user interface to another area of the user interface without moving the corresponding cursor or contact on the touch screen display (e.g., by using the tab key or arrow keys to move the focus from one button to another), and in these implementations, the focus selector moves in accordance with the movement of the focus between various areas of the user interface. Regardless of the specific form the focus selector takes, the focus selector is generally a user interface element (or a contact on a touch screen display) that is controlled by the user to communicate the user's intended interaction with the user interface (e.g., by indicating to the device the user interface element through which the user intends to interact).For example, the position of a focus selector (e.g., a cursor, contact, or selection box) over an individual button while a press input is detected on a touch sensing surface (e.g., a touch pad or touch screen) indicates that the user intends to activate that individual button (as opposed to other user interface elements shown on the device's display).

[0150] As used in this specification and the claims, the term "characteristic strength" of a contact refers to the characteristics of that contact based on one or more strengths of the contact. In some embodiments, the characteristic strength is based on a plurality of strength samples. The characteristic strength is optionally based on a set number of strength samples, or on a set of strength samples collected during a predetermined time (e.g., 0.05, 0.1, 0.2, 0.5, 1, 2, 5, 10 seconds) associated with a predetermined event (e.g., after detecting the contact, before detecting the lift-off of the contact, before or after detecting the start of movement of the contact, before detecting the end of the contact, before or after detecting an increase in the strength of the contact, and / or before or after detecting a decrease in the strength of the contact). The characteristic strength of a contact is optionally based on one or more of the maximum value of the strength of the contact, the mean value of the strength of the contact, the average value of the strength of the contact, the top 10 percentile value of the strength of the contact, the median value of the strength of the contact, the top 90 percent value of the strength of the contact, etc. In some embodiments, the duration of the contact is used when determining the characteristic strength (e.g., when the characteristic strength is the average of the strength of the contact over time). In some embodiments, the characteristic strength is compared to a set of one or more strength thresholds to determine whether an action has been performed by a user. For example, the set of one or more strength thresholds optionally includes a first strength threshold and a second strength threshold. In this example, a contact having a characteristic strength that does not exceed the first threshold results in a first action, a contact having a characteristic strength that exceeds the first strength threshold but does not exceed the second strength threshold results in a second action, and a contact having a characteristic strength that exceeds the second threshold results in a third action. In some embodiments, the comparison between the characteristic strength and one or more thresholds is not used to determine whether to perform a first action or a second action, but is used to determine whether to perform one or more actions (e.g., whether to perform an individual action or to refrain from performing an individual action).

[0151] As used herein, "installed application" refers to a software application that has been downloaded onto an electronic device (e.g., devices 100, 300, and / or 500) and is ready to be launched (e.g., opened) on the device. In some embodiments, the downloaded application becomes an installed application by an installation program that extracts the program portion from the downloaded package and integrates the extracted portion with the operating system of the computer system.

[0152] As used herein, the terms "open application" or "running application" refer to a software application that has retained state information (e.g., as part of device / global internal state 157 and / or application internal state 192). An open or running application is optionally any one of the following types of applications. ● An active application that is currently displayed on the display screen of the device being used by the application, ● A background application (or background process) for which one or more processes are being processed by one or more processors although not currently displayed, and ● An application in an interrupted or suspended state that is not running but is stored in memory (both volatile and non-volatile) and has state information that can be used to resume execution of the application.

[0153] As used herein, the term "closed application" refers to a software application that does not have retained state information (e.g., state information for a closed application is not stored in the device's memory). Thus, closing an application includes stopping and / or removing the application process for the application and removing the state information for the application from the device's memory. Generally, opening a second application within a first application does not close the first application. When the second application is being displayed and the display of the first application is aborted, the first application becomes a background application.

[0154] Next, attention is directed to embodiments of a user interface ("UI") and related processes implemented on an electronic device such as the portable multifunctional device 100, device 300, or device 500.

[0155] Figures 6A - 6AG show exemplary user interfaces for viewing and modifying content items including aggregated content items while continuing to play visual content, according to some embodiments. The user interfaces of these figures are used to illustrate processes described later, including the process of FIG. 7.

[0156] FIG. 6A shows an electronic device 600 that is a smartphone equipped with a touch-sensing display 602. In some embodiments, the electronic device 600 includes one or more features of devices 100, 300, and / or 500. The electronic device 600 shows a media library user interface 604. The media library user interface 604 includes a plurality of tiles representing a plurality of media items (e.g., photos and / or videos) that are part of a media library stored in and / or otherwise associated with the electronic device 600. The media library user interface 604 includes selectable options 606A-606D. Option 606A is selectable to present media items in groups based on the calendar year in which the media items were captured. Option 606B is selectable to present media items in groups based on the calendar month in which the media items were captured. Option 606C is selectable to present media items in groups based on the calendar day in which the media items were captured. The currently selected option 606D in FIG. 6A is selectable to present all media items in the media library (e.g., sorted based on the captured date).

[0157] The media library user interface 604 also includes selectable options 606I and 606J. Option 606I is selectable to switch the aspect ratio in which media items are presented within the media library user interface 604. In FIG. 6A, all media items are presented within the media library user interface 604 in a square aspect ratio. In some embodiments, when the user selects option 606I, the electronic device 600 displays all media items within the media library user interface 604 in their native aspect ratio. Option 606J is selectable to enable the user to select one or more media items presented within the media library user interface 604 such that one or more operations can be performed on the selected media item (e.g., share and / or delete the selected media item).

[0158] The media library user interface 604 also includes selectable options 606E - 606H. Option 606E is selectable to display the media library user interface 604. Option 606F is selectable to display a curated content user interface that presents the user with one or more media items selected and / or curated for the user based on selection criteria. Option 606G is selectable to display one or more collections of media items (e.g., one or more albums). The one or more collections include, in various embodiments, one or more user - defined collections and / or one or more automatically - generated collections. Option 606H is selectable to enable the user to search for media items within the media library (e.g., perform a keyword search of media items).

[0159] In FIG. 6A, the electronic device 600 detects a user input 608 corresponding to the selection of option 606F. In FIG. 6B, in response to detecting the user input 608, the electronic device 600 displays a curated content user interface 610. The curated content user interface 610 presents one or more media items selected and / or curated for the user based on selection criteria. In the illustrated embodiment, the curated content user interface 610 includes a plurality of tiles 612A, 612B representing a plurality of aggregated content items. In some embodiments, the aggregated content item is an automatically generated content item that includes a plurality of media items (e.g., a plurality of photos and / or videos) selected (e.g., by the electronic device 600) from the user's media library based on selection criteria. For example, the plurality of media items selected for the aggregated content item can include media items captured within a particular time frame and associated with a particular geographical location (e.g., Yosemite in October 2020 in FIG. 6B). In some embodiments, the aggregated content item is first automatically generated but can be modified and / or edited by the user, and the modified aggregated content item can be saved and stored (as described in more detail herein). In FIG. 6B, the aggregated content item tile 612A includes a favorite option 613A that can be selected to add the corresponding aggregated content item to a favorite folder, and an option 613B that can be selected to display additional options for the aggregated content item tile 612A. In some embodiments, the tiles 612A, 612B representing the aggregated content items are animated and / or display a preview of the associated aggregated content item (e.g., an animated preview, a moving preview, and / or a video preview) (e.g., playing a preview of a media item having a plurality of frames or panning and / or zooming a still media item).

[0160] The curated content user interface 610 also includes one or more highlighted media items, including the highlighted media item 612C. The highlighted media items are media items (e.g., photos and / or videos) from the user's media library that are selected (e.g., automatically selected) for presentation to the user based on one or more selection criteria. In some embodiments, the highlighted media items presented in the curated content user interface 610 change over time (e.g., change from one day to the next, or from one week to the next). In FIG. 6B, while displaying the curated content user interface 610, the electronic device 600 detects a user input 614 (e.g., a tap input or a long-press input) corresponding to the selection of the highlighted media item 612C.

[0161] In FIG. 6C, in response to detecting user input 614, electronic device 600 displays highlighted media item 612C along with a plurality of selectable options 616A - 616I. Option 616A is selectable to copy highlighted media item 612C. Option 616B is selectable to initiate a process of sharing highlighted media item 612C via one or more communication media (e.g., email, text message, NFC, Bluetooth, and / or upload to a content sharing platform). Option 616C is selectable to favorite highlighted media item 612C (e.g., add highlighted media item 612C to a favorite album). Option 616D is selectable to display highlighted media item 612C within media library user interface 604. Option 616E is selectable to initiate a process of tagging one or more people shown in highlighted media item 612C. Option 616F is selectable to decrease the frequency with which electronic device 600 selects a media item showing a person shown in highlighted media item 612C as a highlighted media item (e.g., within curated content user interface 610), and / or to decrease the frequency with which a media item showing a person shown in highlighted media item 612C is selected to be included in an aggregated content item. Option 616G is selectable to cause electronic device 600 to stop selecting a media item as a highlighted media item if the media item shows a person shown in highlighted media item 612C (e.g., stop selecting the media item for inclusion in curated content user interface 610, and / or stop selecting the media item for inclusion in an aggregated content item). Option 616H is selectable to delete highlighted media item 612C from the media library.Option 616I can be selected to not select the highlighted media item 612C as the highlighted media item (e.g., remove the highlighted media item from the curated content user interface 610). In FIG. 6C, the electronic device 600 detects a user input 618 (e.g., a tap input and / or a non-tap input).

[0162] In FIG. 6D, in response to the user input 618, the electronic device 600 redisplay the curated content user interface 610. In FIG. 6D, while displaying the curated content user interface 610, the electronic device 600 detects a user input 620 (e.g., a tap input) corresponding to the selection of a tile 612A representing a first aggregated content item (e.g., an aggregated content item for Yosemite in October 2020). In the illustrated embodiment, the first aggregated content item includes a set of media items selected from the user's media library corresponding to a particular geographical location (e.g., Yosemite) and a particular period (e.g., October 2020).

[0163] In FIG. 6E, in response to detecting a user input 620, the electronic device 600 displays an aggregated content user interface 622 corresponding to (e.g., uniquely corresponding to) a first aggregated content item. The aggregated content user interface 622 includes an option 624A that is selectable to return to the curated content user interface 610. The aggregated content user interface 622 also includes a tile 624B that represents the first aggregated content item and is selectable to initiate playback of the visual and / or audio content of the first aggregated content item. The aggregated content user interface 622 also includes a plurality of tiles 624C, 624D that represent media items included in the first aggregated content item. The tiles 624C, 624D are selectable to view the corresponding individual media items (e.g., without playing the complete visual content of the first aggregated content item), and the media items included in the first aggregated content item are represented by their respective tiles within the aggregated content user interface 622. In this way, the aggregated content user interface 622 enables the user to play the first aggregated content item (e.g., via the tile 624B) and also enables the user to view the constituent media items that make up the first aggregated content item. In FIG. 6E, while the aggregated content user interface 622 is being displayed, the electronic device 600 detects a user input 626 (e.g., a tap input) corresponding to the selection of the tile 624B.

[0164] In FIG. 6F, in response to detecting user input 626, electronic device 600 displays playback user interface 625 and begins playback of a first aggregated content item. In the illustrated embodiment, beginning playback of the first aggregated content item includes displaying media item 628A, which is a first media item within the first aggregated content item, as well as title information 627 corresponding to the first aggregated content item. In the illustrated embodiment, beginning playback of the first aggregated content item also includes playing audio content (e.g., an audio track and / or one or more audio tracks). In FIG. 6F, electronic device 600 begins playback of audio track 1.

[0165] In FIG. 6G, electronic device 600 continues to play the first aggregated content item. Title information 627 has moved from the first position in FIG. 6F to the second position in FIG. 6G, and one or more other visual characteristics have changed (e.g., size, font, and color have changed). Further, playback of the first aggregated content item from FIG. 6F to FIG. 6G includes zooming in on media item 628A. While displaying playback user interface 625 including media item 628A and playing audio track 1, electronic device 600 detects user input 630 (e.g., a tap input and / or a non-tap input).

[0166] In FIG. 6H, in response to detecting user input 630, while the electronic device 600 continues to play the visual and audio content of the first aggregated content item, it displays a plurality of playback controls and options. The close option 632A is selectable to terminate the display of the playback user interface 625 and to stop the playback of the visual and / or audio content of the first aggregated content item (e.g., selectable to redisplay the aggregated content user interface 622). The share option 632B is selectable to initiate a process of sharing the first aggregated content item via one or more communication media. The menu option 632C is selectable to display one or more options, as described below. The recipe option 632D is selectable to display a recipe user interface in which the user can modify one or more visual and / or audio characteristics of the first aggregated content item, as described in more detail below. The pause option 632E is selectable to pause the playback of the first aggregated content item (e.g., pause the visual and / or audio playback). The grid option 632F is selectable to display a content grid user interface, as described in more detail below with reference to FIGS. 10A - 10S. In FIG. 6H, the electronic device 600 detects user input 634 corresponding to the selection of option 632C.

[0167] In FIG. 6I, in response to detecting a user input 634, while the electronic device 600 maintains the playback of a first aggregated content item (e.g., maintains audio and visual playback), it displays a plurality of options 636A - 636H. Option 636A is selectable to add the first aggregated content item to the user's favorite media item (e.g., add the first aggregated content item to a favorite album). Option 636B is selectable to initiate a process of changing the title information of the first aggregated content item (e.g., to enable the user to input a new title for the first aggregated content item). Option 636C is selectable to delete the first aggregated content item. Option 636D is selectable to cause the electronic device 600 to modify its selection criteria for generating future aggregated content items such that fewer aggregated content items similar to the first aggregated content item are generated.

[0168] Options 636E through 636H correspond to different duration options for the first aggregated content item and are selectable to modify and / or specify the duration of the first aggregated content item. For example, the first aggregated content item currently has a duration corresponding to Option 636F (e.g., an intermediate duration), and the specified duration is the duration of 38 media items. Option 636E is selectable to shorten the duration of the first aggregated content item by reducing the number of media items within the first aggregated content item (e.g., from 38 media items to 24 media items). Option 636G is selectable to increase the duration of the first aggregated content item by increasing the number of media items within the first aggregated content item. In the illustrated embodiment, Option 636G corresponds to a specific duration (e.g., 1 minute 28 seconds), which corresponds to the maximum duration allowed for sharing the first aggregated content item. Option 636H is selectable to increase the duration of the first aggregated content item to match the duration of the audio track applied to the first aggregated content item. In FIG. 6I, audio track 1 is applied to the first aggregated content item and has a duration of 3 minutes 15 seconds. Thus, selection of Option 636H in FIG. 6I causes the first aggregated content item to be modified (e.g., by adding and / or removing one or more media items and / or modifying the display duration of media items within the first aggregated content item) to have a total duration of (e.g., approximately) 3 minutes 15 seconds. However, since this duration is longer than 1 minute 28 seconds, selection of Option 636H prohibits the first aggregated content item from being shared with other users and / or devices.

[0169] In FIG. 6I, the electronic device 600 detects a user input 638 (e.g., a tap input) corresponding to the selection of option 632C. In FIG. 6J, in response to detecting the user input 638, while maintaining the playback (e.g., audio and visual playback) of the first aggregated content item, the electronic device 600 stops displaying options 636A - 636H. In FIG. 6J, while playing back the first aggregated content item, the electronic device 600 detects a user input 640 (e.g., a tap input) corresponding to the selection of recipe option 632D.

[0170] In FIG. 6K, in response to detecting a user input 640, while the electronic device 600 maintains the playback of a first aggregated content item (e.g., audio and visual playback), it displays a recipe user interface 642. FIG. 6K shows an electronic device 600 that displays the recipe user interface 642 while being oriented in both the vertical (left) and horizontal (right) directions. In the recipe user interface 642, while the playback of the first aggregated content item (e.g., visual and audio playback) is maintained, it enables the user to apply different combinations of visual and audio characteristics to the first aggregated content item. For example, in the illustrated embodiment, each "recipe" includes a combination of a visual filter and an audio track, and the recipe user interface 642 enables the user to switch between these different combinations of the visual filter and the audio track while the playback of the first aggregated content item is maintained. In FIG. 6K, there are six different "recipes" or combinations of visual filters and audio tracks that the user can apply to the first aggregated content item. In FIG. 6K, the first aggregated content item is shown with a first visual filter 646B and a first audio track (audio track 1) applied. The first visual filter 646B and the first audio track define a first default combination (e.g., the first "recipe"). In FIG. 6K, the right side of the first aggregated content item is shown with a second visual filter 646C (which is not the first or third visual filter) applied to show that the user can provide a user input (e.g., a right tap input and / or a left swipe input) to view the first aggregated content item with the second visual filter 646C applied. Similarly, the left side of the first aggregated content item is shown with a third visual filter 646A (which is not the first or second visual filter) applied to show that the user can provide a user input (e.g., a left tap input and / or a right swipe input) to view the first aggregated content item with the third visual filter 646A applied.

[0171] In FIG. 6K, the recipe user interface 642 includes a recipe indication 644A indicating that the first visual filter / audio track combination out of six visual filter / audio track combinations is currently applied to the first aggregated content item. The recipe user interface 642 also includes an audio track indication 644B indicating that audio track 1 (by artist 1) is currently applied to the first aggregated content item. The recipe user interface 642 further includes an audio track selection option 644C that is selectable to display an audio track selection user interface, and a visual filter option 644D that is selectable to display a visual filter selection user interface. In FIG. 6K, the electronic device 600 detects a user input 648 that is a left swipe gesture.

[0172] In FIG. 6L, in response to detecting the user input 648, the electronic device 600 shifts the visual filters 646A, 646B, 646C to the left based on the user input 648 (e.g., at a speed corresponding to the speed of the user input and the translation distance, and translating over the translation distance). Thus, in FIG. 6L, the visual filter 646A is no longer visible, the visual filter 646B is shown applied to the left side of the media item 628A, and the visual filter 646C is shown applied to the right side of the media item 628A. During this user input, the playback of the first aggregated content item is maintained by the electronic device 600 (e.g., the electronic device 600 continues to play the visual content of the first aggregated content item and continues to play audio track 1). In FIG. 6L, the electronic device 600 continues to detect the left swipe gesture of the user input 648.

[0173] In FIG. 6M, in response to the continuation of user input 648, the electronic device 600 continues to shift visual filters 646B and 646C such that visual filter 646B now occupies a small portion on the left side of media item 628A and visual filter 646C is applied to a majority of media item 628A. In FIG. 6M, in response to user input 648 exceeding a threshold translation distance, the electronic device 600 updates recipe indication 644A to indicate that a second recipe (e.g., a combination of a second visual filter / audio track) has been applied to the first aggregated content item. Further, in response to user input 648 exceeding a threshold translation distance, the electronic device 600 stops the playback of audio track 1 (which was part of the first recipe (e.g., a combination of a first default visual filter / audio track)) and starts the playback of audio track 2 (which is part of the second recipe (e.g., a combination of a second default visual filter / audio track)). In some embodiments, when switching between different combinations of visual filters / audio tracks, instead of starting the playback of audio track 2 from the beginning of audio track 2, the electronic device 600 starts the playback from the playback position corresponding to the playback progress of the first aggregated content item. For example, in FIG. 6M, if the first aggregated content item has been played for 40 seconds, the electronic device 600 can start the playback of audio track 2 from the 40 - second mark. In FIG. 6M, the electronic device 600 continues to detect a left - swipe gesture of user input 648.

[0174] In FIG. 6N, in response to the continuation of the user input 648, the electronic device 600 continues to shift the visual filters 646B and 646C. In FIG. 6N, the visual filter 646B is applied to the left end region of the visual content of the first aggregated content item (e.g., the media item 628A currently being displayed), the visual filter 646C is applied to the central region of the visual content of the first aggregated content item, and the fourth visual filter 646D (corresponding to the combination of the third visual filter / audio track) is applied to the right end region of the visual content of the first aggregated content item. The electronic device 600 continues to maintain the playback (e.g., visual and / or audio playback) of the first aggregated content item.

[0175] In FIG. 6O, due to the continued playback of the first aggregated content item, the electronic device 600 no longer displays the media item 628A and, while continuing to play the audio track 2, now displays the second media item 628B of the first aggregated content item. In FIG. 6O, the electronic device 600 is shown as detecting two different user inputs 650A, 650B (at separate times and / or non-simultaneously). The user input 650B, which is a left swipe gesture, causes the electronic device 600 to apply a third recipe (e.g., the third default combination of the visual filter 646D and the third audio track) to the first aggregated content item. The user input 650A, which is a right swipe gesture, causes the electronic device to reapply the first recipe of the first audio track and the visual filter 646B.

[0176] In FIG. 6P, in response to detecting the user input 650A, the electronic device 600 shifts the visual filters 646D, 646C, and 646B to the right based on the user input 650A (e.g., at a speed corresponding to the speed of the user input and the translation distance and translating over the translation distance). In FIG. 6P, the electronic device 600 continues to detect the right swipe gesture user input 650A.

[0177] In FIG. 6Q, in response to the continuation of user input 650A, electronic device 600 continues to shift visual filters 646B and 646C. In FIG. 6Q, based on the determination that user input 650A has exceeded a threshold translation distance, electronic device 600 updates recipe indication 644A to indicate that the first recipe has been applied to the first aggregated content item, stops the playback of audio track 2, and resumes the playback of audio track 1. As described above, in some embodiments, audio track 1 is not played back from the beginning of audio track 1 (e.g., is played back from a playback position corresponding to the playback position of the first aggregated content item).

[0178] As described above and as shown in the figures, when the user swipes between different recipes within the recipe user interface 642, the user can switch between combinations of visual filters and audio tracks applied to the first aggregated content item. In some embodiments, in addition to changing the visual filter and audio track applied to the first aggregated content item, when the user swipes between different recipes (e.g., different combinations of visual filters and audio tracks), the electronic device 600 also changes other audio and / or visual characteristics of the playback of the first aggregated content item, such as the type of visual transitions applied between media items presented during the playback of the first aggregated content item. For example, a first recipe (e.g., a first combination of visual filter / audio track) can utilize a first set of visual transitions (e.g., fade in, fade out), and a second recipe can utilize a second set of visual transitions different from the first set (e.g., swipe in, swipe out). In some embodiments, the visual transitions applied between media items are selected based on the audio characteristics of the audio track that is part of the applied combination of visual filter / audio track. For example, a higher energy or faster audio track (e.g., an audio track having a beats per minute value above a threshold) can utilize a first set of visual transitions, and a lower energy or slower audio track (e.g., an audio track having a beats per minute value below a threshold) can utilize a second set of visual transitions.

[0179] In FIG. 6Q, the electronic device 600 detects a user input 652 (e.g., a tap input) corresponding to the selection of the audio track selection option 644C. In FIG. 6R, in response to detecting the user input 652, the electronic device 600 displays an audio track selection user interface 654. In some embodiments, while the audio track selection user interface 654 is being displayed, the electronic device 600 maintains the playback (e.g., visual and / or audio playback) of the first aggregated content item (e.g., in the background). In some embodiments, while the audio track selection user interface 654 is being displayed, the electronic device 600 pauses the playback (e.g., visual and / or audio playback) of the first aggregated content item. The audio track selection user interface 654 includes a cancel option 658A that is selectable to return to the recipe user interface 642 (e.g., without changing the audio track applied to the first aggregated content item), an end option 658B that is selectable to apply the selected audio track to the first aggregated content item, and a search option 658C that is selectable to search for audio tracks within a music catalog. The audio track selection user interface 654 also includes a plurality of selectable options 656A - 656N corresponding to different audio tracks. The user can select individual options for applying the individual corresponding audio tracks to the first aggregated content item (e.g., playing the selected audio track while the visual content of the first aggregated content item is being played). In some embodiments, the selection of options 656A - 656N replaces the audio track within the recipe (e.g., a combination of visual filter / audio track) currently applied to the first aggregated content item (e.g., the recipe applied to the first aggregated content item when the user input 652 was detected) with the selected audio track.For example, in FIG. 6R, the first recipe is currently a combination of audio track 1 and visual filter 646B, but selection of a different audio track in the audio track selection user interface 654 modifies the first recipe such that it becomes a combination of the selected audio track and visual filter 646B (e.g., without audio track 1). In FIG. 6R, the audio track selection user interface indicates that track 1 by artist 1 is currently applied to the first aggregated content item. In FIG. 6R, the electronic device 600 detects a user input 660 corresponding to the selection of option 656D corresponding to audio track 3 by artist 3.

[0180] FIG. 6S shows a first exemplary scenario in which an electronic device 600 is not permitted to apply a selected audio track to a first aggregated content item. For example, a user of the electronic device 600 is not subscribed to a music subscription service and does not have access rights to the selected audio track. In FIG. 6S, in response to detecting a user input 660 and in accordance with a determination that the electronic device 600 is not permitted to apply the selected audio track to the first aggregated content item (e.g., in accordance with a determination that the user is not subscribed to a music subscription service), the electronic device 600 displays a music preview user interface 662. The music preview user interface 662 provides the user with a preview of the selected audio track applied to the first aggregated content item and displays a visual playback of the first aggregated content item while the selected audio track is being played. The music preview user interface 662 also enables the user to swipe between different visual filter options (e.g., 646A, 646B, 646C) while the visual content of the first aggregated content item is being played and the selected audio track is being played. However, the user can only view the preview within the preview user interface 662, and the user cannot save and / or share the first aggregated content item with the selected audio track applied (e.g., no options for the user to save or share the first aggregated content item are provided). The user can select either a cancel option 664A to cancel the selection of the audio track and return to the audio track selection user interface 654, or a free trial option 664B that can be selected to initiate a process of registering the user for a free trial of the music subscription service so that the user can apply the selected audio track to the first aggregated content item.

[0181] Figure 6T shows a second exemplary scenario in which the electronic device 600 is permitted to apply a selected audio track to a first aggregated content item. In Figure 6T, the audio track selection user interface 654 indicates that track 3 is selected, and the electronic device 600 plays audio track 3. In some embodiments, by swiping through different recipes within the recipe user interface 642, different audio tracks are played from different playback positions based on the current playback position of the first aggregated content item (e.g., different audio tracks are played from a playback position that is not the beginning of the audio track), but the selection of the audio track within the audio track selection user interface 654 causes the selected audio track to be played from the beginning. In Figure 6T, the electronic device 600 detects a user input 666 corresponding to the selection of the end option 658B.

[0182] In Figure 6U, in response to detecting the user input 666, the electronic device 600 redisplay the recipe user interface 642. Similar to the case of Figure 6Q, the recipe user interface 642 displays the playback of the first aggregated content item with the first recipe applied thereto (e.g., the recipe indicator 644A indicates "1 out of 6 recipes", and the visual filter 646B is applied to the first aggregated content item). However, due to the user's selection of audio track 3 in the audio track selection user interface 654, audio track 1 has been replaced by audio track 3 in the first recipe (e.g., in the first visual filter / audio track combination), such that while the visual content of the first aggregated content item is being played with the visual filter 646B applied, audio track 3 is applied to the first aggregated content item. In Figure 6U, the electronic device 600 detects a user input 668 (e.g., a tap input) corresponding to the selection of the visual filter selection option 644D.

[0183] In FIG. 6V, in response to detecting a user input 668, the electronic device 600 displays a visual filter selection user interface 670. The visual filter selection user interface 670 includes a plurality of tiles 674A-674, and different tiles correspond to different visual filters. Further, different tiles each display continuous playback of the visual content of the first aggregated content item with the respective visual filter applied to the visual content of the first aggregated content item. The visual filter selection user interface 670 also includes a cancel option 672A that is selectable to return to the recipe user interface 642 (e.g., without applying different visual filters), and an end option 672B that is selectable to apply the selected visual filter to the first aggregated content item. In some embodiments, the selection and / or application of different visual filters within the visual filter selection user interface 670 causes the visual filter within the currently applied recipe to be replaced with the selected visual filter.

[0184] As described above, while the visual filter selection user interface 670 is being displayed, the electronic device 600 maintains playback (e.g., audio and visual playback) of the first aggregated content item, and different ones of the tiles 674A-674O show playback of the visual content of the first aggregated content item with different visual filters applied. In FIG. 6W, the visual content of the first aggregated content item continues to be played such that the visual content of the first aggregated content item transitions from displaying the media item 628B of FIG. 6V to displaying the media item 628C of FIG. 6W. In FIG. 6W, the electronic device 600 detects a user input 676 (e.g., a tap input) corresponding to the selection of the end option 674B.

[0185] In Figure 6X, in response to detecting user input 676, while the electronic device 600 maintains the playback (e.g., audio and / or visual playback) of the first aggregated content item, it pauses the display of the visual filter selection user interface 672 and redisplays the recipe user interface 642. As described above, while the visual filter selection user interface 672 is being displayed, the visual content playback of the first aggregated content item transitions from media item 628B to media item 628C. Thus, the electronic device 600 now displays media item 628C within the recipe user interface 642. In Figure 6X, the electronic device 600 detects a user input 678 (e.g., a tap input) corresponding to the selection of the first recipe (e.g., a combination of the first visual filter / audio track) currently applied to the first aggregated content item in Figure 6X.

[0186] In Figure 6Y, in response to detecting user input 678, the electronic device 600 pauses the display of the recipe user interface 642 and displays the continued playback of the first aggregated content item within the playback user interface 625. Further, in Figure 6Y, the electronic device 600 updates the title information 627 to present title information corresponding to the currently presented media item 628C (e.g., changing the title information from "Yellowstone October 2020" to "Half Dome October 2020"). In Figure 6Y, the electronic device 600 detects a user input 680 (e.g., a tap input) corresponding to the selection of the pause option 632E.

[0187] In FIG. 6Z, in response to detecting user input 680, electronic device 600 pauses the playback of the first aggregated content item (e.g., pauses audio and / or visual playback). In response to detecting user input 680, electronic device 600 also displays a navigation object 682 (e.g., a scrubber). Navigation object 682 includes representations of different media items within the aggregated content item arranged in the order presented within the aggregated content item so that the user can navigate through the various media items while the playback of the aggregated content item is paused. In FIG. 6Z, navigation object 682 indicates that electronic device 600 is currently displaying the third media item within the first aggregated content item. Further, in response to user input 680, electronic device 600 replaces pause option 632E with play option 632H and replaces recipe option 632D with aspect ratio option 632G. Play option 632H is selectable to resume playback (e.g., resume audio and / or visual playback of the first aggregated content item). Aspect ratio option 632G is selectable to switch the display of the displayed media item between a full screen aspect ratio and a native aspect ratio. In FIG. 6Z, electronic device 600 detects a user input 683 (e.g., a tap input) corresponding to the selection of aspect ratio option 632G.

[0188] In FIG. 6AA, in response to detecting user input 683, electronic device 600 stops displaying media item 628C in full screen aspect ratio and then displays media item 628C in native aspect ratio. In FIG. 6AA, the playback of the first aggregated content item remains paused. In FIG. 6AA, electronic device 600 detects a user input 684 (e.g., a tap input) corresponding to the selection of play option 632H.

[0189] In FIG. 6AB, in response to detecting user input 684, electronic device 600 returns the display of media item 628C to the full-screen aspect ratio and resumes playing the first aggregated content item. Further, in response to user input 684, the electronic device replaces aspect ratio option 632G with recipe option 632D and replaces play option 632H with pause option 632E. In FIG. 6AB, while playing the visual content of the aggregated content item and playing audio track 3, electronic device 600 detects a tap-and-hold input 686 that is a continuous tap input (e.g., a tap-and-hold input over a certain duration). In FIGS. 6AB and 6AC, in response to detecting tap-and-hold input 686 (and while continuing to detect tap-and-hold input 686), electronic device 600 maintains the display of media item 628C while continuing to play audio track 3 (e.g., maintains the continuous display and / or pauses visual playback while maintaining the display). In the illustrated scenario, without any user input, the electronic device is moving to display subsequent media items as part of playing the first aggregated content item, but tap-and-hold input 686 causes electronic device 600 to maintain the display of media item 628C while continuing to play audio track 3 (e.g., as long as tap-and-hold input 686 is detected).

[0190] In FIG. 6AC, after the end of the tap-and-hold input 686 (e.g., detecting the lift-off of the input from the touch-sensing surface of the display), the electronic device detects a user input 688 which is a tap input on the left side of the display 602. In the illustrated embodiment, a tap input on the left side of the display 602 (e.g., a predefined area proximate to the left edge of the display 602) causes navigation to the previous media item within the first aggregated content item (e.g., while the playback of the first aggregated content item is continuing and / or while maintaining the playback of the audio track), and a tap input on the right side of the display 602 (e.g., a predefined area proximate to the right edge of the display 602) causes navigation to the subsequent media item within the first aggregated content item (e.g., while maintaining the playback of the first aggregated content item and / or while maintaining the playback of the audio track). In some embodiments, the user input 688 is a swipe input (e.g., a right swipe) instead of a tap input.

[0191] In FIG. 6AD, in response to detecting the user input 688, the electronic device 600 stops displaying the media item 628C and displays the previous media item 628B. In FIG. 6AD, while displaying the media item 628B and while maintaining the playback of the first aggregated content item and the playback of audio track 3, the electronic device 600 detects a user input 690 which is a tap input on the right side of the display 602. In some embodiments, the input 690 is a swipe input (e.g., a left swipe) instead of a tap input.

[0192] In FIG. 6AE, in response to detecting user input 690, electronic device 600 stops displaying media item 628B and displays subsequent media item 628C. The user inputs shown in FIGS. 6AB, 6AC, and 6AD interrupted the normal playback of the first aggregated content item. While the first aggregated content item continues to play through each of these figures (and audio track 3 continues to play through each of these figures), user inputs 686, 688, 690 caused the playback of the visual content of the first aggregated content item to be changed in some way (e.g., maintaining the display for the current media item longer than normal, navigating to a previous or subsequent media item). In some embodiments, in response to these user inputs, electronic device 600 accelerates or decelerates the playback of the first aggregated content item to account for the changes to the playback of the visual content of the first aggregated content item caused by the user input. For example, in response to user input 686 (which caused media item 628C to be displayed longer than normal), electronic device 600 accelerates the playback of subsequent media items so that the playback of the first aggregated content item maintains the target playback duration. Similarly, in response to user input 688 (which navigates in the reverse direction), electronic device 600 accelerates the playback of subsequent media items, and in response to user input 690 (which navigates in the forward direction), electronic device 600 decelerates the playback of subsequent media items to maintain the target playback duration of the first aggregated content item.

[0193] In FIG. 6AF, the playback of the first aggregated content item continues from FIG. 6AE. In FIG. 6AF, the playback of the visual content of the first aggregated content item includes displaying three media items 628D, 628E, 628F in a default arrangement. In some embodiments, the media items 628D, 628E, 628F are selected to be presented together in the default arrangement based on the similarity of the content shown in the media items. Further, in FIG. 6AF, based on the determination that no user input has been received over a threshold duration, the electronic device 600 stops displaying options 632A - 632F.

[0194] In FIG. 6AG, the playback of the first aggregated content item continues from FIG. 6AF, and while the electronic device 600 continues to play audio track 3, it replaces the display of media items 628D, 628E, and 628F with the display of media item 628G.

[0195] FIG. 7 is a flowchart showing a method of browsing and editing content items using a computer system according to some embodiments. Method 700 is executed in a computer system (e.g., 100, 300, 500) (e.g., a smartphone, smartwatch, tablet, digital media player, computer set - top entertainment box, smart TV, and / or a computer system that controls an external display) that communicates with a display generation component (e.g., a display controller, a touch - sensitive display system, and / or a display (e.g., integrated and / or connected)) and one or more input devices (e.g., a touch - sensitive surface (e.g., a touch - sensitive display), a mouse, a keyboard, and / or a remote control). Some operations of method 700 are optionally combined, the order of some operations is optionally changed, and some operations are optionally omitted.

[0196] As will be described below, method 700 provides an intuitive way to view and manage content items. This method reduces the cognitive burden on the user when viewing and managing content items, thereby creating a more efficient human-machine interface. In the case of a battery-operated computing device, power is conserved and the battery charging interval is lengthened by enabling the user to view and edit content items more quickly and efficiently.

[0197] The computer system plays (702) (e.g., displays via a display generation component) the visual content of a first aggregated content item (e.g., media item 628A of FIG. 6K) via a display generation component (e.g., plays the visual content of a first aggregated content item via a display generation component) (e.g., video, and / or a content item automatically generated from a plurality of content items) (in some embodiments, the computer system plays the visual and audio content of a first aggregated content item), and the first aggregated content item includes an ordered sequence of a first plurality of content items selected (e.g., automatically and / or without user input) from a set of content items based on a first set of selection criteria (e.g., the first aggregated content item represents an ordered sequence of a plurality of photos and / or videos and / or an automatically generated collection of photos and / or videos (e.g., a collection of photos and / or videos automatically aggregated and / or selected from a set of content items based on one or more shared characteristics)). In some embodiments, the plurality of photos and / or videos that make up the first plurality of content items are selected from a set of photos and / or videos associated with the computer system (e.g., stored in the computer system, associated with the user of the computer system, and / or associated with a user account associated with the computer system (e.g., signed in)).

[0198] While playing the visual content of the first aggregated content item (e.g., media item 628A in FIG. 6K) (704), the computer system plays audio content separate from the content item (706) (e.g., plays audio track 1 in FIG. 6K) (e.g., outputs and / or causes the output of a separate song or media asset while the visual content of the first aggregated content item is being displayed via a display generation component (e.g., via one or more speakers, one or more headphones, and / or one or more earphones). In some embodiments, the computer system also plays audio content corresponding to and / or being part of the first aggregated content item (e.g., audio from one or more videos incorporated into the aggregated content item) (e.g., an audio track superimposed on the first aggregated content item and / or played while the visual content of the first aggregated content item is being played and / or displayed).

[0199] While playing the visual and audio content of the first aggregated content item (708), the computer system detects user input (e.g., 648) (e.g., gestures (e.g., tap gesture, swipe gesture) and / or voice input (e.g., via a touch-sensitive display and / or touch-sensitive surface)) via one or more input devices (710).

[0200] In response to detecting a user input (712), while the computer system continues to play the visual content of the first aggregated content item, it modifies the audio content being played (e.g., non-volume audio parameters (e.g., audio parameters different from volume)) (714) (e.g., changes the audio content from a first audio track to a second audio track different from the first audio track (e.g., from a first music track to a second music track different from the first music track)) (e.g., FIGS. 6K - 6N, while continuing to play the visual content of the first aggregated content item (e.g., displaying media item 628A), changes audio track 1 to audio track 2) (e.g., without stopping, pausing, and / or otherwise interrupting the playback of the visual content of the first aggregated content item). By modifying the audio content in response to detecting a user input while continuing to play the visual content of the first aggregated content item, it enables the user to quickly modify the audio content applied to the visual content, thereby reducing the number of inputs required to modify the audio content applied to the visual content. By modifying the audio content in response to detecting a user input, it provides feedback to the user regarding the current state of the device (e.g., that the device detected a user input).

[0201] In some embodiments, in response to detecting a user input (e.g., 648), while the computer system continues to play the visual content of the first aggregated content item, the computer system modifies (716) the visual parameters of the playback of the visual content of the first aggregated content item (e.g., brightness, saturation, hue, contrast, color, visual transitions between content items of the first plurality of content items, display duration of each content item of the first plurality of content items, one or more visual transitions used for the first aggregated content item (e.g., between content items presented within the first aggregated content item)) (e.g., FIGS. 6K - 6N, change the visual filter applied to media item 628A in response to user input 648) (e.g., change the visual filter applied to the first aggregated content item (e.g., from a first visual filter to a second visual filter different from the first visual filter)) (e.g., without stopping, pausing, and / or otherwise interrupting the playback of the visual content of the first aggregated content item) (e.g., without changing the order of the ordered sequence of the first plurality of content items). By modifying the visual parameters of the audio content and the visual content of the first aggregated content item in response to detecting a user input while the visual content of the first aggregated content item continues to play, the user is enabled to quickly modify the visual parameters applied to the audio content and the visual content, thereby reducing the number of inputs required to modify the visual parameters applied to the audio content and the visual content. By modifying the visual parameters in response to detecting a user input, feedback regarding the current state of the device (e.g., that the device detected a user input) is provided to the user.

[0202] In some embodiments, playing the visual content of the first aggregated content item (e.g., before detecting user input) includes displaying the visual content in a state where a first visual filter is applied to a first region of the visual content (e.g., a first display region) (e.g., the entire display region of the visual content and / or a portion of the display region of the visual content) (e.g., as shown in FIG. 6K, where visual filter 646B is applied to the first region of media item 628A). In some embodiments, playing the visual content of the first aggregated content item (e.g., before detecting user input) includes displaying the visual content in a state where a second visual filter different from the first visual filter is applied to a second region of the visual content different from the first region (e.g., a second display region) (e.g., while simultaneously displaying the first visual filter applied to the first region).

[0203] In some embodiments, while continuing to play the visual content of the first aggregated content item, modifying the visual parameters of the playback of the visual content of the first aggregated content item includes displaying the visual content in a state where a second visual filter different from the first visual filter is applied to a first region of the visual content (e.g., in FIG. 6N, visual filter 646C is applied to the first region of media item 628A). In some embodiments, modifying the visual parameters includes replacing the display of the visual content with the first visual filter applied to the first region with the display of the visual content with the second visual filter applied to the first region. In some embodiments, modifying the visual parameters includes replacing the display of a second region of the visual content with the second visual filter applied with the display of the second region of the visual content with a third visual filter different from the first and second visual filters applied to the second region. In some embodiments, the visual filter includes a set of two or more of a default exposure setting (e.g., a default exposure value and / or a default exposure adjustment), a default contrast setting (e.g., a default contrast value and / or a default contrast adjustment), a default highlight setting (e.g., a default highlight value and / or a default highlight adjustment), a default shadow setting (e.g., a default shadow value and / or a default shadow adjustment), a default brightness setting (e.g., a default brightness value and / or a default brightness adjustment), a default saturation setting (e.g., a default saturation value and / or a default saturation adjustment), a default warmth setting (e.g., a default warmth value and / or a default warmth adjustment), and / or a default tint setting (e.g., a default tint value and / or a default tint adjustment). In response to detecting user input, modifying the visual filter applied to the visual content of the first aggregated content item enables the user to quickly modify the visual filter applied to the visual content of the first aggregated content item, thereby reducing the number of inputs required to modify the visual filter applied to the visual content.In response to detecting user input, provide the user with feedback regarding the current state of the device (e.g., that the device detected user input) by modifying a visual filter applied to the visual content of a first aggregated content item.

[0204] In some embodiments, playing audio content separate from the content item while playing the visual content of a first aggregated content item includes playing a first audio track separate from the content item while playing the visual content of a first aggregated content item (e.g., playing audio track 1 while displaying media item 628A in FIG. 6K).

[0205] In some embodiments, while playing a first audio track (e.g., audio track 1 in FIG. 6K), the visual content of the first aggregated content item is displayed with a first visual filter (e.g., 646B) applied to a first region of the visual content (e.g., the entire display region of the visual content and / or a portion of the display region of the visual content). In some embodiments, the first audio track (e.g., audio track 1 in FIG. 6K) is part of (e.g., forms and / or defines) a first predetermined combination with the first visual filter (e.g., 646B). In some embodiments, the first predetermined combination does not include any other audio track or visual filter.

[0206] In some embodiments, while continuously playing the visual content of the first aggregated content item, modifying the playing audio content includes playing a second audio track (e.g., audio track 2 in FIG. 6N) that is separate from the content item and different from the first audio track while continuously playing the visual content of the first aggregated content item. In some embodiments, in response to detecting user input, the computer system pauses the playback of the first audio track.

[0207] In some embodiments, while playing the second audio track (e.g., audio track 2 in FIG. 6N), the visual content (e.g., 628A) of the first aggregated content item is displayed with a second filter (e.g., 646C) applied to a first region of the visual content. In some embodiments, in response to detecting user input, the computer system replaces the display of the first region of the visual content with the first visual filter applied with the display of the first region of the visual content with the second visual filter applied. In some embodiments, while the first visual filter is applied to the first region of the visual content, the second visual filter is applied to a second region of the visual content that is different from the first region, and in response to detecting user input, the second visual filter is applied to the first region of the visual content and the first visual filter stops being applied to the first region of the visual content.

[0208] In some embodiments, the second audio track (e.g., audio track 2 in FIG. 6N) is part of (e.g., forms and / or defines) a second predefined combination with the second visual filter (e.g., 646C). In some embodiments, the second predefined combination does not include any other audio track or visual filter.

[0209] In some embodiments, the first predetermined combination (e.g., one of the memory recipes 6 in FIG. 6K) and the second predetermined combination (e.g., two of the memory recipes 6 in FIG. 6N) are part of a plurality of predetermined combinations of filters and audio tracks. The plurality of predetermined combinations of filters and audio tracks are arranged in a certain order (e.g., memory recipes 1-6). The second predetermined combination is selected to be adjacent to the first predetermined combination within the order (e.g., immediately before and / or after the first predetermined combination), and the first audio track (e.g., audio track 1 in FIG. 6K) is different from the second audio track (e.g., audio track 2 in FIG. 6N). By sequentially ordering a first predetermined combination including a first visual filter and a first audio track adjacent to a second predetermined combination including a second visual filter and a second audio track different from the first visual filter and the first audio track, improved feedback is provided to the user by making it clear to the user that both the audio content and the visual parameters are modified in response to user input.

[0210] In some embodiments, the computer system applies a first predetermined combination to a first aggregated content item (e.g., plays a first audio track and displays the visual content of the first aggregated content item with a first visual filter applied to a first region), and while the first predetermined combination is being applied to the first aggregated content item, the computer system detects user input, and in response to detecting the user input, the computer system applies a second predetermined combination to the first aggregated content item (e.g., plays a second audio track and displays the visual content of the first aggregated content item with a second visual filter applied to the first region). In some embodiments, in response to detecting the user input, the computer system aborts the application of the first predetermined combination (e.g., stops playing the first audio track and stops applying the first visual filter to the first region). In some embodiments, the second predetermined combination is applied in response to the user input based on the second predetermined combination being adjacent to the first predetermined combination in the order (e.g., according to a determination that the second predetermined combination is adjacent to the first predetermined combination in the order). In some embodiments, the user input includes a direction, and the direction of the user input indicates a request to apply the next predetermined combination in the order, and the second predetermined combination is applied in response to the user input based on the second predetermined combination being immediately after the first predetermined combination in the order. In some embodiments, the user input includes a (e.g., different) direction, and the direction of the user input indicates a request to apply the previous predetermined combination in the order, and the second predetermined combination is applied in response to the user input based on the second predetermined combination being immediately before the first predetermined combination in the order.

[0211] In some embodiments, a first visual filter (e.g., 646B) is selected to be part of a first predetermined combination with a first audio track (e.g., audio track 1) based on one or more audio characteristics of the first audio track (e.g., beats per minute and / or sound wave characteristics) and one or more visual characteristics of the first visual filter (e.g., exposure, luminance, saturation, hue, and / or contrast). In some embodiments, a second visual filter is selected to be part of a second predetermined combination with a second audio track based on another audio characteristic of the second audio track and one or more visual characteristics of the second visual filter. By selecting the first visual filter to pair with the first audio track based on one or more audio characteristics of the first audio track, the quality of the filter / audio track combinations provided to the user is improved, thereby providing an improved means for selection by the user. Otherwise, additional input would be required to further find the desired combinations of visual filters and audio tracks.

[0212] In some embodiments, playing the visual content of the first aggregated content item (e.g., before detecting user input) includes, via a display generation component, simultaneously displaying the visual content (e.g., 628A) in a state (e.g., FIG. 6K) where a first visual filter (e.g., 646B) is applied to a first region of the visual content including a central display portion of the visual content, and the visual content (e.g., 628A) in a state (e.g., FIG. 6K) where a second visual filter (e.g., 646C) is applied to a second region of the visual content different from the first region, including a first edge (e.g., left edge, right edge, top edge, and / or bottom edge) of the visual content (e.g., a second region that does not overlap with the first region and / or a second region adjacent to the first region). By simultaneously displaying the visual content in a state where the first visual filter is applied to the first region of the visual content and the second visual filter is applied to the second region of the visual content, feedback regarding the current state of the device (e.g., that the second visual filter is ordered adjacent to the first visual filter) is provided to the user.

[0213] In some embodiments, playing the visual content of the first aggregated content item (e.g., before detecting user input) further includes displaying the visual content (e.g., 628A) via a display generation component while the first visual filter (e.g., 646B) is applied to the first region and the second visual filter (e.g., 646C) is applied to the second region, with a third visual filter (e.g., 646A) different from the first and second visual filters applied to a third region of visual content (e.g., a third region that does not overlap the first or second region) (e.g., a third region adjacent to the first region) (e.g., FIG. 6K) that is different from the first and second regions, the third region including a second edge of visual content that is different from the first edge (e.g., a left edge, a right edge, an upper edge, and / or a lower edge) (e.g., an edge opposite the first edge). Displaying the visual content simultaneously with the first visual filter applied to the first region of the visual content, the second visual filter applied to the second region of the visual content, and the third visual filter applied to the third region of the visual content provides feedback to the user regarding the current state of the device (e.g., that the second and third visual filters are ordered adjacent to the first visual filter).

[0214] In some embodiments, playing the visual content (e.g., 628A, 628B, 628C, 628D) of the first aggregated content item (e.g., before detecting user input) includes applying a first visual transition type of transition (e.g., crossfade, fade to black, reveal bleed, pan, scale, and / or rotation) to the visual content of the first aggregated content item (e.g., applying a first type of visual transition between content items of the first aggregated content item), and while continuing to play the visual content of the first aggregated content item, modifying the visual parameters of the playback of the visual content of the first aggregated content item includes modifying the transition to a second visual transition type different from the first visual transition type (e.g., applying a second type of visual transition between content items within the first aggregated content item). In some embodiments, playing the visual content of the first aggregated content item (e.g., before detecting user input) includes displaying a first content item (e.g., a first image and / or a first video) of the first aggregated content item, displaying a transition from the first content item to a second content item of the first aggregated content item, the transition being of the first visual transition type, and after displaying the transition from the first content item to the second content item, displaying the second content item. After detecting user input and modifying the visual parameters of the playback of the visual content of the first aggregated content item (e.g., including modifying the first visual transition type to the second visual transition type), the computer system displays a third content item of the first aggregated content item, and after displaying the third content item, the computer system displays a transition from the third content item to a fourth content item of the first aggregated content item, the transition being of a second visual transition type different from the first visual transition type.In response to detecting user input, by modifying the visual transition applied to the visual content of the first aggregated content item, the user is enabled to quickly modify the visual transition applied to the visual content of the first aggregated content item, thereby reducing the number of inputs required to modify the visual transition applied to the visual content.

[0215] In some embodiments, the first visual transition type is selected from a plurality of visual transition types based on audio content (e.g., track 1 of FIG. 6K) that is being played (e.g., before the start of user input, such as before the start of user input 648) based on, for example, sound wave information and / or beats per minute information. In some embodiments, the second visual transition type is selected from a plurality of visual transition types based on audio content (e.g., track 2 of FIG. 6N) that is being played (e.g., after the end of user input, such as after the end of user input 648) based on, for example, sound wave information and / or beats per minute information. Automatically selecting the transition type based on the audio content being played improves the quality of the visual transitions proposed to the user and enables the user to apply those improved visual transitions without further user input.

[0216] In some embodiments, the first visual transition type is selected from a first set of visual transition types based on the tempo (e.g., beats per minute information) of the audio content (e.g., track 1 of FIG. 6K) being played before detecting a user input (e.g., 648), and the second visual transition type is selected from a second set of visual transition types that are different from the first set based on the tempo (e.g., beats per minute information) of the audio content (e.g., track 2 of FIG. 6N) being played after detecting the user input (e.g., 648) (e.g., a first set of visual transition types (e.g., exposure bleed, pan, scale, and / or rotation) for audio content having a beats per minute value within a first range (e.g., a "high energy" song having a high beats per minute (e.g., above a threshold)), and a second set of visual transition types for audio content having a beats per minute within a second range (e.g., a "low energy" song having a lower beats per minute value (e.g., below a threshold))). Automatically selecting the transition type based on the audio content being played improves the quality of the visual transitions presented to the user and enables the user to apply those improved visual transitions without further user input.

[0217] In some embodiments, playing the visual content of a first aggregated content item (e.g., 628A, 628B, 628C, 628D) (e.g., before detecting user input) includes displaying the visual content (e.g., 628A of FIG. 6K) with a first set of visual parameters (e.g., FIGS. 6K, 646B) applied to a first region of the visual content (e.g., a first display region), and while simultaneously displaying the visual content with a first visual filter applied to the first region, displaying the visual content with a second set of visual parameters different from the first set of visual parameters (e.g., FIGS. 6K, 646C) applied to a second region of the visual content that is different from and adjacent to the first region (e.g., a second region that does not overlap the first region), and displaying a divider (e.g., the blank space between visual filters 646B and 646C in FIG. 6K) between the first region and the second region. In some embodiments, the divider is a visually distinct region between the first region and the second region. In some embodiments, the divider is a visual divider that appears based on visual parameters different from the visual parameters applied to the first and second regions (e.g., the divider is not a distinct region between the first region and the second region) (e.g., the divider is a dividing line between the first region and the second region). By simultaneously displaying the visual content with the first set of visual parameters applied to the first region of the visual content and the second set of visual parameters applied to the second region of the visual content, feedback regarding the current state of the device (e.g., the first set of visual parameters is currently selected and user input would cause the second set of visual parameters to be selected) is provided to the user.

[0218] In some embodiments, in response to detecting a user input (e.g., 648), the computer system shifts a divider simultaneously with the user input (e.g., shifts the blank space between visual filters 646B and 646C in FIGS. 6K - 6N) (e.g., shifts the divider in a direction corresponding to the direction of the user input) while continuing to play the visual content of the first aggregated content item and without shifting the visual content of the first aggregated content item (e.g., 628A in FIGS. 6K - 6N). In some embodiments, shifting the divider simultaneously with the user input includes changing the size of a first region and changing the size of a second region based on the user input and / or based on shifting the divider. In some embodiments, shifting the divider simultaneously with the user input includes increasing the size of the first region (e.g., by a first amount) and decreasing the size of the second region (e.g., by the first amount) based on the user input and / or based on shifting the divider. Shifting the divider simultaneously with the user input provides the user with feedback regarding the current state of the device (e.g., that the device detected the user input and / or that the user input caused the application of a first and / or second set of visual parameters).

[0219] In some embodiments, before detecting a user input (e.g., 648), a first aggregated content item is configured to display a first content item (e.g., 628A) (or, optionally, each content item) of a first plurality of content items over a first duration (e.g., 1 second, or 3 seconds), and modifying the visual parameters of the playback of the visual content of the first aggregated content item comprises configuring the first aggregated content item to display the first content item (e.g., 628A) (or, optionally, each content item) over a second duration different from the first duration (e.g., 2 seconds, or 4 seconds). In some embodiments, before detecting a user input, a first aggregated content item is configured to display a second content item among the first plurality of content items over a third duration, and modifying the visual parameters of the playback of the visual content of the first aggregated content item comprises configuring the first aggregated content item to display the second content item over a fourth duration different from the third duration in response to detecting the user input. In some embodiments, the second duration is shorter than the first duration (e.g., modifying the audio content comprises playing new audio content having a tempo faster than the audio content (e.g., a higher beats per minute value)) based on a determination that the user input causes faster playback of the audio content. In some embodiments, the second duration is longer than the first duration (e.g., modifying the audio content comprises playing new audio content having a tempo slower than the audio content (e.g., a lower beats per minute value)) based on a determination that the user input causes slower playback of the audio content. In response to detecting the user input, modifying the duration for which a content item is displayed enables the user to quickly modify the duration for which the content item is displayed, thereby reducing the number of inputs required to modify the display duration of the content item.

[0220] In some embodiments, user input (e.g., 648, 650A, 650B) includes gestures (e.g., tap gestures, swipe gestures, and / or different gestures) (e.g., touch screen gestures and / or non-touch screen gestures such as mouse clicks or hover gestures) via, for example, a touch-sensitive display and / or a touch-sensitive surface. In response to detecting a gesture, modifying the audio content enables the user to quickly modify the audio content applied to visual content, thereby reducing the number of inputs required to modify the audio content applied to the visual content. In response to detecting a gesture, modifying the audio content provides the user with feedback regarding the current state of the device (e.g., that the device detected a gesture).

[0221] In some embodiments, modifying the audio content being played while continuing to play the visual content of the first aggregated content item includes changing the audio content from a first audio track (e.g., track 1 of FIG. 6K) (e.g., the first music track and / or the first song) to a second audio track different from the first audio track (e.g., track 2 of FIG. 6N) (e.g., the second music track and / or the second song) while continuing to play the visual content of the first aggregated content item (e.g., 628A). In some embodiments, playing audio content separate from the content item before detecting user input includes playing a first audio track, and modifying the audio content being played while continuing to play the visual content of the first aggregated content item includes stopping the playback of the first audio track and playing a second audio track (e.g., replacing the playback of the first audio track with the playback of the second audio track) while continuing to play the visual content of the first aggregated content item. In response to detecting user input, changing the audio content from the first audio track to the second audio track enables the user to quickly modify the audio track applied to the visual content, thereby reducing the number of inputs required to modify the audio track applied to the visual content. Changing the audio content from the first audio track to the second audio track provides the user with feedback regarding the current state of the device (e.g., that the device detected user input).

[0222] In some embodiments, changing the audio content from a first audio track to a second audio track includes pausing the playback of the first audio track at a first playback position of the first audio track that is not the start position of the first audio track (e.g., FIGS. 6K - 6N, pausing the playback of audio track 1) (e.g., pausing the playback of the first audio track during playback of the first audio track (e.g., in the middle of the first audio track)) (e.g., pausing the playback of the first audio track at its current playback position if a user input is detected) (e.g., if a user input is detected 37 seconds after entering the first audio track, pausing the playback of the first audio track at the 37 - second mark), and starting the playback of the second audio track at a second playback position of the second audio track that is not the start position of the second audio track (e.g., FIG. 6N, starting the playback of audio track 2) (e.g., starting the playback of the second audio track in the middle of the second audio track (e.g., from a playback position within the second audio track that is not the start of the second audio track) (e.g., 37 seconds after entering the second audio track, 48 seconds after entering the second audio track)). In some embodiments, the second playback position corresponds to the first playback position (e.g., if the first playback position is 23 seconds after entering the first audio track (e.g., a user input is detected at the 23 - second mark of the first audio track and / or the first audio track is stopped at the 23 - second mark), the second playback position is 23 seconds after entering the second audio track (e.g., the second audio track starts playback from the 23 - second mark)). In some embodiments, the second playback position corresponds to a percentage of completion of the second audio track that corresponds to the percentage of completion of the first playback position within the first audio track (e.g., the first playback position represents that x% of the first audio track has been completed, and the second playback position represents that x% of the second audio track has been completed).In some embodiments, the second playback position is a playback position that enters the audio track and is longer than a predetermined time (for example, longer than 5 seconds after entering the audio track, longer than 10 seconds after entering the audio track, longer than 20 seconds after entering the audio track, or longer than 30 seconds after entering the audio track). By automatically starting the playback of the second audio track at the second playback position of the second audio track that is not the start of the second audio track, a more accurate preview of how the playback of the first aggregated content item is with the second audio track applied is provided to the user without requiring further user input.

[0223] In some embodiments, the computer system detects one or more duration setting inputs (for example, one or more inputs that select options 636E - 636H) (for example, one or more tap inputs and / or one or more non - tap inputs) via one or more input devices (for example, while playing the visual and audio content of the first aggregated content item). In response to detecting one or more duration setting inputs, the computer system modifies the duration (for example, length) of the first aggregated content item (for example, the duration of the visual content of the first aggregated content item) (for example, from a first duration to a second duration). In some embodiments, before detecting one or more duration setting inputs, the speed at which the content of the first aggregated content item is displayed is such that the computer system takes a first duration to play the first aggregated content, and after detecting one or more duration setting inputs, the speed at which the content of the first aggregated content item is displayed is such that the computer system takes a second duration (different from the first duration) to play the first aggregated content. By modifying the duration of the first aggregated content item in response to detecting user input, it enables the user to quickly modify the duration of the first aggregated content item, thereby reducing the number of inputs required to modify the duration of the aggregated content item.

[0224] In some embodiments, while continuing to play the visual content of the first aggregated content item, modifying the playing audio content includes changing the audio content from a first audio track (e.g., a first music track and / or a first song) to a second audio track (e.g., a second music track and / or a second song) different from the first audio track while continuing to play the visual content of the first aggregated content item. The first audio track has a first duration (e.g., length), and the second audio track has a second duration (e.g., length) different from the first duration. In response to detecting user input, the computer system modifies the duration (e.g., length) of the first aggregated content item (e.g., the duration of the visual content of the first aggregated content item) based on the second duration (e.g., option 636H "complete song") (e.g., modifies the duration of the first aggregated content item to be equal to the second duration). In some embodiments, modifying the duration of the first aggregated content item includes modifying the individual duration for which each content item of at least a subset of the first plurality of content items is configured to be displayed (e.g., modifying the duration for which the first content item is displayed, modifying the duration for which the second content item is displayed). In some embodiments, modifying the duration of the first aggregated content item includes modifying the number of content items displayed within the first aggregated content item (e.g., modifying the number of content items within the first plurality of content items). Automatically modifying the duration of the first aggregated content item based on the duration of the second audio track enables the user to quickly modify the duration of the first aggregated content item without further user input.

[0225] In some embodiments, while playing audio content, the computer system detects, via one or more input devices (e.g., while playing visual and audio content of a first aggregated content item), one or more duration-conforming inputs (e.g., one or more inputs that select option 636H) (e.g., one or more tap inputs and / or one or more non-tap inputs). In response to detecting one or more duration-conforming inputs and in accordance with a determination that the audio content has a first duration, the computer system modifies the duration (e.g., length) (e.g., the duration of the visual content of the first aggregated content item) of the first aggregated content item from a second duration different from the first duration to the first duration (e.g., based on the determination that the audio content has the first duration). In some embodiments, in response to detecting one or more duration-conforming inputs and in accordance with a determination that the audio content has a third duration different from the first duration and the second duration, the computer system modifies the duration of the first aggregated content item from the second duration to the third duration. By modifying the duration of the first aggregated content item in response to detecting user input, the user is enabled to quickly modify the duration of the first aggregated content item, thereby reducing the number of inputs required to modify the duration of the aggregated content item.

[0226] In some embodiments, while playing the visual content of the first aggregated content item (e.g., 628A, 628B, 628C, 628D) and audio content separate from the content item (e.g., audio track 1, audio track 2, audio track 3 of FIGS. 6F - 6AG), the computer system displays, via a display generation component, a first selectable object (e.g., 644D) that is selectable to display a plurality of visual filter options (e.g., each visual filter option corresponds to an individual visual filter) corresponding to a plurality of visual filters (e.g., for a plurality of visual filters). While displaying the first selectable object, the computer system detects, via one or more input devices, a first selection input (e.g., 668) (e.g., a tap input and / or a non - tap input) corresponding to the selection of the first selectable object. In response to detecting the first selection input, the computer system displays a visual filter selection user interface (e.g., 670) while continuing to play the visual content of the first aggregated content item (in some embodiments, while continuing to play audio content separate from the content item). Displaying the visual filter selection user interface includes simultaneously displaying a first user interface object (e.g., 674A) (e.g., a first user interface object corresponding to the first visual filter) that includes a continued display of the visual content of the first aggregated content item with the first visual filter applied to the visual content, and a second user interface object (e.g., 674B) (e.g., a second user interface object corresponding to a second visual filter different from the first visual filter) that includes a continued display of the visual content of the first aggregated content item with a second visual filter different from the first visual filter applied to the visual content. Displaying a plurality of visual filter options simultaneously enables the user to quickly browse and select a desired visual filter, thereby reducing the number of inputs required to select a visual filter.

[0227] In some embodiments, the first user interface object is displayed in a first region of the visual filter selection user interface, and the second user interface object is displayed in a second region of the visual filter selection user interface that does not overlap the first region. In some embodiments, displaying the visual filter selection user interface includes displaying continuous playback of the visual content of the first aggregated content item with a third visual filter applied to the visual content that is different from the first and second visual filters (e.g., corresponding to a third visual filter that is different from the first and second visual filters), the third user interface object, simultaneously with the first user interface object and the second user interface object. In some embodiments, the method includes detecting, via one or more input devices, user input corresponding to a selection of the first user interface object while displaying the visual filter selection user interface including the first user interface object and the second user interface object, and in response to detecting the user input, aborting the display of the visual filter selection user interface (e.g., aborting the display of the second user interface object), and displaying continuous playback of the visual content of the first aggregated content item with the first visual filter applied to the visual content. In some embodiments, selection of the first user interface object and / or selection of the second user interface object maintains continuous playback of audio content separate from the content item (e.g., selection of a user interface object within the visual filter selection user interface does not affect the audio content being played). In some embodiments, selection of the first user interface object causes playback of a second audio content different from the audio content (e.g., selection of a user interface object within the visual filter selection user interface changes the audio content being played and / or applied to the first aggregated content item).

[0228] In some embodiments, while playing the visual content of the first aggregated content item and audio content separate from the content item, the computer system, via a display generation component, displays a second selectable object (e.g., 644C) that is selectable to display a plurality of audio track options (e.g., corresponding to a plurality of audio tracks, where each audio track option corresponds to an individual audio track). In some embodiments, while displaying the second selectable object, the computer system detects, via one or more input devices, a second selection input (e.g., 652) (e.g., a tap input) (e.g., a non-tap input) corresponding to a selection of the second selectable object. In response to detecting the second selection input, the computer system displays an audio track selection user interface (e.g., 654) (in some embodiments, while continuing to play the visual content of the first aggregated content item) (in some embodiments, pauses the playback of the visual content of the first aggregated content item in response to detecting the second selection input). The audio track selection user interface includes a third user interface object (e.g., 656A) corresponding to a first audio track, the third user interface object being selectable to initiate a process of applying the first audio track to the first aggregated content item (e.g., playing the first audio track while playing the visual content of the first aggregated content item), and a fourth user interface object (e.g., 656B) corresponding to a second audio track different from the first audio track, the fourth user interface object being selectable to initiate a process of applying the second audio track to the first aggregated content item (e.g., playing the second audio track while playing the visual content of the first aggregated content item).By simultaneously displaying a third user interface object corresponding to a first audio track and a fourth user interface object corresponding to a second audio track, the user is enabled to quickly select a desired audio track, thereby reducing the number of inputs required to select an audio track.

[0229] In some embodiments, the audio track selection user interface further includes a fifth user interface object corresponding to a third audio track different from the first and second audio tracks, and the fifth user interface object is for applying the third audio track to the first aggregated content item (e.g., playing the third audio track while playing the visual content of the first aggregated content item). In some embodiments, the second selection input is detected while the visual content of the first aggregated content item is displayed with the first visual filter applied, and the selection of the third user interface object and / or the selection of the fourth user interface object maintains the application of the first visual filter to the visual content of the first aggregated content item (e.g., the selection of the user interface object within the audio track selection user interface does not affect the visual filter applied to the visual content). In some embodiments, the selection of the third user interface object causes a second visual filter different from the first visual filter to be applied to the visual content (e.g., the selection of the user interface object within the audio track selection user interface changes the visual filter applied to the visual content of the first aggregated content item). In some embodiments, the first track and the second audio track are selected for inclusion in the audio track selection user interface based on the visual content of the first aggregated content item (e.g., the song suggestions are generated and / or provided based on the visual content included in the first aggregated content item) (e.g., mountain-related songs for the first aggregated content item related to a mountain climbing trip, or surfing-related songs for the first aggregated content item related to a surfing trip).

[0230] In some embodiments, a third user interface object (e.g., 656A) includes a display of a track title (e.g., song name) corresponding to a first audio track, and a fourth user interface object (e.g., 656B) includes a display of a track title (e.g., song name) corresponding to a second audio track. In some embodiments, the third user interface object further displays album art corresponding to the first audio track, and the fourth user interface object further displays album art corresponding to the second audio track. By displaying a third user interface object that includes a track title corresponding to a first audio track and a fourth user interface object that includes a track title corresponding to a second audio track, the user is enabled to quickly select a desired audio track, thereby reducing the number of inputs required to select an audio track.

[0231] In some embodiments, while displaying an audio track selection user interface (e.g., 654) that includes a third user interface object (e.g., 656A - 656N) and a fourth user interface object (e.g., 656A - 656N), the computer system detects a third selection input (e.g., 660) (e.g., tap input and / or non - tap input) via one or more input devices. In response to detecting the third selection input, in accordance with the determination that the third selection input corresponds to the selection of the third user interface object, the computer system plays the first audio track from the beginning of the first audio track, and in accordance with the determination that the third selection input corresponds to the selection of the fourth user interface object, the computer system plays the second audio track from the beginning of the second audio track. By playing the first audio track from the beginning of the first audio track or the second audio track from the beginning of the second audio track in response to the third selection input, it enables the user to quickly listen to and select the desired audio track, thereby reducing the number of inputs required to select an audio track.

[0232] In some embodiments, while playing the visual content of the first aggregated content item, the first audio track is played from the beginning of the first audio track and / or the second audio track is played from the beginning of the second audio track. In some embodiments, modifying the audio content in response to user input includes changing the audio content from the first audio track to the second audio track, and the second audio track starts from a playback position that is not the start position of the second audio track (e.g., a particular set of user inputs causes a mid-track switch of the audio track (e.g., a user input corresponding to a change from a first default combination of a first visual filter and a first audio track to a second default combination of a second visual filter and a second audio track causes a mid-track switch of the audio track) (e.g., causes playback of the second audio track to start from a playback position that is not the beginning of the second audio track (e.g., enters the second audio track and is longer than a threshold duration)), in contrast, selection of an audio track from an audio track selection user interface causes the selected audio track to be played from the beginning of the audio track).

[0233] In some embodiments, while displaying an audio track selection user interface (e.g., 654) that includes a third user interface object (e.g., 656A - 656N) and a fourth user interface object (e.g., 656A - 656N), the computer system detects a fourth selection input (e.g., 660) corresponding to the selection of the third user interface object (e.g., 656D) via one or more input devices. In response to detecting the fourth selection input, according to a determination that the user of the computer system (e.g., the user account logged into the computer system) is not subscribed to an audio service (e.g., an audio service that provides access to a first audio track and / or a default audio service), the computer system begins a process of displaying a prompt for the user to subscribe to the audio service (e.g., FIGS. 6S, 664B). In some embodiments, the computer system begins a process of displaying a notification indicating that the user is not subscribed to the audio service and / or begins a process of displaying a notification prompting the user to sign up for a free trial of the audio service. In some embodiments, the computer system displays selectable user interface objects that can be selected to begin the process of subscribing to the audio service. In some embodiments, in response to detecting the fourth selection input, according to a determination that the user of the computer system is subscribed to the audio service, the computer system plays the first audio track (e.g., from the beginning of the first audio track) (in some embodiments, while playing the visual content of the first aggregated content item). In some embodiments, while displaying a prompt for the user to subscribe to the audio service, the computer system receives one or more user inputs corresponding to a request to subscribe to the audio service, and in response to receiving the one or more user inputs, begins a process of subscribing the user to the audio service.In some embodiments, in response to receiving one or more user inputs, the computer system requests authentication to enroll the user in an audio service (e.g., displays a user interface for the user to enter a password and / or passcode and / or collects biometric information for biometric authentication). In accordance with a determination that the user is not enrolled in the audio service, feedback regarding the current state of the device (e.g., the device has determined that the user is not enrolled in the audio service) is provided to the user by initiating a process of displaying a notification prompting the user to enroll in the audio service.

[0234] In some embodiments, while displaying an audio track selection user interface (e.g., 654) that includes a third user interface object (e.g., 656A - 656N) and a fourth user interface object (e.g., 656A - 656N), the computer system detects, via one or more input devices, a fifth selection input (e.g., 660) (e.g., a tap input and / or a non - tap input) corresponding to the selection of the third user interface object (e.g., 656D). In response to detecting the fifth selection input and according to a determination that a user of the computer system (e.g., a user account logged into the computer system) is not subscribed to an audio service (e.g., an audio service that provides access to a first audio track), the computer system begins a process of displaying a preview user interface (e.g., 662). Displaying the preview user interface includes playing a preview of a first aggregated content item in which a first audio track (e.g., track 3 in FIG. 6S) is applied to visual content (e.g., 628B) of the first aggregated content item (e.g., playing a preview of the first aggregated content item includes playing the first audio track while simultaneously playing (e.g., displaying) the visual content of the first aggregated content item). The preview user interface does not permit (e.g., prevents and / or prohibits) the user from sharing the preview and / or saving the preview for later playback until the user subscribes to the audio service (e.g., the preview user interface does not include any selectable options that would enable the user to share the preview and / or save the preview for later playback, or provide any user input). By beginning the process of displaying the preview user interface according to a determination that the user is not subscribed to the audio service, feedback regarding the current state of the device (e.g., the device has determined that the user is not subscribed to the audio service) is provided to the user.

[0235] In some embodiments, while playing the visual content of the first aggregated content item and audio content separate from the content item, the computer system displays a fifth user interface object (e.g., 632D) that is selectable to cause the computer system to enter an editing mode. In some embodiments, entering the editing mode includes displaying an editing user interface.

[0236] In some embodiments, after displaying a fifth user interface object (e.g., while the fifth user interface object is being displayed and / or after the fifth user interface object is no longer being displayed), the computer system detects a second user input (e.g., 648 (e.g., a swipe gesture)) (e.g., a gesture (e.g., a tap gesture, a swipe gesture) and / or a voice input via (e.g., via a touch-sensitive display and / or a touch-sensitive surface)) via one or more input devices. In response to detecting the second user input and in accordance with a determination that the computer system is in an editing mode (e.g., FIGS. 6K-6N), the computer system modifies the audio content being played while continuing to play the visual content of the first aggregated content item. In some embodiments, the computer system also modifies the visual parameters of the playback of the visual content of the first aggregated content item. In some embodiments, in response to detecting the second user input and in accordance with a determination that the computer system is not in an editing mode (e.g., FIG. 6J), the computer system stops modifying the audio content being played. In some embodiments, the computer system stops modifying the visual parameters of the playback of the visual content of the first aggregated content item. Modifying the audio content in response to the second user input and in accordance with a determination that the computer system is in an editing mode provides the user with feedback regarding the current state of the device (e.g., that the computer system is in an editing mode).

[0237] In some embodiments, while playing the visual content of the first aggregated content item (e.g., FIGS. 6Y, 628C) and audio content separate from the content item (e.g., FIGS. 6Y, track 3), and while displaying a fifth user interface object (e.g., 632D) (e.g., before causing the computer system to enter an edit mode and / or while the computer system is not in an edit mode), a sixth user interface object (e.g., 632E) that is selectable to pause the playback of the visual content of the first aggregated content item is displayed via a display generation component. In some embodiments, selection of the sixth selectable user interface object also pauses the playback of audio content separate from the content item.

[0238] While displaying the sixth selectable user interface object, the computer system detects a sixth selection input (e.g., a tap input and / or a non-tap input) corresponding to the selection of the sixth user interface object (e.g., 680) via one or more input devices. In response to detecting the sixth selection input, the computer system pauses the playback of the visual content of the first aggregated content item (e.g., displays the visual content of the first aggregated content item in a paused state). In some embodiments, the computer system also pauses the playback of audio content separate from the content item. In response to detecting the sixth selection input, the computer system replaces the display of the fifth user interface object (e.g., 632D) (e.g., the "Recipe" option) with a seventh user interface object (e.g., 632G) (e.g., the aspect ratio toggle option) that is selectable to modify the aspect ratio of the visual content of the first aggregated content item. While the seventh user interface object is being displayed (e.g., and while the visual content of the first aggregated content item is paused and / or while the visual content of the first aggregated content item is being displayed in a paused state), the computer system detects a seventh selection input (e.g., 683) (e.g., a tap input and / or a non-tap input) corresponding to the selection of the seventh user interface object via one or more input devices.In response to detecting a seventh selection input, the computer system, via a display generation component (e.g., while maintaining the display of the visual content of the first aggregated content item in a paused state), causes the visual content of the first aggregated content item (e.g., 628C, FIG. 6Z) to transition from being displayed in a first aspect ratio to being displayed in a second aspect ratio different from the first aspect ratio (e.g., 628C, FIG. 6AA) (e.g., causes the visual content of the first aggregated content item to transition from being displayed in a full-screen aspect ratio (e.g., an aspect ratio that fills the display area and / or the display) to being displayed in a native aspect ratio (e.g., the native aspect ratio of the content item being displayed)). In response to detecting a sixth selection input, the computer system pauses the playback of the visual content of the first aggregated content item and replaces the display of a fifth user interface object with the display of a seventh user interface object, thereby providing the user with feedback regarding the current state of the device (e.g., that the computer system detected a sixth selection input).

[0239] In some embodiments, while the playback of the visual content of the first aggregated content item is paused, while the seventh user interface object is being displayed, and while the visual content of the first aggregated content item is being displayed in a second aspect ratio, the computer system, via the display generation component, displays an eighth user interface object that is selectable to resume the playback of the visual content of the first aggregated content item. While the eighth user interface object is being displayed, the computer system displays an eighth selection input (e.g., a tap input and / or a non-tap input) corresponding to the selection of the eighth selectable user interface object. In response to detecting the eighth selection input, the computer system, via the display generation component, displays the visual content of the first aggregated content item transitioning from being displayed in the second aspect ratio to being displayed in the first aspect ratio, and resumes the playback of the visual content of the first aggregated content item (e.g., in the first aspect ratio) (in some embodiments, the playback of audio content separate from the content item is also resumed).

[0240] In some embodiments, while playing the visual content of the first aggregated content item, the computer system, via one or more input devices, detects a pause input (e.g., 680) (e.g., one or more tap inputs and / or one or more non-tap inputs) corresponding to a request to pause the playback of the visual content of the first aggregated content item (e.g., selecting a pause option by a tap input). In response to detecting the pause input, the computer system pauses the playback of the visual content of the first aggregated content item (e.g., FIG. 6Z). In some embodiments, pausing the playback of the visual content of the first aggregated content item includes continuously displaying the visual content that was being displayed when the pause input was detected (e.g., continuously displaying until one or more further user inputs (e.g., one or more user inputs to resume the playback of the visual content of the first aggregated content item) are received). In some embodiments, in response to detecting the pause input, the computer system, via a display generation component, displays video navigation user interface elements (e.g., 682) (e.g., a scrubber bar) for navigating through the visual content of the first aggregated content item (e.g., its plurality of frames (e.g., images)). By pausing the playback of the visual content of the first aggregated content item and displaying the video navigation user interface elements in response to detecting the pause input, feedback regarding the current state of the device (e.g., that the computer system detected a pause input) is provided to the user.

[0241] In some embodiments, displaying a visual navigation user interface element (e.g., 682) includes simultaneously displaying a representation of a first content item of a first plurality of content items and a representation of a second content item of the first plurality of content items (e.g., different from the first content item) (e.g., FIGS. 6Z-6AA). In some embodiments, displaying a visual navigation user interface element further includes simultaneously displaying a representation of a third content item of a first plurality of content items different from the first and second content items, along with the representations of the first and second content items. In some embodiments, the visual navigation user interface element is a scrubber bar, and the scrubber bar includes representations of content items aggregated in a first aggregated content item. By simultaneously displaying the representation of the first content item and the representation of the second content item, feedback regarding the current state of the device (e.g., that the first aggregated content item includes the first content item and the second content item) is provided to the user.

[0242] In some embodiments, in response to detecting a pause input (e.g., 1226), the computer system, via a display generation component and simultaneously with a visual navigation user interface element (e.g., 1228) (in some embodiments, while the playback of the visual content of the first aggregated content item is paused), displays a duration control option (e.g., 1232A). While the duration control option is being displayed, the computer system detects, via one or more input devices, a duration control input (e.g., 1242) corresponding to a selection of the duration control option (e.g., one or more remote control inputs and / or one or more non-remote control inputs) (e.g., one or more tap inputs and / or one or more non-tap inputs). In response to detecting the duration control input, the computer system, via the display generation component, simultaneously displays a first playback duration option (e.g., 1233A - 1244E) (e.g., a short playback duration option) corresponding to a first playback duration and a second playback duration option (e.g., 1244A - 1244E) (e.g., a long playback duration option) corresponding to a second playback duration different from the first playback duration. In some embodiments, the selection of the first playback duration option and / or the second playback duration option causes the first aggregated content item to be modified (e.g., increase and / or decrease the number of content items included in the first aggregated content item based on the selected playback duration option). By simultaneously displaying the first playback duration option and the second playback duration option, the user is enabled to quickly set the playback duration of the first aggregated content item, thereby reducing the number of inputs required to set the playback duration.

[0243] In some embodiments, in response to detecting a pause input (e.g., 1226), the computer system, via a display generation component and simultaneously with a visual navigation user interface element (e.g., 1228) (in some embodiments, while the playback of the visual content of the first aggregated content item is paused), displays an audio track control option (e.g., 1232B). While the audio track control option is being displayed, the computer system detects, via one or more input devices, an audio track control input (e.g., 1248) corresponding to the selection of the audio track control option (e.g., one or more remote control inputs and / or one or more non-remote control inputs) (e.g., one or more tap inputs and / or one or more non-tap inputs). In response to detecting the audio track control input, the computer system simultaneously displays, via the display generation component, a first audio track option (e.g., 1250A - 1250E) corresponding to a first audio track and a second audio track option (e.g., 1250A - 1250E) corresponding to a second audio track different from the first audio track. In some embodiments, the selection of the first audio track option causes the first audio track to be applied to the first aggregated content item (e.g., causes the first audio track to be played while the visual content of the first aggregated content item is being played), and the selection of the second audio track option causes the second audio track to be applied to the first aggregated content item (e.g., causes the second audio track to be played while the visual content of the first aggregated content item is being played). By simultaneously displaying the first audio track option and the second audio track option, it enables the user to quickly set the audio track applied to the first aggregated content item, thereby reducing the number of inputs required to set the audio track.

[0244] In some embodiments, playing the visual content of the first aggregated content item includes, via a display generation component, at a first time, displaying a first content item (e.g., 628A in FIG. 6G) among a first plurality of content items within the first aggregated content item; via the display generation component, at the same time as the first content item (e.g., at the first time), displaying first title information (e.g., 627 in FIG. 6G) (e.g., text (e.g., location information and / or date information)) corresponding to the first content item; at a second time after the first time, via the display generation component, displaying a second content item (e.g., 628C in FIG. 6Y) (e.g., different from the first content item) among the first plurality of content items within the first aggregated content item; and via the display generation component, at the same time as the second content item (e.g., at the second time), displaying second title information (e.g., 627 in FIG. 6Y) corresponding to the second content item and different from the first title information (e.g., text (e.g., location information and / or date information)). By displaying the first title information at the same time as the first content item and the second title information at the same time as the second content item, feedback regarding the current state of the device (e.g., that the device has identified the first title information corresponding to the first content item and the second title information corresponding to the second content item) is provided to the user.

[0245] In some embodiments, while playing the visual content (e.g., 628C of FIGS. 6AB) and the audio content (e.g., track 3 of FIGS. 6AB) of the first aggregated content item, the computer system detects, via one or more input devices, one or more visual parameter modification inputs (e.g., 686, 688, 690) (e.g., one or more touch screen inputs, one or more remote control inputs, and / or different inputs). In response to detecting one or more visual parameter modification inputs, according to a determination that the one or more visual parameter modification inputs correspond to a first gesture (e.g., 686, 688, 690) (e.g., long press, tap on the left side of the screen, tap on the right side of the screen, left swipe, and / or right swipe), the computer system modifies the playback of the visual content of the first aggregated content item in a first manner (e.g., FIGS. 6AB - 6AD) (e.g., display the previous content item, display the next content item, and / or maintain the display of the current content item). According to a determination that the one or more visual parameter modification inputs correspond to a second gesture (e.g., 686, 688, 690) (e.g., long press, tap on the left side of the screen, tap on the right side of the screen, left swipe, and / or right swipe) that is different from the first gesture, the computer system modifies the playback of the visual content of the first aggregated content item in a second manner that is different from the first manner (e.g., FIGS. 6AB - 6AD) (e.g., display the previous content item, display the next content item, and / or maintain the display of the current content item). By modifying the playback of the visual content of the first aggregated content item in the first manner according to the first gesture and in the second manner according to the second gesture, it enables the user to quickly modify the playback of the visual content of the first aggregated content item using various gestures, thereby reducing the number of inputs required to modify the playback of the visual content of the first aggregated content item.

[0246] In some embodiments, the first gesture is a long-press gesture (e.g., 686) (e.g., continuous contact with a touch screen display, continuous contact with a touch pad, and / or continuous click of a mouse), and modifying the visual content of the first aggregated content item in the first manner includes maintaining the display of the currently displayed content item during the long-press gesture (e.g., FIGS. 6AB - 6AC) (e.g., during some or all of its duration) (e.g., while the contact with the touch screen display and / or touch pad is maintained and / or while the mouse button is depressed). In some embodiments, the computer system maintains the display of the currently displayed content item during the long-press gesture while the audio content being played is continuing to play. By maintaining the display of the currently displayed content item in response to the long-press gesture, it enables the user to easily maintain the display of the currently displayed content item, thereby reducing the number of inputs required to maintain the display of the currently displayed content item.

[0247] In some embodiments, during a long-press gesture (e.g., 686), while maintaining the display of the currently displayed content item, the computer system detects the end of the long-press gesture via one or more input devices. After detecting the end of the long-press gesture (e.g., in response to detecting the end of the long-press gesture), the computer system modifies the playback duration of one or more subsequent content items (e.g., all subsequent content items) that are displayed after the currently displayed content item (e.g., decreases the playback duration of one or more subsequent content items (e.g., decreases the time each content item of one or more subsequent content items is displayed)). In some embodiments, before detecting the long-press gesture, a first subsequent content item configured to be displayed after the currently displayed content item is configured to be displayed for a first duration during playback of visual content, and after detecting the long-press gesture, the first subsequent content item is configured to be displayed for a second duration different from the first duration (e.g., a second duration shorter than the first duration). By automatically adjusting the playback duration of one or more subsequent content items in response to the end of the long-press gesture that caused the extended display of the content item, it is possible to adjust the playback of visual content so that the user takes into account the extended playback duration of the content item without further user input.

[0248] In some embodiments, the first gesture is a first tap gesture (e.g., 688, 690) (e.g., a tap gesture in a first region of a touch screen display). Modifying the playback of the visual content of the first aggregated content item in a first manner includes navigating to a previous content item in the ordered sequence of content items within the first aggregated content item (e.g., FIGS. 6AC - 6AD) (e.g., replacing the display of the currently displayed content item within the first aggregated content item with a previous content item within the first aggregated content item) (in some embodiments, while the playing audio content continues to play). In some embodiments, the second gesture is a second tap gesture different from the first tap gesture (e.g., 688, 690) (e.g., a tap gesture in a second region of a touch screen display), and modifying the playback of the visual content of the first aggregated content item in a second manner includes navigating to a next content item in the ordered sequence of content items within the first aggregated content item (e.g., FIGS. 6AD - 6AE) (e.g., replacing the display of the currently displayed content item within the first aggregated content item with a next content item within the first aggregated content item) (in some embodiments, while the playing audio content continues to play). Depending on the various tap gestures, navigating between the content items within the first aggregated content item enables the user to easily navigate between the content items, thereby reducing the number of inputs required to navigate between the content items within the first aggregated content item.

[0249] In some embodiments, the first gesture is a first swipe gesture (e.g., a swipe gesture in a first direction), and modifying the playback of the visual content of the first aggregated content item in a first manner includes navigating to a previous content item in the ordered sequence of content items within the first aggregated content item (e.g., FIGS. 6AC - 6AD) (e.g., replacing the display of the currently displayed content item within the first aggregated content item with a previous content item within the first aggregated content item) (in some embodiments, while the playing audio content continues to play). In some embodiments, the second gesture is a second swipe gesture different from the first swipe gesture (e.g., a swipe gesture in a second direction (e.g., a second direction opposite or substantially opposite to the first direction)), and modifying the playback of the visual content of the first aggregated content item in a second manner includes navigating to a next content item in the ordered sequence of content items within the first aggregated content item (e.g., FIGS. 6AD - 6AE) (e.g., replacing the display of the currently displayed content item within the first aggregated content item with a next content item within the first aggregated content item) (in some embodiments, while the playing audio content continues to play). Depending on the various swipe gestures, navigating between the content items within the first aggregated content item enables the user to easily navigate between the content items, thereby reducing the number of inputs required to navigate between the content items within the first aggregated content item.

[0250] In some embodiments, modifying the playback of the visual content of the first aggregated content item in a first manner includes modifying the playback of the visual content of the first aggregated content item in the first manner while continuing to play audio content separate from the content item (e.g., FIGS. 6AB - 6AE), and modifying the playback of the visual content of the first aggregated content item in a second manner includes modifying the playback of the visual content of the first aggregated content item in the second manner while continuing to play audio content separate from the content item (e.g., FIGS. 6AB - 6AE). In some embodiments, providing a user input to advance to the next (or previous) content item of the first aggregated content item does not cause a corresponding skip / forward / change in the playback of the audio content. Thus, in some embodiments, the audio content is played independently of a user input to advance to the next (or previous) content item of the first aggregated content item. Modifying the playback of the visual content of the first aggregated content item while continuing to play audio content separate from the content item provides feedback to the user regarding the current state of the device (e.g., that the visual content of the first aggregated content item is modified while the first aggregated content item continues to play).

[0251] In some embodiments, while displaying a first content item of a first aggregated content item via a display generation component (e.g., during playback of visual content of the first aggregated content item), one or more input devices are used to detect a third user input (e.g., a long-press input, a tap input, a swipe input, and / or a different input), and in response to detecting the third user input (e.g., 614), the computer system simultaneously displays, via the display generation component, a tagging option (e.g., 616D) that is selectable to initiate a process of identifying (e.g., tagging) the person shown in the first content item, and a removal option (e.g., 616E, 616F) that is selectable to initiate a process of removing one or more content items that show the person also shown in the first content item from the first aggregated content item. Displaying the tagging option that is selectable to initiate a process of identifying the person shown in the first content item enables the user to quickly identify the person shown in the first content item, thereby reducing the number of inputs required to tag and / or identify the person shown. Displaying the removal option that is selectable to initiate a process of removing one or more content items that show the person also shown in the first content item from the first aggregated content item enables the user to quickly and easily remove content items that show a particular person, thereby reducing the number of inputs required to remove such content items.

[0252] In some embodiments, the removal option is the option of "not showing this person more" that reduces the number of instances (e.g., the number of content items) within the first aggregated content item in which the person is shown. In some embodiments, the removal option reduces the number of instances (e.g., the number of content items) within the first aggregated content item in which only that person is shown (and no other persons are shown). In some embodiments, the removal option is the option of "never showing this person" in which all instances in which the person is shown (e.g., all content items) are removed from the first aggregated content item.

[0253] In some embodiments, in response to detecting a third user input, the computer system displays a tagging option (e.g., without displaying the removal option). In some embodiments, in response to detecting a third user input, the computer system displays a removal option (e.g., without displaying the tagging option). In some embodiments, the tagging option and / or the removal option are accessible by interacting with a content item within a media library user interface and / or by interacting with a content item within a spotlight photo user interface.

[0254] Note that the details of the process described above with respect to method 700 (e.g., FIG. 7) are also applicable in a similar manner to the methods described later. For example, methods 900 and 1100 optionally include one or more of the various method characteristics described above with reference to method 700. For example, the aggregated content items in each of methods 700, 900, 1100 may be the same aggregated content item. For the sake of brevity, these details are not repeated below.

[0255] Figures 8A-8L illustrate exemplary user interfaces for managing content playback after playing a content item, according to some embodiments. The user interfaces in these figures are used to illustrate processes described below, including the process of FIG. 9.

[0256] FIG. 8A shows an electronic device 600 that is a smartphone having a touch-sensitive display 602. In some embodiments, the electronic device 600 includes one or more features of devices 100, 300, and / or 500. The electronic device 600 shows the playback user interface 625 described above with reference to FIGS. 6A-6AG. In FIG. 8A, the playback user interface 625 displays the playback of a first aggregated content item and displays the last media item 628Z of the first aggregated content item. For example, in FIG. 8A, the first aggregated content item described above with reference to FIGS. 6A-6AG is permitted to continue playing until it reaches the last media item 628Z of FIG. 8A.

[0257] In FIG. 8B, in response to a determination that the playback of the first aggregated content item has met one or more end criteria (e.g., the last media item of the first aggregated content item has been displayed for a threshold duration and / or there is less than a threshold duration remaining in the playback of the first aggregated content item), the electronic device 600 displays the next content item user interface 800. The next content item user interface 800 is overlaid on the playback user interface 625 that continues to display the last media item 628Z of the first aggregated content item. The playback user interface 625 is made visually unobtrusive (e.g., dimmed and / or blurred) while the next content item user interface 800 is overlaid thereon. The next content item user interface 800 includes tiles 804A, 804B, 804C representing other aggregated content items, and the tiles 804A, 804B, 804C are selectable to initiate the playback of the corresponding aggregated content item. Tile 804A corresponds to the "next" or subsequent aggregated content item that automatically starts playback without further user input.

[0258] The next content item user interface 800 includes a countdown timer 802A that indicates to the user, without further user input, that the next aggregated content item (e.g., "Palm Springs 2017") will start playback at the end of the countdown timer 802A. The next content item user interface 800 also includes a replay option 802B selectable to replay the first aggregated content item and a share option 802C selectable to initiate a process of sharing the first aggregated content item via one or more communication media.

[0259] FIG. 8C shows an exemplary scenario where, after the electronic device 600 displays the next content item user interface 800, no user input is received over a threshold duration (e.g., 10 seconds or 20 seconds) and the countdown timer 802A counts down to 0. In FIG. 8C, in accordance with the determination that the next content item user interface 800 is displayed for the threshold duration without any user input, the electronic device 600 automatically starts playing the second aggregated content item. The playback of the second aggregated content item includes displaying the title information 627 of the second aggregated content item, displaying the first media item (media item 806) of the second aggregated content item, and playing an audio track (e.g., audio track 4). In FIG. 8D, the playback of the second aggregated content item continues and the title information 627 moves from the first display position to the second display position.

[0260] FIGS. 8E - 8L show alternative scenarios where one or more user inputs are received while the electronic device 600 is displaying the next content item user interface 800. In FIG. 8E, the electronic device detects user inputs 808A, 808B, 808C, 808D, and 808E, and each of these will be described in turn below.

[0261] In FIG. 8E, the electronic device 600 detects a user input 808A (e.g., a tap input) corresponding to the selection of tile 804A. Tile 804A corresponds to the second aggregated content item. In FIG. 8F, in response to user input 808A, the electronic device 600 stops displaying the first aggregated content item and the next content item user interface 800 and starts playing the second aggregated content item (e.g., Palm Springs 2017).

[0262] In FIG. 8E, the electronic device 600 detects a user input 808B (e.g., a tap input) corresponding to the selection of the replay option 802B. In FIG. 8G, in response to the user input 808B, the electronic device 600 stops displaying the next content item user interface 800 and starts replaying the first aggregated content item within the playback user interface 625. As described above with reference to FIG. 6F, starting the playback of the first aggregated content item includes displaying the title information 627 of the first aggregated content item and the first media item (e.g., media item 628A), and playing the audio track (e.g., audio track 3) applied to the first aggregated content item.

[0263] In FIG. 8E, the electronic device 600 detects a user input 808C (e.g., a tap input) corresponding to the "negative space" disposed on the next content item user interface 800. In other words, the user input 808C does not correspond to the selection of any specific user interface object within the next content item user interface 800. In FIG. 8H, in response to the user input 808C, the electronic device 800C stops displaying the next content item user interface 800 and redisplays the playback user interface 625 in its previous state (e.g., not in its prominent state (e.g., with increased brightness and / or clarity)) (displaying the last media item 628Z of the first aggregated content item).

[0264] In FIG. 8E, the electronic device 600 detects a user input 808D that is a left swipe input at a position corresponding to the next content item user interface 800. In FIG. 8I, in response to the user input 808D, the electronic device 600 shifts the tiles 804A - 804C based on the user input (e.g., at a translation speed corresponding to the translation speed and / or translation distance of the user input, and / or over the translation distance). In FIG. 8I, the tiles 804A, 804B, 804C are shifted to the left, and additional tiles 804D, 804E corresponding to additional aggregated content items appear. The tiles 804A - 804E are selectable to start playback of the corresponding aggregated content items. In some embodiments, the tiles 804A - 804E display an animated preview of their corresponding aggregated content items. Further, in FIG. 8I, in response to the user input 808D, the electronic device 600 stops the display of the timer 802A and cancels the automatic playback of subsequent aggregated content items. In some embodiments, any user input received while the next content item user interface 800 is being displayed (e.g., user inputs 808A, 808B, 808C, 808D, 808E) causes the electronic device 600 to stop the display of the timer 802A and cancel the automatic playback of subsequent aggregated content items.

[0265] In FIG. 8E, the electronic device 600 detects a user input 808E corresponding to the selection of the sharing option 802C. In FIG. 8J, in response to the user input 808E, the electronic device 600 displays a shared user interface 810. The shared user interface 810 includes options 812A-812D. Different ones of the options 812A-812D correspond to different users or different groups of users, and the selection of an individual option 812A-812D initiates a process of sharing a first aggregated content item with the corresponding user or group of users associated with the selected option. The shared user interface 810 also includes options 814A-814D. Different ones of the options 814A-814D correspond to different communication media (e.g., option 814A corresponds to near field communication, option 814B corresponds to SMS messages and / or instant messaging, option 814C corresponds to email, and option 814D corresponds to instant messaging). The selection of an individual option 814A-814D initiates a process of sharing the first aggregated content item via the corresponding communication media associated with the selected option. The shared user interface 810 also includes a close option 816 that is selectable to terminate the display of the shared user interface 810 and optionally redisplay the next content item user interface 800.

[0266] FIG. 8K illustrates a scenario where the electronic device 600 determines that a first aggregated content item is too long to be shared (e.g., the playback duration of the first aggregated content item exceeds a threshold playback duration). In FIG. 8K, in response to user input 808E and in accordance with the determination that the first aggregated content item is too long to be shared, the electronic device 600 displays a notification 818 and options 820A, 820B, 820C. Option 820A is selectable to initiate a process of modifying the playback duration of the first aggregated content item (e.g., by removing and / or adding one or more media items from the first aggregated content item). Option 820B is selectable to initiate a process of modifying the audio track applied to the first aggregated content item (e.g., such that the user can select a shorter audio track that results in a shorter aggregated content item that can be shared). Option 820C is selectable to cancel the sharing operation and optionally redisplay the next content item user interface 800.

[0267] FIG. 8L shows another alternative scenario where the electronic device 600 determines that the first aggregated content item includes one or more media items not stored in the user's media library. For example, the first aggregated content item can include one or more media items that are available on the electronic device 600 and / or available on the electronic device 800 but not stored in the user's media library. In FIG. 8K, in response to the user input 808E and in accordance with the determination that the first aggregated content item includes one or more media items not stored in the user's media library and / or not stored locally on the electronic device 600, the electronic device 600 displays the notification 822, as well as the options 824A, 824B, and 824C. Option 824A is selectable to initiate a process of adding one or more media items to the user's media library and / or storing one or more media items locally on the electronic device 600. Option 824B is selectable to initiate a process of sharing the first aggregated content item without adding one or more media items to the user's media library and / or without storing one or more media items locally on the electronic device 600. Option 824C is selectable to cancel the sharing operation and optionally redisplay the next content item user interface 800.

[0268] FIG. 9 is a flowchart showing a method of managing content playback after playing a content item using a computer system according to some embodiments. Method 900 is executed in a computer system (e.g., 100, 300, 500) (e.g., a smartphone, a smartwatch, a tablet, a digital media player, a computer set-top entertainment box, a smart TV, and / or a computer system that controls an external display) that communicates with a display generation component (e.g., a display controller, a touch-sensing display system, and / or a display (e.g., integrated and / or connected)) and one or more input devices (e.g., a touch-sensing surface (e.g., a touch-sensing display), a mouse, a keyboard, and / or a remote control). Some operations of method 900 are optionally combined, the order of some operations is optionally changed, and some operations are optionally omitted.

[0269] As described below, method 900 provides an intuitive way to navigate and browse content items. This method reduces the user's cognitive burden when navigating and browsing content items, thereby creating a more efficient human-machine interface. In the case of a battery-operated computing device, the user can navigate and browse content items faster and more efficiently, thereby saving power and increasing the battery charging interval.

[0270] The computer system plays (902) the visual content of the first aggregated content item (e.g., media item 826Z of the first aggregated content item in FIG. 8A) via a display generation component (e.g., displays the visual content of the first aggregated content item via a display generation component) (e.g., a first video and / or a first content item automatically generated from a plurality of content items) (in some embodiments, the computer system plays the visual and audio content of the first aggregated content item), and the first aggregated content item includes a first plurality of content items (e.g., 628A, 862B, 862C, 862Z) selected (e.g., automatically and / or without user input) from a media library that includes photos and / or videos taken by a user of the computer system (e.g., 600) (e.g., scanned physical photos taken by the user using one or more cameras of the computer system's camera or other devices associated with the user and / or uploaded from a dedicated digital camera), and the first plurality of content items are selected based on a first set of selection criteria (e.g., the first aggregated content item represents an ordered sequence of a plurality of photos and / or videos and / or an automatically generated set of photos and / or videos (e.g., a set of photos and / or videos automatically aggregated and / or selected from a set of content items based on one or more shared characteristics)). In some embodiments, the plurality of photos and / or videos that make up the first plurality of content items are selected from a set of photos and / or videos associated with the computer system (e.g., stored in the computer system, associated with the user of the computer system, and / or associated with a user account associated with the computer system (e.g., signed in)).

[0271] While playing the visual content of the first aggregated content item (904), the computer system plays (906) audio content (e.g., FIG. 8A, audio track 3) (e.g., audio content separate from the first plurality of content items) (e.g., while the visual content of the first aggregated content item is being displayed via a display generation component, outputting and / or causing the output of an audio track (e.g., via one or more speakers, one or more headphones, and / or one or more earphones)) (e.g., audio content corresponding to and / or being part of the first aggregated content item (e.g., audio from one or more videos incorporated in the aggregated content item) and / or audio content separate from the first aggregated content item (e.g., an audio track superimposed on the first aggregated content item and / or played while the visual content of the first aggregated content item is being played and / or displayed)).

[0272] After playing at least a portion of the visual content of the first aggregated content item (908), the computer system detects (910) that the playing of the visual content of the first aggregated content item meets one or more end criteria (e.g., detecting that the playing of the first aggregated content item is complete, detecting that the playing of the first aggregated content item has exceeded a threshold playback time, and / or detecting that less than a threshold duration remains for the first aggregated content item).

[0273] After detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (912) (e.g., in response to detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria), according to a determination that the first set of playback conditions among one or more playback conditions is satisfied (914) (e.g., according to a determination that a threshold duration has elapsed since detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria, according to a determination that a threshold duration has elapsed since the playback of the visual content of the first aggregated content item was completed, and / or according to a determination that user input corresponding to a request to start the playback of the visual content of the second aggregated content item was received), the computer system plays the visual content of a second aggregated content item different from the first aggregated content item (e.g., FIG. 8C, media item 806 of the second aggregated content item) (e.g., a second video, a first video, and / or a second content item automatically generated from a plurality of content items) (916) (e.g., automatically and / or without user input) (e.g., and stops the playback of the visual content of the first aggregated content item). The second aggregated content item includes an ordered sequence of a second plurality of content items different from the first plurality of content items. Further, the second plurality of content items are selected from a media library including photos and / or videos taken by a user of the computer system, and the second plurality of content items are selected based on a second set of selection criteria (e.g., different from the first set of selection criteria) (e.g., the second aggregated content item represents an ordered sequence of a plurality of photos and / or videos and / or an automatically generated set of photos and / or videos (e.g., a set of photos and / or videos automatically aggregated and / or selected from a set of content items based on one or more shared characteristics)).In some embodiments, the plurality of photos and / or videos that make up the second plurality of content items are selected from a set of photos and / or videos associated with the computer system (e.g., stored in the computer system, associated with a user of the computer system, and / or associated with a user account associated with the computer system (e.g., signed in)) (e.g., from the same set of photos and / or videos from which the first plurality of content items of the first aggregated content item were selected). In accordance with a determination that the playback conditions are met, automatically playing the visual content of the second aggregated content item enables the user to view additional aggregated content items without the need for additional input.

[0274] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets (e.g., ends) one or more end criteria, and in accordance with a determination that a first set of playback conditions among one or more playback conditions is met, and / or while the visual content of the second aggregated content item is being played back, the computer system plays back (e.g., automatically and / or without user input) a second audio content (e.g., different from the audio content), (e.g., and stops the playback of the audio content that was being played back during the visual playback of the visual content of the first aggregated content item) (e.g., outputs and / or causes an output of an audio track (e.g., via one or more speakers, one or more headphones, and / or one or more earphones) while the visual content of the second aggregated content item is being displayed via a display generation component) (e.g., an audio content corresponding to and / or being part of the second aggregated content item (e.g., audio from one or more videos incorporated in the aggregated content item) and / or an audio content separate from the second aggregated content item (e.g., an audio track superimposed on the second aggregated content item and / or played back while the visual content of the second aggregated content item is being played back and / or displayed)).

[0275] In some embodiments, the computer system detects an image capture input (e.g., one or more tap inputs and / or one or more non-tap inputs) corresponding to a request to capture image data using a camera via one or more input devices, and in response to detecting the image capture input, the computer system adds a new content item (e.g., a new photo and / or a new video) (e.g., a new photo and / or a new video captured using the camera in response to detecting the image capture input) to a media library (e.g., media library user interface 604). By automatically adding a new content item to the media library in response to detecting the image capture input, it is possible to save the images captured without the user having to make additional inputs.

[0276] In some embodiments, before playing the visual content of a second aggregated content item (e.g., FIG. 8C) (in some embodiments, and after detecting that the playback of the visual content of a first aggregated content item meets one or more end criteria), the computer system displays, via a display generation component, a timer (e.g., 802A) that indicates a predetermined duration (e.g., counts down or counts up until it reaches 3 seconds, 5 seconds, 10 seconds, or 20 seconds). In some embodiments, the visual content of the second aggregated content item automatically starts playback after the timer counts down the predetermined duration (e.g., according to a determination that the timer has counted down the predetermined duration) (e.g., immediately after the timer has counted down the predetermined duration). By displaying a timer that counts down a predetermined duration before playing the visual content of the second aggregated content item, feedback regarding the current state of the device (e.g., that the visual content of the second aggregated content item will start playback after a predetermined duration) is provided to the user.

[0277] In some embodiments, while the timer (e.g., 802A) is being displayed, the computer system detects a first input (e.g., 808A, 808B, 808C, 808D, 808E) (e.g., a tap input and / or a non-tap input) via one or more input devices, and in response to detecting the first input, the computer system cancels the automatic playback of a second aggregated content item (e.g., FIGS. 8F - 8L) (e.g., determines that a first set of one or more playback conditions has not been met) (and optionally, stops displaying the timer). By canceling the automatic playback of the second aggregated content item in response to detecting the first input, feedback regarding the current state of the device (e.g., that the device detected the first input) is provided to the user.

[0278] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (e.g., in response to the detection), and (in some embodiments, before playing the visual content of the second aggregated content item), the computer system, via a display generation component, displays a first user interface object (e.g., 804A) corresponding to (e.g., uniquely corresponding to) the second aggregated content item (while, in some embodiments, the audio content continues to play). While the first user interface object is being displayed, the computer system detects, via one or more input devices, a second input (e.g., 808A) corresponding to the selection of the first user interface object (e.g., one or more tap inputs and / or one or more non-tap inputs). In response to detecting the second input, the computer system plays the visual content of the second aggregated content item (e.g., FIG. 8F) (e.g., without waiting for a first set of one or more playback conditions to be met and / or without waiting for a displayed countdown timer to expire). By displaying a first user interface object that is selectable to play the visual content of the second aggregated content item, the user is enabled to quickly select the next aggregated content item to be played, thereby reducing the number of inputs required to select the next aggregated content item.

[0279] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (e.g., in response to detecting), the computer system, via a display generation component, displays a first user interface object (e.g., 804A) corresponding to (e.g., uniquely corresponding to) the second aggregated content item. In some embodiments, the computer system displays the first user interface object while continuing to play the audio content. While the first user interface object is being displayed, the computer system, via one or more input devices, detects a third input (e.g., 808B, 808C, 808D, 808E) (e.g., one or more tap inputs and / or one or more non-tap inputs) that does not correspond to the selection of the first user interface object (e.g., at a position on the displayed user interface that does not correspond to the first user interface object) (e.g., that does not correspond to the selection of any user interface object). In response to detecting the third input, the computer system cancels the automatic playback of the visual content of the second aggregated content item (e.g., automatically stops playing) (e.g., FIGS. 8G-8L). In some embodiments, in response to detecting the third input, the computer system stops displaying the first user interface object. By canceling the automatic playback of the second aggregated content item in response to detecting the third input, the user is enabled to easily cancel the automatic playback of the second aggregated content item, thereby reducing the number of inputs required to cancel the automatic playback of the second aggregated content item.

[0280] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (e.g., in response to detecting), the computer system displays a replay user interface object (e.g., 802B) via a display generation component. In some embodiments, the computer system displays the replay user interface object while the audio content is still being played. In some embodiments, the computer system displays a first user interface object corresponding to a second aggregated content item (and selectable to initiate playback of the visual content of the second aggregated content item) simultaneously with the replay user interface object. While the replay user interface object is being displayed, the computer system detects a fourth input (e.g., 808B) (e.g., one or more tap inputs and / or one or more non-tap inputs) corresponding to the selection of the replay user interface object via one or more input devices. In response to detecting the fourth input, the computer system plays the visual content of the first aggregated content item from the beginning of the first aggregated content item (e.g., FIG. 8G) (e.g., replays the first aggregated content item). In some embodiments, the order of the content items of the aggregated content item is maintained between both the first playback and the replay. In some embodiments, additional content items other than the content items of the first aggregated content item are not played during the replay. By playing the visual content of the first aggregated content item in response to detecting the fourth input, the user is enabled to quickly replay the first aggregated content item, thereby reducing the number of inputs required to replay the first aggregated content item.

[0281] In some embodiments, a second aggregated content item (e.g., the Palm Springs 2017 of FIG. 8E) is selected from a plurality of aggregated content items based on selection criteria. In some embodiments, a computer system automatically selects content items included in the second aggregated content item. By automatically selecting the second aggregated content item based on the selection criteria, the quality of the suggestions to the user is improved, thereby providing a means for the user to make a selection. Otherwise, additional input is required to further find the desired content.

[0282] In some embodiments, before playing the visual content of the second aggregated content item (e.g., immediately before playing the visual content of the second aggregated content item), the computer system gradually pauses (e.g., fades) the playback of the audio content (e.g., track 3 of FIG. 8E). In some embodiments, the computer system pauses the playback of the audio content before playing the visual content of the second aggregated content item. By gradually pausing the playback of the audio content before playing the visual content of the second aggregated content item, feedback regarding the current state of the device (e.g., that the device will immediately start playing a new aggregated content item) is provided to the user. By gradually pausing the playback of the audio content, the audio content is played more quietly, reducing power usage and improving the battery life of the device.

[0283] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (e.g., in response to detecting), (e.g., in response to detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria), (in some embodiments, before playing the visual content of the second aggregated content item), the computer system, via a display generation component, displays a first user interface object (e.g., 804A) corresponding to (e.g., uniquely corresponding to) the second aggregated content item. In some embodiments, the computer system displays the first user interface object while continuing to play the audio content. While the first user interface object is being displayed, the computer system detects a fifth input (e.g., 808D) (e.g., one or more swipe inputs and / or one or more non-swipe inputs) via one or more input devices. In response to detecting the fifth input, the computer system, via the display generation component, displays a user interface object (e.g., 804D, 804E) corresponding to (e.g., uniquely corresponding to) a third aggregated content item that is different from the first aggregated content item and the second aggregated content item. The third aggregated content item includes an ordered sequence of a third plurality of content items that is different from the first plurality of content items and the second plurality of content items. Further, the third plurality of content items is selected from a media library that includes photos and / or videos taken by a user of the device, and the third plurality of content items is selected based on a third set of selection criteria (e.g., different from the first set of selection criteria and / or the second set of selection criteria). In some embodiments, the third aggregated content item represents an ordered sequence of multiple photos and / or videos, and / or an automatically generated set of photos and / or videos (e.g., a set of photos and / or videos automatically aggregated and / or selected from a set of content items based on one or more shared characteristics).In some embodiments, the plurality of photographs and / or videos that make up the third plurality of content items are selected from a set of photographs and / or videos associated with the computer system (e.g., stored in the computer system, associated with a user of the computer system, and / or associated with a user account associated with the computer system (e.g., signed in)) (e.g., selected from the same set of photographs and / or videos from which the first plurality of content items of the first aggregated content item were selected). In some embodiments, while displaying the second user interface object, the computer system detects user input corresponding to the selection of the second user interface object via one or more input devices, and in response to detecting user input corresponding to the selection of the second user interface object, the computer system plays the visual content of the third aggregated content item. By displaying a second user interface object corresponding to the third aggregated content item in response to detecting a fifth input, the user is enabled to quickly select the next content item to be played, thereby reducing the number of inputs required to select the next content item.

[0284] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (e.g., in response to the detection) (and optionally, before playing the visual content of the second aggregated content item), the computer system, via a display generation component (and optionally, while the audio content is still being played), simultaneously displays a first user interface object (e.g., 804A) corresponding to (e.g., uniquely corresponding to) the second aggregated content item, and a second user interface object (e.g., 804B - 804E) corresponding to (e.g., uniquely corresponding to) a third aggregated content item different from the first aggregated content item and the second aggregated content item. The third aggregated content item includes an ordered sequence of a third plurality of content items different from the first plurality of content items and the second plurality of content items. Further, the third plurality of content items is selected from a media library including photos and / or videos taken by a user of the device, and the third plurality of content items is selected based on a third set of selection criteria (e.g., different from the first set of selection criteria and / or the second set of selection criteria). By displaying the first user interface object corresponding to the second aggregated content item and the second user interface object corresponding to the third aggregated content item, the user is enabled to quickly select the next content item to be played, thereby reducing the number of inputs required to select the next content item.

[0285] In some embodiments, while simultaneously displaying a first user interface object (e.g., 804A) and a second user interface object (e.g., 804B - 804E), the computer system continues to play audio content (e.g., FIG. 8I, audio track 3). By continuing to play the audio content while displaying the first user interface object and the second user interface object, feedback regarding the current state of the device (e.g., that the selection of the next content item to be played has not yet been detected by the device) is provided to the user.

[0286] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (e.g., in response to the detection), and (in some embodiments, before playing the visual content of the second aggregated content item), the computer system, via the display generation component, at a first time, displays a first user interface object (e.g., 804A) corresponding to (e.g., uniquely corresponding to) the second aggregated content item. While (in some embodiments, while the audio content is still being played) the first user interface object is being displayed, displaying the first user interface object includes simultaneously displaying a first content item (e.g., the image of the user in the water in FIG. 8E) among the second plurality of content items within the second aggregated content item and title information corresponding to the second aggregated content item (e.g., "Palm Springs 2017" in FIG. 8E), (e.g., text information (e.g., a name, location information, and / or time information generated for the second aggregated content item)). In some embodiments, playing the visual content of the second aggregated content item includes, via the display generation component, at a second time after the first time, simultaneously displaying the first content item (e.g., 806) and the title information (e.g., 627) (while (in some embodiments, while no longer displaying the first user interface object)). Further, at the first time, the title information is displayed at a first position with respect to the first content item within the first user interface object, and at the second time, the title information is displayed at a second position with respect to the first content item, and the second position is different from the first position.In some embodiments, at a third time (e.g., a third time between the first time and the second time and / or a third time that is the second time), the computer system begins to cease displaying title information at a first position for a first content item (e.g., begins to gradually fade out the title information at the first position for the first content item) and begins to display title information at a second position for the first content item (e.g., begins to gradually fade in the title information at the second position for the first content item). In some embodiments, the computer system gradually fades out the display of title information at a first position for a first content item and gradually fades in the display of title information at a second position for the first content item (in some embodiments, at least a portion of the gradually fading out display of title information at the first position occurs simultaneously with at least a portion of the gradually fading in display of title information at the second position). By displaying title information at a first position for a first content item at a first time and displaying title information at a second position for the first content item at a second time, feedback regarding the current state of the device (e.g., that the device has started playing a second aggregated content item) is provided to the user.

[0287] In some embodiments, at a second time, the computer system displays title information (e.g., 627) in a first display area via a display generation component (e.g., FIG. 8C) (e.g., simultaneously displaying the first content item and the title information, where the title information is displayed in the first display area), and at a third time after the second time, the computer system displays the title information (e.g., 627) in a second display area different from the first display area via the display generation component (e.g., FIG. 8D) (e.g., displaying the title information in the second display area without displaying the first content item). In some embodiments, at the second time, the title information is displayed using a first set of visual parameters (e.g., font, color, and / or font size), and at the third time, the title information is displayed using a third set of visual parameters different from the first set. By displaying the title information corresponding to the second aggregated content item, feedback regarding the current state of the device (e.g., that the device has identified the title information corresponding to the second aggregated content item) is provided to the user.

[0288] In some embodiments, after detecting that the playback of the visual content of the first aggregated content item meets one or more end criteria (e.g., in response to the detection), the computer system, via a display generation component (and optionally, before playing the visual content of the second aggregated content item and / or while continuing to play the audio content), simultaneously displays a first user interface object (e.g., 804A) corresponding to (e.g., uniquely corresponding to) the second aggregated content item and a shareable user interface object (e.g., 802C) selectable to initiate a process of sharing the first aggregated content item (e.g., sharing the first aggregated content item via one or more communication media such as text messages, emails, short-range wireless communication and / or file transfer, upload to a shared media album, and / or upload to a third-party platform). In some embodiments, while simultaneously displaying the first user interface object and the shareable user interface object, the computer system detects, via one or more input devices, an input corresponding to the selection of the shareable user interface object, and in response to detecting the input, the computer system, via the display generation component, displays a sharing u...

Claims

1. A computer system in communication with a display generation component and one or more input devices, comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item including an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria; playing audio content separate from the first aggregate content item while playing the visual content of the first aggregate content item; detecting user input via the one or more input devices while playing the visual and audio content of the first aggregated content item; In response to detecting the user input, modifying the playing audio content while continuing to play the visual content of the first aggregated content item; A method comprising:

2. In response to detecting the user input, while continuing to play the visual content of the first aggregated content item, modifying visual parameters of the playback of the visual content of the first aggregated content item; The method of claim 1 further comprising:

3. playing the visual content of the first aggregated content item via the display generation component includes displaying the visual content with a first visual filter applied to a first region of the visual content; while continuing to play the visual content of the first aggregated content item, modifying the visual parameters of the playback of the visual content of the first aggregated content item includes displaying the visual content with a second visual filter applied to the first region of the visual content, the second visual filter being different from the first visual filter. The method of claim 2.

4. playing audio content separate from the first aggregated content item while playing the visual content of the first aggregated content item comprises playing a first audio track separate from the content item while playing the visual content of the first aggregated content item; while playing the first audio track, the visual content of the first aggregated content item is displayed with the first visual filter applied to the first region of the visual content; the first audio track is part of a first predefined combination with the first visual filter; modifying the playing audio content while continuing to play the visual content of the first aggregated content item comprises playing a second audio track separate from the content item and different from the first audio track while continuing to play the visual content of the first aggregated content item; while playing the second audio track, the visual content of the first aggregated content item is displayed with the second filter applied to the first region of the visual content; the second audio track is part of a second predefined combination with the second visual filter; the first predefined combination and the second predefined combination are part of a plurality of predefined combinations of filters and audio tracks; the plurality of predefined combinations of filters and audio tracks are arranged in an order; and the second predefined combination is selected to be adjacent to the first predefined combination in the order, and the first audio track is different from the second audio track; The method according to claim 3.

5. 5. The method of claim 4, wherein the first visual filter is selected to be part of the first predefined combination with the first audio track based on one or more audio characteristics of the first audio track and one or more visual characteristics of the first visual filter.

6. playing the visual content of the first aggregated content item; via said display generation component, the visual content with the first visual filter applied to the first region of the visual content, the first region including a central displayed portion of the visual content; the visual content with the second visual filter applied to a second region of the visual content that is different from the first region, the second region including a first edge of the visual content; 6. The method of claim 3, further comprising simultaneously displaying

7. playing the visual content of the first aggregated content item; while simultaneously displaying the visual content with the first visual filter applied to the first region and the second visual filter applied to the second region, displaying, via the display generation component, the visual content with a third visual filter different from the first visual filter and the second visual filter applied to a third region of the visual content different from the first region and the second region, the third region including a second edge of the visual content different from the first edge; The method of claim 6 further comprising:

8. playing the visual content of the first aggregated content item includes applying a transition of a first visual transition type to the visual content of the first aggregated content item; while continuing to play the visual content of the first aggregated content item, modifying the visual parameters of the playback of the visual content of the first aggregated content item includes modifying the transition to a second visual transition type different from the first visual transition type.

8. The method according to any one of claims 2 to 7.

9. the first visual transition type is selected from a plurality of visual transition types based on the audio content being played prior to detecting the user input; the second visual transition type is selected from the plurality of visual transition types based on audio content being played after detecting the user input. The method according to claim 8.

10. the first visual transition type is selected from a first set of visual transition types based on a tempo of the audio content being played prior to detecting the user input; the second visual transition type is selected from a second set of visual transition types different from the first set based on a tempo of the audio content being played after detecting the user input.

10. The method according to claim 8 or 9.

11. playing the visual content of the first aggregated content item; displaying the visual content with a first set of visual parameters applied to a first region of the visual content; while simultaneously displaying the visual content with the first visual filter applied to the first region, displaying the visual content with a second set of visual parameters applied to a second region of the visual content distinct from and adjacent to the first region, the second set of visual parameters being distinct from the first set of visual parameters; displaying a divider between the first region and the second region; The method of any one of claims 2 to 10, comprising:

12. In response to detecting the user input, shifting the divider contemporaneously with the user input while continuing to play the visual content of the first aggregated content item and without shifting the visual content of the first aggregated content item; The method of claim 11 further comprising:

13. prior to detecting the user input, the first aggregated content item is configured to display a first content item of the first plurality of content items for a first duration; modifying the visual parameters of the playback of visual content of the first aggregated content item includes configuring the first aggregated content item to display the first content item for a second duration different from the first duration.

13. The method according to any one of claims 2 to 12.

14. The method of claim 1 , wherein the user input comprises a gesture.

15. 15. The method of claim 1, wherein modifying the audio content being played while continuing to play the visual content of the first aggregated content item comprises changing the audio content from a first audio track to a second audio track different from the first audio track while continuing to play the visual content of the first aggregated content item.

16. changing the audio content from the first audio track to the second audio track; pausing playback of the first audio track at a first playback position of the first audio track that is not a start position of the first audio track; beginning playback of the second audio track at a second playback position of the second audio track that is not a start position of the second audio track; 16. The method of claim 15, comprising:

17. detecting one or more duration setting inputs via the one or more input devices; modifying a duration of the first aggregated content item in response to detecting the one or more duration setting inputs; 17. The method of claim 1, further comprising:

18. Modifying the playing audio content while continuing to play the visual content of the first aggregated content item includes changing the audio content from a first audio track to a second audio track different from the first audio track while continuing to play the visual content of the first aggregated content item, the first audio track having a first duration and the second audio track having a second duration different from the first duration; The method further comprising: modifying a duration of the first aggregated content item based on the second duration in response to detecting the user input; 18. The method of claim 1, further comprising:

19. detecting one or more duration-matching inputs via the one or more input devices while playing the audio content; in response to detecting the one or more duration matching inputs and in accordance with determining that the audio content has a first duration, modifying a duration of the first aggregated content item from a second duration different from the first duration to the first duration; 19. The method of any one of claims 1 to 18, further comprising:

20. displaying, via the display generation component, a first selectable object selectable to display a plurality of visual filter options while playing the visual content of the first aggregated content item and the audio content separate from the content item; detecting, while displaying the first selectable object, a first selection input via the one or more input devices corresponding to a selection of the first selectable object; while continuing to play visual content of the first aggregated content item in response to detecting the first selection input; a first user interface object including a display of the continued playback of the visual content of the first aggregated content item with the first visual filter applied to the visual content; a second user interface object including a display of the continued playback of the visual content of the first aggregated content item with a second visual filter applied to the visual content, the second visual filter being different from the first visual filter; and displaying a visual filter selection user interface, including simultaneously displaying 20. The method of any one of claims 1 to 19, further comprising:

21. displaying, via the display generation component, a second selectable object selectable to display a plurality of audio track options while playing the visual content of the first aggregate content item and the audio content separate from the content item; detecting, while displaying the second selectable object, a second selection input via the one or more input devices corresponding to a selection of the second selectable object; in response to detecting the second selection input, an audio track selection user interface, a third user interface object corresponding to the first audio track, the third user interface object being selectable to initiate a process of applying the first audio track to the first aggregated content item; and a fourth user interface object corresponding to a second audio track different from the first audio track, the fourth user interface object being selectable to initiate a process of applying the second audio track to the first aggregated content item; and displaying an audio track selection user interface, 21. The method of any one of claims 1 to 20, further comprising:

22. the third user interface object includes a representation of a track title corresponding to the first audio track; the fourth user interface object includes a representation of a track title corresponding to the second audio track; 22. The method of claim 21.

23. detecting a third selection input via the one or more input devices while displaying the audio track selection user interface including the third user interface object and the fourth user interface object; In response to detecting the third selection input, playing the first audio track from a beginning of the first audio track in response to a determination that the third selection input corresponds to a selection of the third user interface object; playing the second audio track from a beginning of the second audio track in response to a determination that the third selection input corresponds to a selection of the fourth user interface object; 23. The method of claim 21 or 22, further comprising:

24. detecting, while displaying the audio track selection user interface including the third user interface object and the fourth user interface object, a fourth selection input via the one or more input devices corresponding to a selection of the third user interface object; In response to detecting the fourth selection input, pursuant to determining that a user of the computer system has not subscribed to an audio service, initiating a process for displaying a prompt for the user to subscribe to the audio service; 24. The method of any one of claims 21 to 23, further comprising:

25. detecting, while displaying the audio track selection user interface including the third user interface object and the fourth user interface object, a fifth selection input via the one or more input devices, corresponding to a selection of the third user interface object; In response to detecting the fifth selection input, initiating a process of displaying a preview user interface in accordance with a determination that a user of the computer system does not subscribe to an audio service, where displaying the preview user interface includes playing a preview of the first aggregated content item with the first audio track applied to the visual content of the first aggregated content item, where the preview user interface does not allow the user to share the preview and / or save the preview for later playback until the user subscribes to the audio service; 25. The method of any one of claims 21 to 24, further comprising:

26. displaying a fifth user interface object selectable to cause the computer system to enter an edit mode while playing the visual content of the first aggregate content item and the audio content separate from the content item; detecting a second user input via the one or more input devices after displaying the fifth user interface object; In response to detecting the second user input, modifying the playing audio content while continuing to play visual content of the first aggregate content item in accordance with determining that the computer system is in the edit mode; ceasing to modify the playing of the audio content in response to determining that the computer system is not in the edit mode; and 26. The method of any one of claims 1 to 25, further comprising:

27. displaying, via the display generation component, a sixth user interface object selectable to pause playback of the visual content of the first aggregated content item while playing the visual content of the first aggregated content item and the audio content separate from the content item and while displaying the fifth user interface object; detecting, while displaying the sixth selectable user interface object, a sixth selection input via the one or more input devices corresponding to a selection of the sixth user interface object; In response to detecting the sixth selection input, pausing playback of the visual content of the first aggregated content item; replacing display of the fifth user interface object with a seventh user interface object selectable to modify an aspect ratio of the visual content of the first aggregated content item; detecting, while displaying the seventh user interface object, a seventh selection input via the one or more input devices corresponding to a selection of the seventh user interface object; displaying, via the display generation component, the visual content of the first aggregated content item transitioning from being displayed in a first aspect ratio to being displayed in a second aspect ratio different from the first aspect ratio in response to detecting the seventh selection input; 27. The method of claim 26, further comprising:

28. detecting, while playing the visual content of the first aggregated content item, a pause input via the one or more input devices corresponding to a request to pause playing the visual content of the first aggregated content item; In response to detecting the pause input, pausing playback of the visual content of the first aggregated content item; displaying, via the display generation component, a video navigation user interface element for navigating through the visual content of the first aggregated content item; and 28. The method of any one of claims 1 to 27, further comprising:

29. Displaying the visual navigation user interface element comprises: a representation of a first content item of the first plurality of content items; and a representation of a second content item of the first plurality of content items; and 30. The method of claim 28, comprising simultaneously displaying:

30. In response to detecting the pause input, displaying a duration control option via said display generation component and concurrently with said visual navigation user interface element; detecting, while displaying the duration control options, a duration control input via the one or more input devices corresponding to a selection of the duration control option; in response to detecting the duration control input, via the display generation component, a first playback duration option corresponding to a first playback duration; a second playback duration option corresponding to a second playback duration different from the first playback duration; and 30. The method of claim 28 or 29, further comprising:

31. In response to detecting the pause input, displaying audio track control options via said display generation component and concurrently with said visual navigation user interface elements; detecting, while displaying the audio track control options, an audio track control input via the one or more input devices corresponding to a selection of the audio track control option; in response to detecting the audio track control input, via the display generation component, a first audio track option corresponding to the first audio track; a second audio track option corresponding to a second audio track different from the first audio track; and 31. The method of any one of claims 28 to 30, further comprising:

32. playing the visual content of the first aggregated content item; displaying, via the display generation component, a first content item of the first plurality of content items in the first aggregated content item at a first time; displaying, via the display generation component, first title information corresponding to the first content item concurrently with the first content item; displaying, via the display generation component, a second content item of the first plurality of content items in the first aggregated content item at a second time after the first time; displaying, via the display generation component, second title information corresponding to the second content item and different from the first title information, concurrently with the second content item; 32. The method of any one of claims 1 to 31, comprising:

33. detecting, during playback of the visual and audio content of the first aggregated content item, one or more visual parameter modification inputs via the one or more input devices; in response to detecting the one or more visual parameter modifying inputs; modifying the playback of the visual content of the first aggregated content item in a first manner in accordance with determining that the one or more visual parameter modification inputs correspond to a first gesture; modifying the playback of the visual content of the first aggregated content item in a second manner different from the first manner in accordance with a determination that the one or more visual parameter modification inputs correspond to a second gesture different from the first gesture; 33. The method of any one of claims 1 to 32, further comprising:

34. the first gesture is a long press gesture; modifying the playback of the visual content of the first aggregated content item in the first manner includes maintaining a display of a currently displayed content item during the long press gesture.

34. The method of claim 33.

35. detecting, via the one or more input devices, an end of the long press gesture while maintaining display of the currently displayed content item during the long press gesture; modifying a playback duration of one or more subsequent content items that are displayed after the currently displayed content item after detecting an end of the long press gesture; 35. The method of claim 34, further comprising:

36. the first gesture is a first tap gesture; modifying the playback of the visual content of the first aggregated content item in the first manner includes navigating to a previous content item in the ordered sequence of content items in the first aggregated content item; the second gesture is a second tap gesture different from the first tap gesture; modifying the playback of the visual content of the first aggregated content item in the second manner includes navigating to a next content item in the ordered sequence of content items in the first aggregated content item.

36. The method of claim 34 or 35.

37. the first gesture is a first swipe gesture; modifying the playback of the visual content of the first aggregated content item in the first manner includes navigating to a previous content item in the ordered sequence of content items in the first aggregated content item; the second gesture is a second swipe gesture different from the first swipe gesture; modifying the playback of the visual content of the first aggregated content item in the second manner includes navigating to a next content item in the ordered sequence of content items in the first aggregated content item.

37. The method of any one of claims 34 to 36.

38. modifying the playback of the visual content of the first aggregated content item in the first manner includes modifying the playback of the visual content of the first aggregated content item in the first manner while continuing to play the audio content separate from the content item; modifying the playback of the visual content of the first aggregated content item in the second manner includes modifying the playback of the visual content of the first aggregated content item in the second manner while continuing to play the audio content separate from the content item.

38. The method of any one of claims 34 to 37.

39. while displaying a first content item of the first aggregated content item via the display generation component; and in response to detecting the third user input, via the display generation component: a tagging option selectable to initiate a process of identifying people depicted in the first content item; a remove option selectable to initiate a process of removing, from the first aggregated content item, one or more content items depicting people that are also depicted in the first content item; simultaneously displaying 39. The method of any one of claims 1 to 38, further comprising:

40. 40. A non-transitory computer readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generating component and one or more input devices, the one or more programs including instructions for performing the method of any one of claims 1 to 39.

41. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 40. A computer system comprising: said one or more programs comprising instructions for performing the method of any one of claims 1 to 39.

42. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: Means for carrying out the method according to any one of claims 1 to 39, A computer system comprising:

43. 40. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs including instructions for performing the method of any one of claims 1 to 39.

44. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from the set of content items based on a first set of selection criteria; while playing the visual content of the first aggregate content item, playing audio content separate from the content item; detecting user input via the one or more input devices while playing the visual and audio content of the first aggregate content item; In response to detecting the user input, modifying the playing audio content while continuing to play the visual content of the first aggregated content item; A non-transitory computer-readable storage medium containing instructions.

45. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; wherein the one or more programs include: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from the set of content items based on a first set of selection criteria; while playing the visual content of the first aggregate content item, playing audio content separate from the content item; detecting user input via the one or more input devices while playing the visual and audio content of the first aggregate content item; In response to detecting the user input, modifying the playing audio content while continuing to play the visual content of the first aggregated content item; A computer system including instructions.

46. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: means for playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria; means for playing audio content separate from the first aggregate content item while playing the visual content of the first aggregate content item; means for detecting user input via the one or more input devices during playback of the visual and audio content of the first aggregated content item; In response to detecting the user input, means for modifying the playing audio content while continuing to play the visual content of the first aggregated content item; A computer system comprising:

47. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more sensors, the one or more programs comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from the set of content items based on a first set of selection criteria; while playing the visual content of the first aggregate content item, playing audio content separate from the content item; detecting user input via the one or more input devices while playing the visual and audio content of the first aggregate content item; In response to detecting the user input, modifying the playing audio content while continuing to play the visual content of the first aggregated content item; A computer program product comprising instructions.

48. A computer system in communication with a display generation component and one or more input devices, comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from a media library including photos and / or videos taken by a user of the computer system, the first plurality of content items selected based on a first set of selection criteria; playing audio content while playing the visual content of the first aggregated content item; detecting, after playing back at least a portion of the visual content of the first aggregated content item, that the playing back of the visual content of the first aggregated content item satisfies one or more termination criteria; after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; in accordance with a determination that a first set of playback conditions of one or more playback conditions are satisfied, playing visual content of a second aggregated content item different from the first aggregated content item, the second plurality of content items different from the first plurality of content items, the second plurality of content items further comprising an ordered sequence of the second plurality of content items selected from the media library including photos and / or videos taken by a user of the computer system and selected based on a second set of selection criteria; A method comprising:

49. detecting an image capture input via the one or more input devices corresponding to a request to capture image data using a camera; adding a new content item to the media library in response to detecting the image capture input; 49. The method of claim 48, further comprising:

50. displaying, via the display generation component, a timer indicating the progression to a predetermined duration prior to playing the visual content of the second aggregated content item; 50. The method of claim 48 or 49, further comprising:

51. detecting a first input via the one or more input devices while displaying the timer; canceling automatic playback of the second aggregated content item in response to detecting the first input; 51. The method of claim 50, further comprising:

52. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; displaying, via the display generation component, a first user interface object corresponding to the second aggregated content item; detecting, while displaying the first user interface object, a second corresponding to a selection of the first user interface object via the one or more input devices; in response to detecting the second input, playing visual content of the second aggregated content item; and 52. The method of any one of claims 48 to 51, further comprising:

53. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; displaying, via the display generation component, a first user interface object corresponding to the second aggregated content item; detecting, while displaying the first user interface object, a third input via the one or more input devices that does not correspond to a selection of the first user interface object; canceling automatic playback of visual content of the second aggregated content item in response to detecting the third input; and 53. The method of any one of claims 48 to 52, further comprising:

54. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; displaying, via said display generation component, a replay user interface object; detecting, while displaying the replay user interface object, a fourth input via the one or more input devices corresponding to a selection of the replay user interface object; in response to detecting the fourth input, playing visual content of the first aggregated content item from a beginning of the first aggregated content item; 54. The method of any one of claims 48 to 53, further comprising:

55. 55. The method of any one of claims 48 to 54, wherein the second aggregated content item is selected from a plurality of aggregated content items based on a selection criterion.

56. gradually ceasing playback of the audio content before playing visual content of the second aggregated content item; 56. The method of any one of claims 48 to 55, further comprising:

57. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; displaying, via the display generation component, a first user interface object corresponding to the second aggregated content item; detecting a fifth input via the one or more input devices while displaying the first user interface object; in response to detecting the fifth input, displaying, via the display generation component, a user interface object corresponding to a third aggregated content item distinct from the first aggregated content item and the second aggregated content item, the third plurality of content items distinct from the first plurality of content items and the second plurality of content items, the third plurality of content items further comprising an ordered sequence of the third plurality of content items selected from the media library including photos and / or videos taken by a user of the device and selected based on a third set of selection criteria; 57. The method of any one of claims 48 to 56, further comprising:

58. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; via said display generation component, a first user interface object corresponding to the second aggregated content item; and a second user interface object corresponding to a third aggregated content item distinct from the first aggregated content item and the second aggregated content item, the third plurality of content items distinct from the first plurality of content items and the second plurality of content items, the third plurality of content items further comprising an ordered sequence of the third plurality of content items selected from the media library including photos and / or videos taken by a user of the device, the third plurality of content items selected based on a third set of selection criteria; simultaneously displaying 58. The method of any one of claims 48 to 57, further comprising:

59. continuing to play the audio content while simultaneously displaying the first user interface object and the second user interface object; 60. The method of claim 58, further comprising:

60. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; via the display generation component, at a first time, a first user interface object corresponding to the second aggregated content item; a first content item of the second plurality of content items in the second aggregated content item; and title information corresponding to the second aggregated content item; and displaying a first user interface object, including simultaneously displaying Further comprising: playing visual content of the second aggregated content item includes simultaneously displaying, via the display generation component, the first content item and the title information at a second time after the first time; and At the first time, the title information is displayed in a first position relative to the first content item within the first user interface object; At the second time, the title information is displayed in a second location relative to the first content item, the second location being different from the first location.

60. The method of any one of claims 48 to 59.

61. playing visual content of the second aggregated content item; displaying, via the display generation component, the title information in a first display area at the second time; displaying, via the display generation component, the title information in a second display area distinct from the first display area at a third time subsequent to the second time; 61. The method of claim 60, further comprising:

62. via the display generation component after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; a first user interface object corresponding to the second aggregated content item; and a sharing user interface object selectable to initiate a process of sharing the first aggregated content item; simultaneously displaying 62. The method of any one of claims 48 to 61, further comprising:

63. detecting a sixth input via the one or more input devices while simultaneously displaying the first user interface object and the shared user interface object, the sixth input corresponding to a selection of the shared user interface object; In response to detecting the sixth input, displaying, via the display generation component, an indication that the audio content applied to the first aggregated content item is not authorized to be shared by the user of the computer system in accordance with a determination that the audio content applied to the first aggregated content item is not authorized to be shared by the user; 63. The method of claim 62, further comprising:

64. detecting a seventh input via the one or more input devices while simultaneously displaying the first user interface object and the shared user interface object, the seventh input corresponding to a selection of the shared user interface object; In response to detecting the seventh input, displaying, via the display generation component, selectable playback duration options to initiate a process of reducing a playback duration of the first aggregated content item in accordance with a determination that the audio content applied to the first aggregated content item is not authorized to be shared by a user of the computer system; 64. The method of claim 62 or 63, further comprising:

65. detecting, while simultaneously displaying the first user interface object and the shared user interface object, an eighth input via the one or more input devices corresponding to a selection of the shared user interface object; In response to detecting the eighth input, displaying, via the display generation component, selectable audio content options to initiate a process of selecting different audio content to be applied to the first aggregated content item in accordance with a determination that the audio content applied to the first aggregated content item is not authorized to be shared by a user of the computer system; 65. The method of any one of claims 62 to 64, further comprising:

66. detecting a ninth input via the one or more input devices while simultaneously displaying the first user interface object and the shared user interface object, the ninth input corresponding to a selection of the shared user interface object; In response to detecting the ninth input, displaying, via the display generation component, a selectable synchronization option to initiate a process of saving the first content item to the media library in accordance with a determination that the first plurality of content items in the first aggregated content item includes a first content item that is not stored locally on the computer system; 66. The method of any one of claims 62 to 65, further comprising:

67. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; displaying, via the display generation component, a preview object that displays an animated preview of visual content of the second aggregated content item; 67. The method of any one of claims 48 to 66, further comprising:

68. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; displaying, via the display generation component, a location object corresponding to a geographic location and selectable for displaying one or more aggregated content item options corresponding to the geographic location; 68. The method of any one of claims 48 to 67, further comprising:

69. after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; displaying, via the display generation component, a first person object corresponding to the first person and selectable for displaying one or more aggregate content item options corresponding to the first person; 69. The method of any one of claims 48 to 68, further comprising:

70. displaying, via the display generation component, a media library user interface; In response to determining that a first setting is enabled, the media library user interface: a plurality of aggregated content items including the first aggregated content item; the media library including photos and / or videos taken by the user of the computer system; Provides access to in accordance with determining that the first setting is disabled, the media library user interface provides access to the plurality of aggregated content items without providing access to the media library including photos and / or videos taken by the user of the computer system.

70. The method of any one of claims 48 to 69.

71. 71. A non-transitory computer readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generating component and one or more input devices, the one or more programs including instructions for performing the method of any one of claims 48 to 70.

72. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 71. A computer system comprising: said one or more programs comprising instructions for performing the method of any one of claims 48 to 70.

73. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: Means for carrying out the method according to any one of claims 48 to 70, A computer system comprising:

74. 71. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs including instructions for performing the method of any one of claims 48 to 70.

75. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from a media library including photos and / or videos taken by a user of the computer system, the first plurality of content items selected based on a first set of selection criteria; playing audio content while playing the visual content of the first aggregated content item; detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria after playing at least a portion of the visual content of the first aggregated content item; after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; in accordance with a determination that a first set of playback conditions of one or more playback conditions are satisfied, playing visual content of a second aggregated content item different from the first aggregated content item, the second plurality of content items different from the first plurality of content items, the second plurality of content items further comprising an ordered sequence of the second plurality of content items selected from the media library including photos and / or videos taken by a user of the computer system and selected based on a second set of selection criteria; A non-transitory computer-readable storage medium containing instructions.

76. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; wherein the one or more programs include: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from a media library including photos and / or videos taken by a user of the computer system, the first plurality of content items selected based on a first set of selection criteria; playing audio content while playing the visual content of the first aggregated content item; detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria after playing at least a portion of the visual content of the first aggregated content item; after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; in accordance with a determination that a first set of playback conditions of one or more playback conditions are satisfied, playing visual content of a second aggregated content item different from the first aggregated content item, the second plurality of content items different from the first plurality of content items, the second plurality of content items further comprising an ordered sequence of the second plurality of content items selected from the media library including photos and / or videos taken by a user of the computer system and selected based on a second set of selection criteria; A computer system including instructions.

77. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: means for playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from a media library including photos and / or videos taken by a user of the computer system, the first plurality of content items selected based on a first set of selection criteria; means for playing audio content while playing the visual content of the first aggregated content item; means for detecting, after playing back at least a portion of the visual content of the first aggregated content item, that the playing of the visual content of the first aggregated content item satisfies one or more termination criteria; after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; means for playing visual content of a second aggregated content item different from the first aggregated content item in accordance with a determination that a first set of playback conditions of the one or more playback conditions are satisfied, the second aggregated content item comprising an ordered sequence of a second plurality of content items different from the first plurality of content items, the second plurality of content items being selected from the media library including photos and / or videos taken by a user of the computer system and selected based on a second set of selection criteria; A computer system comprising:

78. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from a media library including photos and / or videos taken by a user of the computer system, the first plurality of content items selected based on a first set of selection criteria; playing audio content while playing the visual content of the first aggregated content item; detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria after playing at least a portion of the visual content of the first aggregated content item; after detecting that the playback of the visual content of the first aggregated content item satisfies one or more termination criteria; in accordance with a determination that a first set of playback conditions of one or more playback conditions are satisfied, playing visual content of a second aggregated content item different from the first aggregated content item, the second plurality of content items different from the first plurality of content items, the second plurality of content items further comprising an ordered sequence of the second plurality of content items selected from the media library including photos and / or videos taken by a user of the computer system and selected based on a second set of selection criteria; A computer program product comprising instructions.

79. A computer system in communication with a display generation component and one or more input devices, comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item including an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria; detecting user input via the one or more input devices while playing the visual content of the first aggregated content item; In response to detecting the user input, pausing playback of the visual content of the first aggregated content item; via said display generation component, a first representation of a first content item of the first plurality of content items; a second representation of a second content item of the first plurality of content items; and displaying a user interface including simultaneously displaying multiple representations of content items in the first plurality of content items; A method comprising:

80. the first content item corresponds to a first playback position of the first aggregated content item; the second content item corresponds to a second playback position of the first aggregated content item that is different from the first playback position; The method further comprising: detecting a selection input via the one or more input devices while simultaneously displaying the first representation of the first content item and the second representation of the second content item; In response to detecting the selection input, playback of visual content of the first aggregated content item from the first playback position in response to a determination that the selection input corresponds to a selection of the first representation of the first content item; in response to determining that the selection input corresponds to a selection of the second representation of the second content item, playing visual content of the first aggregated content item from the second playback position; 80. The method of claim 79, further comprising:

81. displaying, via the display generation component, a selectable add content option for initiating a process of adding one or more content items to the first aggregated content item; detecting, while displaying the add content option, a second selection input via the one or more input devices corresponding to a selection of the add content option; in response to detecting the second selection input, via the display generation component, a third representation of the third content item; and a fourth representation of the fourth content item; and displaying representations of a plurality of content items not included in the first aggregated content item, including simultaneously displaying detecting, while simultaneously displaying the third representation of the third content item and the fourth representation of the fourth content item, a first set of inputs via the one or more input devices corresponding to a request to add the third content item to the first aggregated content item; modifying the first aggregated content item to include the third content item in response to detecting the first set of inputs; 81. The method of claim 79 or 80, further comprising:

82. Displaying the third representation of the third content item displaying the representation of the third content item in a first manner in accordance with a determination that the third content item satisfies one or more relevance criteria with respect to the first aggregated content item; displaying the representation of the third content item in a second manner, different from the first manner, in accordance with a determination that the third content item does not satisfy the one or more relevance criteria for the first aggregated content item; Including, Displaying the fourth representation of the fourth content item displaying the representation of the fourth content item in the first manner in accordance with a determination that the fourth content item satisfies the one or more relevance criteria with respect to the first aggregated content item; displaying the representation of the fourth content item in the second manner in accordance with a determination that the fourth content item does not satisfy the one or more relevance criteria with respect to the first aggregated content item; and Including, 82. The method of claim 81.

83. displaying, via the display generation component, related content options selectable to initiate a process of displaying additional content related to the first aggregated content item; detecting, while displaying the related content options, a third selection input via the one or more input devices corresponding to a selection of the related content option; In response to detecting the third selection input, displaying, via the display generation component, a representation of a fifth content item of a media library in accordance with a determination that the fifth content item is not included in the first aggregated content item and satisfies one or more relevance criteria for the first aggregated content item; withdrawing from displaying the representation of the fifth content item in the media library in accordance with a determination that the fifth content item does not satisfy the one or more relevance criteria with respect to the first aggregate content item; 83. The method of any one of claims 79 to 82, further comprising:

84. detecting, while displaying the representations of content items in the first plurality of content items, a fourth selection input via the one or more input devices corresponding to a selection of one or more content items of the first plurality of content items that includes the first content item; responsive to detecting the fourth selection input, displaying, via the display generation component, selectable sharing options for initiating a process of sharing the selected one or more content items via one or more communication media; 84. The method of any one of claims 79 to 83, further comprising:

85. detecting, while displaying the representations of content items in the first plurality of content items, a fifth selection input via the one or more input devices corresponding to a selection of one or more content items of the first plurality of content items that includes the first content item; in response to detecting the fifth selection input, displaying, via the display generation component, a selectable remove option to initiate a process of removing the selected one or more content items from the first aggregated content item; 85. The method of any one of claims 79 to 84, further comprising:

86. prior to displaying the user interface, the first content item is positioned in a first consecutive position within the ordered sequence of the first plurality of content items; displaying the user interface includes displaying the first representation of the first content item at a first display location corresponding to the first sequential location; The method further comprising: detecting, while displaying the representations of a content item in the first plurality of content items, a gesture corresponding to the first representation of the first content item via the one or more input devices; In response to detecting the gesture, moving the first representation of the first content item from the first display location to a second display location different from the first display location, the second display location corresponding to a second consecutive location in the ordered sequence of the first plurality of content items; reordering the ordered sequence of the first plurality of content items, including moving the first content item from the first consecutive location to the second consecutive location; modifying the first aggregated content item based on the reordering of the ordered sequence of the first plurality of content items; 86. The method of any one of claims 79 to 85, further comprising:

87. detecting a set of user inputs via the one or more input devices while displaying the user interface; In response to detecting said set of user inputs, via said display generation component, a first content length option corresponding to a first number of content items, wherein displaying the first content length option includes displaying the first number of content items; and a second content length option corresponding to a second number of content items different from the first number of content items, where displaying the second content length option includes displaying the second number of content items; and and 87. The method of any one of claims 79 to 86, further comprising:

88. detecting a second set of user inputs via the one or more input devices while displaying the user interface; In response to detecting the second set of user inputs, via said display generation component, a third content length option corresponding to the first number of content items; and a fourth content length option corresponding to a second number of content items different from the first number of content items; and and detecting a sixth selection input via the one or more input devices while simultaneously displaying the third content length option and the fourth content length option; In response to detecting the sixth selection input, modifying the user interface to display representations of the first number of content items in accordance with a determination that the sixth selection input corresponds to a selection of the third content length option; modifying the user interface to display representations of the second number of content items in accordance with a determination that the sixth selection input corresponds to a selection of the fourth content length option; 88. The method of any one of claims 79 to 87, further comprising:

89. detecting, after displaying the user interface, a third set of inputs via the one or more input devices corresponding to requests to add a first set of one or more additional content items to the first aggregated content item and / or remove a first set of one or more removed content items from the first aggregated content item; modifying the first aggregated content item to include the one or more additional content items and / or modifying the first aggregated content item to exclude one or more removed content items in response to detecting the third set of inputs; detecting a fourth set of inputs via the one or more input devices corresponding to a request to modify a duration of the first aggregated content item after modifying the first aggregated content item to include the one or more additional content items; modifying a duration of the first aggregated content item, including adding a second set of one or more additional content items to the first aggregated content item based on the modification of a duration of the first aggregated content item and / or removing a second set of one or more removed content items from the first aggregated content item based on the modification of a duration of the first aggregated content item in response to detecting the fourth set of inputs; 89. The method of any one of claims 79 to 88, further comprising:

90. 90. The method of claim 89, comprising: in response to detecting the fourth set of inputs, in accordance with a determination that the fourth set of inputs includes a request to decrease the duration of the first aggregated content item, reducing the duration of the first aggregated content item by removing the second set of removed content items without removing any of the first set of one or more additional content items.

91. in response to detecting the fourth set of inputs, in accordance with a determination that the fourth set of inputs includes a request to increase the duration of the first aggregated content item, increasing the duration of the first aggregated content item by adding the second set of additional content items without adding any of the first set of one or more removed content items; 91. The method of claim 89 or 90, comprising:

92. playing the visual content of the first aggregated content item; displaying the first content item via the display generation component; a transition from the first content item to the second content item via the display generation component after displaying the first content item, in accordance with a determination that the second content item satisfies one or more similarity criteria with respect to the first content item, the transition from the first content item to the second content item is of a first visual transition type; in accordance with a determination that the second content item does not satisfy the one or more similarity criteria with respect to the first content item, the transition from the first content item to the second content item is of a second visual transition type different from the first visual transition type. displaying a transition from the first content item to the second content item; Including, 92. The method of any one of claims 79 to 91.

93. 93. The method of claim 92, wherein the one or more similarity criteria include one or more of a time-based similarity criterion, a location-based similarity criterion, and / or a content-based similarity criterion.

94. the transition from the first content item to the subsequent content item is of the first visual transition type; playing the visual content of the first aggregated content item; displaying the transition from the first content item to the second content item via the display generation component; and a transition from the second content item to a third content item, different from the first and second content items, via the display generation component after displaying the second content item; in accordance with a determination that the third content item satisfies one or more similarity criteria with respect to the second content item, the transition from the second content item to the third content item being of the first visual transition type. Displaying the transitions; and Further comprising:

94. The method of claim 92 or 93.

95. Displaying the transition from the second content item to the third content item includes: in accordance with a determination that the third content item does not satisfy the one or more similarity criteria with respect to the second content item, the transition from the first content item to the second content item being of a third visual transition type different from the first visual transition type; 95. The method of claim 94, further comprising:

96. playing audio content separate from the first plurality of content items while playing the visual content of the first aggregated content item; 96. The method of any one of claims 79 to 95, further comprising:

97. Displaying the user interface via the display generation component comprises: a paused visual content of the first aggregated content item; and a video navigation user interface element for navigating through the visual content of the first aggregated content item, wherein the first representation of the first content item and the second representation of the second content item are displayed simultaneously as part of the video navigation user interface element; including simultaneously displaying 97. The method of any one of claims 79 to 96.

98. detecting a first set of navigation inputs via the one or more input devices while displaying the video navigation user interface element, the first representation of the first content item and the second representation of the second content item being simultaneously displayed; in response to detecting the first set of navigation inputs; at a first time, via the display generation component; a first representation of the first content item in a first manner; a second representation of the second content item in a second manner different from the first manner; and and at a second time after the first time, via the display generation component; the first representation of the first content item in the second format; and the second representation of the second content item in the first format; and and 98. The method of claim 97, further comprising:

99. prior to detecting the one or more navigation inputs, via the display generation component; the first representation of the first content item in the first format; the second representation of the second content item in the first format; and simultaneously displaying 99. The method of claim 98, further comprising:

100. ceasing display of the video navigation user interface element in accordance with a determination that one or more fade criteria have been met after simultaneously displaying the paused visual content of the first aggregated content item and the video navigation user interface element; 100. The method of any one of claims 97 to 99, further comprising:

101. detecting a second set of one or more navigation inputs via the one or more input devices while simultaneously displaying the paused visual content of the first aggregated content item and the video navigation user interface element; in response to detecting the second set of one or more navigation inputs; replacing the display of the paused visual content of the first aggregated content item with a display of the first content item in response to detecting a first portion of the second set of one or more navigation inputs; in response to detecting a second portion of the second set of one or more navigation inputs, replacing a display of the first content item with a display of the second content item; 101. The method of any one of claims 97 to 100, further comprising:

102. 102. A non-transitory computer readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generating component and one or more input devices, the one or more programs including instructions for performing the method of any one of claims 79 to 101.

103. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 102. A computer system comprising: said one or more programs comprising instructions for performing the method of any one of claims 79 to 101.

104. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: Means for carrying out the method according to any one of claims 79 to 101, A computer system comprising:

105. 102. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs including instructions for performing the method of any one of claims 79 to 101.

106. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from the set of content items based on a first set of selection criteria; detecting user input via the one or more input devices while playing the visual content of the first aggregated content item; In response to detecting the user input, pausing the playback of the visual content of the first aggregated content item; via said display generation component, a first representation of a first content item of the first plurality of content items; a second representation of a second content item of the first plurality of content items; and displaying a user interface including simultaneously displaying multiple representations of content items in the first plurality of content items. A non-transitory computer-readable storage medium containing instructions.

107. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; wherein the one or more programs include: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from the set of content items based on a first set of selection criteria; detecting user input via the one or more input devices while playing the visual content of the first aggregated content item; In response to detecting the user input, pausing the playback of the visual content of the first aggregated content item; via said display generation component, a first representation of a first content item of the first plurality of content items; a second representation of a second content item of the first plurality of content items; and displaying a user interface including simultaneously displaying multiple representations of content items in the first plurality of content items. Including instructions, Computer system.

108. 1. A computer system configured to communicate with a display generation component and one or more input devices, comprising: means for playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from a set of content items based on a first set of selection criteria; means for detecting user input via the one or more input devices while playing the visual content of the first aggregated content item; In response to detecting the user input, pausing the playback of the visual content of the first aggregated content item; via said display generation component, a first representation of a first content item of the first plurality of content items; a second representation of a second content item of the first plurality of content items; and displaying a user interface including simultaneously displaying multiple representations of content items in the first plurality of content items. Means, A computer system comprising:

109. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more sensors, the one or more programs comprising: playing, via the display generation component, visual content of a first aggregated content item, the first aggregated content item comprising an ordered sequence of a first plurality of content items selected from the set of content items based on a first set of selection criteria; detecting user input via the one or more input devices while playing the visual content of the first aggregated content item; In response to detecting the user input, pausing the playback of the visual content of the first aggregated content item; via said display generation component, a first representation of a first content item of the first plurality of content items; a second representation of a second content item of the first plurality of content items; and displaying a user interface including simultaneously displaying multiple representations of content items in the first plurality of content items. A computer program product comprising instructions.

Citation Information

Patent Citations

  • System and method for defining frame-accurate images for media asset management

    JP2010524124A

  • Display device, display method and display program

    JP2011003977A