User interfaces for managing visual content in media

The method addresses inefficiencies in managing visual content by integrating text detection and capture affordances in media interfaces, enhancing user experience and power conservation.

JP2025121968AActive Publication Date: 2025-08-20APPLE INC
View PDF 5 Cites 0 Cited by

Patent Information

Application Number
JP2025076835
Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
Priority Date
2022-03-10
Filing Date
2025-05-02
Publication Date
2025-08-20
Estimated Expiration
2042-04-15

Smart Images

  • Figure 2025121968000001_ABST
    Figure 2025121968000001_ABST
Patent Text Reader

Abstract

To provide electronic devices with faster, more efficient methods and interfaces for managing a visual content in media.SOLUTION: A method includes: displaying a camera user interface that includes concurrently displaying a representation of media and a media capture affordance; in accordance with a determination that an individual set of criteria is satisfied including a criterion that is satisfied when an individual text is detected in the representation of the media, displaying a first user interface object corresponding to one or more text management operations; detecting a first input directed to the camera user interface; in accordance with a determination that the first input corresponds to selection of the media capture affordance, initiating capture of media to be added to a media library; and displaying a plurality of options to manage the individual text.SELECTED DRAWING: Figure 6A
Need to check novelty before this filing date? Find Prior Art

Description

[Technical Field]

[0001] (CROSS-REFERENCE TO RELATED APPLICATIONS) This application is a continuation of U.S. patent application Ser. No. 63 / 176,847, filed April 19, 2021, entitled "USER INTERFACES FOR MANAGING VISUAL CONTENT IN MEDIA," U.S. patent application Ser. No. 63 / 197,497, filed June 6, 2021, entitled "USER INTERFACES FOR MANAGING VISUAL CONTENT IN MEDIA," U.S. patent application Ser. No. 17 / 484,844, filed September 24, 2021, entitled "USER INTERFACES FOR MANAGING VISUAL CONTENT IN MEDIA," U.S. patent application Ser. No. 17 / 484,714, filed September 24, 2021, entitled "USER INTERFACES FOR MANAGING VISUAL CONTENT IN MEDIA," and U.S. patent application Ser. No. 17 / 484,714, filed September 24, 2021, entitled "USER INTERFACES FOR MANAGING VISUAL CONTENT IN MEDIA." This application claims priority to U.S. patent application Ser. No. 17 / 484,856, entitled "USER INTERFACES FOR MANAGING VISUAL CONTENT IN MEDIA," filed on March 10, 2022, and U.S. patent application Ser. No. 63 / 318,677, entitled "USER INTERFACES FOR MANAGING VISUAL CONTENT IN MEDIA," filed on March 10, 2022, the contents of which are incorporated herein by reference in their entireties.

[0002] TECHNICAL FIELD This disclosure relates generally to computer user interfaces, and more particularly to techniques for managing visual content in media. [Background technology]

[0003] Smartphones and other personal electronic devices allow users to capture and view content in media. Users can capture various types of media, including video and image data. Users can store the captured media on their smartphones or other personal electronic devices. Summary of the Invention

[0004] However, some techniques for managing content in media using computer systems are generally cumbersome and inefficient. For example, some existing techniques use complex and time-consuming user interfaces that may involve multiple key presses or keystrokes. Existing techniques take more time than necessary, wasting the user's time and the device's energy. This latter consideration is particularly important in battery-operated devices.

[0005] The present technology thus provides electronic devices with faster and more efficient methods and interfaces for managing visual content in media. Such methods and interfaces optionally supplement or replace other methods for managing visual content in media. Such methods and interfaces reduce the cognitive burden on users and create more efficient human-machine interfaces. For battery-operated computing devices, such methods and interfaces conserve power and extend the time between battery charges.

[0006] According to some embodiments, a method is described that is executed on a computer system in communication with a display generation component. The method includes: displaying, via the display generation component, a camera user interface, the camera user interface including simultaneously displaying a representation of media and a media capture affordance; displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied while simultaneously displaying the representation of media and the media capture affordance, the first user interface object corresponding to one or more text management operations; and withholding display of the first user interface object in accordance with a determination that the respective set of criteria is not satisfied; detecting, while displaying the representation of media, a first input directed at the camera user interface; and, in response to detecting the first input directed at the camera user interface, initiating capture of media to be added to a media library associated with the computer system in accordance with a determination that the first input corresponds to selection of the media capture affordance; and displaying, via the display generation component, a plurality of options for managing the respective text in accordance with a determination that the first input corresponds to selection of the first user interface object.

[0007] According to some embodiments, a non-transitory computer-readable storage medium is described that stores one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs configured to: display, via the display generation component, a camera user interface, the camera user interface including simultaneously displaying a representation of media and a media capture affordance; and perform, via the display generation component, one or more text management operations in accordance with a determination that a respective set of criteria is satisfied while simultaneously displaying the representation of media and the media capture affordance, the respective set of criteria including criteria satisfied when a respective text is detected within the representation of media. The method includes instructions for displaying a corresponding first user interface object, withholding display of the first user interface object in accordance with a determination that the respective set of criteria is not met, and detecting a first input directed at the camera user interface while displaying the representation of the media; in response to detecting the first input directed at the camera user interface, initiating capture of media to be added to a media library associated with the computer system in accordance with a determination that the first input corresponds to selection of a media capture affordance; and displaying, via the display generation component, a plurality of options for managing the respective text in accordance with a determination that the first input corresponds to selection of the first user interface object.

[0008] According to some embodiments, a transient computer-readable storage medium is described, the transient computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs configured to: display, via the display generation component, a camera user interface, the camera user interface including simultaneously displaying a representation of media and a media capture affordance; and perform, via the display generation component, one or more text management operations in response to a determination that a respective set of criteria is satisfied while simultaneously displaying the representation of media and the media capture affordance, the respective set of criteria including criteria satisfied when a respective set of text is detected within the representation of media. and, in accordance with a determination that the respective set of criteria is not met, forgoing display of the first user interface object; and, while displaying the representation of the media, detecting a first input directed at the camera user interface; in response to detecting the first input directed at the camera user interface, initiating capture of media to be added to a media library associated with the computer system in accordance with a determination that the first input corresponds to a selection of a media capture affordance; and, in accordance with a determination that the first input corresponds to a selection of the first user interface object, displaying, via the display generation component, a plurality of options for managing the respective text.

[0009] According to some embodiments, a computer system configured to communicate with a display generation component is described, the computer system comprising one or more processors and a memory storing one or more programs configured to be executed by the one or more processors, the one or more programs being configured to: display, via the display generation component, a camera user interface, the camera user interface including simultaneously displaying a representation of media and a media capture affordance; and, pursuant to a determination that a respective set of criteria is satisfied while simultaneously displaying the representation of media and the media capture affordance, generate, via the display generation component, a first user interface object corresponding to one or more text management operations, the first user interface object corresponding to one or more text management operations. and displaying, via the display generation component, a plurality of options for managing the respective text, wherein the respective set of criteria is met; withholding display of the first user interface object in accordance with a determination that the respective set of criteria is not met; detecting a first input directed at the camera user interface while displaying the representation of the media; and in response to detecting the first input directed at the camera user interface, initiating capture of media to be added to a media library associated with the computer system in accordance with a determination that the first input corresponds to selection of a media capture affordance; and displaying, via the display generation component, a plurality of options for managing the respective text in accordance with a determination that the first input corresponds to selection of the first user interface object.

[0010] According to some embodiments, a computer system configured to be in communication with a display generation component is described. The computer system includes one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; means for displaying, via a display generation component, a camera user interface, the camera user interface including simultaneously displaying a representation of media and a media capture affordance; means for displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied while simultaneously displaying the representation of media and the media capture affordance, the first user interface object corresponding to one or more text management operations, the first user interface object including criteria that are satisfied when respective text is detected within the representation of media, and for withholding display of the first user interface object in accordance with a determination that the respective set of criteria is not satisfied; means for detecting, while displaying the representation of media, a first input directed at the camera user interface; and means for, in response to detecting the first input directed at the camera user interface, initiating capture of media to be added to a media library associated with the computer system in accordance with a determination that the first input corresponds to selection of a media capture affordance, and displaying, via the display generation component, a plurality of options for managing the respective text in accordance with a determination that the first input corresponds to selection of the first user interface object.

[0011] According to some embodiments, a computer program product is described, comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component. the one or more programs include instructions for: displaying, via a display generation component, a camera user interface, the camera user interface including simultaneously displaying a representation of media and a media capture affordance; displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied while simultaneously displaying the representation of media and the media capture affordance, the first user interface object corresponding to one or more text management operations in accordance with a determination that the respective set of criteria is not satisfied; detecting, while displaying the representation of media, a first input directed at the camera user interface; and in response to detecting the first input directed at the camera user interface, initiating capture of media to be added to a media library associated with the computer system in accordance with a determination that the first input corresponds to selecting a media capture affordance; and displaying, via the display generation component, a plurality of options for managing the respective text in accordance with a determination that the first input corresponds to selecting the first user interface object.

[0012] According to some embodiments, a method is described. The method is executed on a computer system in communication with a display generation component and one or more input devices. The method includes: displaying, via the display generation component, a first representation of a previously captured media item; detecting, while displaying the first representation of the previously captured media item, input, via the one or more input devices, corresponding to a request to display a second representation of the previously captured media item; displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to the request to display the second representation of the previously captured media item; and, while displaying the second representation of the previously captured media item, displaying, via the display generation component, a visual indication corresponding to a portion of text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed in accordance with a determination that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria.

[0013] According to some embodiments, a non-transitory computer-readable storage medium is described. The non-transitory computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs including instructions to: display, via the display generation component, a first representation of a previously captured media item; detect, while displaying the first representation of the previously captured media item, via the one or more input devices, an input corresponding to a request to display a second representation of the previously captured media item; display, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to the request to display the second representation of the previously captured media item; and, while displaying the second representation of the previously captured media item, display, via the display generation component, a visual indication corresponding to a portion of text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed in accordance with a determination that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria.

[0014] According to some embodiments, a temporary computer-readable storage medium is described that stores one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs including instructions to: display, via the display generation component, a first representation of a previously captured media item; detect, while displaying the first representation of the previously captured media item, via the one or more input devices, an input corresponding to a request to display a second representation of the previously captured media item; display, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to the request to display the second representation of the previously captured media item; and, while displaying the second representation of the previously captured media item, display, via the display generation component, a visual indication corresponding to a portion of text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed in accordance with a determination that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria.

[0015] According to some embodiments, a computer system configured to communicate with a display generation component and one or more input devices is described, the computer system comprising one or more processors and a memory storing one or more programs configured to be executed by the one or more processors, the one or more programs including instructions to: display, via the display generation component, a first representation of a previously captured media item; detect, while displaying the first representation of the previously captured media item, an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; display, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to the request to display the second representation of the previously captured media item; and, while displaying the second representation of the previously captured media item, display, via the display generation component, a visual indication corresponding to a portion of text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed in accordance with a determination that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria.

[0016] According to some embodiments, a computer system configured to communicate with a display generation component and one or more input devices is described, the computer system comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; means for displaying, via the display generation component, a first representation of a previously captured media item; means for detecting, while displaying the first representation of the previously captured media item, an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; means for displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to the request to display the second representation of the previously captured media item; and means for displaying, while displaying the second representation of the previously captured media item, a visual indication, via the display generation component, corresponding to a portion of text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed, in accordance with determining that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria.

[0017] According to some embodiments, a computer program product is described, comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices. The one or more programs include instructions for: displaying, via the display generation component, a first representation of a previously captured media item; detecting, while displaying the first representation of the previously captured media item, an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to the request to display the second representation of the previously captured media item; and, while displaying the second representation of the previously captured media item, displaying, via the display generation component, a visual indication corresponding to a portion of text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed in accordance with a determination that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria.

[0018] According to some embodiments, a method is described in a computer system in communication with one or more cameras, one or more input devices, and a display generation component. The method includes: displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; displaying, via the display generation component, a camera user interface including a representation of a field of view of one or more cameras in response to detecting the request to display the camera user interface; displaying, in response to determining that the representation of the field of view of the one or more cameras includes detected text that meets one or more criteria, a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area; detecting, via the one or more input devices while simultaneously displaying the representation of the field of view and the text insertion user interface object, an input corresponding to selection of the text insertion user interface object; and inserting, in response to detecting the input corresponding to selection of the text insertion user interface object, at least a portion of the detected text into the text entry area.

[0019] According to some embodiments, a non-transitory computer-readable storage medium is described that stores one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, one or more input devices, and a display generation component, the one or more programs including instructions for: displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; displaying, via the display generation component, a camera user interface including a representation of a field of view of one or more cameras; and, in response to detecting the request to display the camera user interface, displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with determining that the representation of the field of view of the one or more cameras includes detected text that satisfies one or more criteria; detecting, via the one or more input devices, input corresponding to selection of the text insertion user interface object while simultaneously displaying the representation of the field of view and the text insertion user interface object; and inserting at least a portion of the detected text into the text entry area in response to detecting the input corresponding to selection of the text insertion user interface object.

[0020] According to some embodiments, a temporary computer-readable storage medium is described. The temporary computer-readable storage medium stores one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, one or more input devices, and a display generation component, the one or more programs including instructions for: displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; displaying, via the display generation component, a camera user interface including a representation of a field of view of one or more cameras; and, in response to detecting the request to display the camera user interface, displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with a determination that the representation of the field of view of the one or more cameras includes detected text that satisfies one or more criteria; detecting, via the one or more input devices, input corresponding to selection of the text insertion user interface object while simultaneously displaying the representation of the field of view and the text insertion user interface object; and inserting at least a portion of the detected text into the text entry area in response to detecting the input corresponding to selection of the text insertion user interface object.

[0021] According to some embodiments, a computer system configured to communicate with one or more cameras, one or more input devices, and an output generation component is described, the computer system comprising one or more processors and a memory storing one or more programs configured to be executed by the one or more processors, the one or more programs including instructions to: display a first user interface including a text entry area; detect a request to display a camera user interface while displaying the first user interface including the text entry area; display, via the display generation component, a camera user interface including a representation of a field of view of one or more cameras in response to detecting the request to display the camera user interface; display a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with determining that the representation of the field of view of the one or more cameras includes detected text that satisfies one or more criteria; detect, via the one or more input devices, input corresponding to selection of the text insertion user interface object while simultaneously displaying the representation of the field of view and the text insertion user interface object; and insert at least a portion of the detected text into the text entry area in response to detecting the input corresponding to selection of the text insertion user interface object.

[0022] According to some embodiments, a computer system configured in communication with one or more cameras, one or more input devices, and an output generation component is described, the computer system comprising: a memory storing one or more programs configured to be executed by one or more processors; means for displaying a first user interface including a text entry area; means for detecting a request to display a camera user interface while displaying the first user interface including the text entry area; means for displaying, via the output generation component, a camera user interface including a representation of a field of view of the one or more cameras, and, in response to determining that the representation of the field of view of the one or more cameras includes detected text that satisfies one or more criteria, displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area; means for detecting, via the one or more input devices while simultaneously displaying the representation of the field of view and the text insertion user interface object, an input corresponding to selection of the text insertion user interface object; and means for inserting at least a portion of the detected text into the text entry area in response to detecting input corresponding to selection of the text insertion user interface object.

[0023] According to some embodiments, a computer program product is described, comprising one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, one or more input devices, and a display generation component. The one or more programs include instructions for: displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; displaying, via the display generation component, a camera user interface including a representation of a field of view of one or more cameras in response to detecting the request to display the camera user interface; displaying, in accordance with determining that the representation of the field of view of the one or more cameras includes detected text that satisfies one or more criteria, a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area; detecting, via the one or more input devices, input corresponding to selection of the text insertion user interface object while simultaneously displaying the representation of the field of view and the text insertion user interface object; and inserting at least a portion of the detected text into the text entry area in response to detecting the input corresponding to selection of the text insertion user interface object.

[0024] According to some embodiments, a method is described. The method is executed on a computer system in communication with a display generation component. The method includes: displaying, via the display generation component, a media user interface including a representation of media; receiving, while displaying the media user interface including the representation of the media, a request to display additional information about a plurality of detected features within the representation of the media; and, in response to receiving the request to display the additional information about the plurality of detected features, displaying, while displaying the media user interface including the representation of the media, one or more indications of a plurality of detected features within the media, the one or more indications of the plurality of detected features including a first indication of a first detected feature displayed at a first location within the representation of the media, the first location corresponding to the location of the first detected feature within the representation of the media, wherein, in response to determining that the first detected feature is a first type of feature, the first indication has a first appearance, and in response to determining that the first detected feature is a second type of feature different from the first type of feature, the first indication has a second appearance different from the first appearance.

[0025] According to some embodiments, a non-transitory computer-readable storage medium is described, the non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs displaying, via the display generation component, a media user interface including a representation of media, receiving, while displaying the media user interface including the representation of media, a request to display additional information related to a plurality of detected characteristics within the representation of media, and, in response to receiving the request to display the additional information related to the plurality of detected characteristics, while displaying the media user interface including the representation of media. A non-transitory computer-readable storage medium comprising instructions for displaying one or more indications of a plurality of detected characteristics in media, wherein the one or more indications of the plurality of detected characteristics include a first indication of a first detected characteristic displayed at a first location within a representation of the media, the first location corresponding to a location of the first detected characteristic within the representation of the media, and in accordance with a determination that the first detected characteristic is a first type of characteristic, the first indication has a first appearance, and in accordance with a determination that the first detected characteristic is a second type of characteristic different from the first type of characteristic, the first indication has a second appearance different from the first appearance.

[0026] According to some embodiments, a transient computer-readable storage medium is described, the transient computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs being configured to: display a media user interface including a representation of media via the display generation component; receive a request to display additional information related to a plurality of detected characteristics within the representation of media while displaying the media user interface including the representation of media; and, in response to receiving the request to display the additional information related to the plurality of detected characteristics, while displaying the media user interface including the representation of media. A transient computer-readable storage including instructions for displaying one or more indications of a plurality of detected characteristics in media, the one or more indications of the plurality of detected characteristics including a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to the location of the first detected characteristic within the representation of the media, wherein in accordance with a determination that the first detected characteristic is a first type of characteristic, the first indication has a first appearance, and in accordance with a determination that the first detected characteristic is a second type of characteristic different from the first type of characteristic, the first indication has a second appearance different from the first appearance.

[0027] According to some embodiments, a computer system configured to be in communication with a display generation component is described. The computer system includes one or more processors and a memory storing one or more programs configured to be executed by the one or more processors, the one or more programs including instructions for displaying, via a display generation component, a media user interface including a representation of the media, receiving a request to display additional information about a plurality of detected characteristics within the representation of the media while displaying the media user interface including the representation of the media, and displaying one or more indications of the plurality of detected characteristics within the media while displaying the media user interface including the representation of the media in response to receiving the request to display the additional information about the plurality of detected characteristics, wherein the one or more indications of the plurality of detected characteristics include a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to the location of the first detected characteristic within the representation of the media, and in accordance with a determination that the first detected characteristic is a first type of characteristic, the first indication has a first appearance, and in accordance with a determination that the first detected characteristic is a second type of characteristic different from the first type of characteristic, the first indication has a second appearance different from the first appearance.

[0028] According to some embodiments, a computer system configured to be in communication with a display generation component is described. The computer system includes one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; a display generation component for displaying a media user interface including a representation of the media; a display generation component for receiving a request to display additional information about a plurality of detected characteristics within the representation of the media while displaying the media user interface including the representation of the media; and a display component for displaying one or more indications of the plurality of detected characteristics within the media while displaying the media user interface including the representation of the media in response to receiving the request to display the additional information about the plurality of detected characteristics, wherein the one or more indications of the plurality of detected characteristics include a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to the location of the first detected characteristic within the representation of the media, and wherein, in accordance with a determination that the first detected characteristic is a first type of characteristic, the first indication has a first appearance, and in accordance with a determination that the first detected characteristic is a second type of characteristic different from the first type of characteristic, the first indication has a second appearance different from the first appearance.

[0029] According to some embodiments, a computer program product is described. The computer program product comprises one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs including instructions for: displaying, via the display generation component, a media user interface including a representation of the media; receiving, while displaying the media user interface including the representation of the media, a request to display additional information regarding a plurality of detected characteristics within the representation of the media; and, in response to receiving the request to display the additional information regarding the plurality of detected characteristics, displaying, while displaying the media user interface including the representation of the media, one or more indications of the plurality of detected characteristics within the media, the one or more indications of the plurality of detected characteristics including a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to the location of the first detected characteristic within the representation of the media; in accordance with a determination that the first detected characteristic is a first type of characteristic, the first indication has a first appearance; and in accordance with a determination that the first detected characteristic is a second type of characteristic different from the first type of characteristic, the first indication has a second appearance different from the first appearance.

[0030] According to some embodiments, a method is described, the method being performed in a computer system in communication with one or more cameras, a display generation component, and one or more input devices. The method includes receiving a request to display a representation of a field of view of one or more cameras; in response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via a display generation component, a representation of the field of view of the one or more cameras, the representation including text within the field of view of the one or more cameras; automatically displaying, via the display generation component, a plurality of indications of the translated text including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; receiving, via the display generation component, a request to select, via one or more input devices, an individual indication of the plurality of translated portions while displaying the first indication and the second indication; in response to receiving the request to select the individual indication, in accordance with determining that the request is a request to select the first indication, displaying, via the display generation component, a first translation user interface object including the first portion of text and the translation of the first portion of text without including a translation of the second portion of text.

[0031] According to some embodiments, a non-transitory computer-readable storage medium is described, the non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, a display generation component, and one or more input devices, the one or more programs receiving a request to display a representation of a field of view of the one or more cameras, and in response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via the display generation component, a representation of the field of view of the one or more cameras, the representation including text within the field of view of the one or more cameras, and displaying, via the display generation component, a first indication of a translation of a first portion of the text and the text. and a second indication of a translation of a second portion of the text; and, while displaying the first indication and the second indication, receive, via the display generation component, a request to select, via one or more input devices, an individual indication of the plurality of translated portions; and, in response to receiving the request to select an individual indication, in accordance with a determination that the request is a request to select the first indication, display, via the display generation component, a first translation user interface object that includes the first portion of the text and the translation of the first portion of the text without the translation of the second portion of the text.

[0032] According to some embodiments, a transient computer-readable storage medium is described, the transient computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, a display generation component, and one or more input devices, the one or more programs receiving a request to display a representation of a field of view of the one or more cameras, and in response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via the display generation component, a representation of the field of view of the one or more cameras, the representation including text within the field of view of the one or more cameras, and displaying, via the display generation component, a first indication of a translation of a first portion of the text and the text. and a second indication of a translation of the second portion of the text; and, while displaying the first indication and the second indication, receive, via the display generation component, a request to select an individual indication of the plurality of translated portions via one or more input devices; and, in response to receiving the request to select the individual indication, in accordance with a determination that the request is a request to select the first indication, display, via the display generation component, a first translation user interface object including the first portion of the text and the translation of the first portion of the text without the translation of the second portion of the text.

[0033] According to some embodiments, a computer system configured in communication with one or more cameras, a display generation component, and one or more input devices is described. The computer system includes one or more processors and a memory storing one or more programs configured to be executed by the one or more processors, the one or more programs including instructions to: receive a request to display a representation of a field of view of one or more cameras; in response to receiving the request to display the representation of the field of view of the one or more cameras, display, via a display generation component, a representation of the field of view of the one or more cameras, the representation including text within the field of view of the one or more cameras; automatically display, via the display generation component, a plurality of indications of translated text including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; receive, via the display generation component, a request to select, via one or more input devices, individual indications of the plurality of translated portions while displaying the first indication and the second indication; and in response to receiving the request to select the individual indication, in accordance with determining that the request is a request to select the first indication, display, via the display generation component, a first translation user interface object including the first portion of text and the translation of the first portion of text without including a translation of the second portion of text.

[0034] According to some embodiments, a computer system configured in communication with one or more cameras, a display generation component, and one or more input devices is described. The computer system includes one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; a means for receiving a request to display a representation of a field of view of one or more cameras; a means for, in response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via a display generation component, a representation of the field of view of the one or more cameras, the representation including text within the field of view of the one or more cameras; a means for automatically displaying, via the display generation component, a plurality of indications of translated text including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; a means for receiving, via the display generation component, a request to select, via one or more input devices, individual indications of the plurality of translated portions while displaying the first indication and the second indication; and a means for, in response to receiving the request to select the individual indication, displaying, via the display generation component, a first translation user interface object including the first portion of text and a translation of the first portion of text without including a translation of the second portion of text, in accordance with determining that the request is a request to select the first indication.

[0035] According to some embodiments, a computer program product is described, comprising one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, a display generating component, and one or more input devices. The one or more programs include instructions for receiving a request to display a representation of a field of view of one or more cameras; in response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via a display generation component, a representation of the field of view of the one or more cameras, the representation including text within the field of view of the one or more cameras; automatically displaying, via the display generation component, a plurality of indications of the translated text including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; receiving, via the display generation component, a request to select, via one or more input devices, individual indications of the plurality of translated portions while displaying the first indication and the second indication; and in response to receiving the request to select the individual indication, in accordance with determining that the request is a request to select the first indication, displaying, via the display generation component, a first translation user interface object including the first portion of text and the translation of the first portion of text without including a translation of the second portion of text.

[0036] According to some embodiments, a method is described that is performed in a computer system in communication with a display generation component, the method including: detecting a request to display additional information corresponding to the media representation while displaying a user interface that includes the media representation; in response to detecting the request to display the additional information corresponding to the media representation, displaying, via the display generation component, a first user interface object in accordance with a determination that detected text in the media representation has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first action based on the detected text; and in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, displaying, via the display generation component, a second user interface object in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, the second user interface object, when selected, causing the computer system to perform a second action based on the detected text, different from the first action.

[0037] According to some embodiments, a non-transitory computer-readable storage medium is described, the non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs including instructions for: detecting a request to display additional information corresponding to a media representation while displaying a user interface including the media representation; in response to detecting the request to display the additional information corresponding to the media representation, displaying, via the display generation component, a first user interface object in accordance with a determination that detected text in the media representation has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first operation based on the detected text; and in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, displaying, via the display generation component, a second user interface object in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, the second user interface object, when selected, causing the computer system to perform a second operation different from the first operation based on the detected text.

[0038] According to some embodiments, a transient computer-readable storage medium is described, the transient computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs including instructions for: detecting a request to display additional information corresponding to a media representation while displaying a user interface including the media representation; in response to detecting the request to display the additional information corresponding to the media representation, displaying, via the display generation component, a first user interface object in accordance with a determination that detected text in the media representation has a first set of properties, the first user interface object that, when selected, causes the computer system to perform a first operation based on the detected text; and in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, displaying, via the display generation component, a second user interface object that, when selected, causes the computer system to perform a second operation based on the detected text, the second user interface object being different from the first operation.

[0039] According to some embodiments, a computer system is described that is configured to be in communication with a display generation component, the computer system comprising one or more processors and a memory that stores one or more programs configured to be executed by the one or more processors, the one or more programs including instructions to detect, while displaying a user interface that includes a media representation, a request to display additional information corresponding to the media representation, and in response to detecting the request to display the additional information corresponding to the media representation, display, via the display generation component, a first user interface object that, in accordance with a determination that detected text in the media representation has a first set of properties, causes the computer system, when selected, to perform a first operation based on the detected text, and display, via the display generation component, a second user interface object that, in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, causes the computer system, when selected, to perform a second operation based on the detected text, different from the first operation.

[0040] According to some embodiments, a computer system is described, the computer system being configured to be in communication with a display generation component, the computer system comprising: means for detecting a request to display additional information corresponding to a representation of media while the computer system is displaying a user interface including the representation of media; and means for, in response to detecting the request to display the additional information corresponding to the representation of media, displaying, via the display generation component, a first user interface object in accordance with a determination that detected text in the representation of media has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first action based on the detected text; and means for displaying, via the display generation component, a second user interface object in accordance with a determination that the detected text in the representation of media has a second set of properties different from the first set of properties, the second user interface object, when selected, causing the computer system to perform a second action based on the detected text, the second action being different from the first action.

[0041] According to some embodiments, a computer program product is described, comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a generation component, the one or more programs including instructions for: detecting, while displaying a user interface including the media representation, a request to display additional information corresponding to the media representation; in response to detecting the request to display the additional information corresponding to the media representation, displaying, via the display generation component, a first user interface object in accordance with a determination that detected text in the media representation has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first operation based on the detected text; and in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, displaying, via the display generation component, a second user interface object in accordance with a determination that the detected text in the media representation has a second set of properties different from the first set of properties, the second user interface object, when selected, causing the computer system to perform a second operation different from the first operation based on the detected text.

[0042] Executable instructions to perform these functions are optionally contained in a non-transitory computer-readable storage medium or other computer program product configured for execution by one or more processors. Executable instructions to perform these functions are optionally contained in a transitory computer-readable storage medium or other computer program product configured for execution by one or more processors.

[0043] Thus, devices are provided with faster and more efficient methods and interfaces for managing visual content in media, thereby increasing the effectiveness, efficiency, and user satisfaction with such devices. Such methods and interfaces may complement or replace other methods for managing visual content in media. [Brief explanation of the drawings]

[0044] For a better understanding of the various described embodiments, reference should be made to the following Detailed Description of the Invention in conjunction with the following drawings, in which like reference numerals refer to corresponding parts throughout:

[0045] [Figure 1A] FIG. 1 is a block diagram illustrating a portable multifunction device with a touch-sensitive display in accordance with some embodiments.

[0046] [Figure 1B] FIG. 2 is a block diagram illustrating exemplary components for event processing according to some embodiments.

[0047] [Figure 2] FIG. 1 illustrates a portable multifunction device with a touch screen in accordance with some embodiments.

[0048] [Figure 3] FIG. 1 is a block diagram of an exemplary multifunction device having a display and a touch-sensitive surface in accordance with some embodiments.

[0049] [Figure 4A] 1 illustrates an exemplary user interface for a menu of applications on a portable multifunction device in accordance with some embodiments.

[0050] [Figure 4B]1 illustrates an exemplary user interface for a multifunction device having a touch-sensitive surface that is separate from the display in accordance with some embodiments.

[0051] [Figure 5A] 1 illustrates a personal electronic device according to some embodiments.

[0052] [Figure 5B] FIG. 1 is a block diagram illustrating a personal electronic device according to some embodiments.

[0053] [Figure 6A] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6B] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6C] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6D] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6E] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6F] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6G] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6H] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6I] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6J]1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6K] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6L] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6M] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6N] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6O] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6P] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6Q] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6R] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6S] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6T] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6U] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6V] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6W]1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6X] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6Y] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments. [Figure 6Z] 1 illustrates an exemplary user interface for managing visual content within a medium, according to some embodiments.

[0054] [Figure 7A] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7B] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7C] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7D] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7E] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7F] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7G] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7H]1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7I] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7J] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7K] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments. [Figure 7L] 1 illustrates an exemplary user interface for managing visual indicators for visual content in media, according to some embodiments.

[0055] [Figure 8] FIG. 1 is a flow diagram illustrating a method for managing visual content in a medium, according to some embodiments.

[0056] [Figure 9] FIG. 1 is a flow diagram illustrating managing visual indicators for visual content in media, according to some embodiments.

[0057] [Figure 10A] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10B] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10C] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10D] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10E] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10F] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10G] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10H] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10I] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10J] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10K] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10L] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10M] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10N] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10O] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10P] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10Q] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10R]1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10S] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10T] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10U] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10V] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10W] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10X] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10Y] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10Z] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10AA] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10AB] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10AC] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments. [Figure 10AD] 1 illustrates an exemplary user interface for inserting visual content in media, according to some embodiments.

[0058] [Figure 11] FIG. 1 is a flow diagram illustrating a user interface for inserting visual content in media, according to some embodiments.

[0059] [Figure 12A] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12B] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12C] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12D] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12E] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12F] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12G] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12H] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12I] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12J] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12K] 1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments. [Figure 12L]1 illustrates an exemplary user interface for identifying visual content within media, according to some embodiments.

[0060] [Figure 13] FIG. 1 is a flow diagram illustrating a method for identifying visual content in media, according to some embodiments.

[0061] [Figure 14A] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14B] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14C] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14D] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14E] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14F] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14G] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14H] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14I] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14J] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14K]1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14L] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14M] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments. [Figure 14N] 1 illustrates an exemplary user interface for translating visual content in media, according to some embodiments.

[0062] [Figure 15] FIG. 1 is a flow diagram illustrating a method for translating visual content in media, according to some embodiments.

[0063] [Figure 16A] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16B] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16C] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16D] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16E] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16F] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16G] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16H] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16I] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16J] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16K] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16L] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16M] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16N] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments. [Figure 16O] 1 illustrates an exemplary user interface for managing user interface objects for visual content in media, according to some embodiments.

[0064] [Figure 17] FIG. 1 is a flow diagram illustrating a method for managing user interface objects for visual content in media, according to some embodiments. DETAILED DESCRIPTION OF THE INVENTION

[0065] The following description sets forth example methods, parameters, etc. However, it should be recognized that such description is not intended as a limitation on the scope of the present disclosure, but rather is provided as a description of example embodiments.

[0066] There is a need for electronic devices that provide efficient methods and interfaces for managing visual content. For example, there is a need for electronic devices and / or computer systems that allow users to manage visual content contained in objects captured by one or more cameras of the computer system, such as signs and restaurant menus. Such techniques can reduce the cognitive burden placed on users managing visual content, thereby increasing productivity. Furthermore, such techniques can reduce processor and battery power that would otherwise be wasted on redundant user input.

[0067] 1A-1B, 2, 3, 4A-4B, and 5A-5B below provide descriptions of exemplary devices that perform techniques for managing visual content.

[0068]

[0023] Figures 6A-6Z illustrate exemplary user interfaces for managing visual content in media. Figure 8 is a flow diagram illustrating a method for managing visual content in media, according to some embodiments. The user interfaces in Figures 6A-6Z are used to illustrate processes described below, including the process of Figure 8.

[0069] 7A-7L illustrate exemplary user interfaces for managing visual indicators for visual content in media. FIG. 9 is a flow diagram illustrating a method for managing visual indicators for visual content in media, according to some embodiments. The user interfaces in FIGS. 7A-7L are used to illustrate processes described below, including the process of FIG. 9.

[0070] 10A-10AD illustrate exemplary user interfaces for inserting visual content into media. FIG. 11 is a flow diagram illustrating a method for inserting visual content into media. The user interfaces in FIGS. 10A-10AD are used to illustrate processes described below, including the process of FIG. 11.

[0071] 12A-12L illustrate exemplary user interfaces for identifying visual content in media. FIG. 13 is a flow diagram illustrating a method for identifying visual content in media. The user interfaces of FIGS. 12A-12L are used to illustrate processes described below, including the process of FIG. 13.

[0072] 14A-14N illustrate exemplary user interfaces for translating visual content in media. FIG. 15 is a flow diagram illustrating a method for translating visual content in media, according to some embodiments. The user interfaces of FIGS. 14A-14N are used to illustrate processes described below, including the process of FIG. 15.

[0073] 16A-16O illustrate exemplary user interfaces for managing user interface objects for visual content in media, according to some embodiments. FIG. 17 is a flow diagram illustrating a method for managing user interface objects for visual content in media, according to some embodiments. The user interfaces in FIGS. 16A-16O are used to illustrate processes described below, including the process of FIG. 17.

[0074] The processes described below enhance the usability of a device and streamline the user-device interface (e.g., by helping the user provide appropriate inputs when operating / interacting with the device and reducing user errors) through various techniques, including providing improved visual feedback to the user, reducing the number of inputs required to perform an operation, providing additional control options without cluttering the user interface with additional controls that are displayed, performing an operation without requiring further user input when a set of conditions is met, and / or other techniques. These techniques also reduce power usage and improve the device's battery life by allowing the user to use the device more quickly and efficiently.

[0075] Furthermore, for methods described herein in which one or more steps are conditioned on one or more conditions being satisfied, it should be understood that the described method can be repeated in multiple iterations, such that over the course of the iterations, all of the conditions on which the method steps are conditioned are satisfied in different iterations of the method. For example, if a method requires performing a first step if a condition is satisfied and a second step if the condition is not satisfied, one skilled in the art will understand that the steps recited in the claim are repeated in a particular order until the conditions are met and are no longer met. Thus, a method described with one or more steps that depend on one or more conditions being satisfied can be rewritten as a method that is repeated until each condition recited in the method is met. However, this is not required for system or computer-readable medium claims in which the system or computer-readable medium includes instructions for performing a conditional action based on the satisfaction of the corresponding one or more conditions, and thus can determine whether a contingency is met without explicitly repeating the method steps until all conditions on which the method steps are conditioned are satisfied. Those skilled in the art will also understand that, as with methods having conditional steps, the system or computer-readable storage medium may repeat the steps of the method as many times as necessary to ensure that all of the conditional steps have been performed.

[0076] In the following description, terms such as "first" and "second" are used to describe various elements, but these elements should not be limited by these terms. These terms are used only to distinguish one element from another. For example, a first touch can be referred to as a second touch, and similarly, a second touch can be referred to as a first touch, without departing from the scope of the various embodiments described. Although a first touch and a second touch are both touches, they are not the same touch.

[0077] The terminology used in the description of the various embodiments set forth herein is for the purpose of describing particular embodiments only and is not intended to be limiting. In the description of the various embodiments set forth and in the appended claims, the singular forms "a," "an," and "the" are intended to include the plural forms as well, unless the context clearly dictates otherwise. Also, as used herein, the term "and / or" should be understood to refer to and include any and all possible combinations of one or more of the associated listed items. It will be further understood that the terms "includes," "including," "comprises," and / or "comprising," as used herein, specify the presence of stated features, integers, steps, operations, elements, and / or components, but do not exclude the presence or addition of one or more other features, integers, steps, operations, elements, components, and / or groups thereof.

[0078] The term "if" is interpreted, optionally, according to the context, to mean "when" or "upon," or "in response to determining" or "in response to detecting." Similarly, the phrases "if it is determined" or "if [a stated condition or event] is detected" are interpreted, optionally, according to the context, to mean "upon determining" or "in response to determining," or "upon detecting [the stated condition or event]" or "in response to detecting [the stated condition or event]."

[0079] Embodiments of electronic devices, user interfaces for such devices, and associated processes for using such devices are described. In some embodiments, the device is a portable communication device, such as a mobile phone, that also includes other functions, such as PDA and / or music player functions. Exemplary embodiments of portable multifunction devices include, but are not limited to, the iPhone®, iPod Touch®, and iPad® devices from Apple Inc. of Cupertino, California. Optionally, other portable electronic devices, such as a laptop computer or tablet computer having a touch-sensitive surface (e.g., a touchscreen display and / or touchpad), are also used. It should also be understood that in some embodiments, the device is not a portable communication device, but a desktop computer having a touch-sensitive surface (e.g., a touchscreen display and / or touchpad). In some embodiments, the electronic device is a computer system in communication (e.g., via wired communication, via wireless communication) with a display generation component. The display generation component is configured to provide a visual output, such as a display via a CRT display, a display via an LED display, or a display via image projection. In some embodiments, the display generation component is integrated with the computer system. In some embodiments, the display generation component is separate from the computer system. As used herein, "displaying" content includes causing content (e.g., video data rendered or decoded by display controller 156) to be displayed by transmitting data (e.g., image data or video data) over a wired or wireless connection to an integrated or external display generation component to visually generate the content.

[0080] In the following discussion, electronic devices are described that include a display and a touch-sensitive surface, however, it should be understood that the electronic device optionally includes one or more other physical user-interface devices, such as a physical keyboard, a mouse, and / or a joystick.

[0081] The device typically supports a variety of applications such as one or more of a drawing application, a presentation application, a word processing application, a website creation application, a disc authoring application, a spreadsheet application, a gaming application, a telephone application, a video conferencing application, an email application, an instant messaging application, a training support application, a photo management application, a digital camera application, a digital video camera application, a web browsing application, a digital music player application, and / or a digital video player application.

[0082] Various applications running on the device optionally use at least one common physical user-interface device, such as a touch-sensitive surface. One or more features of the touch-sensitive surface and corresponding information displayed on the device are optionally adjusted and / or changed for each application and / or within individual applications. In this way, the common physical architecture of the device (such as the touch-sensitive surface) optionally supports various applications with user interfaces that are intuitive and transparent to the user.

[0083] Attention now turns to embodiments of portable devices with touch-sensitive displays. FIG. 1A is a block diagram illustrating portable multifunction device 100 having touch-sensitive display system 112, according to some embodiments. Touch-sensitive display 112 may conveniently be referred to as a "touch screen" and may also be known or referred to as a "touch-sensitive display system." Device 100 includes memory 102 (optionally including one or more computer-readable storage media), memory controller 122, one or more processing units (CPUs) 120, peripherals interface 118, RF circuitry 108, audio circuitry 110, speaker 111, microphone 113, input / output (I / O) subsystem 106, other input control devices 116, and external port 124. Device 100 optionally includes one or more optical sensors 164. Device 100 optionally includes one or more contact intensity sensors 165 that detect the intensity of a contact on device 100 (e.g., a touch-sensitive surface, such as touch-sensitive display system 112 of device 100). Device 100 optionally includes one or more tactile output generators 167 that generate tactile output on device 100 (e.g., generate tactile output on a touch-sensitive surface such as touch-sensitive display system 112 of device 100 or touchpad 355 of device 300). These components optionally communicate via one or more communication buses or signal lines 103.

[0084] As used herein and in the claims, the term “intensity” of a contact on a touch-sensitive surface refers to the force or pressure (force per unit area) of a contact (e.g., a finger contact) on the touch-sensitive surface, or a proxy for the force or pressure of a contact on the touch-sensitive surface. The intensity of a contact has a range of values that includes at least four distinct values and more typically includes hundreds (e.g., at least 256) distinct values. The intensity of a contact is optionally determined (or measured) using various techniques and various sensors or combinations of sensors. For example, one or more force sensors under or adjacent to the touch-sensitive surface are optionally used to measure force at various points on the touch-sensitive surface. In some implementations, force measurements from multiple force sensors are combined (e.g., weighted averaged) to determine an estimated force of the contact. Similarly, a pressure-sensitive tip of a stylus is optionally used to determine the pressure of the stylus on the touch-sensitive surface. Alternatively, the size and / or change in the contact area detected on the touch-sensitive surface, the capacitance and / or change in the capacitance of the touch-sensitive surface proximate the contact, and / or the resistance and / or change in the capacitance of the touch-sensitive surface proximate the contact are optionally used as a surrogate for the force or pressure of the contact on the touch-sensitive surface. In some implementations, the surrogate measure of the force or pressure of the contact is used directly to determine whether an intensity threshold is exceeded (e.g., the intensity threshold is described in units corresponding to the surrogate measure). In some implementations, the surrogate measure of the contact force or pressure is converted to an estimate of the force or pressure, and the estimate of the force or pressure is used to determine whether an intensity threshold is exceeded (e.g., the intensity threshold is a pressure threshold measured in units of pressure).Using contact intensity as an attribute of user input allows users to access additional device functionality (e.g., on a touch-sensitive display) and / or receive user input (e.g., via a touch-sensitive display, touch-sensitive surface, or physical / mechanical controls such as knobs or buttons) that may not otherwise be accessible to users on devices of reduced size that have limited footprint for displaying affordances.

[0085] As used herein and in the claims, the term “tactile output” refers to a physical displacement of a device relative to a previous position of the device, a physical displacement of a component of the device (e.g., a touch-sensitive surface) relative to another component of the device (e.g., a housing), or a displacement of a component relative to the center of mass of the device, that will be detected by a user with the user's sense of touch. For example, in a situation where a device or a component of a device is in contact with a touch-sensitive surface of a user (e.g., the fingers, palm, or other part of the user's hand), the tactile output produced by the physical displacement will be interpreted by the user as a tactile sensation corresponding to a perceived change in a physical property of the device or a component of the device. For example, movement of a touch-sensitive surface (e.g., a touch-sensitive display or trackpad) is optionally interpreted by the user as a “downclick” or “upclick” of a physical actuator button. In some cases, a user feels a tactile sensation such as a “downclick” or “upclick” even when there is no movement of a physical actuator button associated with the touch-sensitive surface that is physically pressed (e.g., displaced) by the user's action. As another example, movement of a touch-sensitive surface is optionally interpreted or perceived by a user as "roughness" of the touch-sensitive surface, even when there is no change in the smoothness of the touch-sensitive surface. While such user interpretation of touch depends on the user's personal sensory perception, there are many sensory perceptions of touch that are common to the majority of users. Thus, when a tactile output is described as corresponding to a particular sensory perception of a user (e.g., "upclick," "downclick," "roughness"), unless otherwise specified, the generated tactile output corresponds to a physical displacement of the device, or a component of the device, that produces the described sensory perception for a typical (or average) user.

[0086] It should be understood that device 100 is only one example of a portable multifunction device, and that device 100 optionally has more or fewer components than those shown, optionally combines two or more components, or optionally has a different configuration or arrangement of its components. The various components shown in FIG. 1A are implemented in hardware, software, or a combination of both hardware and software, including one or more signal processing circuits and / or application specific integrated circuits.

[0087] Memory 102 optionally includes high-speed random access memory, and optionally includes non-volatile memory, such as one or more magnetic disk storage devices, flash memory devices, or other non-volatile solid-state memory devices. Memory controller 122 optionally controls access to memory 102 by other components of device 100.

[0088] Peripheral interface 118 may be used to couple input and output peripherals of the device to CPU 120 and memory 102. One or more processors 120 operate or execute various software programs and / or instruction sets stored in memory 102 to perform various functions and process data for device 100. In some embodiments, peripheral interface 118, CPU 120, and memory controller 122 are optionally implemented on a single chip, such as chip 104. In some other embodiments, they are optionally implemented on separate chips.

[0089] RF (radio frequency) circuitry 108 transmits and receives RF signals, also called electromagnetic signals. RF circuitry 108 converts electrical signals to or from electromagnetic signals and communicates with communication networks and other communication devices via electromagnetic signals. RF circuitry 108 optionally includes well-known circuitry for performing these functions, including, but not limited to, an antenna system, an RF transceiver, one or more amplifiers, a tuner, one or more oscillators, a digital signal processor, a CODEC chipset, a subscriber identity module (SIM) card, memory, etc. RF circuitry 108 optionally communicates via wireless communication with networks, such as the Internet, also known as the World Wide Web (WWW), an intranet, and / or wireless networks, such as cellular telephone networks, wireless local area networks (LANs) and / or metropolitan area networks (MANs), and with other devices. RF circuitry 108 optionally includes well-known circuitry for detecting near field communication (NFC) fields, such as by short-range radios. Wireless communication is optionally supported by, but is not limited to, Global System for Mobile Communications (GSM), Enhanced Data GSM Environment (EDGE), high-speed downlink packet access (HSDPA), high-speed uplink packet access (HSUPA), Evolution, Data-Only (EV-DO), HSPA, HSPA+, Dual-Cell HSPA (DC-HSPA), Long Term Evolution (LTE), and other standards.evolution (LTE), near field communications (NFC), wideband code division multiple access (W-CDMA), code division multiple access (CDMA), time division multiple access (TDMA), Bluetooth, Bluetooth Low Energy (BTLE), Wireless Fidelity (Wi-Fi) (e.g., IEEE 802.11a, IEEE 802.11b, IEEE 802.11g, IEEE 802.11n, and / or IEEE 802.11ac), voice over Internet Protocol (VoIP), Wi-MAX, protocols for email (e.g., Internet message access protocol (IMAP) and / or post office protocol (POP)), instant messaging (e.g., extensible messaging and presence protocol), The present invention may use any of a number of communication standards, protocols, and technologies, including the Session Initiation Protocol for Instant Messaging and Presence Leveraging Extensions (XMPP), the Session Initiation Protocol for Instant Messaging and Presence Leveraging Extensions (SIMPLE), the Instant Messaging and Presence Service (IMPS), and / or the Short Message Service (SMS), or any other suitable communication protocol, including communication protocols not yet developed as of the filing date of this application.

[0090] Audio circuit 110, speaker 111, and microphone 113 provide an audio interface between a user and device 100. Audio circuit 110 receives audio data from peripherals interface 118, converts the audio data into electrical signals, and transmits the electrical signals to speaker 111. Speaker 111 converts the electrical signals into sound waves audible to humans. Audio circuit 110 also receives electrical signals converted from sound waves by microphone 113. Audio circuit 110 converts the electrical signals into audio data and transmits the audio data to peripherals interface 118 for processing. The audio data is optionally retrieved from and / or transmitted to memory 102 and / or RF circuit 108 by peripherals interface 118. In some embodiments, audio circuit 110 also includes a headset jack (e.g., 212 in FIG. 2 ). The headset jack provides an interface between audio circuitry 110 and a detachable audio input / output peripheral, such as an output-only headphone or a headset with both an output (e.g., mono or binaural headphones) and an input (e.g., a microphone).

[0091] I / O subsystem 106 couples input / output peripherals on device 100, such as touchscreen 112 and other input control devices 116, to peripheral interface 118. I / O subsystem 106 optionally includes display controller 156, optical sensor controller 158, depth camera controller 169, intensity sensor controller 159, haptic feedback controller 161, and one or more input controllers 160 for other input or control devices. One or more input controllers 160 receive / send electrical signals from / to other input control devices 116. Other input control devices 116 optionally include physical buttons (e.g., push buttons, rocker buttons, etc.), dials, slider switches, joysticks, click wheels, etc. In some embodiments, input controller(s) 160 are optionally coupled to any (or none) of a keyboard, an infrared port, a USB port, and a pointer device such as a mouse. The one or more buttons (e.g., 208 in FIG. 2 ) optionally include up / down buttons for volume control of speaker 111 and / or microphone 113. The one or more buttons optionally include push buttons (e.g., 206 in FIG. 2 ). In some embodiments, the electronic device is a computer system in communication with one or more input devices (e.g., via wireless communication over wired communication). In some embodiments, the one or more input devices include a touch-sensitive surface (e.g., a trackpad as part of a touch-sensitive display). In some embodiments, the one or more input devices include one or more camera sensors (e.g., one or more optical sensors 164 and / or one or more depth camera sensors 175), such as for tracking user gestures (e.g., hand gestures) as input. In some embodiments, the one or more input devices are integrated with the computer system. In some embodiments, the one or more input devices are separate from the computer system.

[0092] A quick press of a push button optionally unlocks the touchscreen 112 or optionally initiates the process of unlocking the device using gestures on the touchscreen, as described in U.S. Patent Application Serial No. 11 / 322,549, filed December 23, 2005, "Unlocking a Device by Performing Gestures on an Unlock Image," U.S. Patent No. 7,657,849, which is incorporated herein by reference in its entirety. A longer press of a push button (e.g., 206) optionally turns power on or off to the device 100. The functionality of one or more of the buttons is optionally customizable by the user. The touchscreen 112 is used to implement virtual or soft buttons and one or more soft keyboards.

[0093] Touch-sensitive display 112 provides an input and output interface between the device and a user. Display controller 156 receives and / or sends electrical signals to touchscreen 112. Touchscreen 112 displays visual output to the user. This visual output optionally includes graphics, text, icons, animation, and any combination thereof (collectively "graphics"). In some embodiments, some or all of the visual output optionally corresponds to user interface objects.

[0094] Touchscreen 112 has a touch-sensitive surface, sensor, or set of sensors that accepts input from a user based on haptic and / or tactile contact. Touchscreen 112 and display controller 156 (along with any associated modules and / or instruction sets in memory 102) detects contacts (and any movement or cessation of contact) on touchscreen 112 and translates the detected contacts into interactions with user interface objects (e.g., one or more softkeys, icons, web pages, or images) displayed on touchscreen 112. In an exemplary embodiment, the point of contact between touchscreen 112 and the user corresponds to the user's finger.

[0095] Touchscreen 112 optionally uses LCD (liquid crystal display), LPD (light emitting polymer display), or LED (light emitting diode) technology, although other display technologies are used in other embodiments. Touchscreen 112 and display controller 156 optionally use any of a number of now known or later developed touch sensing technologies to detect contact and any movement or disruption thereof, including, but not limited to, capacitive, resistive, infrared, and surface acoustic wave technologies, as well as other proximity sensor arrays or other elements that determine one or more points of contact with touchscreen 112. In an exemplary embodiment, projected mutual capacitance sensing technology is used, such as that found in the iPhone® and iPod Touch® from Apple Inc. of Cupertino, California.

[0096] The touch-sensitive display in some embodiments of touchscreen 112 is optionally similar to the multi-touch-sensing touchpad described in U.S. Patent Nos. 6,323,846 (Westerman et al.), 6,570,557 (Westerman et al.), and / or 6,677,932 (Westerman), and / or U.S. Patent Application Publication No. 2002 / 0015024 A1, each of which is incorporated by reference in its entirety. However, touchscreen 112 displays visual output from device 100, whereas touch-sensitive touchpads do not provide visual output.

[0097] The touch-sensitive display in some embodiments of touch screen 112 is described in the following applications: (1) U.S. patent application Ser. No. 11 / 381,313, filed May 2, 2006, entitled "Multipoint Touch Surface Controller," (2) U.S. patent application Ser. No. 10 / 840,862, filed May 6, 2004, entitled "Multipoint Touchscreen," (3) U.S. patent application Ser. No. 10 / 903,964, filed July 30, 2004, entitled "Gestures For Touch Sensitive Input Devices," (4) U.S. patent application Ser. No. 11 / 048,264, filed January 31, 2005, entitled "Gestures For Touch Sensitive Input Devices," and (5) U.S. patent application Ser. No. 11 / 038,590, filed January 18, 2005, entitled "Mode-Based Graphical User Interfaces For Touch Sensitive Input Devices." No. 11 / 228,758, filed September 16, 2005, entitled "Virtual Input Device Placement On A Touch Screen User Interface," (7) U.S. Patent Application No. 11 / 228,700, filed September 16, 2005, entitled "Operation Of A Computer With A Touch Screen Interface," (8) U.S. Patent Application No. 11 / 228,737, filed September 16, 2005, entitled "Activating Virtual Keys Of A Touch-Screen Virtual Keyboard," and (9) U.S. Patent Application No. 11 / 367,749, filed March 3, 2006, entitled "Multi-Functional Hand-Held Device," all of which are incorporated herein by reference in their entireties.

[0098] Touchscreen 112 optionally has a video resolution greater than 100 dpi. In some embodiments, the touchscreen has a video resolution of approximately 160 dpi. A user optionally contacts touchscreen 112 using any suitable object or accessory, such as a stylus, a finger, or the like. In some embodiments, the user interface is designed to operate primarily using finger-based contact and gestures, which may not be as precise as stylus-based input due to the larger contact area of a finger on the touchscreen. In some embodiments, the device translates coarse finger input into precise pointer / cursor positions or commands to perform actions desired by the user.

[0099] In some embodiments, in addition to the touchscreen, device 100 optionally includes a touchpad for activating or deactivating certain functions. In some embodiments, the touchpad is a touch-sensitive area of the device that, unlike the touchscreen, does not display visual output. The touchpad is optionally a touch-sensitive surface separate from touchscreen 112 or an extension of the touch-sensitive surface formed by the touchscreen.

[0100] Device 100 also includes a power system 162 that provides power to the various components. Power system 162 optionally includes a power management system, one or more power sources (e.g., battery, alternating current (AC)), a recharging system, power failure detection circuitry, power converters or inverters, power status indicators (e.g., light emitting diodes (LEDs)), and any other components associated with generating, managing, and distributing electrical power within a portable device.

[0101] Device 100 also optionally includes one or more optical sensors 164. FIG. 1A shows an optical sensor coupled to optical sensor controller 158 in I / O subsystem 106. Optical sensor 164 optionally includes a charge-coupled device (CCD) or a complementary metal-oxide semiconductor (CMOS) phototransistor. Optical sensor 164 receives light from the environment projected through one or more lenses and converts the light into data representing an image. Optical sensor 164 optionally works in conjunction with imaging module 143 (also called a camera module) to capture still images or video. In some embodiments, the optical sensor is located on the back side of device 100 opposite touchscreen display 112 on the front of the device, so that the touchscreen display can be used as a viewfinder for capturing still images and / or video. In some embodiments, the optical sensor is located on the front of the device so that an image of a user is optionally captured for video conferencing while the user views other video conference participants on the touchscreen display. In some embodiments, the position of the optical sensor 164 can be changed by the user (e.g., by rotating the lens and sensor within the device housing), so that a single optical sensor 164 is used for both video conferencing and capturing still images and / or video, along with a touchscreen display.

[0102] Device 100 also optionally includes one or more depth camera sensors 175. FIG. 1A shows a depth camera sensor coupled to depth camera controller 169 in I / O subsystem 106. Depth camera sensor 175 receives data from the environment and creates a three-dimensional model of an object (e.g., a face) in a scene from a viewpoint (e.g., the depth camera sensor). In some embodiments, in conjunction with imaging module 143 (also referred to as a camera module), depth camera sensor 175 is optionally used to determine a depth map of different portions of an image captured by imaging module 143. In some embodiments, a depth camera sensor is located on the front of device 100 to obtain images of the user with depth information for videoconferences and to capture selfie images with depth map data while the user views other videoconference participants on a touchscreen display. In some embodiments, depth camera sensor 175 is located on the back of the device, or on the back and front of device 100. In some embodiments, the position of the depth camera sensor 175 can be changed by the user (e.g., by rotating the lens and sensor within the device housing), so that the depth camera sensor 175 is used for both video conferencing and capturing still images and / or video, in conjunction with a touchscreen display.

[0103] Device 100 also optionally includes one or more contact intensity sensors 165. FIG. 1A shows contact intensity sensors coupled to intensity sensor controller 159 in I / O subsystem 106. Contact intensity sensors 165 optionally include one or more piezoresistive strain gauges, capacitive force sensors, electric force sensors, piezoelectric force sensors, optical force sensors, capacitive touch-sensitive surfaces, or other intensity sensors (e.g., sensors used to measure the force (or pressure) of a contact on a touch-sensitive surface). Contact intensity sensors 165 receive contact intensity information (e.g., pressure information, or a proxy for pressure information) from the environment. In some embodiments, at least one contact intensity sensor is juxtaposed with or proximate to the touch-sensitive surface (e.g., touch-sensitive display system 112). In some embodiments, at least one contact intensity sensor is located on the back of device 100, opposite touchscreen display 112 located on the front of device 100.

[0104] Device 100 also optionally includes one or more proximity sensors 166. Figure 1A shows proximity sensor 166 coupled to peripherals interface 118. Alternatively, proximity sensor 166 is optionally coupled to input controller 160 within I / O subsystem 106. Proximity sensor 166 optionally functions as described in U.S. patent application Ser. Nos. 11 / 241,839, "Proximity Detector In Handheld Device," 11 / 240,788, "Proximity Detector In Handheld Device," 11 / 620,702, "Using Ambient Light Sensor To Augment Proximity Sensor Output," 11 / 586,862, "Automated Response To And Sensing Of User Activity In Portable Devices," and 11 / 638,251, "Methods And Systems For Automatic Configuration Of Peripherals," which are incorporated herein by reference in their entireties. In some embodiments, the proximity sensor turns off and disables touchscreen 112 when the multifunction device is placed near the user's ear (e.g., when the user is making a phone call).

[0105] Device 100 also optionally includes one or more tactile output generators 167. FIG. 1A shows tactile output generators coupled to haptic feedback controller 161 in I / O subsystem 106. Tactile output generator 167 optionally includes one or more electroacoustic devices, such as speakers or other audio components, and / or electromechanical devices that convert energy into linear motion, such as motors, solenoids, electroactive polymers, piezoelectric actuators, electrostatic actuators, or other tactile output generating components (e.g., components that convert electrical signals into tactile output on the device). Contact intensity sensor 165 receives tactile feedback generation instructions from haptic feedback module 133 and generates a tactile output on device 100 that can be sensed by a user of device 100. In some embodiments, at least one tactile output generator is juxtaposed with or proximate to a touch-sensitive surface (e.g., touch-sensitive display system 112) and generates a tactile output, optionally by moving the touch-sensitive surface vertically (e.g., in / out of the surface of device 100) or horizontally (e.g., back and forth in the same plane as the surface of device 100). In some embodiments, at least one tactile output generator sensor is located on the back of device 100, opposite touchscreen display 112, which is located on the front of device 100.

[0106] Device 100 also optionally includes one or more accelerometers 168. FIG. 1A shows accelerometer 168 coupled to peripherals interface 118. Alternatively, accelerometer 168 is optionally coupled to input controller 160 in I / O subsystem 106. Accelerometer 168 optionally functions as described in U.S. Patent Application Publication No. 20050190059, "Acceleration-based Theft Detection System for Portable Electronic Devices," and U.S. Patent Application Publication No. 20060017692, "Methods And Apparatuses For Operating A Portable Device Based On An Accelerometer," both of which are incorporated by reference herein in their entireties. In some embodiments, information is displayed on the touchscreen display in portrait or landscape orientation based on an analysis of data received from the one or more accelerometers. In addition to accelerometer(s) 168, device 100 optionally includes a magnetometer and a GPS (or GLONASS or other global navigation system) receiver for obtaining information about the location and orientation (e.g., vertical or horizontal) of device 100.

[0107] In some embodiments, software components stored in memory 102 include operating system 126, communications module (or instruction set) 128, touch / motion module (or instruction set) 130, graphics module (or instruction set) 132, text input module (or instruction set) 134, Global Positioning System (GPS) module (or instruction set) 135, and applications (or instruction set) 136. Additionally, in some embodiments, memory 102 (FIG. 1A) or 370 (FIG. 3) stores device / global internal state 157, as shown in FIGS. 1A and 3. Device / global internal state 157 includes one or more of: active application state indicating which applications, if any, are currently active; display state indicating which applications, views, or other information occupy various regions of touchscreen display 112; sensor state including information obtained from the device's various sensors and input control devices 116; and location information regarding the device's location and / or orientation.

[0108] Operating system 126 (e.g., Darwin, RTXC, LINUX, UNIX, OS X, iOS, WINDOWS, or an embedded operating system such as VxWorks) includes various software components and / or drivers that control and manage general system tasks (e.g., memory management, storage device control, power management, etc.) and facilitate communication between various hardware and software components.

[0109] Communications module 128 facilitates communication with other devices via one or more external ports 124 and also includes various software components for processing data received by RF circuitry 108 and / or external port 124. External port 124 (e.g., Universal Serial Bus (USB), FIREWIRE, etc.) is adapted to couple to other devices directly or indirectly via a network (e.g., the Internet, wireless LAN, etc.). In some embodiments, the external port is a multi-pin (e.g., 30-pin) connector that is the same as, similar to, and / or compatible with the 30-pin connector used on iPod® (trademark of Apple Inc.) devices.

[0110] Contact / motion module 130, optionally in cooperation with display controller 156, detects contact with touchscreen 112 and other touch-sensing devices (e.g., a touchpad or physical click wheel). Contact / motion module 130 includes various software components for performing various operations related to contact detection, such as determining whether contact occurs (e.g., detecting a finger-down event), determining the intensity of the contact (e.g., the force or pressure of the contact, or a surrogate for the force or pressure of the contact), determining whether there is contact movement and tracking the movement across the touch-sensitive surface (e.g., detecting one or more finger-drag events), and determining whether the contact has ceased (e.g., detecting a finger-up event or an interruption of the contact). Contact / motion module 130 receives contact data from the touch-sensitive surface. Determining the movement of the contact point, as represented by the series of contact data, optionally includes determining the speed (magnitude), velocity (magnitude and direction), and / or acceleration (change in magnitude and / or direction) of the contact point. These actions are optionally applied to a single contact (e.g., one finger contact) or multiple simultaneous contacts (e.g., "multi-touch" / multiple finger contacts). In some embodiments, contact / motion module 130 and display controller 156 detect contacts on the touchpad.

[0111] In some embodiments, contact / motion module 130 uses a set of one or more intensity thresholds to determine whether an action has been performed by a user (e.g., to determine whether a user has “clicked” on an icon). In some embodiments, at least a subset of the intensity thresholds are determined according to software parameters (e.g., the intensity thresholds are not determined by the activation threshold of a particular physical actuator, but can be adjusted without modifying the physical hardware of device 100). For example, the mouse “click” threshold of a trackpad or touchscreen display can be set to any of a wide range of predefined thresholds without modifying the trackpad or touchscreen display hardware. Additionally, in some implementations, a user of the device is provided with a software setting to adjust one or more of the set of intensity thresholds (e.g., by adjusting individual intensity thresholds and / or by adjusting multiple intensity thresholds at once via a system-level click “intensity” parameter).

[0112] Contact / motion module 130 optionally detects gesture input by a user. Different gestures on the touch-sensitive surface have different contact patterns (e.g., different movements, timing, and / or intensities of detected contacts). Thus, gestures are optionally detected by detecting particular contact patterns. For example, detecting a finger tap gesture includes detecting a finger down event, followed by detecting a finger up (lift off) event at the same position (or substantially the same position) as the finger down event (e.g., the position of an icon). As another example, detecting a finger swipe gesture on the touch-sensitive surface includes detecting a finger down event, followed by one or more finger drag events, followed by detecting a finger up (lift off) event.

[0113] Graphics module 132 includes various known software components that render and display graphics on touchscreen 112 or other display, including components that modify the visual impact (e.g., brightness, transparency, saturation, contrast, or other visual properties) of the displayed graphics. As used herein, the term "graphics" includes any object that can be displayed to a user, including, but not limited to, characters, web pages, icons (such as user interface objects including soft keys), digital images, video, animation, etc.

[0114] In some embodiments, graphics module 132 stores data representing graphics to be used. Each graphic is optionally assigned a corresponding code. Graphics module 132 receives one or more codes specifying the graphics to be displayed, including coordinate data and other graphic property data, as needed, from an application or the like, and then generates screen image data to output to display controller 156.

[0115] The tactile feedback module 133 includes various software components for generating instructions used by the tactile output generator(s) 167 to generate tactile outputs at one or more locations on the device 100 in response to a user's interaction with the device 100.

[0116] Text input module 134 is optionally a component of graphics module 132 and provides a soft keyboard for entering text in various applications (e.g., contacts 137, email 140, IM 141, browser 147, and any other application requiring text input).

[0117] The GPS module 135 determines the location of the device and provides this information for use within various applications (e.g., to the phone 138 for use in location-based dialing, to the camera 143 as picture / video metadata, and to applications that provide location-based services such as weather widgets, local yellow pages widgets, and map / navigation widgets).

[0118] Application 136 optionally includes the following modules (or instruction sets), or a subset or superset thereof: • a contacts module 137 (sometimes called an address book or contact list); ●Telephone module 138, ●Videoconferencing module 139, ● an email client module 140; ● Instant messaging (IM) module 141, ●Training support module 142, camera module 143 for still images and / or video; ● Image management module 144; ●Video player module, ●Music player module, ● Browser module 147, ●Calendar module 148, • A widget module 149 optionally including one or more of a weather widget 149-1, a stock price widget 149-2, a calculator widget 149-3, an alarm clock widget 149-4, a dictionary widget 149-5, and other widgets obtained by the user, as well as user-created widgets 149-6; a widget creator module 150 for creating user-created widgets 149-6; ● Search module 151, A video and music player module 152 that integrates a video player module and a music player module; ● Memo module 153, Map module 154, and / or ●Online video module 155.

[0119] Examples of other applications 136 optionally stored in memory 102 include other word processing applications, other image editing applications, drawing applications, presentation applications, JAVA-enabled applications, encryption, digital rights management, voice recognition, and voice duplication.

[0120] Contacts module 137, along with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, is optionally used to manage an address book or contact list (e.g., stored in memory 102 or in the application internal state 192 of contacts module 137 in memory 370), including adding name(s) to the address book, deleting name(s) from the address book, associating phone number(s), email address(es), street address(es), or other information with names, associating images with names, categorizing and sorting names, providing phone numbers or email addresses to initiate and / or facilitate communication by phone 138, videoconferencing module 139, email 140, or IM 141, etc. used to manage the address book or contact list.

[0121] Telephone module 138, in conjunction with RF circuitry 108, audio circuitry 110, speaker 111, microphone 113, touchscreen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, is optionally used to enter a series of characters corresponding to a telephone number, access one or more telephone numbers in contacts module 137, modify entered telephone numbers, dial individual telephone numbers, place calls, and disconnect and hang up when the call is complete. As previously mentioned, wireless communication optionally uses any of a number of communication standards, protocols, and technologies.

[0122] Videoconferencing module 139 includes executable instructions to cooperate with RF circuitry 108, audio circuitry 110, speaker 111, microphone 113, touchscreen 112, display controller 156, optical sensor 164, optical sensor controller 158, contact / motion module 130, graphics module 132, text input module 134, contact module 137, and telephone module 138 to initiate, conduct, and end a videoconference between a user and one or more other participants according to the user's commands.

[0123] Email client module 140, in conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, includes executable instructions for composing, sending, receiving, and managing emails in response to user commands. In conjunction with image management module 144, email client module 140 greatly facilitates the creation and sending of emails with still or video images captured by camera module 143.

[0124] Instant messaging module 141, in cooperation with RF circuitry 108, touchscreen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, includes executable instructions for entering a series of characters corresponding to an instant message, modifying previously entered characters, sending individual instant messages (e.g., using Short Message Service (SMS) or Multimedia Message Service (MMS) protocols for telephony-based instant messaging, or XMPP, SIMPLE, or IMPS for Internet-based instant messaging), receiving instant messages, and viewing received instant messages. In some embodiments, sent and / or received instant messages optionally include graphics, photos, audio files, video files, and / or other attachments, such as those supported by MMS and / or Enhanced Messaging Service (EMS). As used herein, "instant messaging" refers to both telephony-based messages (e.g., messages sent using SMS or MMS) and Internet-based messages (e.g., messages sent using XMPP, SIMPLE, or IMPS).

[0125] In conjunction with the RF circuitry 108, touchscreen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, GPS module 135, map module 154, and music player module, the training support module 142 includes executable instructions to create workouts (e.g., with time, distance, and / or calorie burn goals), communicate with training sensors (sports devices), receive training sensor data, calibrate sensors used to monitor workouts, select and play music for workouts, and display, store, and transmit workout data.

[0126] Camera module 143, in conjunction with touchscreen 112, display controller 156, optical sensor(s) 164, optical sensor controller 158, contact / motion module 130, graphics module 132, and image management module 144, includes executable instructions to capture and store still images or video (including video streams) in memory 102, modify characteristics of the still images or video, or delete the still images or video from memory 102.

[0127] Image management module 144, in conjunction with touch screen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, and camera module 143, includes executable instructions for arranging, modifying (e.g., editing), or otherwise manipulating, labeling, deleting, presenting (e.g., in a digital slideshow or album), and storing still and / or video images.

[0128] Browser module 147, in conjunction with RF circuitry 108, touch screen 112, display controller 156, contact / motion module 130, graphics module 132, and text input module 134, contains executable instructions for browsing the Internet according to user commands, including retrieving, linking to, receiving, and displaying web pages or portions thereof, as well as attachments and other files linked to web pages.

[0129] The calendar module 148 includes executable instructions to cooperate with the RF circuitry 108, the touch screen 112, the display controller 156, the contact / motion module 130, the graphics module 132, the text input module 134, the email client module 140, and the browser module 147 to create, display, modify, and store calendars and data associated with the calendars (e.g., calendar items, to-do lists, etc.) according to the user's instructions.

[0130] Widget module 149, in conjunction with RF circuitry 108, touchscreen 112, display controller 156, touch / motion module 130, graphics module 132, text input module 134, and browser module 147, optionally provides mini-applications (e.g., weather widget 149-1, stock price widget 149-2, calculator widget 149-3, alarm clock widget 149-4, and dictionary widget 149-5) downloaded and used by a user, or mini-applications created by a user (e.g., user-created widget 149-6). In some embodiments, a widget includes an HTML (Hypertext Markup Language) file, a CSS (Cascading Style Sheets) file, and a JavaScript file. In some embodiments, a widget includes an XML (Extensible Markup Language) file and a JavaScript file (e.g., Yahoo! Widgets).

[0131] The widget creator module 150, in conjunction with the RF circuitry 108, the touch screen 112, the display controller 156, the contact / motion module 130, the graphics module 132, the text input module 134, and the browser module 147, is optionally used by a user to create a widget (e.g., turn a user-specified portion of a web page into a widget).

[0132] The search module 151 includes executable instructions for working in conjunction with the touch screen 112, the display controller 156, the contact / motion module 130, the graphics module 132, and the text input module 134 to search for text, music, sound, images, video, and / or other files in the memory 102 that match one or more search criteria (e.g., one or more user-specified search terms) in accordance with a user's commands.

[0133] Video and music player module 152 includes executable instructions that, in conjunction with touchscreen 112, display controller 156, contact / motion module 130, graphics module 132, audio circuitry 110, speaker 111, RF circuitry 108, and browser module 147, enable a user to download and play pre-recorded music and other sound files stored in one or more file formats, such as MP3 or AAC files, as well as executable instructions for displaying, presenting, or otherwise playing videos (e.g., on touchscreen 112 or on an external display connected via external port 124). In some embodiments, device 100 optionally includes the functionality of an MP3 player, such as an iPod (a trademark of Apple Inc.).

[0134] The notes module 153 includes executable instructions for cooperating with the touch screen 112, the display controller 156, the contact / motion module 130, the graphics module 132, and the text input module 134 to create and manage notes, to-do lists, and the like according to the user's commands.

[0135] Map module 154, in conjunction with RF circuitry 108, touchscreen 112, display controller 156, contact / motion module 130, graphics module 132, text input module 134, GPS module 135, and browser module 147, is used to receive, display, modify, and store maps and data associated with maps (e.g., driving directions, data regarding businesses and other points of interest at or near a particular location, and other location-based data), optionally in accordance with user instructions.

[0136] Online video module 155, in conjunction with touchscreen 112, display controller 156, contact / motion module 130, graphics module 132, audio circuitry 110, speaker 111, RF circuitry 108, text input module 134, email client module 140, and browser module 147, contains instructions that enable a user to access, browse for, receive (e.g., by streaming and / or downloading), and play (e.g., on the touchscreen or on an external display connected via external port 124) particular online videos, send emails with links to particular online videos, and otherwise manage online videos in one or more file formats, such as H.264. In some embodiments, instant messaging module 141 is used to send links to particular online videos, rather than email client module 140. For additional description of online video applications, see U.S. Provisional Patent Application No. 60 / 936,562, filed June 20, 2007, entitled "Portable Multifunction Device, Method, and Graphical User Interface for Playing Online Videos," and U.S. Patent Application No. 11 / 968,067, filed December 31, 2007, entitled "Portable Multifunction Device, Method, and Graphical User Interface for Playing Online Videos," the contents of which are incorporated herein by reference in their entireties.

[0137] The above-identified modules and applications each correspond to sets of executable instructions that perform one or more of the functions previously described and methods described herein (e.g., the computer-implemented methods and other information processing methods described herein). These modules (e.g., sets of instructions) need not be implemented as separate software programs, procedures, or modules; thus, in various embodiments, various subsets of these modules are optionally combined or otherwise reconfigured. For example, a video player module is optionally combined with a music player module into a single module (e.g., video and music player module 152 of FIG. 1A). In some embodiments, memory 102 optionally stores a subset of the above-identified modules and data structures. Additionally, memory 102 optionally stores additional modules and data structures not described above.

[0138] In some embodiments, device 100 is a device in which operation of a predetermined set of functions on the device is performed solely via a touchscreen and / or touchpad. Using the touchscreen and / or touchpad as the primary input control device for operation of device 100 optionally reduces the number of physical input control devices (push buttons, dials, etc.) on device 100.

[0139] The set of predefined functions performed solely through the touchscreen and / or touchpad optionally includes navigation between user interfaces. In some embodiments, the touchpad, when touched by a user, navigates device 100 to a main menu, home menu, or root menu from any user interface displayed on device 100. In such embodiments, a "menu button" is implemented using the touchpad. In some other embodiments, the menu button is a physical push button or other physical input control device rather than a touchpad.

[0140] 1B is a block diagram illustrating exemplary components for event processing, according to some embodiments. In some embodiments, memory 102 (FIG. 1A) or 370 (FIG. 3) includes an event sorter 170 (e.g., within operating system 126) and a separate application 136-1 (e.g., any of applications 137-151, 155, 380-390 described above).

[0141] Event sorter 170 receives the event information and determines which application 136-1 to deliver the event information to and application view 191 for application 136-1. Event sorter 170 includes event monitor 171 and event dispatcher module 174. In some embodiments, application 136-1 includes application internal state 192 that indicates the current application view(s) that are displayed on touch-sensitive display 112 when the application is active or running. In some embodiments, device / global internal state 157 is used by event sorter 170 to determine which application(s) are currently active, and application internal state 192 is used by event sorter 170 to determine which application(s) is / are currently active.

[0142] In some embodiments, application internal state 192 includes additional information such as one or more of resume information to be used when application 136-1 resumes execution, user interface state information indicating or ready to display information being displayed by application 136-1, state cues that allow the user to return to a previous state or view of application 136-1, and redo / undo cues for previous actions taken by the user.

[0143] Event monitor 171 receives event information from peripherals interface 118. The event information includes information about a sub-event (e.g., a user touch as part of a multi-touch gesture on touch-sensitive display 112). Peripherals interface 118 transmits information it receives from I / O subsystem 106 or sensors such as proximity sensor 166, accelerometer(s) 168, and / or microphone 113 (via audio circuitry 110). The information that peripherals interface 118 receives from I / O subsystem 106 includes information from touch-sensitive display 112 or the touch-sensitive surface.

[0144] In some embodiments, event monitor 171 sends requests to peripherals interface 118 at predetermined intervals. In response, peripherals interface 118 transmits event information. In other embodiments, peripherals interface 118 transmits event information only when there is a significant event (e.g., receipt of an input above a predetermined noise threshold and / or for more than a predetermined duration).

[0145] In some embodiments, the event sorter 170 also includes a hit view determination module 172 and / or an active event recognizer determination module 173 .

[0146] Hit view determination module 172 provides software procedures that determine where a sub-event occurred within one or more views when touch-sensitive display 112 is displaying more than one view. A view consists of the controls and other elements that a user can see on the display.

[0147] Another aspect of a user interface associated with an application is the set of views, sometimes referred to herein as application views or user interface windows, in which information is displayed and touch-based gestures occur. The application views (of individual applications) in which touches are detected optionally correspond to programmatic levels within the application's programmatic or view hierarchy. For example, the lowest-level view in which a touch is detected is optionally referred to as a hit view, and the set of events that are recognized as appropriate inputs is optionally determined based at least in part on the hit view of the initial touch that initiates the touch gesture.

[0148] Hit view determination module 172 receives information related to sub-events of a touch-based gesture. When an application has multiple views organized in a hierarchy, hit view determination module 172 identifies the hit view as the lowest view in the hierarchy that should process the sub-events. In most situations, the hit view is the lowest-level view in which an initiating sub-event occurs (e.g., the first sub-event in a series of sub-events that form an event or potential event). Once a hit view is identified by hit view determination module 172, the hit view typically receives all sub-events related to the same touch or input source as the touch or input source identified as the hit view.

[0149] Active event recognizer determination module 173 determines which view(s) in the view hierarchy should receive a particular sequence of sub-events. In some embodiments, active event recognizer determination module 173 determines that only the hit view should receive a particular sequence of sub-events. In other embodiments, active event recognizer determination module 173 determines that all views that contain the physical location of the sub-events are actively participating views, and therefore determines that all actively participating views should receive a particular sequence of sub-events. In other embodiments, even if a touch sub-event is completely confined to the area associated with one particular view, views higher in the hierarchy still remain actively participating views.

[0150] Event dispatcher module 174 dispatches event information to event recognizers (e.g., event recognizer 180). In embodiments that include active event recognizer determination module 173, event dispatcher module 174 delivers event information to the event recognizers determined by active event recognizer determination module 173. In some embodiments, event dispatcher module 174 stores event information in an event queue, which is retrieved by individual event receivers 182.

[0151] In some embodiments, operating system 126 includes event sorter 170. Alternatively, application 136-1 includes event sorter 170. In still other embodiments, event sorter 170 is a stand-alone module or is part of another module stored in memory 102, such as contact / motion module 130.

[0152] In some embodiments, application 136-1 includes multiple event handlers 190 and one or more application views 191, each containing instructions for processing touch events that occur within a separate view of the application's user interface. Each application view 191 of application 136-1 includes one or more event recognizers 180. Typically, an individual application view 191 includes multiple event recognizers 180. In other embodiments, one or more of the event recognizers 180 are part of a separate module, such as a user interface kit or a higher-level object from which application 136-1 inherits methods and other properties. In some embodiments, individual event handlers 190 include one or more of data updater 176, object updater 177, GUI updater 178, and / or event data 179 received from event sorter 170. Event handler 190 optionally utilizes or invokes data updater 176, object updater 177, or GUI updater 178 to update application internal state 192. Alternatively, one or more of the application views 191 include one or more respective event handlers 190. Also, in some embodiments, one or more of the data updater 176, the object updater 177, and the GUI updater 178 are included in individual application views 191.

[0153] A separate event recognizer 180 receives event information (e.g., event data 179) from event sorter 170 and identifies events from the event information. Event recognizer 180 includes an event receiver 182 and an event comparator 184. In some embodiments, event recognizer 180 also includes metadata 183 and at least a subset of event delivery instructions 188 (optionally including sub-event delivery instructions).

[0154] The event receiver 182 receives event information from the event sorter 170. The event information includes information about a sub-event, e.g., a touch or a movement of a touch. Depending on the sub-event, the event information also includes additional information, such as the location of the sub-event. When the sub-event involves a movement of a touch, the event information also optionally includes the speed and direction of the sub-event. In some embodiments, the event includes a rotation of the device from one orientation to another (e.g., from portrait to landscape or vice versa), and the event information includes corresponding information about the device's current orientation (also called the device's posture).

[0155] The event comparator 184 compares the event information with predefined event or sub-event definitions and determines the event or sub-event, or determines or updates the state of the event or sub-event, based on the comparison. In some embodiments, the event comparator 184 includes an event definition 186. The event definition 186 includes definitions of events (e.g., a predefined set of sub-events), such as Event 1 (187-1) and Event 2 (187-2). In some embodiments, sub-events within an event (187) include, for example, touch start, touch end, touch movement, touch cancellation, and multiple touches. In one example, the definition for Event 1 (187-1) is a double tap on a displayed object. The double tap includes, for example, a first touch on a displayed object relative to a predetermined phase (touch start), a first lift-off (touch end) relative to the predetermined phase, a second touch on a displayed object relative to the predetermined phase (touch start), and a second lift-off (touch end) relative to the predetermined phase. In another example, a definition of event 2 (187-2) is a drag on a displayed object. Drag includes, for example, a touch (or contact) on the displayed object to a predetermined stage, a movement of the touch across the touch-sensitive display 112, and a lift-off of the touch (touch end). In some embodiments, the event also includes information about one or more associated event handlers 190.

[0156] In some embodiments, event definition 187 includes definitions of events for individual user interface objects. In some embodiments, event comparator 184 performs a hit test to determine which user interface objects are associated with the sub-event. For example, if a touch is detected on touch-sensitive display 112 in an application view in which three user interface objects are displayed on touch-sensitive display 112, event comparator 184 performs a hit test to determine which of the three user interface objects is associated with the touch (sub-event). If each displayed object is associated with a separate event handler 190, event comparator 184 uses the results of the hit test to determine which event handler 190 to activate. For example, event comparator 184 selects the event handler associated with the sub-event and object that triggers the hit test.

[0157] In some embodiments, the definition of an individual event 187 also includes a delay action that delays delivery of the event information until it is determined whether a set of sub-events corresponds to the event recognizer's event type.

[0158] If the individual event recognizer 180 determines that the sequence of sub-events does not match any of the events in the event definition 186, the individual event recognizer 180 enters an event disabled, event failed, or event finished state, after which it ignores the next sub-event of the touch-based gesture. In this situation, any other event recognizers that remain active for the hit view continue to track and process sub-events of the ongoing touch gesture.

[0159] In some embodiments, individual event recognizers 180 include metadata 183 with configurable properties, flags, and / or lists that indicate to actively participating event recognizers how the event delivery system should perform sub-event delivery. In some embodiments, metadata 183 includes configurable properties, flags, and / or lists that indicate how event recognizers interact with each other or how event recognizers are allowed to interact with each other. In some embodiments, metadata 183 includes configurable properties, flags, and / or lists that indicate how sub-events are delivered to various levels in a view or programmatic hierarchy.

[0160] In some embodiments, an individual event recognizer 180 activates an event handler 190 associated with an event when one or more specific sub-events of the event are recognized. In some embodiments, the individual event recognizer 180 delivers event information associated with the event to the event handler 190. Activating the event handler 190 is separate from sending (and deferring sending) sub-events to the individual hit view. In some embodiments, the event recognizer 180 pops a flag associated with the recognized event, and the event handler 190 associated with the flag catches the flag and performs a predetermined process.

[0161] In some embodiments, the event delivery instructions 188 include sub-event delivery instructions that deliver event information about a sub-event without activating an event handler. Instead, the sub-event delivery instructions deliver the event information to an event handler associated with a set of sub-events or to an actively participating view. The event handler associated with the set of sub-events or the actively participating view receives the event information and performs a predetermined process.

[0162] In some embodiments, data updater 176 creates and updates data used by application 136-1. For example, data updater 176 updates phone numbers used in contacts module 137 or stores video files used in a video player module. In some embodiments, object updater 177 creates and updates objects used by application 136-1. For example, object updater 177 creates new user interface objects or updates the positions of user interface objects. GUI updater 178 updates the GUI. For example, GUI updater 178 prepares display information and sends the display information to graphics module 132 for display on the touch-sensitive display.

[0163] In some embodiments, event handler(s) 190 include or have access to data updater 176, object updater 177, and GUI updater 178. In some embodiments, data updater 176, object updater 177, and GUI updater 178 are included in a single module of an individual application 136-1 or application view 191. In other embodiments, they are included in two or more software modules.

[0164] It should be understood that the foregoing description of event processing of a user's touch on a touch-sensitive display also applies to other forms of user input for operating multifunction device 100 using input devices, although not all of them are initiated on the touchscreen. For example, mouse movements and mouse button presses, contact movements such as tapping, dragging, scrolling on a touchpad, optionally coordinated with single or multiple keyboard presses or holds, pen stylus input, device movement, verbal commands, detected eye movements, biometric input, and / or any combination thereof, optionally utilize as inputs corresponding to sub-events that define the recognized event.

[0165] FIG. 2 illustrates portable multifunction device 100 having touchscreen 112, according to some embodiments. The touchscreen optionally displays one or more graphics within user interface (UI) 200. In this embodiment, as well as other embodiments described below, a user may select one or more of the graphics by performing a gesture on the graphics, for example, using one or more fingers 202 (not drawn to scale) or one or more styluses 203 (not drawn to scale). In some embodiments, selection of one or more graphics is performed when the user breaks contact with the one or more graphics. In some embodiments, the gesture optionally includes one or more taps, one or more swipes (left to right, right to left, upward and / or downward), and / or rolling (right to left, left to right, upward and / or downward) of a finger in contact with device 100. In some implementations or situations, accidental contact with a graphic does not select the graphic, for example, if the gesture corresponding to selection is a tap, a swipe gesture sweeping over an application icon optionally does not select the corresponding application.

[0166] Device 100 also optionally includes one or more physical buttons, such as a "home" button or menu button 204. As previously mentioned, menu button 204 is optionally used to navigate to any application 136 within a set of applications running on device 100. Alternatively, in some embodiments, the menu button is implemented as a soft key within a GUI displayed on touchscreen 112.

[0167] In some embodiments, device 100 includes touchscreen 112, menu button 204, pushbutton 206 for powering the device on / off and locking the device, volume control button(s) 208, subscriber identity module (SIM) card slot 210, headset jack 212, and external docking / charging port 124. Pushbutton 206 is optionally used to power the device on / off by pressing and holding the button down for a predetermined period of time, to lock the device by pressing and releasing the button before the predetermined time has elapsed, and / or to unlock the device or initiate the unlocking process. In alternative embodiments, device 100 also accepts verbal input via microphone 113 for activating or deactivating certain functions. Device 100 also optionally includes one or more contact intensity sensors 165 for detecting the intensity of a contact on touchscreen 112 and / or one or more tactile output generators 167 for generating a tactile output for a user of device 100.

[0168] FIG. 3 is a block diagram of an exemplary multifunction device having a display and a touch-sensitive surface, according to some embodiments. Device 300 need not be portable. In some embodiments, device 300 is a laptop computer, a desktop computer, a tablet computer, a multimedia player device, a navigation device, an educational device (such as a child's learning toy), a gaming system, or a control device (e.g., a home or commercial controller). Device 300 typically includes one or more processing units (CPUs) 310, one or more network or other communication interfaces 360, memory 370, and one or more communication buses 320 interconnecting these components. Communication bus 320 optionally includes circuitry (sometimes called a chipset) that interconnects and controls communication between system components. Device 300 includes input / output (I / O) interface 330, including display 340, which is typically a touchscreen display. I / O interface 330 also optionally includes a keyboard and / or mouse (or other pointing device) 350 and a touchpad 355, a tactile output generator 357 that generates tactile output on device 300 (e.g., similar to tactile output generator(s) 167 described above with reference to FIG. 1A ), and sensors 359 (e.g., light, acceleration, proximity, touch-sensing, and / or contact intensity sensors similar to contact intensity sensor(s) 165 described above with reference to FIG. 1A ). Memory 370 includes high-speed random-access memory such as DRAM, SRAM, DDR RAM, or other random-access solid-state memory devices, and optionally includes non-volatile memory such as one or more magnetic disk storage devices, optical disk storage devices, flash memory devices, or other non-volatile solid-state storage devices. Memory 370 optionally includes one or more storage devices located remotely from CPU(s) 310.In some embodiments, memory 370 stores programs, modules, and data structures similar to, or a subset of, programs, modules, and data structures stored in memory 102 of portable multifunction device 100 (FIG. 1A). Additionally, memory 370 optionally stores additional programs, modules, and data structures not present in memory 102 of portable multifunction device 100. For example, memory 370 of device 300 optionally stores drawing module 380, presentation module 382, word processing module 384, website creation module 386, disc authoring module 388, and / or spreadsheet module 390, whereas memory 102 of portable multifunction device 100 (FIG. 1A) optionally does not store these modules.

[0169] Each of the above-identified elements of FIG. 3 is optionally stored in one or more of the memory devices mentioned above. Each of the above-identified modules corresponds to an instruction set that performs the function described above. The above-identified modules or programs (e.g., instruction sets) need not be implemented as separate software programs, procedures, or modules; thus, in various embodiments, various subsets of these modules are optionally combined or otherwise reconfigured. In some embodiments, memory 370 optionally stores a subset of the above-identified modules and data structures. Additionally, memory 370 optionally stores additional modules and data structures not described above.

[0170] Attention is now optionally directed to user interface embodiments, for example, as implemented on portable multifunction device 100.

[0171] 4A shows an exemplary user interface for a menu of applications on portable multifunction device 100, according to some embodiments. A similar user interface is optionally implemented on device 300. In some embodiments, user interface 400 includes the following elements, or a subset or superset thereof: ● signal strength indicator(s) 402 for wireless communication(s), such as cellular and Wi-Fi signals; ●Time 404, ●Bluetooth indicator 405, ● Battery status indicator 406, Tray 408 with icons of frequently used applications, such as: an icon 416 for the phone module 138, labeled "Phone," that optionally includes an indicator 414 of the number of missed calls or voicemail messages; icon 418 of the email client module 140, labeled "Mail," optionally including an indicator 410 of the number of unread emails; ○ An icon 420 for the Browser module 147, labeled "Browser"; and ○ An icon 422 for the video and music player module 152, also referred to as the iPod (trademark of Apple Inc.) module 152, labeled "iPod", and ● Icons of other applications, such as: ○ Icon 424 of IM module 141, labeled "Messages"; ○ Icon 426 of the calendar module 148, labeled "Calendar" ○ Icon 428 of the image management module 144, labeled "Photos" ○ An icon 430 of the camera module 143, labeled "camera"; ○ Icon 432 of the online video module 155, labeled "Online Video"; Icon 434 of Stock Price Widget 149-2, labeled "Stock Price" ○ Icon 436 of the map module 154, labeled "Map"; Icon 438 of weather widget 149-1, labeled "Weather" ○ Icon 440 of alarm clock widget 149-4, labeled "Clock" ○ Icon 442 of Training Support Module 142, labeled "Training Support"; icon 444 of the Notes module 153, labeled "Notes"; and A settings application or module icon 446 labeled "Settings" that provides access to settings for the device 100 and its various applications 136.

[0172] Note that the icon labels shown in FIG. 4A are merely exemplary. For example, icon 422 of video and music player module 152 is labeled "Music" or "Music Player." Other labels are optionally used for various application icons. In some embodiments, the label for an individual application icon includes the name of the application that corresponds to the individual application icon. In some embodiments, the label of a particular application icon is different from the name of the application that corresponds to that particular application icon.

[0173] 4B shows an exemplary user interface on a device (e.g., device 300 of FIG. 3) that has a touch-sensitive surface 451 (e.g., tablet or touchpad 355 of FIG. 3) that is separate from display 450 (e.g., touchscreen display 112). Device 300 also optionally includes one or more contact intensity sensors (e.g., one or more of sensors 359) that detect the intensity of a contact on touch-sensitive surface 451, and / or one or more tactile output generators 357 that generate a tactile output to a user of device 300.

[0174] Although some of the following examples are given with reference to input on touchscreen display 112 (which combines a touch-sensitive surface and a display), in some embodiments, the device detects input on a touch-sensitive surface that is separate from the display, as shown in FIG. 4B . In some embodiments, the touch-sensitive surface (e.g., 451 in FIG. 4B ) has a primary axis (e.g., 452 in FIG. 4B ) that corresponds to a primary axis (e.g., 453 in FIG. 4B ) on the display (e.g., 450). According to these embodiments, the device detects contact with touch-sensitive surface 451 (e.g., 460 and 462 in FIG. 4B ) at locations that correspond to respective locations on the display (e.g., in FIG. 4B , 460 corresponds to 468 and 462 corresponds to 470). In this way, user input (e.g., contacts 460 and 462, and their movement) detected by the device on the touch-sensitive surface (e.g., 451 in FIG. 4B ) is used by the device to operate a user interface on the display (e.g., 450 in FIG. 4B ) of the multifunction device when the touch-sensitive surface is separate from the display. It should be understood that similar methods are optionally used for the other user interfaces described herein.

[0175] Additionally, while the following examples are given primarily with reference to finger input (e.g., finger contact, finger tap gesture, finger swipe gesture), it should be understood that in some embodiments, one or more of the finger inputs are replaced with input from another input device (e.g., mouse-based input or stylus input). For example, a swipe gesture is optionally replaced by a mouse click (e.g., instead of a contact) followed by movement of a cursor along the path of the swipe (e.g., instead of movement of the contact). As another example, a tap gesture is optionally replaced by a mouse click (e.g., instead of detecting a contact and then ceasing contact detection) while the cursor is positioned over the location of the tap gesture. Similarly, it should be understood that when multiple user inputs are detected simultaneously, multiple computer mice are optionally used simultaneously, or a mouse and finger contacts are optionally used simultaneously.

[0176] FIG. 5A shows an exemplary personal electronic device 500. Device 500 includes a main body 502. In some embodiments, device 500 can include some or all of the functionality described with respect to devices 100 and 300 (e.g., FIGS. 1A-4B ). In some embodiments, device 500 has a touch-sensitive display screen 504, hereafter touchscreen 504. Alternatively, or in addition to touchscreen 504, device 500 has a display and a touch-sensitive surface. Similar to devices 100 and 300, in some embodiments, touchscreen 504 (or the touch-sensitive surface) optionally includes one or more intensity sensors that detect the intensity of contact (e.g., touches) being applied. The one or more intensity sensors of touchscreen 504 (or the touch-sensitive surface) can provide output data representing the intensity of the touch. The user interface of device 500 can respond to touches based on their intensity, meaning that touches of different intensities can invoke different user interface actions on device 500.

[0177] For exemplary techniques for detecting and processing touch intensity, see, for example, related applications International Patent Application No. PCT / US2013 / 040061, filed May 8, 2013, entitled "Device, Method, and Graphical User Interface for Displaying User Interface Objects Corresponding to an Application," published as International Publication No. WO / 2013 / 169849, and International Patent Application No. PCT / US2013 / 069483, filed November 11, 2013, entitled "Device, Method, and Graphical User Interface for Transitioning Between Touch Input to Display Output Relationships," published as International Publication No. WO / 2014 / 105276, each of which is incorporated herein by reference in its entirety.

[0178] In some embodiments, device 500 has one or more input mechanisms 506 and 508. Input mechanisms 506 and 508, if included, may be physical. Examples of physical input mechanisms include push buttons and rotatable mechanisms. In some embodiments, device 500 has one or more attachment mechanisms. Such attachment mechanisms, if included, may allow device 500 to be attached to, for example, hats, eyewear, earrings, necklaces, shirts, jackets, bracelets, watch bands, chains, pants, belts, shoes, wallets, backpacks, etc. These attachment mechanisms allow device 500 to be worn by a user.

[0179] FIG. 5B shows an exemplary personal electronic device 500. In some embodiments, device 500 can include some or all of the components described with respect to FIGS. 1A, 1B, and 3. Device 500 has a bus 512 that operably couples an I / O section 514 to one or more computer processors 516 and a memory 518. The I / O section 514 can be connected to a display 504, which can have touch-sensing components 522 and, optionally, an intensity sensor 524 (e.g., a contact intensity sensor). Additionally, the I / O section 514 can be connected to a communication unit 530 that receives application and operating system data using Wi-Fi, Bluetooth, near field communication (NFC), cellular, and / or other wireless communication technologies. Device 500 can include input mechanisms 506 and / or 508. Input mechanism 506 is optionally, for example, a rotatable input device or a depressible and rotatable input device. In some embodiments, input mechanism 508 is optionally a button.

[0180] In some embodiments, the input mechanism 508 is optionally a microphone. The personal electronic device 500 optionally includes various sensors, such as a GPS sensor 532, an accelerometer 534, an orientation sensor 540 (e.g., a compass), a gyroscope 536, a motion sensor 538, and / or combinations thereof, all of which may be operably connected to the I / O section 514.

[0181] The memory 518 of the personal electronic device 500 may include one or more non-transitory computer-readable storage media that store computer-executable instructions that, when executed by one or more computer processors 516, can cause the computer processors to perform the techniques described below, including, for example, processes 800, 900, 1100, 1300, 1500, and 1700. A computer-readable storage medium may be any medium that can tangibly contain or store computer-executable instructions used by or in connection with an instruction execution system, apparatus, or device. In some embodiments, the storage medium is a transient computer-readable storage medium. In some embodiments, the storage medium is a non-transitory computer-readable storage medium. Non-transitory computer-readable storage media may include, but are not limited to, magnetic storage, optical storage, and / or semiconductor storage. Examples of such storage include magnetic disks, optical disks based on CD, DVD, or Blu-ray technology, as well as persistent solid-state memory such as flash, solid-state drives, and the like. Personal electronic device 500 is not limited to the components and configuration of FIG. 5B and may include other or additional components in multiple configurations.

[0182] As used herein, the term "affordance" refers to a user-interactive graphical user interface object, optionally displayed on a display screen of device 100, 300, and / or 500 (FIGS. 1A, 3, and 5A-5B). For example, images (e.g., icons), buttons, and text (e.g., hyperlinks) each, optionally, constitute an affordance.

[0183] As used herein, the term “focus selector” refers to an input element that indicates the current portion of the user interface with which the user is interacting. In some implementations involving a cursor or other location marker, the cursor acts as the “focus selector,” such that when input (e.g., a press input) is detected on a touch-sensitive surface (e.g., touchpad 355 of FIG. 3 or touch-sensitive surface 451 of FIG. 4B) while the cursor is positioned over a particular user interface element (e.g., a button, window, slider, or other user interface element), the particular user interface element is adjusted according to the detected input. In some implementations involving a touchscreen display that allows direct interaction with user interface elements on the touchscreen display (e.g., touch-sensitive display system 112 of FIG. 1A or touchscreen 112 of FIG. 4A), a detected contact on the touchscreen acts as the “focus selector,” such that when input (e.g., a press input by contact) is detected at the location of a particular user interface element (e.g., a button, window, slider, or other user interface element) on the touchscreen display, the particular user interface element is adjusted according to the detected input. In some implementations, focus is moved from one region of the user interface to another region of the user interface without a corresponding cursor movement or contact movement on the touchscreen display (e.g., by using the tab key or arrow keys to move focus from one button to another), and in these implementations, the focus selector moves to follow the movement of focus between various regions of the user interface. Regardless of the specific form the focus selector takes, the focus selector is generally a user interface element (or contact on a touchscreen display) that is controlled by the user to communicate the user's intended interaction with the user interface (e.g., by indicating to the device the element of the user interface through which the user intends to interact).For example, the position of a focus selector (e.g., a cursor, touch, or selection box) over an individual button while a press input is detected on a touch-sensitive surface (e.g., a touchpad or touchscreen) indicates that the user intends to activate that individual button (and not other user interface elements shown on the device's display).

[0184] As used herein and in the claims, the term "characteristic intensity" of a contact refers to a characteristic of that contact based on one or more intensities of the contact. In some embodiments, the characteristic intensity is based on a plurality of intensity samples. The characteristic intensity is optionally based on a predetermined number of intensity samples, i.e., a set of intensity samples collected during a predetermined time (e.g., 0.05, 0.1, 0.2, 0.5, 1, 2, 5, 10 seconds) associated with a predetermined event (e.g., after detecting the contact, before detecting lift-off of the contact, before or after detecting the start of contact movement, before detecting the end of the contact, before or after detecting an increase in the intensity of the contact, and / or before or after detecting a decrease in the intensity of the contact). The characteristic intensity of the contact is optionally based on one or more of the maximum intensity of the contact, the mean intensity of the contact, the average intensity of the contact, the top 10 percentile intensity of the contact, half the maximum intensity of the contact, 90 percent of the maximum intensity of the contact, etc. In some embodiments, the duration of the contact is used in determining the characteristic intensity (e.g., when the characteristic intensity is an average of the intensity of the contact over time). In some embodiments, the characteristic intensity is compared to a set of one or more intensity thresholds to determine whether an action is performed by the user. For example, the set of one or more intensity thresholds optionally includes a first intensity threshold and a second intensity threshold. In this example, a contact having a characteristic intensity that does not exceed the first threshold results in a first action, a contact having a characteristic intensity above the first intensity threshold but not above the second intensity threshold results in a second action, and a contact having a characteristic intensity above the second threshold results in a third action. In some embodiments, the comparison between the characteristic intensity and the one or more thresholds is not used to determine whether to perform the first action or the second action, but rather to determine whether to perform one or more actions (e.g., whether to perform an individual action or to forgo performing an individual action).

[0185] In some embodiments, a portion of the gesture is identified for purposes of determining the characteristic intensity. For example, the touch-sensitive surface optionally receives successive swipe contacts that transition from a start position to an end position, where the intensity of the contact increases. In this example, the characteristic intensity of the contact at the end position is optionally based on only a portion of the successive swipe contacts (e.g., only the portion of the swipe contact at the end position), rather than the entire swipe contact. In some embodiments, a smoothing algorithm is optionally applied to the intensity of the swipe contacts before determining the characteristic intensity of the contacts. For example, the smoothing algorithm optionally includes one or more of an unweighted moving average smoothing algorithm, a triangular smoothing algorithm, a median filter smoothing algorithm, and / or an exponential smoothing algorithm. In some situations, these smoothing algorithms eliminate narrow spikes or dips in the intensity of the swipe contact for purposes of determining the characteristic intensity.

[0186] The intensity of the contact on the touch-sensitive surface is optionally characterized relative to one or more intensity thresholds, such as a contact-detection intensity threshold, a light pressure intensity threshold, a deep pressure intensity threshold, and / or one or more other intensity thresholds. In some embodiments, the light pressure intensity threshold corresponds to an intensity at which the device performs an action normally associated with clicking a physical mouse button or trackpad. In some embodiments, the deep pressure intensity threshold corresponds to an intensity at which the device performs an action different from an action normally associated with clicking a physical mouse button or trackpad. In some embodiments, when a contact is detected having a characteristic intensity below the light pressure intensity threshold (e.g., and above a nominal contact-detection intensity threshold below which the contact is not detected), the device moves the focus selector according to movement of the contact on the touch-sensitive surface without performing an action associated with the light pressure intensity threshold or the deep pressure intensity threshold. In general, unless otherwise specified, these intensity thresholds are consistent across various sets of values for a user interface.

[0187] An increase in the characteristic intensity of a contact from an intensity below the light pressure intensity threshold to an intensity between the light pressure intensity threshold and the deep pressure intensity threshold may be referred to as inputting a "light press." An increase in the characteristic intensity of a contact from an intensity below the deep pressure intensity threshold to an intensity above the deep pressure intensity threshold may be referred to as inputting a "deep press." An increase in the characteristic intensity of a contact from an intensity below the contact-detection intensity threshold to an intensity between the contact-detection intensity threshold and the light pressure intensity threshold may be referred to as detecting a contact on the touch surface. A decrease in the characteristic intensity of a contact from an intensity above the contact-detection intensity threshold to an intensity below the contact-detection intensity threshold may be referred to as detecting a lift-off of the contact from the touch surface. In some embodiments, the contact-detection intensity threshold is zero. In some embodiments, the contact-detection intensity threshold is greater than zero.

[0188] In some embodiments described herein, one or more actions are performed in response to detecting a gesture including an individual pressure input or in response to detecting an individual pressure input performed by an individual contact (or multiple contacts), where the individual pressure input is detected based at least in part on detecting an increase in intensity of the contact (or multiple contacts) above a pressure input intensity threshold. In some embodiments, the individual action is performed in response to detecting an increase in intensity of the individual contact above the pressure input intensity threshold (e.g., a "downstroke" of the individual pressure input). In some embodiments, the pressure input includes an increase in intensity of the individual contact above the pressure input intensity threshold followed by a decrease in intensity of the contact below the pressure input intensity threshold, and the individual action is performed in response to detecting a subsequent decrease in intensity of the individual contact below the pressure input threshold (e.g., an "upstroke" of the individual pressure input).

[0189] In some embodiments, the device employs intensity hysteresis to avoid accidental input, sometimes referred to as “jitter,” and the device defines or selects a hysteresis intensity threshold that has a predetermined relationship to the pressure input intensity threshold (e.g., the hysteresis intensity threshold is X intensity units below the pressure input intensity threshold, or the hysteresis intensity threshold is 75%, 90%, or some reasonable percentage of the pressure input intensity threshold). Thus, in some embodiments, the pressure input includes an increase in the intensity of a discrete contact above the pressure input intensity threshold followed by a decrease in the intensity of the contact below the hysteresis intensity threshold corresponding to the pressure input intensity threshold, and a discrete action is performed in response to detecting a subsequent decrease in the intensity of the discrete contact below the hysteresis intensity threshold (e.g., an “upstroke” of the discrete pressure input). Similarly, in some embodiments, a pressure input is detected only when the device detects an increase in the intensity of the contact from an intensity below the hysteresis intensity threshold to an intensity above the pressure input intensity threshold, and optionally a subsequent decrease in the intensity of the contact to an intensity below the hysteresis intensity, and a distinct action is performed in response to detecting the pressure input (e.g., an increase in the intensity of the contact or a decrease in the intensity of the contact, as the case may be).

[0190] For ease of explanation, descriptions of operations performed in response to a pressure input associated with a pressure input intensity threshold, or a gesture including a pressure input, are optionally triggered in response to detecting any of: an increase in the intensity of the contact above the pressure input intensity threshold; an increase in the intensity of the contact from an intensity below a hysteresis intensity threshold to an intensity above the pressure input intensity threshold; a decrease in the intensity of the contact below the pressure input intensity threshold; and / or a decrease in the intensity of the contact below a hysteresis intensity threshold corresponding to the pressure input intensity threshold. Further, in examples where an operation is described as being performed in response to detecting a decrease in the intensity of the contact below a pressure input intensity threshold, the operation is optionally performed in response to detecting a decrease in the intensity of the contact below a hysteresis intensity threshold corresponding to and lower than the pressure input intensity threshold.

[0191] As used herein, an "installed application" refers to a software application that has been downloaded onto an electronic device (e.g., device 100, 300, and / or 500) and is ready to launch (e.g., opened) on the device. In some embodiments, a downloaded application becomes an installed application by an installation program that extracts program portions from a downloaded package and integrates the extracted portions with the computer system's operating system.

[0192] As used herein, the terms "open application" or "running application" refer to a software application that has retained state information (e.g., as part of device / global internal state 157 and / or application internal state 192). An open or running application is, optionally, any one of the following types of application: ● the active application currently displayed on the display screen of the device on which the application is being used; Background applications (or background processes) that are not currently displayed, but for which one or more processes are being processed by one or more processors; and A suspended or hibernated application that is not running but has state information stored in memory (volatile and non-volatile, respectively) and that can be used to resume execution of the application.

[0193] As used herein, the term "closed application" refers to a software application that does not have retained state information (e.g., state information for a closed application is not stored in the device's memory). Thus, closing an application includes stopping and / or removing the application process for the application and removing state information for the application from the device's memory. Generally, opening a second application during a first application does not close the first application. When the second application is displayed and the first application is stopped from being displayed, the first application becomes a background application.

[0194] Attention is now directed to embodiments of user interfaces (“UIs”) and related processes implemented on an electronic device such as portable multifunction device 100, device 300, or device 500.

[0195] 6A-6Z illustrate exemplary user interfaces for managing visual content in media, according to some embodiments. The user interfaces in these figures are used to illustrate processes described below, including the process in FIG.

[0196] 6A shows computer system 600 (e.g., an electronic device) displaying a camera user interface, optionally including a live preview 630 that extends from the top of the display of computer system 600 to the bottom of the display of computer system 600. In some embodiments, computer system 600 optionally includes one or more characteristics of device 100, device 300, or device 500. In some embodiments, computer system 600 is a tablet, a phone, a laptop, a desktop, etc.

[0197] Live preview 630 is a representation of the field of view ("FOV") of one or more cameras of computer system 600. In some embodiments, live preview 630 is a representation of a partial FOV. In some embodiments, live preview 630 is based on images detected by one or more camera sensors. In some embodiments, computer system 600 captures images using multiple camera sensors and combines them to display live preview 630. In some embodiments, computer system 600 captures images using a single camera sensor and displays live preview 630.

[0198] 6A includes an indicator region 602 and a control region 606 that are positioned relative to the live preview 630 such that the indicators and controls may be displayed simultaneously with the live preview 630. The camera viewing region 604 is substantially free of indicators and / or controls overlaid. As shown in FIG. 6A , the camera user interface includes a visual boundary 608 that indicates the boundary between the indicator region 602 and the camera viewing region 604, and the boundary between the camera viewing region 604 and the control region 606.

[0199] As shown in FIG. 6A , indicator area 602 includes indicators such as a flash indicator 602a and an animated image indicator 602b. Flash indicator 602a indicates whether the flash is on (e.g., active), off (e.g., inactive), or in another mode (e.g., automatic mode). In FIG. 6A , flash indicator 602a indicates to the user that flash mode is off and that no flash operation will be used when computer system 600 is capturing media. Animated image indicator 602b indicates whether the camera is configured to capture a single image or multiple images (e.g., in response to detecting a request to capture media). In some embodiments, indicator area 602 is superimposed on live preview 630 and optionally includes a colored (e.g., gray, semi-transparent) overlay.

[0200] 6A, the camera viewing area 604 includes a live preview 630 and zoom controls (e.g., affordances) 622. The zoom controls 622 include a 0.5× zoom control 622a, a 1× zoom control 622b, and a 2× zoom control 622c. As shown in FIG. 6A, the 1× zoom control 622b is bolded and enlarged relative to the other zoom controls, indicating that the 1× zoom control 622b is selected and that the computer system 600 is displaying the live preview 630 at the “1×” zoom level.

[0201] As shown in FIG. 6A , control area 606 includes a camera mode control (e.g., control) 620, a shutter control 610, a camera switcher control 614, and a representation of a media collection 612. In FIG. 6A , camera mode controls 620a-620e are displayed, and “Photo” camera mode 620c is shown in bold, indicating that computer system 600 is configured to capture photographic media when shutter control 610 is active. Thus, when activated, shutter control 610 causes computer system 600 to capture media (e.g., a photo when shutter control 610 is activated in FIG. 6A ) using one or more camera sensors based on the current state of live preview 630 and the current state of the camera application (e.g., which camera mode is selected). The captured media is stored locally on computer system 600 and / or transmitted to a remote server for storage. The camera switcher control 614, when enabled, causes the computer system 600 to switch between showing different camera views in the live preview 630, for example, by switching between a rear-facing camera sensor and a front-facing camera sensor. The representation of the media collection 612 shown in FIG. 6A is a representation of media (e.g., images, videos) most recently captured by the computer system 600. In some embodiments, in response to detecting a gesture directed at the media collection 612, the computer system 600 displays a user interface similar to the user interface shown in FIG. 7B (described below). In some embodiments, the indicator area 602 is superimposed on the live preview 630 and optionally includes a colored (e.g., gray, semi-transparent) overlay. In FIG. 6A, the computer system detects a tap input 650a on (and / or directed at) the shutter control 610.

[0202] As shown in Figure 6B, in response to detecting the tap input 650a, the computer system 600 begins capturing media to capture the live preview 630 of Figure 6A and displays the new representation in the media collection 612. In Figure 6B, the new representation is the representation of the live preview 630 of Figure 6A (e.g., captured in response to detecting the tap input 650a on the shutter control 610), and is displayed at the top of the media collection 612 in Figure 6B because the new representation corresponds to the most recently captured representation of media.

[0203] As shown in Figure 6B, live preview 630 includes a representation showing person 640 standing behind a tree, with person 640's head and any portion of person 640's body not obscured by the tree. Placed on top of the tree is sign 642, which includes text portion 642a (e.g., "LOST DOG") and text portion 642b (e.g., a paragraph of text beginning with "LOVEABLE"). In Figure 6B, the text within text portions 642a-642b is not visually prominent, and in the embodiment shown in Figure 6B, the text within text portions 642a and 642b is small and cannot be easily read by a user viewing computer system 600. In Figure 6B, computer system 600 detects an unpinch input 650b on live preview 630.

[0204] As shown in Figure 6C, in response to detecting the pinch-unpinch input 650b, the computer system 600 replaces the display of the 2x zoom control 622c of Figure 6B with a display of a 2.5x zoom control. The computer system 600 also updates the live preview 630 to reflect the change in zoom level, such that objects within the field of view of one or more cameras are displayed at a "2.5x" zoom level (e.g., as indicated by the newly displayed and selected (e.g., enlarged and bolded) 2.5x zoom control 622d) instead of the "1x" zoom level of Figure 6B.

[0205] In comparison to Figure 6B, text portions 642a-642b in Figure 6C are visually more prominent (e.g., larger and more readable) than text portions 642a-642b in Figure 6B. In Figure 6C, a determination has been made that text portions 642a-642b (and / or the text included in text portions 642a-642b) each meet a set of prominence criteria. Text portions 642a-642b in Figure 6C meet the set of prominence criteria because each text portion occupies a larger-than-threshold portion (e.g., 10%) of live preview 630 and / or each text portion includes text larger than a threshold size (e.g., larger than 6 pt font). In some embodiments, one or more text portions meet a set of prominence criteria based on other criteria, such as whether the individual text portions include one or more types of text (e.g., email, phone number, quick response (“QR”) code, etc.), whether the individual text portions are displayed at or near a particular location (e.g., a central location) in the live preview 630, whether the individual text portions are relevant based on the context of the media displayed as the live preview 630, etc. (as described in more detail in connection with Figures 7A-7L, 8, and 9).

[0206] As shown in FIG. 6C, upon determining that text portions 642a-642b meet the set of prominence criteria (and / or because at least a portion of the text meets the set of prominence criteria), computer system 600 displays brackets 636a around text portions 642a-642b and displays text management controls 680 to the right of zoom control 622d within camera viewing area 604.

[0207] Returning to Figure 6B, bracket 636a and text management control 680 were not displayed in Figure 6B because a determination was made that text portions 642a-642b did not meet the set of prominence criteria (and / or because no portions of the text met the set of prominence criteria). In Figure 6B, a determination was made that text portions 642a-642b did not meet the set of prominence criteria because they did not occupy a portion of live preview 630 that exceeded the threshold and did not contain text greater than a threshold size. In some embodiments (as shown in Figures 6B-6C), the determination of whether an individual text portion meets the set of prominence criteria is made based on how / when the text portion is currently displayed within live preview 630, rather than solely on whether live preview 630 contains text (and / or text portion).

[0208] Returning to FIG. 6C , brackets 636a are placed around the image of the dog on sign 642 because the image of the dog is placed between text portions 642a-642b. In some embodiments, multiple brackets are displayed, such that a bracket is displayed around text portion 642a and another bracket is displayed around text portion 642b. In some embodiments, multiple brackets are displayed because a determination has been made that multiple text portions (e.g., "multiple portions of text") meet a set of prominence criteria and an object is placed between the text portions. In some embodiments, when an object is not placed between the multiple portions of text, only one bracket is displayed around the multiple portions of text. In some embodiments, if text portion 642a meets a set of prominence criteria but text portion 642b does not meet the set of prominence criteria, brackets are displayed around text portion 642a and no brackets are displayed around text portion 642b (or vice versa). In some embodiments, the computer system 600 indicates that an individual text portion (e.g., a portion of text) meets a set of prominence criteria by highlighting the individual text portion in other ways, such as highlighting, bolding, sizing, and a box around the individual text portion, in addition to and / or instead of displaying parentheses around the individual text portion.

[0209] As shown in FIG. 6C , computer system 600 displays text type indications 638a-638b (e.g., underlining) to indicate that a particular type of text (e.g., email, address, phone number, QR code, etc.) has been detected (e.g., a data detector) within text portion 642b. In FIG. 6C , text type indication 638a is displayed below "123 Main Street" to indicate that an address has been detected, and text type indication 638b is displayed below "123-4567" to indicate that a phone number has been detected. In some embodiments, when a text type indicator is displayed below a portion of text, a user can perform an action by selecting the portion of text and / or the text type indicator (e.g., as further described below in connection with FIGS. 6M-6N).

[0210] 6C-6D illustrate an exemplary embodiment in which computer system 600 is moved in a physical environment. Figures 6C-6D include a graphical representation 660 illustrating an original position 660a of computer system 600 (e.g., in Figures 6C-6D) relative to a changed position 660b of computer system 600 (e.g., in Figure 6D) in the physical environment. As shown in Figure 6C, computer system 600 is in original position 660a. In Figure 6C, the position of computer system 600 has been changed.

[0211] As shown in FIG. 6D , in response to a change in the position of computer system 600 (e.g., from original position 660a to modified position 660b), computer system 600 translates the live preview upward. In FIG. 6D , live preview 630 is translated upward such that the top portion of live preview 630 in FIG. 6C (e.g., the portion including text portion 642a) ceases to be displayed and a new bottom portion of live preview 630 (as shown in FIG. 6D ) is newly displayed. In FIG. 6D , a determination is made that text portion 642a does not satisfy the set of prominence criteria, but that text portion 642b (continues to) satisfy the set of prominence criteria. Here, because text portion 642a is no longer displayed as part of live preview 630 (e.g., in the camera viewing area) in FIG. 6D , a determination is made that text portion 642a does not satisfy the set of prominence criteria. As shown, because text portion 642a does not meet the set of prominence criteria and text portion 642b does meet the set of prominence criteria, computer system 600 displays brackets 636b around text portion 642b (rather than text portion 642a) and ceases displaying brackets 636a. In other words, computer system 600 dynamically changes brackets 636a to brackets 636b in accordance with changes in a determination of whether one or more text portions (e.g., text portions currently displayed as part of live preview 630) meet and / or do not meet the set of prominence criteria. Thus, one or more determinations of whether one or more text portions meet a set of prominence criteria are dynamic and may change when the live preview 630 changes in response to a request to zoom in (e.g., an unpinch input) / zoom out (e.g., a pinch input) or pan (e.g., a swipe input right, left, up, or down), and / or when the live preview 630 changes in response to a movement (e.g., forward, back, up, or down) of the computer system 600 and / or one or more cameras of the computer system 600.In some embodiments, if one or more determinations of whether one or more text portions meet a set of prominence criteria change, the display of one or more brackets (e.g., brackets 636a-636b) and / or the display of text management controls 680 change (as described further below in connection with FIGS. 7A-7L, 8, and 9). In some embodiments, while computer system 600 displays brackets 636a around text portion 642b (and / or in response to detecting text in live preview 630), computer system 600 darkens and / or decreases the saturation (e.g., saturation, hue, and / or hue) of portions of live preview 630 that do not have text (e.g., a picture of a dog) while maintaining the saturation and / or brightness of text portion 642b (and / or other portions of the text). In some embodiments, as part of darkening portions of live preview 630 that do not have text and maintaining the brightness of text portion 642b, computer system 600 displays text portion 642b at a greater brightness than portions of live preview 630 that do not have text.

[0212] As shown in Figure 6D, computer system 600 continues to display text management control 680 because a determination has been made that text portion 642b (continues to) satisfy the set of prominence criteria. In Figure 6D, text management control 680 is displayed because at least one determination has been made that the currently displayed text portion (e.g., of live preview 630) satisfies the set of prominence criteria, regardless of whether another text portion (e.g., text portion 642a) may (or does not) continue to satisfy the set of prominence criteria. In Figure 6D, computer system 600 is returned to original position 660a.

[0213] As shown in Figure 6E, in response to the computer system 600 being in the original position 660a, the computer system 600 re-displays the live preview 630 using one or more techniques such as those described above in connection with Figure 6C. In Figure 6E, the computer system 600 detects a tap input 650e on the text management control 680.

[0214] As shown in Figure 6F, in response to detecting tap input 650e, computer system 600 changes the display of text management control 680. Specifically, computer system 600 displays text management control 680 of Figure 6F in an active and / or selected state (e.g., as indicated by the bolding of text management control 680 in Figure 6F) and ceases to display text management control 680 in an inactive and / or deselected state (e.g., as indicated by the lack of bolding of text management control 680 in Figure 6E).

[0215] As shown in FIG. 6F , in response to detecting tap input 650e, computer system 600 highlights text portions 642a-642b and dims other portions of live preview 630 (and / or other objects in the field of view of one or more cameras), e.g., person 640, the image of the dog on sign 642, and the tree displayed in live preview 630. In conjunction with dimming other portions of live preview 630, computer system 600 ceases displaying one or more controls in camera viewing area 604 (e.g., zoom control 622 in FIG. 6E ). Computer system 600 also dims (or ceases displaying) portions of the camera user interface, such as indicators in indicator area 602 and controls in camera control area 606. In some embodiments, some of the dimmed indicators and / or controls in the camera user interface of FIG. 6F are not selectable (e.g., do not cause computer system 600 to perform an action when selected). In some embodiments, some of the indicators and / or controls remain selectable and / or are not dimmed in response to detecting tap input 650e. In some embodiments, computer system 600 maintains the display of some controls in camera viewing area 604 in response to detecting tap input 650e. In some embodiments, computer system 600 emphasizes portions 642a-642b by increasing the size of the text within text portions 642a-642b, highlighting the text within text portions 642a-642b, displaying boxes around text portions 642a-642b, etc. In some embodiments, dimming portions of live preview 630 includes decreasing the saturation of portions of live preview 630 that do not have text (e.g., photos of dogs) while maintaining the saturation of text portions 642a and 642b (e.g., using similar techniques as described above in connection with FIG. 6D ).

[0216] In particular, in FIG. 6F , the portion of text highlighted in response to detecting input 650e is the portion of text that was surrounded by parentheses (e.g., parentheses 636a) when input 650e was received in FIG. 6E . In some embodiments, the parentheses around the portion of text indicate to the user which text will be highlighted and / or managed by the user when a selection of the text management control 680 is made. In some embodiments, one or more portions of text that are displayed via live preview 630 when input is received on the text management control 680 but that do not have parentheses surrounding the text are not highlighted in response to a selection of the text management control 680 (e.g., in FIG. 7F below, “BRAND” is not highlighted when the text management control 680 is selected in FIG. 7F ). In some embodiments, one or more portions of text that are displayed via live preview 630 when input is received on the text management control 680 but that do not have parentheses surrounding the text are highlighted in response to a selection of the text management control 680 (e.g., when a determination is made that one or more portions of text meet a set of prominence criteria).

[0217] As shown in FIG. 6F , in response to detecting tap input 650e, computer system 600 also displays text management options 682 and instructions 684 (e.g., "Swipe or tap to select text") that indicate one or more inputs / gestures that can be used to select a subset of text from within text portions 642a-642b. In FIG. 6F , text management options 682 are options for managing text portions 642a-642b. In particular, text management options 682 include copy option 682a, select all option 682b, search option 682c, and share option 682d. In some embodiments, in response to receiving input directed to copy option 682a, computer system 600 copies the selected text (e.g., the text within text portions 642a-642b in FIG. 6F ) and / or saves the selected text in a copy / paste buffer, which allows the selected text to be pasted in response to receiving a request to paste the selected text. In some embodiments, in response to receiving input directed to select all option 682b, computer system 600 selects all of the highlighted text on computer system 600. In some embodiments, when computer system 600 selects all of the text within the selected text, computer system 600 highlights the selected text. In some embodiments, in response to receiving input directed to search option 682c, computer system 600 searches for the selected text (e.g., the highlighted text portion in FIG. 6F ) via a search application (e.g., a web application, a dictionary application, a personal assistant application) and / or displays one or more definitions and resources for the highlighted and / or selected text.In some embodiments, in response to receiving input directed to share option 682d, computer system 600 initiates a process to share the selected text via one or more applications (e.g., email, text messaging, word processing, social media applications) (e.g., one or more predetermined applications). In some embodiments, as part of initiating the process to share the selected text, computer system 600 displays a scrollable list of applications, and selecting an application from the scrollable list of applications causes computer system 600 to share the selected text using the selected application. In some embodiments, the scrollable list of applications is displayed simultaneously with a portion of live preview 630 (e.g., including one or more of text portions 642a-642b). In FIG. 6F , computer system 600 detects tap input 650f on a portion of live preview 630 (e.g., a portion within a shaded area of live preview 630 and / or a portion of live preview 630 that does not include text portions 642a-642b and / or text management controls 680).

[0218] As shown in FIG. 6G, in response to detecting tap input 650f, computer system 600 displays text management control 680 in an inactive state, de-emphasizes text portions 642a-642b, illuminates the remainder of live preview 630 and the camera user interface, and ceases displaying text management options 682 and instructions 684. Also, in response to detecting tap input 650f, computer system 600 re-displays brackets 636a because a determination has been made that text portions 642a-642b (continue to) meet the set of prominence criteria. Effectively, in response to detecting tap input 650f, the camera user interface is returned to the state it was in before tap input 650e was detected on text management control 680. In FIG. 6G, computer system 600 detects tap input 650g on text management control 680.

[0219] As shown in Figure 6H, in response to detecting tap input 650g, computer system 600 displays the camera interface of Figure 6H using one or more techniques such as those described above in connection with Figure 6F. In particular, in Figure 6H (and Figure 6F), computer system 600 highlights text portions 642a-642b because they meet a set of prominence criteria. In some embodiments, computer system 600 dims one or more portions of the text that do not meet the set of prominence criteria in response to detecting tap input 650g. In Figure 6H, computer system 600 detects tap input 650h on text portion 642b.

[0220] As shown in FIG. 6I, in response to detecting tap input 650h, computer system 600 selects text portion 642b and repositions text management options 682 so that text management options 682 appear over text portion 642b in FIG. 6I instead of appearing over text portion 642a (e.g., as shown in FIG. 6H). Text management options 682 are repositioned to indicate that text management options can be used to manage the text in text portion 642b, but not the text in text portion 642a. Thus, in other words, computer system 600 changes the text selected to be managed using text management options 682 in response to detecting an input (e.g., a swipe or tap) that selects a particular portion of text.

[0221] Notably, the live preview 630 in FIG. 6I does not include the person 640 that was included in the live preview 630 in FIG. 6H because the person 640 has moved behind the tree in the live preview 630 in FIG. 6I and is therefore not within the field of view of one or more cameras of the computer system 600. As shown in FIG. 6I, the live preview 630 continues to update to reflect changes in the field of view of one or more cameras of the computer system 600 while the text management control 680 is actively displayed and / or while the text management options 682 are displayed. In some embodiments, the live preview 630 does not continue to update while the text management control 680 is actively displayed and / or while the text management options 682 are displayed. Thus, in embodiments in which the live preview 630 is not updated, the computer system 600 maintains a display of the portion of the person 640 poking out from behind the tree in the live preview 630 of FIG. 6I. In FIG. 6I, the computer system 600 detects an unpinch input 650i.

[0222] As shown in FIG. 6J , in response to detecting an unpinch input 650i, computer system 600 displays live preview 630 at the increased zoom level and maintains the display of text portion 642b and text management options 682. In some embodiments, because text portion 642b is selected, computer system 600 continues to display at least a subset of text portion 642b in response to a request to zoom in (e.g., an unpinch input) (and / or zoom out, pan, and / or move computer system 600 and / or one or more cameras of computer system 600). In some embodiments, the display of the selected text portion (e.g., text portion 642b) is static. Thus, in some embodiments in which the selected text portion is static, computer system 600 continues to display the selected text portion (e.g., while the camera user interface is still displayed) regardless of whether the selected text portion remains within the field of view of one or more cameras (e.g., as computer system 600 is moved, panned, and / or zoomed) (e.g., as further described below in connection with FIGS. 6L-6M ). In FIG. 6J, computer system 600 detects a tap input 650j on the word "Fluffy" contained in text portion 642b.

[0223] As shown in Figure 6K, in response to detecting tap input 650j, computer system 600 selects and highlights the word "Fluffy." In Figure 6K, only the selected word "Fluffy" can be managed with text management options 682 displayed in Figure 6K. In Figure 6K, computer system 600 detects left swipe input 650k beginning with the word "Fluffy."

[0224] As shown in Figure 6L, in response to detecting leftward swipe input 650k, computer system 600 selects and highlights multiple words included in text portion 642b based on the direction of swipe input 650k. As shown in Figure 6L, the words "THE NAME FLUFFY" are highlighted to indicate that "THE NAME FLUFFY" has been selected based on swipe input 650k. In Figure 6L, only the selected words "THE NAME FLUFFY" can be managed with text management options 682 displayed in Figure 6L.

[0225] 6L-6M illustrate an exemplary embodiment in which computer system 600 is moved in a physical environment while computer system 600 continues to display the selected text portion (or a subset of the selected text portion), regardless of whether the selected text portion remains within the field of view of one or more cameras (e.g., as described further below in connection with FIGS. 6L-6M). Figures 6L-6M include a graphical representation 660 showing an original position 660a of computer system 600 (e.g., in FIGS. 6L-6M) relative to a changed position 660c of computer system 600 (e.g., in FIG. 6M).

[0226] As shown in Figure 6L, tree symbol 646 represents the static portion of the tree displayed in live preview 630 of Figures 6L-6M. In Figure 6L, tree symbol 646 is displayed below text portion 642b. In Figure 6L, the position of computer system 600 has been changed.

[0227] As shown in FIG. 6M , in response to a change in the position of computer system 600 (e.g., as indicated by changed position 660c relative to original position 660a), computer system 600 updates the live preview so that tree symbol 646 appears above text portion 642b. In particular, in FIG. 6M , text portion 642b of computer system 600 is no longer within the field of view of one or more cameras, and text portion 642b is positioned where text portion 642b appears in live preview 630 of FIG. 6M (e.g., as evidenced by tree symbol 646 moving to a higher position in live preview 630). However, because a subset of text portion 642b (e.g., "THE NAME FLUFFY") has been selected, computer system 600 continues to display text portion 642b in live preview 630 of FIG. 6M . In some embodiments, computer system 600 displays only the selected subset of text portion 642b without displaying other portions of text portion 642b that are not selected. In some embodiments, computer system 600 does not update live preview 630 when text is selected and the camera is moved (and / or zoomed / panned) in the physical environment. In some embodiments (e.g., as further described below in connection with Figures 7A-7L, 8, and 9), computer system 600 selects different portions of text (e.g., if text appears toward the bottom of a tree in live preview 630) in response to computer system 600 and / or the camera of computer system 600 being moved (e.g., and / or zoomed / panned). In Figure 6M, computer system 600 detects input 650m on "123-4567" with text type indication 638b displayed below.

[0228] 6N, in response to detecting input 650m, and because a determination has been made that input 650m is a tap input and that "123-4567" corresponds to a telephone number, computer system 600 displays a telephone dialer user interface and automatically (e.g., without requiring user input on a keypad and / or contact information card) initiates a telephone call to "123-4567." In some embodiments, a confirmation screen is displayed before computer system 600 initiates the telephone call.

[0229] As shown in FIG. 6O, in response to detecting input 650m, and because a determination has been made that input 650m is a long press input and that "123-4567" corresponds to a phone number, computer system 600 displays phone number management options 692, including call option 692a, send message option 692b, add contact option 692c, and copy option 692d. As shown in FIG. 6O, computer system 600 displays options for managing certain types of text (e.g., email, phone number, QR code) that differ from managing other types of text (as shown by the display of text management option 682 in FIG. 6L when "THE NAME FLUFFY" is selected, as opposed to the display of phone number management option 692 when "123-4567" is selected in FIG. 6O). In some embodiments, in response to detecting input directed at call option 692a, computer system 600 initiates a phone call to "123-4567" (e.g., using similar techniques as described above in connection with FIG. 6N). In some embodiments, in response to detecting input directed toward send message option 692b, computer system 600 initiates a process to send a message to "123-4567" (e.g., display a text management application). In some embodiments, in response to detecting input directed toward add contact option 692c, computer system 600 initiates a process to add a contact to a contact list that has "123-4567" as a phone number in the contact's information. In some embodiments, in response to detecting input directed toward copy option 692d, computer system 600 copies "123-4567" using one or more techniques such as those described above in connection with copy option 682a in FIG. 6F.

[0230] 6P-6T show an example embodiment in which a QR code is displayed in a live preview 630. In some embodiments, the QR code can be replaced with other types of matrices and / or barcodes.

[0231] As shown in FIG. 6P , computer system 600 displays QR code 668 concurrently with QR code identifier 670 (e.g., “CAFE32.COM”) in live preview 630. In some embodiments, the QR code identifier identifies one or more of a website, a contact, a cellular plan, an email address, a calendar invite / event, a location (e.g., a GPS location), text, a video, a phone number, a WiFi network, an application, and / or an instance of an application. QR code identifier 670 includes an indication of the information identified by the QR code. In FIG. 6P , QR code 668 is within the field of view of one or more cameras of computer system 600, and QR code identifier 670 is not within the field of view of one or more cameras of computer system 600. Because a determination has been made that QR code 668 corresponds to (e.g., identifies) a destination website belonging to “CAFE32.COM,” computer system 600 displays QR code identifier 670. In FIG. 6P, the computer system 600 detects an input 650 p 1 and / or an input 650 p 2 in the camera viewing area 604 .

[0232] As shown in FIG. 6Q , in response to detecting input 650p1 and / or input 650p2 (and based on determining that at least one of the inputs is a tap input and / or a long press input), computer system 600 displays notification 674 including a preview of a website (e.g., a “CAFE32.COM” address). In some embodiments, the preview of the web address includes the full web address (e.g., “http:\\cafe32.com\menu”) and / or an image from the web address. In some embodiments, computer system 600, in response to detecting one or more inputs, displays notification 674 instead of navigating to the web address corresponding to QR code 668 to minimize the possibility that the user unintentionally navigates to a website site corresponding to QR code 668. In some embodiments, in response to detecting input 650p1 on QR code 668, computer system 600 displays notification 674 (e.g., without automatically navigating to the website). In some embodiments, in response to detecting input 650p2 on the QR code identifier, computer system 600 automatically navigates (e.g., without displaying notification 674) to a website corresponding to QR code 668 (e.g., using one or more similar techniques as described below in connection with computer system 600's response to tap input 650q). In FIG. 6Q, computer system 600 detects tap input 650q on notification 674.

[0233] As shown in FIG. 6R, in response to detecting tap input 650q, computer system 600 automatically navigates to (and / or opens) the web address corresponding to the QR code via web application 678.

[0234] As shown in Figure 6S, the computer system 600 simultaneously displays a QR code 668 with a QR code identifier 670 using one or more techniques such as those described above in connection with Figure 6P. In Figure 6S, the computer system 600 detects a tap input 650s on a text management control 680.

[0235] As shown in FIG. 6T, in response to detecting tap input 650s, computer system 600 displays QR code management options 672, including share option 672a, copy link option 672b, add to reading list option 672c, and open link option 672d. As described above in connection with FIG. 6O, computer system 600 displays different options for managing some specific types of text than for managing other types of text. In some embodiments, in response to detecting input directed at share option 672a, computer system 600 initiates a process to share the web address and / or link corresponding to the QR code (e.g., using one or more similar techniques described with respect to input directed at share option 682d of FIG. 6F). In some embodiments, in response to detecting input directed at copy link option 672b, computer system 600 copies the web address and / or link corresponding to the QR code (e.g., using one or more techniques as described above with respect to copy option 682a of FIG. 6F). In some embodiments, in response to detecting input directed toward add to reading list option 672c, computer system 600 initiates a process to add the web address and / or link corresponding to the QR code to a list of items (e.g., one or more articles, books, websites, etc.). In some embodiments, in response to detecting input directed toward open link option 672d, computer system 600 navigates to (and / or opens) the web address corresponding to the QR code via web application 678 (e.g., using similar techniques as described above in connection with FIG. 6R).

[0236] In some embodiments, QR code management options 672 include one or more options that are dynamically selected based on the type of resource the QR code represents (e.g., the QR code displayed when text management control 680 is selected). For example, the type of resource represented by the QR code may include one or more of: a link to a website, a contact, a cellular plan, an email address, a calendar invite / event, a location (e.g., a GPS location), text, a video, a phone number, a WiFi network, an application, and / or an instance of an application. In some embodiments, QR code management options 672 include a first set of controls when the QR code represents a first type of resource and a second set of controls when the QR code represents a second type of resource that is different from the first type. In some embodiments, the first set of controls has a different amount of controls than the second set of controls. In some embodiments, a preview of the resource represented by the QR code is included in QR code management options 672 (e.g., when the QR code represents a string of text).

[0237] In some embodiments, QR code management options 672 include a different set of controls based on whether computer system 600 is in a locked or unlocked state. In some embodiments, if computer system 600 is in a locked state and the QR code represents a link to an application, a control option to install and / or open the application is displayed. In some embodiments, if computer system 600 is in an unlocked state, even if the application is installed, the link to open the application is not displayed (e.g., is suppressed) to avoid conveying information about which applications are installed on the device to an unauthorized user of the device. Optionally, instead of displaying a link to open the application, the device displays an option to use a portion of an available application without downloading the entire application. In some embodiments, computer system 600 displays a different set of controls (e.g., based on whether computer system 600 is in a locked or unlocked state) to limit the information given to an unauthorized user (e.g., information that can be used to determine whether an application represented by the QR code is installed and / or not installed on computer system 600).

[0238] 6U-6W illustrate an exemplary scenario in which computer system 600 displays a selection indicator around selected text separated into columns. In FIGS. 6U-6W, computer system 600 is oriented such that text in the environment is aligned with the field of view of one or more cameras of computer system 600. FIG. 6U shows computer system 600 displaying a live preview 630 including a representation of text portion 648 (e.g., a soccer player roster). In some embodiments, computer system 600 displays a representation of previously captured media including a representation of text portion 648, and one or more techniques described below in connection with FIGS. 6U-6W are used to select words within text portion 648.

[0239] As shown in FIG. 6U, text portion 648 includes a name column 648a, a position column 648b, a state column 648c, and a grade column 648d. The individual columns contain text detected by computer system 600 (e.g., using one or more techniques such as those described above in connection with FIGs. 6A-6F). As shown in FIG. 6U, computer system 600 highlights text portion 648 while reducing the visual prominence of portions of live preview 630 (e.g., the soccer ball) that do not contain text (e.g., using one or more techniques such as those described above in connection with FIGs. 6A-6F). Also, because computer system 600 detected text portion 648, computer system 600 places a box around text 648 to highlight text 648. As shown in FIG. 6U, computer system 600 displays text management control 680 as active (e.g., as indicated by the bolding of text management control 680) and displays text management options 682 (e.g., as described above in connection with FIG. 6F). In FIG. 6U, computer system 600 detects a first portion of swipe input 650u on name column 648a that moves from the header "name" of name column 648a to the header "position" of position column 648b.

[0240] As shown in FIG. 6V , in response to detecting the first portion of the swipe input 650u, the computer system 600 displays selection indicators 696 (e.g., a “gray highlight”) around all of the words in the name column 648a (“Name,” “Maria,” “Kate,” “Sarah,” and “Ashley”) and the header “position” in the position column 648b. The selection indicators 696 are positioned based on the location of the swipe input 650u. Because the first portion of the computer system 600’s swipe input 650u ends at the header “position” in the position column 648b, the computer system 600 displays selection indicators 696 around all of the words up to (e.g., including) the words in the name column 648a and including the header “position.” In some embodiments, because the first portion of the computer system 600’s swipe input 650u ends at the header “position” in the position column 648b, the computer system 600 does not include the header “position” in the position column 648b. In some embodiments where the end of the input ends at the location of the word “DEFENDER” in position column 648b (e.g., the third line of position column 648b), computer system 600 highlights all words up to the word “DEFENDER,” including all words in name column 648a, the header “position” in position column 648b (e.g., the first line of position column 648b), and the word “Forward” in the second line of position column 648b.

[0241] The shape of selection indicator 696 depends on whether the selected text (e.g., the text that selection indicator 696 surrounds) is aligned with computer system 600. In FIG. 6V, computer system 600 displays selection indicator 696 as a polygon with right angles (e.g., a shape with all right angles, referred to herein as a rectangle-based selection indicator). Selection indicator 696 is a rectangle-based selection indicator because a determination is made that the selected text (e.g., the text of selection indicator 696) is aligned with computer system 600 (e.g., and / or aligned with the field of view of one or more cameras of computer system 600) (e.g., as described in additional detail below in connection with FIGS. 6X-6Z). In FIG. 6V, computer system 600 detects a second portion of swipe input 650u, which is a rightward swipe input that moves from the header “position” in position column 648b to the header “state” in state column 648c.

[0242] As shown in FIG. 6W , because the computer system recognized the words in name column 648a as being in the same column, in response to detecting a second portion of swipe input 650u, computer system 600 expands selection indicator 696 to the right (e.g., using one or more techniques such as those described above in connection with FIGS. 6U-6W ) so that selection indicator 696 is displayed around the words (e.g., all of the words) in name column 648a and position column 648b and also around the header “state” in state column 648c. As shown in FIG. 6W , selection indicator 696 continues to be a rectangle-based selection indicator because the text portion continues to be aligned with the field of view of one or more cameras. In FIG. 6W , computer system 600 no longer detects input swipe input 650u. However, computer system 600 continues to display selection indicator 696 around the portion of the text.

[0243] 6X-6Z illustrate an exemplary scenario in which, when computer system 600 is oriented (e.g., oriented differently relative to an individual text portion than how computer system 600 of FIGS. 6U-6V was oriented relative to the individual text portion), computer system 600 displays selection indicators around selected text such that the text in the environment is not aligned with the field of view of one or more cameras of computer system 600. FIG. 6X shows computer system 600 displaying a live preview 630 including a representation of text portion 652 (e.g., a paragraph of text related to soccer). Text portion 652 is on paper in the environment, as captured by the field of view of one or more cameras of computer system 600. In some embodiments, computer system 600 displays a representation of previously captured media, including a representation of text portion 648, and one or more techniques described below in connection with FIGS. 6U-6W are used to select words within text portion 648.

[0244] In Figure 6X, text portion 652 is not aligned with the field of view of one or more cameras. In Figure 6X, computer system 600 is oriented in a location such that computer system 600 is not parallel to text portion 652 and / or rotated / tilted along an axis (z-axis) in the environment (e.g., the user is holding the phone at an angle and / or tilt such that the field of view of one or more cameras is not aligned with text portion 652). In Figure 6X, computer system 600 detects a diagonal swipe input 650x that moves from the word "while" in text portion 652 to the last period (".") in text portion 652.

[0245] As shown in FIG. 6Y , in response to detecting swipe input 650x, computer system 600 displays selection indicator 696 around a subset of text portion 652 from the word “while” in text portion 652 to the last period in text portion 652. As shown in FIG. 6Y , selection indicator 696 is a polygon with some angles that are not right angles (e.g., a shape with some acute angles and some obtuse angles, referred to herein as a non-rectangular-based selection indicator). The non-rectangular-based selection indicator is drawn by the computer system to match or appear to match (or substantially match or appear to substantially match) the orientation of text portion 652 in live preview 630 (e.g., if selection indicator 696 were a rectangular-based selection indicator on the surface including text portion 642, but as if the surface including text portion 642 were viewed from the same perspective as is viewed in FIGS. 6X-6Z ). As explained above, because a determination was made that text portion 652 is not aligned with computer system 600, selection indicator 696 in FIG. 6Y is non-rectangle-based (e.g., in contrast to selection indicator 696 in FIGS. 6V-6U, which is a rectangle-based selection indicator (e.g., relative to the orientation of the display of computer system 600)). In FIG. 6Y, computer system 600 detects swipe input 650y that moves from the word "while" to the word "synthetic" in text portion 652. Notably, swipe input 650y moves diagonally relative to computer system 600, but moves along the lines of the words in text portion 652. In some embodiments, even though the edges of selection indicator 696 are displayed diagonally relative to the edges of the computer system's display area, some or all of the edges of selection indicator 696 are positioned where the computer system determines they are parallel or perpendicular to the lines of text in text portion 652.In some embodiments, as the angle of the camera changes relative to the surface containing the text portion 642, the angle of the edge of the selection indicator 696 shifts within the viewing area to maintain the edge where it is determined to be parallel or perpendicular to the lines of text in the text portion 652.

[0246] 6Z , in response to detecting swipe input 650y, computer system 600 expands selection indicator 696 in the direction of swipe input 650y so that selection indicator 696 surrounds a subset of text portion 652 from the word “synthetic” to the final period within text portion 652 (e.g., if “while” is included in the portion of text). Even though selection indicator 696 has expanded, it still appears as a selection indicator with a non-rectangular base 696. Additionally, selection indicator 696 continues to be displayed around the portion of text after computer system 600 no longer detects swipe input 650y.

[0247] 7A-7L illustrate exemplary user interfaces for managing visual indicators for visual content in media using a computer system, according to some embodiments. The user interfaces in these figures are used to illustrate processes described below, including the process in FIG.

[0248] 7A shows a computer system 600 simultaneously displaying a media gallery user interface 710 that includes a thumbnail media representation 712 and a gallery area 702. The thumbnail media representation 712 includes thumbnail media representations 712a-712c, each representing a different media item (e.g., a media item captured at a different instance in time). The gallery area 702 includes a library control 702a (e.g., that, when selected, causes the computer system 600 to display the thumbnail media representation 712), a "for you" control 702b (e.g., that, when selected, causes the computer system 600 to display dynamically generated thumbnail representations of media items based on user preferences), an album control 702c (e.g., that, when selected, causes the computer system 600 to display thumbnail album representations, each representing a collection of media items), and a search control 702d (e.g., that, when selected, causes the computer system 600 to display a search user interface including one or more controls for searching for media items). In Figure 7A, the library control 702a is selected (e.g., as indicated by the bolding of the library control 702a), and the computer system 600 detects a tap input 750a on the thumbnail media representation 712a.

[0249] 7B , in response to detecting a tap input 750a, the computer system 600 displays a media viewer user interface 720 and ceases displaying the media gallery user interface 710. The media viewer user interface 720 includes a media viewer area 724 disposed between an application control area 722 and an application control area 726. The media viewer area 724 includes an enlarged representation 724a of the same media item as the thumbnail media representation 712a. The media viewer user interface 720 is substantially free of controls overlaid, while the application control area 722 and the application control area 726 are substantially overlaid with controls.

[0250] Enlarged representation 724a includes sign 642, which includes text portion 642a (e.g., "LOST DOG") and text portion 642b (e.g., a paragraph of text beginning with "LOVEABLE"), as described above in connection with FIG. 6B. The text in text portions 642a-642b is not visually prominent, and the text in text portions 642a-642b is small and cannot be easily read by a user viewing computer system 600. Additionally, enlarged representation 724a includes person 740 standing in front of a tree. Person 740 is wearing a hat that includes the word "BRAND" (e.g., text portion 742).

[0251] Application control area 722 optionally includes an indicator of the time (e.g., "7:54" in FIG. 7B) when the currently displayed magnified representation of media (e.g., magnified representation 724a) was captured, a cellular signal status indicator 720a that indicates the status of the cellular signal, and a battery level status indicator 720b that indicates the status of the remaining battery life of computer system 600. Application control area 722 also includes a back control 722a (e.g., that, when selected, causes computer system 600 to redisplay media gallery user interface 710) and an edit control 722b (e.g., that, when selected, causes computer system 600 to display a media editing user interface including one or more controls for editing a representation of the media item represented by currently displayed magnified representation 724a).

[0252] The application control area 726 includes a portion of the thumbnail media representations 712 (e.g., 712a-712c) displayed in a single row. The thumbnail media representation 712a is displayed as selected because the enlarged representation 724a is displayed in the media viewer area 724. In particular, the thumbnail media representation 712a is displayed as selected in FIG. 7B by being displayed with space from the other thumbnails (e.g., 712b and 712c). The application control area 726 also includes a send control 726b (e.g., that, when selected, causes the computer system 600 to initiate the process of sending the media item represented by the enlarged media representation), a favorite control 726c (e.g., that, when selected, causes the computer system 600 to mark / unmark the media item represented by the enlarged representation 724a as favorite media), and a trash control 726d (e.g., that, when selected, causes the computer system 600 to delete (or initiate the process of deleting) the media item represented by the enlarged representation 724a). In FIG. 7B, computer system 600 detects an unpinch input 750b on media viewer area 724 (e.g., at a location on the display of computer system 600 corresponding to media viewer area 724 and / or directed at a location on the display of computer system 600 corresponding to media viewer area 710).

[0253] As shown in Figure 7C, in response to detecting an unpinch input 750b, the computer system 600 updates the magnified representation 724a to reflect the change in zoom level, such that the display of the magnified representation 724a in Figure 7C is displayed at a greater zoom level than the display of the magnified representation 724a in Figure 7B. At the increased zoom level, the text portions 642a-642b in Figure 7C are larger and visually more prominent (e.g., larger, more readable) than the text portions 642a-642b in Figure 7B. In addition to updating the magnified representation 724a, the computer system 600 also expands the media viewer area 724 in Figure 7B so that the magnified representation 724a in Figure 7B occupies the portion of the display that the application control areas 722 and 726 previously occupied in Figure 7A.

[0254] 7C, a determination has been made (e.g., using one or more similar techniques as described above in connection with FIGS. 6A-6C) that the text in text portion 642a and the text in text portion 642b in FIG. 7C, respectively, do not meet a set of prominence criteria. Accordingly, computer system 600 does not display parentheses corresponding to (e.g., surrounding) text portions 642a-642b in FIG. 7B. Furthermore, because the text in text portions 642a-642b do not meet a set of prominence criteria (e.g., as described above in connection with FIG. 6B), computer system 600 does not display text management control 680.

[0255] In some embodiments, the set of prominence criteria includes criteria that are met when a determination is made that one or more of the text portions 642a-642b includes text that occupies a predetermined amount of space (e.g., 10%-100%) of the enlarged representation 724a. In some embodiments, the set of prominence criteria includes criteria that are met when a determination is made that one or more of the portions 642a-642b includes text that is located in or near a predetermined location (e.g., a central location) of the enlarged representation 724a. In some embodiments (e.g., as described above in connection with FIGS. 6M-6T), the set of prominence criteria includes criteria that are met when a determination is made that one or more of the text portions 642a-642b includes text of a particular type (e.g., email, phone number, address, QR code, etc.). In some embodiments, the set of prominence criteria includes criteria that are met when a determination is made that one or more of the text portions 642a-642b contain text that is relevant to the context of the expanded representation 724a (e.g., the text meets a relevance threshold (e.g., the computer system 600 determines that the text is 90%, 95%, 99% relevant)).

[0256] In Figure 7C, a determination is made that the primary subject of enlarged representation 724a is sign 642. That is, the context of enlarged representation 724a is the content displayed within sign 642. In Figure 7C, a further determination is made that text portion 742 (e.g., "BRAND") is irrelevant because it appears above the hat of person 740 and is therefore not relevant to the context of what is displayed in enlarged representation 724a. In some embodiments, because text portion 742 appears on the person or something above the person in enlarged representation 724a, computer system 600 determines that text portion 742 is not relevant to the context of what is displayed in enlarged representation 724a.

[0257] Because a determination has been made that text portion 742 is irrelevant, a determination has been made that text portion 742 does not meet the set of prominence criteria. In particular, even though text portion 742 has more text than text portions 642a-642b, a determination has been made that text portion 742 does not meet the set of prominence criteria. As shown in FIG. 7C , because text portion 742 does not meet the set of prominence criteria (e.g., due to a determination that text portion 742 is not relevant to the context of enlarged representation 724a), computer system 600 does not display one or more parentheses around text portion 742 ("BRAND"). In FIG. 7C , computer system 600 detects tap input 750c on text portion 742.

[0258] As shown in Figure 7D, in response to detecting tap input 750c, computer system 600 maintains the display of enlarged representation 724a as shown in Figure 7C. In Figure 7D, because a determination has been made that text portion 742 does not meet the set of prominence criteria (e.g., as described above in connection with Figure 7C), computer system 600 does not update the display of enlarged representation 724a to indicate that text portion 742 is selected. Also, because the text management controls are not displayed and selected, computer system 600 does not update the display of enlarged representation 724a to indicate that text portion 742 is selected (e.g., as opposed to computer system 600 updating the representations of the media in Figures 6J-6L, as described above). Also, because computer system 600 does not update the display of enlarged representation 724a in Figure 7D, text portions 642a-642b continue to not meet the set of prominence criteria. Thus, as shown in Figure 7D, computer system 600 does not display brackets corresponding to text portions 642a or 642b. In Figure 7D, computer system 600 detects an unpinch 750d input in media viewer area 724. In some embodiments, instead of an unpinch input 750d, computer system 600 detects a directional swipe corresponding to a request to pan (e.g., translate) enlarged representation 724a shown in Figure 7D.

[0259] As shown in FIG. 7E , in response to detecting an unpinch input 750d, computer system 600 updates magnified representation 724a to reflect the change in zoom level, such that the display of magnified representation 724a in FIG. 7E is displayed at a greater zoom level than the display of magnified representation 724a in FIG. 7D . In FIG. 7E , a determination is made that the text of text portion 642a meets a set of prominence criteria, but the text of text portion 642b does not meet the set of prominence criteria. As a result, computer system 600 displays brackets 736a at a location corresponding to the location of text portion 642a (e.g., surrounding text portion 642a). However, computer system 600 does not display brackets 736a or any other brackets at a location corresponding to the location of text portion 642b (e.g., because the text of text portion 642b does not meet the set of prominence criteria). In particular, a determination was made that even though text portion 742 (e.g., "BRAND") has more text than text portions 642a-642b (e.g., due to text portion 742 being unrelated), text portion 742 still does not meet the set of prominence criteria. In some embodiments, when computer system 600 detects a directional swipe instead of an unpinch input 750d, computer system 600 pans magnified representation 724a such that a different portion of magnified representation 724a is displayed in response to receiving unpinch input 750d.

[0260] As shown in Figure 7E, because a determination has been made that the text in text portion 642a meets a set of prominence criteria, computer system 600 displays text management control 680. Because text management control 680 is not selected (e.g., no input directed at the text management control has been detected), text management control 680 is displayed in an inactive state (e.g., as indicated by the text management control 680 not being bolded). In Figure 7E, computer system 600 detects an unpinch input 750e in media viewer area 724.

[0261] As shown in FIG. 7F , in response to detecting an unpinch input 750e, the computer system 600 updates the display of the magnified representation 724a to reflect the change in zoom level, such that the display of the magnified representation 724a in FIG. 7F is displayed at a greater zoom level than the display of the magnified representation 724a in FIG. 7E . In FIG. 7F , a determination is made that the text in text portion 642a meets a set of prominence criteria and that the text in text portion 642b meets a set of prominence criteria. Accordingly, brackets 636a are displayed around the entirety of both text portions 642a-642b, as described above in connection with FIG. 6C . Notably, in FIG. 7F , a determination is made that even though text portion 742 (e.g., “BRAND”) has more text than text portions 642a-642b (e.g., due to the text portions 742 being unrelated), text portion 742 continues to not meet the set of prominence criteria. As shown in FIG. 7F, computer system 600 displays text type indication 638a below "123 MAIN STREET" to indicate that an address has been detected and text type indication 638b below "123-4567" to indicate that a phone number has been detected (e.g., using one or more techniques such as those described above in connection with FIG. 6C). In some embodiments, computer system 600 displays multiple brackets, including one bracket around text portion 642a and another bracket around text portion 642b, and / or other combinations of brackets (e.g., using one or more techniques such as those described above in connection with FIG. 6A-6M). In some embodiments (e.g., referring again to FIG. 7E), computer system 600 displays a text type indicator below a text portion regardless of whether the text portion to which the text type indicator belongs meets a set of prominence criteria. In FIG. 7F, computer system 600 detects an unpinch input 750f in media viewer area 724.

[0262] As shown in Figure 7G, in response to detecting an unpinch input 750f, the computer system 600 updates the magnified representation 724a to reflect the change in zoom level, such that the display of the magnified representation 724a in Figure 7G is displayed at a greater zoom level than the display of the magnified representation 724a in Figure 7F. In some embodiments, the input that displays the magnified representation 724a, as shown in Figure 7G, corresponds to a directional swipe input.

[0263] As shown in FIG. 7G, expanded representation 724a includes a subset of text portion 642a and a subset of text portion 642b. As a result of a determination that text portions 642a-642b in their entirety no longer meet a set of prominence criteria (e.g., and / or expanded representation 724a includes only a subset of text portions 642a and 642b), computer system 600 ceases displaying brackets 636a around text portions 642a and 642b in their entirety. In FIG. 7G, a determination has been made that a subset of the text in text portion 642b (e.g., the phone number "123-4567") meets a set of prominence criteria (e.g., while another subset of the text in text portion 642b does not meet the criteria). In some embodiments, a determination is made that a subset of text portion 642b meets a set of prominence criteria because a determination is made that the user intends to interact with or view the phone number based on input previously detected by computer system 600 (e.g., when viewing Figures 7A-7G, computer system 600 continues to zoom in close to the phone number).

[0264] In some embodiments, a determination is made that Figure 7G includes a subset of text portion 642a (e.g., "DOG") that meets a set of prominence criteria. In response to this determination, computer system 600 displays a set of brackets around the subset of text portion 642a simultaneously with bracket 736c.

[0265] As shown in FIG. 7G, computer system 600 displays only a portion of the address "123 MAIN STREET." As a result, computer system 600 ceases displaying text type indication 638a. In some embodiments, computer system 600 continues to display text type indication 638a below the portion of the address "123 MAIN STREET" displayed in FIG. 7G. In some embodiments, because other portions of the address are not displayed, computer system 600 determines that the portion of the address does not meet the set of prominence criteria. In FIG. 7G, computer system 600 detects right swipe 750g in media viewer area 724.

[0266] As shown in FIG. 7H, in response to detecting rightward swipe 750g, computer system 600 pans enlarged representation 724a to the right. Enlarged representation 724a is panned such that computer system 600 ceases displaying the rightmost portions of text portions 642a-642b shown in FIG. 7G and re-displays the leftmost portions of text portions 642a-642b in FIG. 7H. As shown in FIG. 7H, computer system 600 does not display the entire telephone number (e.g., 123-4567) and ceases displaying brackets 736c and text type indication 638b. In some embodiments, a portion of text type indication 638b remains displayed below the portion of the telephone number (e.g., "12") still displayed in FIG. 7H. In FIG. 7H, the computer system 600 displays more of the address (e.g., 123 MAIN STREET) in FIG. 7H and re-displays the text type indication 638a below "123 MAIN STREET" to indicate to the user that the address has been found.

[0267] In Figure 7H, a determination is made that another subset of text portion 642b (e.g., "$1000 REWARD") meets the set of prominence criteria (e.g., without any other subset of text portion 642b meeting the set of prominence criteria). Because a determination is made that the other subset of text portion 642b meets the set of prominence criteria, computer system 600 displays parentheses 736d around the other subset of text portion 642b, "$1000 REWARD." In some embodiments, a determination is made that the other subset of text portion 642b (e.g., "$1000 REWARD") is the most relevant text displayed based on the context of the displayed content of enlarged representation 724a. In some embodiments, a determination is made that Figure 7H includes a subset of text portion 642a (e.g., "LOST") that meets the set of prominence criteria. In some embodiments, in response to this determination, the computer system 600 displays a separate set of brackets around the subset of the text portion 642a simultaneously with the brackets 736e. In Figure 7H, the computer system 600 detects a tap input 750h on the text management control 680.

[0268] As shown in FIG. 7I, in response to detecting tap input 750h, computer system 600 displays text management options 682, including copy option 682a (e.g., that, when selected, causes computer system 600 to copy the text enclosed in brackets 736d), select all option 682b (e.g., that, when selected, causes computer system 600 to select all of the text enclosed in brackets 736d), search option 682c (e.g., that, when selected, causes computer system 600 to search for the text enclosed in brackets 736d via a search (e.g., a web search, a dictionary search)), and share option 682d (e.g., that, when selected, causes computer system 600 to initiate a process to share the text enclosed in brackets 736d). In some embodiments, the various components of text management options 682 function as described above in connection with FIGS. 6A-6M. In some embodiments, computer system 600 displays multiple text management options, each corresponding to a separate portion of the text enclosed in a separate pair of brackets. In some embodiments, selection of an individual text management option allows the user to manage the portion of the text that corresponds to the individual text management option.

[0269] As shown in Figure 7I, the computer system 600 displays the text management control 680 as enabled (e.g., as indicated by the text management control 680 being bolded). In Figure 7I, the computer system 600 detects a tap input 750i on the text management control 680.

[0270] As shown in Figure 7J, in response to detecting a tap input 750i, the computer system 600 re-displays the enlarged representation 724a using one or more techniques such as those described above in connection with Figure 7H. In Figure 7J, the computer system 600 detects a downward swipe input 750j within the media viewer area 724.

[0271] As shown in Figure 7K, in response to detecting a downward swipe input 750j, the computer system 600 pans the media viewer area 724 downward (e.g., based on the swipe input) such that the display of the text portion 642b ceases and the computer system displays only the subset of the text portion 642a. In Figure 7K, a determination is made that the subset of the text portion 642a (e.g., "LOST") meets a set of prominence criteria. Because the subset of the text portion 642a meets the set of prominence criteria, the computer system 600 displays brackets 736e surrounding the subset of the text portion 642a.

[0272] In Figure 7K, computer system 600 does not display bracket 736d because text portion 642b is not displayed as part of enlarged representation 724a in Figure 7K. In Figure 7K, computer system 600 detects a pinch input 750k within media viewer area 724.

[0273] As shown in FIG. 7L , in response to detecting pinch input 750k, computer system 600 updates magnified representation 724a to reflect the change in zoom level (e.g., a decrease in zoom level) such that the display of magnified representation 724a in FIG. 7L is displayed at a decreased zoom level compared to the zoom level of the display of magnified representation 724a in FIG. 7K. In FIG. 7L , a determination is made that text portions 642a and 642b do not meet a set of prominence criteria. Accordingly (e.g., because a determination is made that text portions 642a and 642b do not meet a set of prominence criteria), computer system 600 does not display (and / or ceases to display) text management control 680 and / or any parentheses surrounding text portions 642a-642b.

[0274] 7A-7L are described in the context of computer system 600 displaying a representation of previously captured media and a media viewer user interface, one or more techniques such as those described above in connection with Figures 6A-6Z can also be applied while computer system 600 is displaying previously captured media and a media viewer user interface. The techniques described above in connection with Figures 7A-7L can also be applied in the context of computer system 600 displaying a live preview (e.g., a representation of the field of view of one or more cameras before the media was captured), such as live preview 630 of Figures 6A-6M, and a camera user interface.

[0275] 6A-6Z are described in the context of computer system 600 displaying a live preview and a camera user interface, one or more of the techniques described above in connection with Figures 7A-7L may also be applied while computer system 600 is displaying a live preview and a camera user interface. The techniques described in connection with Figures 6A-6Z may also be applied in the context of computer system 600 displaying previously captured media, such as magnified media representation 724a, and a media viewer user interface.

[0276] 8 is a flow diagram illustrating a method for managing visual content in media using a computer system, according to some embodiments. Method 800 is performed in a computer system (e.g., 100, 300, 500) in communication with a display generation component. Some operations in method 800 are optionally combined, the order of some operations is optionally changed, and some operations are optionally omitted.

[0277] As described below, method 800 provides an intuitive way to manage visual content within a medium. The method reduces the cognitive burden on a user managing visual content within a medium, thereby creating a more efficient human-machine interface. For battery-operated computing devices, allowing a user to manage visual content within a medium more quickly and efficiently conserves power and extends the time between battery charges.

[0278] Method 800 is performed in a computer system (e.g., 600) (e.g., smartphone, desktop computer, laptop, tablet) in communication with a display generation component (e.g., display control, touch-sensitive display system). In some embodiments, the computer system is in communication with one or more input devices (e.g., touch-sensitive surfaces) and / or a first camera of one or more cameras (e.g., one or more cameras (e.g., dual cameras, triple cameras, quad cameras, etc.) on the same or different sides of the computer system (e.g., front camera, back camera)).

[0279] The computer system, via the display generation component, displays (802) a camera user interface (e.g., a media capture user interface, a media viewing user interface, a media editing user interface), which includes simultaneously displaying representations (e.g., 630) of media (e.g., photographic media, video media) (e.g., live media; live previews (e.g., media corresponding to representations of one or more camera fields of view (e.g., current fields of view) that have not been captured in response to detecting a request to capture media (e.g., detecting selection of a shutter affordance); previously captured media (e.g., media corresponding to representations of one or more camera fields of view (e.g., previous fields of view) that have been captured); saved media items that can be accessed by the user at a later time; representations of media displayed in response to receiving a gesture on a thumbnail representation of the media (e.g., in a media gallery)) and media capture affordances (e.g., 610) (e.g., user interface objects).

[0280] While simultaneously displaying (804) a representation of the media (e.g., 630) and a media capture affordance (e.g., 610) (e.g., a user interface object), in accordance with a determination that a respective set of criteria is satisfied, the computer system, via the display generation component, displays (804) (e.g., simultaneously with the representation of the media) (e.g., in the user interface) a first user interface object (e.g., 680) corresponding to one or more text management operations (e.g., simultaneously with the representation of the media and / or the first user interface object), wherein the respective set of criteria is satisfied when a respective text (e.g., 642a, 642b) (e.g., one or more characters represented in the media) is detected in the representation of the media (e.g., 630). In some embodiments, the plurality of options (e.g., 672, 682, 692) include one or more options to copy the individual text (e.g., 682a), select the individual text (e.g., 682b), search the individual text (e.g., 682c), share the individual text (e.g., 682d), and translate the individual text.

[0281] While simultaneously displaying (804) a representation of the media (e.g., 630) and a media capture affordance (e.g., 610) (e.g., a user interface object), the computer system refrains from displaying (808) a first user interface object pursuant to a determination that a respective set of criteria is not met.

[0282] While displaying a representation of media (e.g., 630) (e.g., while simultaneously displaying the representation of media, a media capture affordance, and a first user interface object), the computer system detects (810) a first input (e.g., 650a, 650e, 650g, 650u) (e.g., mouse / trackpad click / activation, keyboard input, scroll wheel input, hover gesture, tap gesture, swipe gesture) directed at the camera user interface (e.g., 602, 604, 606). In some embodiments, the first input is a non-tap gesture (e.g., a rotate gesture and / or a long press gesture).

[0283] In response to detecting (812) a first input (650a, 650e, 650g, 650u) (e.g., a first gesture) directed toward the camera user interface, and in accordance with determining that the first input (e.g., 650a) corresponds to selecting a media capture affordance (e.g., 610) (e.g., a gesture directed toward the media capture affordance, a gesture at a location corresponding to the media capture affordance), the computer system initiates (814) capture of media to be added to a media library (e.g., 612) associated with the computer system (e.g., 600) (e.g., without displaying options to manage individual text).

[0284] In response to detecting (812) a first input (650a, 650e, 650g, 650u) (e.g., a first gesture) directed at the camera user interface, and following a determination that the first input (e.g., 650e, 650g, 650u) corresponds to a selection of a first user interface object (e.g., 680), the computer system displays (816) via the display generation component a plurality of options for managing the respective text (e.g., 672, 682, 692) (e.g., without initiating capture of media to be added to a media library (e.g., 624) associated with the computer system (e.g., 600). In some embodiments, the plurality of options are displayed adjacent to the respective text (e.g., included in a representation of the media). In some embodiments, the plurality of options (e.g., 672, 682, 692) (e.g., as described above in connection with FIG. 6F) include one or more options to copy the individual text (e.g., 682a), select the individual text (e.g., 682b), search the individual text (e.g., 682c), share the individual text (e.g., 682d), and translate the individual text. In some embodiments, following a determination (e.g., as described above in connection with FIG. 6F) that a first input (e.g., 650e, 650g, 650u) corresponds to a selection of a first user interface object (e.g., 680), the first user interface object is in an active state (e.g., transitioned to an active state from being displayed in an inactive state), and the first user interface object displayed in the active state (e.g., 680 in FIG. 6F) (e.g., bolded and pressed state / appearance) has a different appearance than when the first user interface object is displayed in an inactive state (e.g., 680 in FIG. 6G) (e.g., not bolded and released state / appearance).In some embodiments, in response to a determination that the first input (e.g., 650a) corresponds to a selection of the media capture affordance (e.g., 610), the first user interface object (e.g., 680) is in an inactive state (e.g., transitioned from being displayed in an inactive state to an active state). In some embodiments, in response to a determination that the first input (e.g., 650e, 650g, 650u) corresponds to a selection of the first user interface object (e.g., 680) and the first user interface object is in an inactive state (e.g., 680 in FIG. 6E), the computer system displays a plurality of options (e.g., 672, 682, 692) for managing the individual text. In some embodiments, pursuant to a determination that the first input corresponds to a selection of a first user interface object (e.g., 680) and that the first user interface object is in an active state (e.g., 680 in FIG. 6F) (e.g., as described above with reference to FIG. 6F), the computer system refrains from displaying the plurality of options (e.g., 672, 682, 692) for managing the individual text. In some embodiments, pursuant to a determination that the first input (650a) corresponds to a selection of a media capture affordance, the first user interface object (e.g., 610) continues to be displayed. In some embodiments, following a determination that the first input (e.g., 650a) corresponds to a selection of a media capture affordance (e.g., 610) or a selection of a first user interface object (e.g., 680), one or more interface objects (e.g., the media capture affordance (e.g., 610), the camera settings affordance(s), the camera mode affordance(s) (e.g., 620)) are either hidden in the camera user interface or are displayed as inactive (e.g., dimmed) (e.g., not responsive to user input for the respective objects).Displaying multiple options for managing the individual text in accordance with a determination that the first input corresponds to selecting the first user interface object allows the user to quickly and efficiently manage the individual text without cluttering the user interface with additional user interface objects. Providing additional control of the system without cluttering the UI with additional displayed controls improves system usability and makes the user-system interface more efficient (e.g., by assisting the user in providing appropriate input when operating / interacting with the system and reducing user errors), and additionally reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently. Displaying multiple options for managing the individual text when a predetermined condition is met (e.g., based on whether the first input corresponds to selecting the first user interface object) automatically provides the user with a variety of options for different ways to manage the individual text. Taking action when a set of conditions is met without requiring further user input enhances the operability of the system, makes the user-system interface more efficient (e.g., by assisting the user in providing appropriate inputs when operating / interacting with the system and reducing user errors), and also reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0285] In some embodiments, the first input (e.g., 650e, 650g, 650u) is a tap gesture (e.g., a tap gesture) directed at a first user interface object (e.g., 672, 682, 692) (e.g., a gesture at a location corresponding to the first user interface object).

[0286] In some embodiments, the representation of the media (e.g., 630) includes the discrete text (e.g., if the discrete text is displayed when the representation of the media is displayed). In some embodiments, after detecting a first input (e.g., 650e, 650g, 650u) (and while not displaying an indication that the text is selected, and / or after detecting an input / gesture corresponding to selecting the first user interface object, and / or while the first user interface object is displayed as being in an active state, and / or while displaying multiple options for managing the discrete text), the computer system detects a second input (e.g., 650j) (e.g., a tap gesture and / or a swipe gesture) directed at the camera user interface. In some embodiments, the second input is a non-tap gesture (e.g., a rotate gesture and / or a press and hold gesture). In some embodiments, the first input is a non-swipe gesture (e.g., a rotate gesture, a long press gesture, a mouse / trackpad click / activate, a keyboard input, a scroll wheel input, a hover gesture, and / or a tap gesture). In some embodiments, in response to detecting a second input (e.g., 650j) directed at the camera user interface, and in accordance with determining that the second input corresponds to selecting a first one or more portions of the individual text, the computer system displays an indication (e.g., 642b in FIG. 6K) that the first one or more portions (e.g., 642b) of the individual text (e.g., 642a, 642b) are selected. In some embodiments, the indication is displayed around the individual text. In some embodiments, as part of displaying the indication that the one or more portions of the individual text are selected, the computer system emphasizes (e.g., highlights, underlines, bolds, increases the size) one or more portions of the individual text.In some embodiments, while displaying an indication that a first portion of the respective text is selected, the computer system does not display an indication that a second portion of the respective text (e.g., different from the first portion) is selected. In some embodiments, pursuant to a determination that the second input (650j) corresponds to a selection of one or more portions of the respective text (e.g., 642a), and while the first user interface (e.g., 680) object is displayed as being in an active state (e.g., 680 as described above in connection with FIG. 6F), the computer system displays an indication that the first one or more portions of the respective text (e.g., 642a) are selected (e.g., as described above in connection with FIGS. 6K and 6L). In some embodiments, pursuant to a determination that the second input (e.g., 650j) corresponds to a selection of one or more portions of the respective text (e.g., 642), and while the first user interface object (e.g., 680) is displayed as inactive (e.g., as described above in connection with FIG. 6G) and / or is not displayed (e.g., as described above in connection with FIGS. 7C, 7G), the computer system does not display (e.g., refrains from displaying) an indication that the first one or more portions of the respective text (e.g., 642a) are selected (e.g., as discussed above in connection with FIGS. 7C-7G). Displaying an indication that the first one or more portions of the respective text are selected provides the user with visual feedback regarding whether text is selected and which text is currently selected.Providing improved visual feedback to the user improves operability of the computer system (e.g., by assisting the user in providing appropriate input when operating / interacting with the computer system and reducing user errors), makes the computer system interface more efficient, and reduces power usage and improves battery life by allowing the user to use the computer system more quickly and efficiently. In response to detecting a second input and in accordance with determining that the second input corresponds to selecting the first one or more portions of the discrete text, displaying an indication that the first one or more portions of the discrete text have been selected provides the user with additional control for selecting text without cluttering the user interface with additional user interface objects. Providing additional control of the system without cluttering the UI with additional displayed controls improves operability of the system (e.g., by assisting the user in providing appropriate input when operating / interacting with the system and reducing user errors), makes the user-system interface more efficient, and reduces power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0287] In some embodiments, the second input (e.g., 650j) (e.g., second gesture) is a tap gesture (e.g., directed at one or more portions of the individual text) or a swipe gesture (e.g., directed at one or more portions of the individual text). In some embodiments, the first input is a first type of input and the second input is a second type of input that is different from the first type of input.

[0288] In some embodiments, in response to detecting a first input (e.g., 650e, 650g, 650u) directed at the camera user interface, and in accordance with a determination that the first input (e.g., 650e, 650g, 650u) corresponds to selecting a first user interface object (e.g., 680), the computer system displays an indication (e.g., 684) (e.g., instructions) (e.g., instructions indicating one or more inputs that cause the computer system to display the text as selected) (e.g., not previously displayed before the first input was detected) regarding selecting (e.g., how to select) text included in the representation of media (e.g., 630). In some embodiments, in response to detecting a first input directed at the camera user interface, and in accordance with a determination that the first input corresponds to selecting a media capture affordance, the computer system does not display an indication regarding selecting (e.g., how to select) text included in the representation of media. In some embodiments, an indication (e.g., 684) regarding selecting (e.g., how to select) text included in a media representation is displayed simultaneously with multiple options (e.g., 682a, 682b, 682c, 682d) for managing individual text (e.g., 642b). In some embodiments, an indication (e.g., 684) regarding text selection is displayed when a first user interface object (e.g., 680) is displayed in an active state (e.g., 680 as described above in connection with FIG. 6F), and an indication (e.g., 684) regarding text selection is not displayed when the first user interface object (e.g., 680 as described above in connection with FIG. 6G). Displaying an indication regarding how to select text included in a media representation provides a user with visual feedback regarding the steps required to select the desired text.Providing improved visual feedback to a user improves the usability of the computer system (e.g., by assisting the user in providing appropriate input when operating / interacting with the computer system and reducing user errors), makes the computer system interface more efficient, and additionally reduces the power usage and improves battery life of the computer system by allowing the user to use the computer system more quickly and efficiently.

[0289] In some embodiments, prior to detecting a first input (e.g., 650e, 650g, 650u), the representation of the media (e.g., 630) is displayed with a first appearance (e.g., 630 in FIG. 6E ) (e.g., a first blur value, a first dim value). In some embodiments, in response to detecting a first input (e.g., 650e, 650g, 650u) directed at the camera user interface, and in accordance with determining that the first input (e.g., 650e, 650g, 650u) corresponds to selecting a first user interface object (e.g., 680), the computer system displays the representation of the media (e.g., 630) with a second appearance (e.g., 630 in FIG. 6F ) (e.g., a second blur value, a second dim value) that is different from the first appearance (e.g., 630 in FIG. 6E ). In some embodiments, as part of displaying the representation of the media in a second appearance different from the first appearance, the computer system blurs and / or darkens at least a portion of the representation of the media. In some embodiments, in response to detecting a first input directed at the camera user interface and in accordance with determining that the first input corresponds to a selection of a media capture affordance, the computer system displays the representation of the media in a third appearance different from the second appearance. In some embodiments, the third appearance is the first appearance. In some embodiments, the third appearance (e.g., black, solid color) is different from the first appearance (e.g., a blurred version of the field of view of one or more cameras). In some embodiments, the representation of the media having the third appearance is displayed for a predetermined period of time (e.g., less than one second) that is not based on whether the first user interface object is displayed in an active state. In some embodiments, the representation of the media having the second appearance is displayed when the first user interface object is displayed in an active state and is not displayed when the first user interface object is displayed in an inactive state.In some embodiments, the representation of the media having the first appearance is not displayed when the first user interface object is displayed in an active state and is displayed when the first user interface object is displayed in an inactive state. Displaying the representation of the media having a second appearance different from the representation's first appearance in response to detecting the first input provides visual feedback to the user regarding whether text has been selected by the user by de-emphasizing less relevant portions of the representation of the media. Providing the user with improved visual feedback improves operability of the computer system (e.g., by assisting the user in providing appropriate inputs when operating / interacting with the computer system and reducing user errors), makes the computer system interface more efficient, and additionally reduces power usage and improves battery life of the computer system by allowing the user to use the computer system more quickly and efficiently.

[0290] In some embodiments, the media representation (e.g., 630) includes the individual text (e.g., 642a, 642b) (e.g., if the individual text is displayed when the media representation is displayed). In some embodiments, pursuant to a determination that the individual set of criteria is met, the computer system emphasizes (e.g., highlights, displays an object (e.g., a shape, brackets around (e.g., yellow brackets)), underlines, magnifies) a second one or more portions of the individual text (e.g., 642a, 642b). In some embodiments, pursuant to a determination that the individual set of criteria is met, the computer system emphasizes the second one or more portions of the individual text without emphasizing other portions of the individual text and / or other portions of the media representation that do not include the second one or more portions of the individual text. Highlighting the second one or more portions of the individual text provides the user with improved visual feedback regarding whether a particular portion of the individual text included in the media meets the individual set of criteria. Providing improved visual feedback to a user improves the usability of the computer system (e.g., by assisting the user in providing appropriate input when operating / interacting with the computer system and reducing user errors), makes the computer system interface more efficient, and additionally reduces the power usage and improves battery life of the computer system by allowing the user to use the computer system more quickly and efficiently.

[0291] In some embodiments, as part of highlighting the second one or more portions of the distinct text, the computer system displays an indication (e.g., 636a, 636b, 736c, 736d) that the distinct text has been detected. In some embodiments, while the second one or more portions of the distinct text (e.g., 642a, 642b) are being highlighted, the computer system receives a request to display a second representation (e.g., 630 in FIG. 6F) of media (e.g., the same or different media as that represented by the representation of the media). In some embodiments, the request to display the second representation of the media is detected when one or more changes in the field of view of one or more cameras in communication with the computer system are detected. In some embodiments, the request to display the second representation of the media is detected when a request to zoom out / zoom in and / or pan the representation of the media is detected. In some embodiments, the request to display the second representation of the media is detected when the computer system is moved.

[0292] In some embodiments, in response to receiving a request to display a second representation of the media (e.g., 630 in FIG. 6F ) (e.g., including a portion of the distinct text and / or a second distinct text different from the distinct text), the computer system translates (e.g., moves) an indication that distinct text has been detected (e.g., 636a, 636b, 736c, 736d) from a first position within the camera user interface to a second position within the camera user interface. In some embodiments, in response to receiving a request to display the second representation of the media, the indication that distinct text has been selected is modified to enclose a portion of text that is different from the portion of text that it enclosed before the request to display the second representation of the media was received. By translating the indication that distinct text has been detected from a first position within the camera user interface to a second position within the camera user interface in response to receiving a request to display the second representation of the media, a user can maintain view of the indication while the system moves between the first and second positions. Taking action when a set of conditions is met without requiring further user input enhances the operability of the system, makes the user-system interface more efficient (e.g., by assisting the user in providing appropriate inputs when operating / interacting with the system and reducing user errors), and also reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0293] In some embodiments, after detecting a first input (e.g., 650e, 650g, 650u) and pursuant to a determination (first determination) that the first input (e.g., 650e, 650g, 650u) corresponds to a selection of a first user interface object (e.g., 680) (and / or while the first user interface object is displayed as being in an active state and / or while displaying multiple options managing the respective text), the representation of the media (e.g., 630) includes the respective text (e.g., 642a, 642b) and an indication that a third one or more portions of the respective text (e.g., 642a, 642b) are selected. In some embodiments, the computer system receives a request to display a third representation of the media (e.g., 630) (e.g., the same or different media as the media represented by the representation of the media). In some embodiments, the request to display the third representation of the media is detected when one or more changes in the field of view of one or more cameras in communication with the computer system are detected. In some embodiments, the request to display the third representation of the media is detected when a request to zoom out / in and / or pan the representation of the media is detected. In some embodiments, the request to display the third representation of the media is detected when the computer system is moved. In some embodiments, in response to receiving a request (e.g., 650c, 650d, 750e, 750f, 750g) to display the third representation of the media (e.g., 630), the computer system displays an indication that at least a portion of the text included in the third representation of the media (e.g., 630) has been selected, wherein the indication (e.g., 636a, 636b, 736c, 736d) that at least a portion of the text (e.g., 642a, 642b) included in the third representation of the media has been selected is different from the indication that a third one or more portions of the individual text (e.g., 642a, 642b) have been selected.In some embodiments, the portion of text included in the third representation of the media includes at least a portion of the text within the third one or more portions of text. In response to receiving a request to display the third representation, displaying an indication that at least a portion of the text included in the third representation of the media is selected provides the user with an additional, efficient way of controlling which portions of the text are selected without cluttering the user interface. Reducing the number of inputs required to perform operations improves usability of the computer system (e.g., by assisting the user in providing appropriate inputs when operating / interacting with the computer system and reducing user errors), makes the user-system interface more efficient, and additionally reduces power usage and improves battery life of the computer system by allowing the user to use the computer system more quickly and efficiently.

[0294] In some embodiments, after detecting a first input (e.g., 650e, 650g, 650u) and pursuant to a determination (e.g., a first determination) that the first input corresponds to a selection of a first user interface object (and / or while the first user interface object is displayed as being in an active state and / or while displaying multiple options managing the respective text), the representation of the media (e.g., 630) includes the respective text (e.g., 642b) and an indication that a fourth one or more portions of the respective text (e.g., 642b) have been selected and that the fourth one or more portions of the respective text (e.g., 642b) are displayed in a third position within the camera user interface (and / or on the display). In some embodiments, the computer system detects changes in the physical environment within the field of view of one or more cameras in communication with the computer system. In some embodiments, in response to detecting a change in the physical environment (e.g., 660a, 660b) within the field of view of one or more cameras, the computer system continues to display a fourth one or more portions of the distinct text (e.g., 642b) at a third position within the camera user interface (and / or on the display). In some embodiments, the selected text is frozen. In some embodiments, at least a portion of a fourth representation of the media is displayed (e.g., newly displayed in response to detecting a change in the physical environment) while maintaining the display of the fourth one or more portions of the distinct text. In some embodiments, the computer system freezes the selected text (e.g., and / or displays the selected text in the same location and / or at the same size) while updating the representation of the media (e.g., live preview) to reflect the change in the physical environment. By continuing to display the fourth one or more portions of the distinct text at a third position within the camera user interface, the user can maintain view of the text selected by the user while the system moves between the first and second points.Performing an action when a set of conditions is met without requiring further user input improves the usability of the system (e.g., by assisting the user in providing appropriate input when operating / interacting with the system and reducing user errors), makes the user-system interface more efficient, and additionally reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0295] In some embodiments, prior to detecting a first input (e.g., 650e, 650g, 650u) directed at the camera user interface, the computer system (e.g., 600) is in communication with one or more cameras, and the representation (e.g., 630) of media is a representation (e.g., 630) of one or more objects in a physical environment (e.g., a physical space) within a field of view of the one or more cameras (e.g., a live camera preview). In some embodiments, receiving a request to display a fourth representation of media (e.g., a representation of an updated field of view of the camera) includes detecting a change in the field of view of the camera. In some embodiments, the fourth representation of media includes a change in the field of view of the camera. In some embodiments, when one or more objects (e.g., non-text objects) within the field of view are moving, the representation of the media is updated to indicate that the one or more objects are moving. In some embodiments, the representation of the media is a live representation of the field of view of the camera. Displaying a representation of media that is a representation of one or more objects in physical space within the field of view of one or more cameras (e.g., a live camera preview) provides the user with greater control over the computer system (e.g., changing the field of view of the system's cameras) to determine whether one or more objects in physical space can be captured without cluttering the user interface. Providing additional control of the system without cluttering the UI with additional displayed controls improves usability of the system (e.g., by assisting the user in providing appropriate input when operating / interacting with the system and reducing user errors), makes the user-system interface more efficient, and additionally reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0296] In some embodiments, the representation of the media (e.g., 630) is a first representation of the media. In some embodiments, while displaying the first user interface object, the computer system detects a request (e.g., 750k) to display a fifth representation of the media (e.g., 630) (e.g., the same or different media as the media represented by the first representation of the media). In some embodiments, the request to display the fifth representation of the media is detected when one or more changes in the field of view of one or more cameras in communication with the computer system are detected. In some embodiments, the request to display the fifth representation of the media is detected when a request to zoom out / zoom in and / or pan the representation of the media is detected. In some embodiments, the request to display the fifth representation of the media is detected when the computer system is moved. In some embodiments, in response to detecting a request to display a fifth representation of the media (e.g., 750k) and in accordance with a determination that a respective set of criteria is not met (e.g., no distinct text is detected in the fifth representation of the media, or the distinct text is detected but not sufficiently prominent), the computer system ceases displaying the first user interface object (e.g., 680). In some embodiments, in response to detecting a request to display a fifth representation of the media and in accordance with a determination that distinct text is detected in the fifth representation of the media, the computer system continues to display the first user interface object. By ceasing display of the first user interface object when a predetermined condition is met (e.g., in response to detecting a request to display a fifth representation of the media and in accordance with a determination that a respective set of criteria is not met), an indication is automatically provided to the user of whether the representation of the media does not include the text detected by the computer system.Taking action when a set of conditions is met without requiring further user input enhances the operability of the system, makes the user-system interface more efficient (e.g., by assisting the user in providing appropriate inputs when operating / interacting with the system and reducing user errors), and also reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0297] In some embodiments, each criterion includes criteria that are met when the individual text meets a predetermined prominence criterion (e.g., the text is at a size or location within the media representation that indicates the text is important and / or relevant) (e.g., the individual text indicates a first user interface object when on a sign and does not indicate a first user object when detected on clothing) (e.g., based on the context of the media representation (e.g., the context of the media representation is important / relevant based on the context of the image); based on the individual text occupying a certain amount of space on the displayed media representation when in a particular location (e.g., center) on the displayed media representation; based on the individual text belonging to a particular type of text (e.g., email, phone number, QR code, uniform access code location, etc.)), or a determination is made that the individual text meets a predetermined prominence criterion (e.g., the text is at a size or location within the media representation that indicates the text is important and / or relevant) (determined to be relevant based on one or more techniques such as those described below in connection with Figures 7C, 7E-7J, and 9).

[0298] In some embodiments, while displaying a representation of media (e.g., 630) (and, in some embodiments, after detecting input corresponding to a selection of the first user interface object and / or while the first user interface object is displayed as being in an active state and / or while displaying multiple options for managing the respective text), and pursuant to a determination that the respective text (e.g., 642a-642b) includes a portion of text that has been determined to be a respective type of text (e.g., phone number, email) (e.g., based on one or more regular expression patterns corresponding to different types of text), the computer system displays an indication (e.g., 638a-638b) (e.g., a data detector indication) that the respective type of text has been detected. In some embodiments, as part of displaying the indication that the respective type of text has been detected, the computer system emphasizes (e.g., highlights, underlines, brackets) the portion of the text. In some embodiments, the indication that the respective type of text has been detected is displayed adjacent to, around, etc. the portion of text that belongs to the respective type of text. In some embodiments, pursuant to a determination that the respective text does not include a portion of text belonging to a respective type of text (e.g., a phone number, an email), the computer system does not display (e.g., refrains from displaying) an indication that a respective type of text has been detected. Displaying an indication that a respective type of text has been detected in the representation of media provides visual feedback to the user as to whether the representation of media includes a certain type of text.Providing improved visual feedback to the user improves the usability of the computer system, makes the user-system interface more efficient (e.g., by assisting the user in providing appropriate input when operating / interacting with the computer system and reducing user errors), and also reduces power usage and improves the battery life of the computer system by allowing the user to use the computer system more quickly and efficiently.

[0299] In some embodiments, while displaying the plurality of options managing the respective text (e.g., 680), the computer system receives a third input (e.g., 650h) (e.g., a tap input) directed at a portion of the camera user interface that does not include the respective text (e.g., a dimmed or blurred portion of the representation of the media (e.g., a portion of the representation of the media that does not include the text) (and / or a dimmed portion of the camera user interface)). In some embodiments, in response to receiving the third input (e.g., 650h), the computer system ceases displaying the plurality of options managing the respective text (e.g., 680). In some embodiments, in response to receiving the third input, one or more interface objects (e.g., media capture affordance, camera settings affordance(s), camera mode affordance(s)) are displayed (e.g., undisplayed) and / or displayed as active (e.g., undimmed) within the camera user interface (e.g., in response to user input on the respective objects). By ceasing to display multiple options that manage individual text in response to receiving input directed at a portion of the camera user interface, the user is provided with more control over the system without cluttering the user interface with additional user interface objects. Providing additional control of the system without cluttering the UI with additional displayed controls improves usability of the system (e.g., by assisting the user in providing appropriate input when operating / interacting with the system and reducing user errors), makes the user-system interface more efficient, and additionally reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0300] In some embodiments, while simultaneously displaying a representation of media (e.g., 630) and a media capture affordance (e.g., 610) (e.g., before displaying a first user interface object), pursuant to determining that the representation of media (e.g., 630) includes a first machine-readable code (e.g., a linear barcode, a matrix barcode, or a QR code), the computer system displays the first user interface object (e.g., 680) and displays a representation of a uniform resource locator (e.g., 668) corresponding to the first machine-readable code. Displaying the first user interface object and displaying the representation of the uniform resource location improves security by informing the user of the location of the resource corresponding to the QR code before providing input to navigate to the resource. Providing improved security makes the user interface more secure and reduces unauthorized execution of secure operations, as well as reducing power usage and improving battery life of the computer system by allowing users to use the computer system more safely and efficiently. Displaying the first user interface object and displaying the representation of the uniform resource location when a predetermined condition is met (e.g., pursuant to determining that the representation of the media includes machine-readable code) informs the user of the resource associated with the machine-readable code before the user selects the machine-readable code and provides the user with a uniform resource locator corresponding to the first machine-readable code. Performing an action when a set of conditions is met without requiring further user input enhances system usability and makes the user-system interface more efficient (e.g., by assisting the user in providing appropriate inputs when operating / interacting with the system and reducing user errors), as well as reducing system power usage and improving battery life by allowing the user to use the system more quickly and efficiently.

[0301] In some embodiments, while the media representation (e.g., 630) includes the second machine-readable code (and while the machine-readable code is selected), in accordance with a determination that the first input (e.g., 650u) corresponds to selecting a first user interface object, the plurality of options (e.g., 672) for managing the discrete text include one or more options for managing information (e.g., a uniform resource location) corresponding to the second machine-readable code. In some embodiments, while the media representation does not include machine-readable code (and / or while the machine-readable code is not selected), in accordance with a determination that the first input corresponds to selecting a first user interface object, the plurality of options for managing the discrete text do not include one or more options for managing information. In some embodiments, one or more of the plurality of options for managing the discrete text that are displayed when the machine-readable code is selected are different from the one or more options for managing the discrete text that are displayed when text that does not include machine-readable code is selected. In accordance with determining that the first input corresponds to a selection in the first user interface, including one or more options for managing information corresponding to the machine-readable code in the plurality of options provides the user with more control options (e.g., additional text management options) without cluttering the user interface. Providing additional control of the system without cluttering the UI with additional displayed controls improves usability of the system (e.g., by assisting the user in providing appropriate input when operating / interacting with the system and reducing user errors), makes the user-system interface more efficient, and additionally reduces system power usage and improves battery life by allowing the user to use the system more quickly and efficiently.

[0302] In some embodiments, the camera user interface includes multiple selectable camera setting affordances (e.g., 620a-620e) that change one or more camera settings (e.g., flash affordance, timer affordance, filter effect affordance, f-stop affordance, aspect ratio affordance, live photo affordance, etc.) (e.g., multiple user interface objects that access individual camera settings). In some embodiments, the camera user interface includes multiple camera mode affordances (e.g., 620) (e.g., multiple user interface objects that set individual camera modes). In some embodiments, the multiple camera setting affordances (e.g., 602a, 602b) are displayed simultaneously with the media capture affordance (e.g., 610) and / or the multiple camera mode affordances (e.g., 620). In some embodiments, each camera mode (e.g., video (e.g., 620b), photo (e.g., 620c), portrait (e.g., 620d), slow motion (e.g., 620a), and panorama (e.g., 620e) modes) (e.g., 620) has multiple settings (e.g., for portrait camera mode: studio lighting setting, contour lighting setting, and stage lighting setting) with multiple values (e.g., light levels for each setting) for the mode (e.g., portrait mode) in which the camera (e.g., camera sensor) is operating to capture media (including post-processing that is automatically performed after capture). In this way, for example, camera modes differ from modes that do not affect how the camera operates when capturing media or do not include multiple settings (e.g., flash mode, which has one setting with multiple values (e.g., inactive, active, auto)).In some embodiments, camera modes allow a user to capture different types of media (e.g., photos or videos) and can optimize (e.g., through post-processing) the settings for each mode to capture the particular type of media corresponding to the particular mode with particular features (e.g., shape (e.g., square, rectangular), speed (e.g., slow motion, time lapse), audio, video). For example, if a computer system is configured to operate in ...

Claims

1. 1. A method comprising: a computer system in communication with a display generation component, displaying, via the display generation component, a camera user interface that includes simultaneously displaying a representation of the media and a media capture affordance; While simultaneously displaying a representation of the media and the media capture affordance, displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied, the criteria including criteria that are satisfied when a respective text is detected within the representation of the media; withholding display of the first user interface object in accordance with a determination that a respective set of criteria has not been met; Detecting a first input directed at the camera user interface while displaying the representation of the media; in response to detecting the first input directed at the camera user interface; Initiating capture of media to be added to a media library associated with the computer system according to determining that the first input corresponds to a selection of the media capture affordance; displaying, via the display generation component, a plurality of options for managing the individual pieces of text in accordance with determining that the first input corresponds to a selection of the first user interface object; A method comprising:

2. The method of claim 1 , wherein the first input is a tap gesture directed toward the first user interface object.

3. the media representation includes the individual pieces of text; The method comprises: detecting a second input directed at the camera user interface after detecting the first input; in response to detecting the second input directed at the camera user interface; displaying an indication that the first one or more portions of the individual text are selected in response to determining that the second input corresponds to selecting a first one or more portions of the individual text; and The method of claim 1 or 2, further comprising:

4. The method of claim 3 , wherein the second input is a tap gesture or a swipe gesture.

5. 5. The method of claim 1, further comprising: in response to detecting the first input directed at the camera user interface, displaying an indication regarding a selection of text included in the representation of the media in accordance with the determination that the first input corresponds to a selection of the first user interface object.

6. prior to detecting the first input, the representation of the media is displayed in a first appearance; The method comprises:

6. The method of claim 1, further comprising: in response to detecting the first input directed at the camera user interface, displaying the representation of the media in a second appearance different from the first appearance in accordance with the determination that the first input corresponds to a selection of the first user interface object.

7. the media representation includes the individual pieces of text; The method comprises: The method of claim 1 , further comprising highlighting a second one or more portions of the respective text in accordance with the determination that the respective set of criteria is satisfied.

8. highlighting the second one or more portions of the distinct text includes displaying an indication that distinct text has been detected; The method comprises: receiving a request to display a second representation of media while the second one or more portions of the individual text are highlighted; responsive to receiving the request to display a second representation of the media, translating the indication that distinct text has been detected from a first position within the camera user interface to a second position within the camera user interface; The method of claim 7 further comprising:

9. after detecting the first input, in accordance with the first input corresponding to a selection of the first user interface object, the media representation includes the respective text and an indication that a third one or more portions of the respective text are selected; The method comprises: receiving a request to display a third representation of the media; 9. The method of claim 1, further comprising: in response to receiving the request to display a third representation of the media, displaying an indication that at least a portion of the text included in the third representation of the media has been selected, wherein the indication that the at least a portion of the text included in the third representation of the media has been selected is different from the indication that the third one or more portions of the individual text have been selected.

10. after detecting the first input, in accordance with a determination that the first input corresponds to a selection of the first user interface object, the representation of the media includes the individual text and an indication that a fourth one or more portions of the individual text are selected, the fourth one or more portions of the individual text being displayed in a third position within the camera user interface; The method comprises: detecting changes in a physical environment within the field of view of one or more cameras in communication with the computer system; continuing to display the fourth one or more portions of the individual text in the third position within the camera user interface in response to detecting the change in the physical environment within the field of view of the one or more cameras; 10. The method of claim 1, further comprising:

11. prior to detecting the first input directed at the camera user interface; the computer system is in communication with one or more cameras; The method of claim 1 , wherein the representation of the media is a representation of one or more objects in a physical environment within the field of view of the one or more cameras.

12. the representation of media is a first representation of media; The method comprises: detecting a request to display a fifth representation of media while displaying the first user interface object; In response to detecting the request to display a fifth representation of the media, ceasing display of the first user interface object in accordance with a determination that the respective set of criteria is not satisfied; and 12. The method of claim 1, further comprising:

13. The method of claim 1 , wherein the respective criteria include criteria that are met when a determination is made that the respective text meets a predetermined prominence criterion.

14. 14. The method of claim 1, further comprising, while displaying the representation, displaying an indication that the distinct type of text has been detected pursuant to a determination that the distinct text includes a portion of text that has been determined to be a distinct type of text.

15. receiving a third input directed at a portion of the camera user interface that does not include the individual text while displaying the plurality of options managing the individual text; ceasing to display the plurality of options managing the individual text in response to receiving the third input; and 15. The method of any one of claims 1 to 14, further comprising:

16. While simultaneously displaying the representation of the media and the media capture affordance, in response to determining that the representation of the media includes first machine-readable code, displaying the first user interface object; displaying a representation of a uniform resource locator corresponding to the first machine-readable code; 16. The method of any one of claims 1 to 15, further comprising:

17. 17. The method of claim 1, wherein, in accordance with a determination that the first input corresponds to a selection of the first user interface object while the media representation includes second machine-readable code, the plurality of options for managing the individual text include one or more options for managing information corresponding to the second machine-readable code.

18. The method of claim 1 , wherein the camera user interface includes a plurality of selectable camera setting affordances for changing one or more camera settings.

19. The method of claim 1 , wherein the camera user interface includes an affordance that, when selected, causes a representation of one or more previously captured media to be displayed.

20. 20. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with a display generation component, the one or more programs including instructions for performing the method of any one of claims 1 to 19.

21. a computer system configured in communication with a display generation component, one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 20. A computer system comprising: said one or more programs comprising instructions for performing the method of any one of claims 1 to 19.

22. a computer system configured in communication with a display generation component, A computer system comprising means for carrying out the method of any one of claims 1 to 19.

23. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with a display generation component, the one or more programs comprising: and displaying, via the display generation component, a camera user interface that includes simultaneously displaying a representation of the media and a media capture affordance. While simultaneously displaying a representation of the media and the media capture affordance, displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied, the first set including criteria that are satisfied when a respective text is detected within the representation of the media; withholding display of the first user interface object in response to a determination that a respective set of criteria is not satisfied; Detecting a first input directed at the camera user interface while displaying the representation of the media; in response to detecting the first input directed at the camera user interface; Initiating capture of media to be added to a media library associated with the computer system in accordance with determining that the first input corresponds to a selection of the media capture affordance; and a non-transitory computer-readable storage medium comprising instructions for, in accordance with determining that the first input corresponds to a selection of the first user interface object, displaying, via the display generation component, a plurality of options for managing the individual pieces of text.

24. a computer system configured in communication with a display generation component, one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 1. A computer system comprising: displaying, via the display generation component, a camera user interface that includes simultaneously displaying a representation of the media and a media capture affordance; While simultaneously displaying a representation of the media and the media capture affordance, displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied, the first set including criteria that are satisfied when a respective text is detected within the representation of the media; withholding display of the first user interface object in response to a determination that a respective set of criteria is not satisfied; Detecting a first input directed at the camera user interface while displaying the representation of the media; in response to detecting the first input directed at the camera user interface; Initiating capture of media to be added to a media library associated with the computer system in accordance with determining that the first input corresponds to a selection of the media capture affordance; and instructions for displaying, via the display generation component, a plurality of options for managing the individual pieces of text in accordance with determining that the first input corresponds to a selection of the first user interface object.

25. a computer system configured in communication with a display generation component, one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; means for displaying, via the display generation component, a camera user interface that includes simultaneously displaying a representation of the media and a media capture affordance; While simultaneously displaying a representation of the media and the media capture affordance, displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied, the first set including criteria that are satisfied when a respective text is detected within the representation of the media; means for withholding display of the first user interface object in accordance with a determination that a respective set of criteria has not been met; means for detecting a first input directed at the camera user interface while displaying the representation of the media; in response to detecting the first input directed at the camera user interface; Initiating capture of media to be added to a media library associated with the computer system in accordance with determining that the first input corresponds to a selection of the media capture affordance; means for displaying, via the display generation component, a plurality of options for managing the individual pieces of text in accordance with a determination that the first input corresponds to a selection of the first user interface object.

26. 1. A method comprising: A computer system in communication with a display generation component and one or more input devices, comprising: displaying, via the display generation component, a first representation of a previously captured media item; While displaying the first representation of the previously captured media item, detecting an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to a request to display the second representation of the previously captured media item; while displaying the second representation of the previously captured media item; displaying, via the display generation component, a visual indication corresponding to a portion of the text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed, in accordance with determining that the portion of the text included in the second representation of the previously captured media item satisfies a respective set of criteria; A method comprising:

27. while displaying the second representation of the previously captured media item; 27. The method of claim 26, further comprising withholding display of the visual indication pursuant to a determination that the portion of text included in the second representation of the previously captured media item does not meet the respective set of criteria.

28. While displaying the second representation of the previously captured media item, detecting an input via the one or more input devices corresponding to a request to display a third representation of the previously captured media item; displaying, via the display generation component, the third representation of the previously captured media item in response to detecting the input corresponding to the request to display the third representation of the previously captured media item; while displaying the third representation of the previously captured media item; displaying, via the display generation component, a visual indication corresponding to the portion of text included in the third representation of the previously captured media item in accordance with determining that the portion of text included in the third representation satisfies the respective set of criteria; 28. The method of claim 26 or 27, further comprising:

29. 29. The method of any one of claims 26 to 28, wherein the first representation of the previously captured media item is a representation of the media item displayed at a first zoom level, and the second representation of the previously captured media item is a representation of the media item displayed at a second zoom level different from the first zoom level.

30. 30. The method of any one of claims 26 to 29, wherein the first representation of the previously captured media item is a representation of the media item displayed with a first translation amount, and the second representation of the previously captured media item is a representation of the media item displayed with a second translation amount different from the first translation amount.

31. the first representation of the previously captured media item includes a portion of the text; 31. The method of any one of claims 26 to 30, wherein the portion of text included in the first representation does not satisfy the respective set of criteria.

32. 32. The method of claim 26, wherein the input corresponding to the request to display the second representation of the previously captured media item is an input detected on the display generation component.

33. The second representation of the previously captured media item is displayed at a third zoom level, and the method further comprises: detecting, while displaying the second representation of the previously captured media item at the third zoom level, an input via the one or more input devices corresponding to a request to change the zoom level of the second representation of the previously captured media item; In response to detecting the input corresponding to the request to change the zoom level of the second representation of the previously captured media item, displaying, via the display generation component, a fourth representation of the previously captured media item at a fourth zoom level different from the third zoom level; while displaying the fourth representation of the previously captured media item at the fourth zoom level; withholding display of the visual indication in accordance with a determination that a first portion of text included in the fourth representation of the previously captured media item does not satisfy the respective set of criteria; 33. The method of any one of claims 26 to 32, further comprising:

34. detecting, while displaying the second representation of the previously captured media item, an input via the one or more input devices corresponding to a request to translate the second representation of the previously captured media item; In response to detecting the input corresponding to a request to translate the second representation of the previously captured media item, displaying a fifth representation of the previously captured media item that includes a portion of the media item that was not included in the second representation of the previously captured media item; while displaying the fifth representation of the previously captured media item; withholding display of the visual indication in accordance with a determination that a second portion of text included in the fifth representation of the previously captured media item does not satisfy the respective set of criteria; and 34. The method of any one of claims 26 to 33, further comprising:

35. While displaying the second representation of the previously captured media item, detecting an input via the one or more input devices, the input being a different type of input than the input corresponding to a request to display the second representation of the previously captured media item; withholding display of the visual indication in response to detecting the input being a different type of input than the input corresponding to a request to display the second representation of the previously captured media item; 35. The method of any one of claims 26 to 34, further comprising:

36. 36. The method of any one of claims 26 to 35, wherein the respective sets of criteria include criteria that are met when a determination is made that the amount of prominence of a respective portion of text included in a respective representation of a respective previously captured media item exceeds a prominence threshold.

37. 37. The method of claim 36, wherein the amount of prominence above the prominence threshold is based on distinct portions of the text that occupy more than the distinct expression threshold.

38. 38. A method according to claim 36 or 37, wherein the amount of prominence above the prominence threshold is based on the respective portion of text being displayed in a particular location within the respective representation.

39. 39. The method of any one of claims 36 to 38, wherein the amount of prominence above the prominence threshold is based on the individual portion of text being a particular type of text.

40. 40. The method of any one of claims 36 to 39, wherein the amount of prominence above the prominence threshold is based on a relevance score of the portion of the text relative to the content of the respective previously captured media item that meets a relevance score threshold.

41. while displaying the second representation of the previously captured media item; displaying a first user interface object corresponding to one or more text management operations in accordance with a determination that the portion of text included in the second representation of the previously captured media item satisfies the respective set of criteria; withholding display of the first user interface object corresponding to one or more text management operations in accordance with a determination that the portion of the text included in the second representation of the previously captured media item does not satisfy the respective set of criteria; and 41. The method of any one of claims 26 to 40, further comprising:

42. detecting, while displaying the second representation of the previously captured media item and displaying the first user interface object corresponding to one or more text management options, an input via the one or more input devices corresponding to a request to modify the second representation of the previously captured media item; displaying a twelfth representation of the previously captured media item in response to detecting the input corresponding to a request to modify the second representation of the previously captured media item; and while displaying the twelfth representation of the previously captured media item; withholding display of the first user interface object corresponding to one or more text management operations in accordance with a determination that a respective portion of text included in the twelfth representation of the previously captured media item does not satisfy the respective set of criteria; and 42. The method of claim 41 further comprising:

43. The second representation of the previously captured media item is displayed at a fifth zoom level, and the method further comprises: detecting, while displaying the second representation of the previously captured media item at the fifth zoom level and the visual indication, an input via the one or more input devices corresponding to a request to zoom in on the second representation of the previously captured media item; In response to detecting the input corresponding to the request to zoom in on the second representation of the previously captured media item, displaying, via the display generation component, a seventh representation of the previously captured media item at a sixth zoom level greater than the fifth zoom level, the seventh representation including a second portion of text included in the second representation of the previously captured media item that differs from the portion of the text included in the second representation; while displaying the seventh representation of the previously captured media item; displaying, via the display generation component, a visual indication corresponding to the second portion of the text that is different from the visual indication corresponding to the portion of the text in accordance with determining that the second portion of the text satisfies the respective criteria; and 43. The method of any one of claims 26 to 42, further comprising:

44. The second representation of the previously captured media is displayed at a seventh zoom level, and the method further comprises: detecting, while displaying the second representation of the previously captured media item at the seventh zoom level, an input via the one or more input devices corresponding to a request to zoom out the second representation of the previously captured media item; In response to detecting the input corresponding to the request to zoom out the second representation of the previously captured media item, displaying, via the display generation component, an eighth representation of the previously captured media item at an eighth zoom level that is less than the seventh zoom level; while displaying the eighth representation of the previously captured media item at the eighth zoom level; ceasing to display the visual indication in accordance with a determination that a first respective portion of text included in the eighth representation of the previously captured media item does not satisfy the respective set of criteria; and 44. The method of any one of claims 26 to 43, further comprising:

45. 45. The method of any one of claims 26 to 44, wherein the visual indication surrounds the portion of the text included in the second representation.

46. 46. The method of claim 26-45, wherein the location of the display of the visual indication coincides with the location of the portion of the text included in the second representation.

47. The second representation of the previously captured media item is displayed at a ninth zoom level, and the method further comprises: detecting, while displaying the second representation of the previously captured media item at the ninth zoom level and the visual indication, an input via one or more input devices corresponding to a request to zoom in on the second representation of the previously captured media item; In response to detecting the input corresponding to a request to zoom in on the second representation of the previously captured media item, displaying, via the display generation component, a ninth representation of the previously captured media item that includes a distinct portion of text included in the second representation of the previously captured media, the distinct portion of text being distinct from the portion of text included in the second representation, at a tenth zoom level that is greater than the ninth zoom level; while displaying the ninth representation of the previously captured media item; ceasing to display the visual indication corresponding to the portion of the text; and displaying, via the display generation component, a visual indication corresponding to the second distinct portion of the text that is different from the visual indication corresponding to the portion of the text in accordance with determining that the second distinct portion of the text satisfies the respective criteria; 47. The method of any one of claims 26 to 46, further comprising:

48. 48. The method of any one of claims 26 to 47, wherein a first subset of at least the portions of text is selectable.

49. The second representation of the previously captured media item includes a third portion of non-selectable text, and the method further comprises: While displaying the second representation of the previously captured media item, detecting an input via one or more input devices corresponding to a request to display a tenth representation of the previously captured media item; 49. The method of any one of claims 26 to 48, further comprising: in response to detecting an input corresponding to a request to display the tenth representation of the previously captured media item, displaying the tenth representation of the previously captured media item including the portion of the text, wherein the third portion of the text included in the tenth representation of the previously captured media item is selectable.

50. 50. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with a display generation component and one or more input devices, the one or more programs including instructions for performing the method of any one of claims 26 to 49.

51. 1. A computer system configured in communication with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; wherein the one or more programs contain instructions for performing the method of any one of claims 26 to 49. Computer system.

52. a computer system configured in communication with a display generation component and one or more inputs, 50. A computer system comprising means for carrying out the method of any one of claims 26 to 49.

53. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with a display generation component and one or more input devices, the one or more programs comprising: displaying, via the display generation component, a first representation of a previously captured media item; While displaying the first representation of the previously captured media item, detect an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to a request to display the second representation of the previously captured media item; while displaying the second representation of the previously captured media item; pursuant to determining that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria, displaying, via the display generation component, a visual indication corresponding to a portion of the text included in the second representation that was not displayed when the first representation of the previously captured media item was displayed. A non-transitory computer-readable storage medium containing instructions.

54. 1. A computer system configured in communication with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 1. A computer system comprising: displaying, via the display generation component, a first representation of a previously captured media item; While displaying the first representation of the previously captured media item, detect an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to a request to display the second representation of the previously captured media item; while displaying the second representation of the previously captured media item; pursuant to determining that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria, displaying, via the display generation component, a visual indication corresponding to a portion of the text included in the second representation that was not displayed when the first representation of the previously captured media item was displayed. A computer system including instructions.

55. 1. A computer system configured in communication with a display generation component and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; means for displaying a first representation of a previously captured media item via the display generation component; means for detecting, while displaying the first representation of the previously captured media item, an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; means for displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to a request to display the second representation of the previously captured media item; while displaying the second representation of the previously captured media item; means for displaying, via the display generation component, a visual indication corresponding to a portion of the text included in the second representation of the previously captured media item that was not displayed when the first representation of the previously captured media item was displayed, in accordance with a determination that the portion of the text included in the second representation of the previously captured media item satisfies a respective set of criteria; A computer system comprising:

56. 1. A method comprising: A computer system in communication with one or more cameras, one or more input devices, and a display generation component, comprising: Displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; In response to detecting the request to display the camera user interface, via the display generation component, a camera user interface, the camera user interface comprising: displaying a camera user interface including a representation of the field of view of the one or more cameras; displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with a determination that the representation of the field of view of the one or more cameras includes detected text that meets one or more criteria; detecting, while simultaneously displaying the representation of the field of view and the text insertion user interface object, an input via the one or more input devices corresponding to a selection of the text insertion user interface object; In response to detecting the input corresponding to a selection of the text insertion user interface object, inserting at least a portion of the detected text into the text entry area; A method comprising:

57. displaying the camera user interface in response to detecting the request to display the camera user interface, 57. The method of claim 56, comprising withholding display of the text insertion user interface object pursuant to a determination that the representation of the field of view of the one or more cameras does not include detected text that meets one or more criteria.

58. 58. The method of claim 56 or 57, wherein displaying the camera user interface in response to detecting the request to display the camera user interface includes: displaying the text insertion user interface object in response to a determination that the representation of the field of view of the one or more cameras includes detected text that meets one or more criteria, and that the representation of the field of view of the one or more cameras does not include detected text that meets one or more criteria, wherein the text insertion user interface object is not selectable.

59. displaying the camera user interface in response to detecting the request to display the camera user interface, 59. The method of claim 58, comprising, pursuant to determining that the text insertion user interface object is not selectable, displaying the text insertion user interface object with an appearance that indicates that the text insertion user interface object is disabled.

60. 60. A method according to any one of claims 56 to 59, wherein the camera user interface is not displayed before the request to display the camera user interface has been detected.

61. 61. The method of any one of claims 56 to 60, wherein the user interface includes an input entry user interface element, the input entry user interface element including a user interface object displayed at a location within the input entry user interface element.

62. while displaying the first user interface including a text entry area, detecting input directed at the text entry area via the one or more input devices prior to detecting the request to display the camera user interface; displaying, via the display generation component, a third user interface object in response to detecting the input directed at the text entry area; and 62. The method of any one of claims 56 to 61, further comprising:

63. 63. The method of claim 62, wherein a fourth user interface object is displayed simultaneously with the third user interface object, the fourth user interface object being selected to cause the computer system to display the copied text.

64. 64. The method of any one of claims 56 to 63, wherein before detecting the request to display the camera user interface, the first user interface includes a keyboard displayed at a first location within the first user interface, and displaying the camera user interface includes replacing the display of the keyboard at the first location with the display of the camera user interface at the first location.

65. The camera user interface is displayed at a first size, and the method includes: detecting an input directed at the camera user interface via the one or more input devices while displaying the camera user interface at a first size; Responsive to detecting the input directed at the camera user interface, changing a size of the camera user interface from a first size to a second size different from the first size; 65. The method of any one of claims 56 to 64, further comprising:

66. the detected text includes a first portion of text and a second portion of text; At least the inserted portion of the detected text 66. The method of any one of claims 56 to 65, including the first portion of the text and excluding the second portion of the text in accordance with a determination that the first portion of the text is more prominent than the second portion of the text in the representation of the field of view of the one or more cameras.

67. 67. The method of any one of claims 56 to 66, wherein the text entry area is associated with a first type of text, and the one or more criteria include a respective criterion that is satisfied when a respective portion of the detected text is detected as being text of the first type.

68. 68. The method of claim 67, wherein, in response to detecting the input corresponding to a selection of the text insertion user interface object, in accordance with a determination that the representation of the field of view of the one or more cameras includes detected text that satisfies one or more criteria, a third portion of the detected text satisfies the respective criteria, a fourth portion of the detected text does not satisfy the respective criteria, and the at least a portion of the detected text includes the third portion of the detected text but does not include the fourth portion of the detected text.

69. the text entry area is associated with a second particular type of text; the representation of the field of view includes the detected text; The method comprises:

69. The method of any one of claims 56 to 68, further comprising refraining from displaying the selectable text insertion user interface in accordance with a determination that the detected text does not satisfy the one or more criteria, wherein the one or more criteria include a criterion that is satisfied when a portion of the detected text is text of the second particular type.

70. the detected text includes a fifth portion of text and a sixth portion of text, and the method further comprises: detecting, while simultaneously displaying the representation of the field of view and the text insertion user interface object, a request corresponding to a selection of a fifth portion of the text via the one or more input devices before detecting the input corresponding to a selection of the text insertion user interface object; In response to detecting the request to select the fifth portion of the text, selecting the fifth portion of the text without selecting the sixth portion of the text; 70. The method of any one of claims 56 to 69, further comprising:

71. the detected text includes a seventh portion of the text and an eighth portion of the text; Inserting the portion of the detected text into the text entry area includes: inserting the seventh portion of the text into the text entry area in accordance with a determination that the seventh portion of the text satisfies a set of text selection criteria and the eighth portion of the text does not satisfy a set of text selection criteria; inserting the eighth portion of the text into the text entry area in accordance with a determination that the seventh portion of the text does not satisfy a set of text selection criteria and the eighth portion of the text satisfies the set of text selection criteria; 71. The method of any one of claims 56 to 70, comprising:

72. 72. The method of claim 71 , wherein the determination that the second distinct portion of text satisfies the set of text selection criteria is based on the location of the one or more cameras and an orientation of the one or more cameras relative to an external environment.

73. the detected text includes a ninth portion of text and a tenth portion of text, and the method further comprises: displaying a first visual indication corresponding to a ninth portion of the text and a second visual indication corresponding to a tenth portion of the text while simultaneously displaying the representation of the field of view and the text insertion user interface object; 73. The method of any one of claims 56 to 72, further comprising:

74. the detected text displayed within the representation of the field of view of the one or more cameras has a first appearance; 74. The method of any one of claims 56 to 73, wherein detected text within the field of view of the one or more cameras has a second appearance different from the first appearance.

75. detecting a request to move a fifth user interface object while displaying the representation of the field of view of the one or more cameras and the fifth user interface object; In response to detecting the request to move the fifth user interface object, displaying, via the display generation component, a sixth user interface object different from the fifth user interface object in accordance with determining that the fifth user interface object is within a predetermined distance from a location of the detected text that satisfies the one or more criteria; 75. The method of any one of claims 56 to 74, further comprising:

76. the detected text includes an eleventh portion of text, and the method further comprises: after inserting the at least a portion of the detected text into the text entry area, while simultaneously displaying the representation of the field of view and the text insertion user interface object, detecting input via one or more input devices directed at an eleventh portion of the text in the representation of the field of view; responsive to detecting the input directed at an eleventh portion of the text within the representation of the field of view, inserting the eleventh portion of the text into the text entry area; 76. The method of any one of claims 56 to 75, further comprising:

77. the detected text includes a twelfth portion of text, and the method further comprises: after inserting the at least a portion of the detected text into the text entry area, detecting input directed at a twelfth portion of the text via one or more input devices while simultaneously displaying the representation of the field of view and the text insertion user interface object; in response to detecting the input directed to a twelfth portion of the text; selecting the twelfth portion of text in accordance with a determination that the twelfth portion of text exceeds a threshold size; forgoing selection of the twelfth portion of text in accordance with determining that the twelfth portion of text does not exceed the threshold size; and 77. The method of any one of claims 56 to 76, further comprising:

78. the detected text includes a thirteenth portion of text that is not selectable, and the method further comprises: detecting, after inserting the at least a portion of the detected text into the text entry area, a first request to modify the representation of the field of view of the one or more cameras via the one or more input devices while simultaneously displaying the representation of the field of view and the text insertion user interface object; selectably modifying a thirteenth portion of the text in response to detecting the first request to display the second camera user interface; 78. The method of any one of claims 56 to 77, further comprising:

79. after inserting the at least a portion of the detected text into the text entry area, detecting a second request to modify the representation of the field of view of the one or more cameras via the one or more input devices while simultaneously displaying the representation of the field of view and the text insertion user interface object; withholding display of the text insertion user interface object in response to detecting the second request to change the representation of the field of view of the one or more cameras; 79. The method of any one of claims 56 to 78, further comprising:

80. 80. The method of any one of claims 56 to 79, wherein the representation of the field of view of the one or more cameras is displayed simultaneously with a portion of the first user interface that includes the text entry area.

81. 81. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with one or more cameras, one or more input devices, and a display generation component, the one or more programs including instructions for performing the method of any one of claims 56 to 80.

82. 1. A computer system configured in communication with one or more cameras, one or more input devices, and a display generation component, the computer system comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; Equipped with 81. A computer system, wherein the one or more programs include instructions for performing the method of any one of claims 56 to 80.

83. 1. A computer system configured in communication with one or more cameras, one or more input devices, and a display generation component, comprising:

81. A computer system comprising means for carrying out the method of any one of claims 56 to 80.

84. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with one or more cameras, one or more input devices, and a display generation component, the one or more programs comprising: Displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; In response to detecting the request to display the camera user interface, via the display generation component, a camera user interface, the camera user interface comprising: displaying a camera user interface including a representation of the field of view of the one or more cameras; displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with a determination that the representation of the field of view of the one or more cameras includes detected text that meets one or more criteria; detecting, while simultaneously displaying the representation of the field of view and the text insertion user interface object, an input via the one or more input devices corresponding to a selection of the text insertion user interface object; a non-transitory computer-readable storage medium comprising instructions for, in response to detecting the input corresponding to a selection of the text insertion user interface object, inserting at least a portion of the detected text into the text entry area;

85. 1. A computer system configured in communication with one or more cameras, one or more input devices, and a display generation component, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 1. A computer system comprising: Displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; In response to detecting the request to display the camera user interface, via the display generation component, a camera user interface, the camera user interface comprising: displaying a camera user interface including a representation of the field of view of the one or more cameras; displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with a determination that the representation of the field of view of the one or more cameras includes detected text that meets one or more criteria; detecting, while simultaneously displaying the representation of the field of view and the text insertion user interface object, an input via the one or more input devices corresponding to a selection of the text insertion user interface object; responsive to detecting the input corresponding to a selection of the text insertion user interface object, inserting at least a portion of the detected text into the text entry area.

86. 1. A computer system configured in communication with one or more cameras, one or more input devices, and a display generation component, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; means for displaying a first user interface including a text entry area; means for detecting a request to display a camera user interface while displaying the first user interface including the text entry area; In response to detecting the request to display the camera user interface, via the display generation component, a camera user interface, the camera user interface comprising: displaying a camera user interface including a representation of the field of view of the one or more cameras; means for displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with a determination that the representation of the field of view of the one or more cameras includes detected text that meets one or more criteria; means for detecting, via the one or more input devices, an input corresponding to a selection of the text insertion user interface object while simultaneously displaying the representation of the field of view and the text insertion user interface object; means for inserting at least a portion of the detected text into the text entry area in response to detecting the input corresponding to a selection of the text insertion user interface object; A computer system comprising:

87. 1. A method comprising: a computer system in communication with a display generation component, displaying, via the display generation component, a media user interface including a representation of the media; receiving, while displaying the media user interface including the representation of the media, a request to display additional information regarding a plurality of detected characteristics within the representation of the media; in response to receiving the request to display additional information about the plurality of detected characteristics, while displaying the media user interface including the representation of the media, displaying one or more indications of the plurality of detected characteristics in the media, wherein the one or more indications of the plurality of detected characteristics include a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to a location of the first detected characteristic within the representation of the media, and wherein displaying the one or more indications includes: In response to determining that the first detected characteristic is a first type of characteristic, the first indication has a first appearance; In response to a determination that the first detected characteristic is a second type characteristic different from the first type characteristic, the first indication has a second appearance different from the first appearance. A method comprising:

88. the one or more indications of a plurality of detected characteristics in the media include a second indication of a second detected characteristic displayed at a second location in the representation of the media, the second location corresponding to a location of the second detected characteristic in the representation of the media; In response to determining that the second detected characteristic is a characteristic of the first type, the second indication has the first appearance; 88. The method of claim 87, wherein, in accordance with a determination that the second detected characteristic is a characteristic of the second type that is different from the characteristic of the first type, the second indication has the second appearance that is different from the first appearance.

89. the first indication of the first detected characteristic is displayed simultaneously with the second indication of the second detected characteristic; 89. The method of claim 88, wherein the first detected characteristic is a characteristic of the first type and the second characteristic is a characteristic of the second type.

90. the first indication of the first detected characteristic is displayed simultaneously with the second indication of the second detected characteristic; the first detected characteristic is different from the second detected characteristic; 89. The method of claim 88, wherein the first detected characteristic is a characteristic of the first type and the second characteristic is a characteristic of the second type.

91. the first indication having the first appearance is displayed in a first color; 91. The method of any one of claims 88 to 90, wherein the first indication having the second appearance is not displayed in the first color.

92. the first indication having the first appearance is displayed with a first graphical representation of the first type of characteristic; 92. The method of any one of claims 88 to 91, wherein the first indication having the second appearance is displayed with a second graphical representation of the second type of characteristic that is different from the first graphical representation.

93. the one or more indications of detected characteristics in the media include a third indication of a third detected characteristic that is a characteristic of the first type, a fourth indication of a fourth detected characteristic that is a characteristic of the first type, and a fifth indication of a fifth detected characteristic that is a characteristic of the second type; the third indication is displayed in a similar appearance to the fourth indication; 93. The method of any one of claims 87 to 92, wherein the third indication is displayed in a different appearance than the fifth indication.

94. 94. The method of any one of claims 87 to 93, wherein receiving the request to display additional information regarding the plurality of detected characteristics of the representation of the media comprises detecting input directed to a media library.

95. detecting a first input directed toward the first indication of the first detected characteristic while displaying the first indication of the first detected characteristic; In response to detecting the first input directed to the first indication of the first detected characteristic, displaying via the display generation component a first user interface object including information related to the first detected characteristic; 95. The method of any one of claims 87 to 94, further comprising:

96. 96. The method of claim 95, wherein the information regarding the first detected characteristic includes a representation of a portion of the media that corresponds to the first detected characteristic.

97. the one or more indications of detected characteristics in the media include a sixth indication of a sixth detected characteristic, and the method further comprises: detecting an input directed toward the sixth indication of the sixth detected characteristic while displaying the first user interface object including information about the first detected characteristic and the sixth indication of the sixth detected characteristic; in response to detecting the input directed toward the sixth indication of the sixth detected characteristic; displaying, via the display generation component, a second user interface object including information regarding the sixth detected characteristic; and ceasing, via the display generation component, to display the first user interface object including information related to the first detected characteristic; and 97. The method of claim 95 or 96, further comprising:

98. 98. The method of any one of claims 95 to 97, wherein the information about the first detected characteristic includes an option to perform an action.

99. the one or more indications include a seventh indication of a seventh detected characteristic; Displaying the one or more indications of the detected characteristics in the media via the display generation component includes displaying an animation of the first indication being displayed before the seventh indication of the seventh characteristic is displayed; 99. The method of any one of claims 87 to 98, wherein after displaying the animation, the first indication is displayed simultaneously with the seventh indication.

100. detecting a second input directed toward the first indication of the first detected characteristic while displaying the first indication of the first detected characteristic; displaying, via the display generation component, a third graphical representation of the first type of characteristic in response to detecting the second input directed to the first indication of the first detected characteristic; 100. The method of any one of claims 87 to 99, further comprising:

101. the one or more indications of detected characteristics in the media include a ninth indication of a ninth detected characteristic, and the method further comprises: detecting an input directed toward the ninth indication of the ninth detected characteristic while displaying the third graphical representation of the first type of characteristic and the ninth indication of the ninth detected characteristic; ceasing, via the display generation component, to display the third graphical representation of the first type of characteristic in response to detecting the input directed to the ninth indication of the ninth detected characteristic; and 101. The method of any one of claims 87 to 100, further comprising:

102. detecting a third input directed toward the first indication of the first detected characteristic while displaying the first indication of the first detected characteristic; In response to detecting the third input directed to the first indication of the first detected characteristic, displaying, via the display generation component, a first user interface object including information related to the first detected characteristic and information corresponding to the representation of the media and not corresponding to the first detected characteristic; 102. The method of any one of claims 87 to 101, further comprising:

103. 103. The method of claim 102, wherein the information corresponding to the representation of the media and not corresponding to the first detected characteristic comprises metadata corresponding to the representation of the media.

104. 104. The method of claim 102 or 103, wherein the information corresponding to the representation of the media and not corresponding to the first detected characteristic includes one or more options for applying effects.

105. 105. The method of any one of claims 102 to 104, wherein the information corresponding to the representation of the media and not corresponding to the first detected characteristic comprises one or more links to related content in a media library.

106. In response to receiving the request to display additional information regarding the plurality of detected characteristics, the first indication is displayed at a first location on the display generation component; the representation of the media is displayed at a first zoom level; The method comprises: displaying the first indication of the first detected characteristic at the first location and detecting a fourth input directed toward the first indication of the first detected characteristic while the representation of the media is displayed at a second zoom level; in response to detecting the fourth input directed at the first indication of the first detected characteristic, enlarging the representation of the media and displaying the first indication at a second location closer to a center of the display generation component than the first location; 106. The method of any one of claims 87 to 105, further comprising:

107. the plurality of detected characteristics includes a tenth detected characteristic that is a tenth type of detected characteristic; Displaying the one or more indications via the display generation component includes:

107. The method of any one of claims 87 to 106, comprising, in accordance with a determination that a tenth location within the representation of the media corresponding to a location of the tenth detected characteristic cannot be determined, displaying, via the display generation component, a tenth indication corresponding to the tenth detected characteristic at a predetermined location on the media user interface.

108. 108. The method of claim 107, wherein the tenth indication displayed at the predetermined location is displayed simultaneously with the first indication displayed at the first location.

109. the plurality of detected characteristics includes an eleventh detected characteristic; Displaying the one or more indications via the display generation component includes: In response to a determination of being unable to determine an eleventh location within the representation of the media corresponding to the tenth detected characteristic location and being unable to determine a twelfth location corresponding to the eleventh detected characteristic location, pursuant to a determination that the tenth detected characteristic and the eleventh detected characteristic are different types of detected characteristics, displaying, via the display generation component, an eleventh indication at a second predetermined location within the media user interface, the eleventh indication corresponding to the characteristic type of the eleventh detected characteristic; pursuant to a determination that the tenth detected characteristic and the eleventh detected characteristic are detected characteristics of the same type, withholding display of the eleventh indication via the display generation component; 109. The method of claim 107 or 108, comprising:

110. while displaying the tenth indication, if the tenth detected characteristic and the eleventh detected characteristic are detected characteristics of the same type, detecting an input directed toward the tenth indication; In response to detecting the input directed to the tenth indication, displaying, via the display generation component, a user interface object including information regarding the tenth detected characteristic and information regarding the eleventh detected characteristic; 110. The method of claim 109, further comprising:

111. 111. The method of any one of claims 87 to 110, wherein displaying the media user interface includes simultaneously displaying a first user interface object that displays additional information with a user interface object that corresponds to one or more text management operations.

112. receiving a request to display a second representation of a second media different from the media while displaying the representation of the media; In response to receiving the request to display the second representation of a second media different from the media, displaying, via the display generation component, a second user interface object displaying additional information in accordance with a determination that the representation of the second media includes one or more detected characteristics; and and forgoing, via the display generation component, displaying the second user interface object that displays additional information in accordance with a determination that the representation of the media does not include the one or more detected characteristics.

112. The method of any one of claims 87 to 111, further comprising:

113. 113. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with a display generation component, the one or more programs including instructions for performing the method of any one of claims 87 to 112.

114. 1. A computer system configured in communication with a display generation component, the computer system comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; wherein the one or more programs contain instructions for performing the method of any one of claims 87 to 112. Computer system.

115. a computer system configured in communication with a display generation component, 113. A computer system comprising means for carrying out the method of any one of claims 87 to 112.

116. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with a display generation component, the one or more programs comprising: displaying, via the display generation component, a media user interface including a representation of the media; receiving, while displaying the media user interface including the representation of the media, a request to display additional information regarding a plurality of detected characteristics within the representation of the media; 1. A non-transitory computer-readable storage medium comprising instructions for, in response to receiving the request to display additional information about the plurality of detected characteristics, displaying one or more indications of a plurality of detected characteristics in the media while displaying the media user interface including the representation of the media, the one or more indications of the plurality of detected characteristics including a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to a location of the first detected characteristic within the representation of the media, and displaying the one or more indications includes: In response to determining that the first detected characteristic is a first type of characteristic, the first indication has a first appearance; In response to determining that the first detected characteristic is a second type characteristic different from the first type characteristic, the first indication has a second appearance different from the first appearance.

1. A non-transitory computer-readable storage medium comprising:

117. a computer system configured in communication with a display generation component, one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 1. A computer system comprising: displaying, via the display generation component, a media user interface including a representation of the media; receiving, while displaying the media user interface including the representation of the media, a request to display additional information regarding a plurality of detected characteristics within the representation of the media; 11. A computer system, comprising: instructions for, in response to receiving the request to display additional information about the plurality of detected characteristics, displaying one or more indications of a plurality of detected characteristics in the media while displaying the media user interface including the representation of the media, the one or more indications of the plurality of detected characteristics including a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to a location of the first detected characteristic within the representation of the media, and displaying the one or more indications includes: In response to determining that the first detected characteristic is a first type of characteristic, the first indication has a first appearance; In accordance with a determination that the first detected characteristic is a second type characteristic that is different from the first type characteristic, the first indication includes having a second appearance that is different from the first appearance.

118. a computer system configured in communication with a display generation component, one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; means for displaying, via said display generation component, a media user interface including a representation of the media; means for receiving, while displaying the media user interface including the representation of the media, a request to display additional information regarding a plurality of detected characteristics within the representation of the media; means for displaying one or more indications of a plurality of detected characteristics in the media while displaying the media user interface including the representation of the media in response to receiving the request to display additional information about the plurality of detected characteristics, the one or more indications of the plurality of detected characteristics including a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to a location of the first detected characteristic within the representation of the media, and wherein displaying the one or more indications includes: In response to determining that the first detected characteristic is a first type of characteristic, the first indication has a first appearance; pursuant to determining that the first detected characteristic is a second type characteristic different from the first type characteristic, the first indication has a second appearance different from the first appearance; 2. A computer system comprising:

119. 1. A method comprising: A computer system in communication with one or more cameras, a display generating component, and one or more input devices, comprising: receiving a request to display a representation of the field of view of the one or more cameras; In response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via the display generation component, the representation of the field of view of the one or more cameras, the representation including text that is within the field of view of the one or more cameras; automatically displaying, via the display generation component, a plurality of indications of the translated text, including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; receiving, via the one or more input devices, a request to select a respective indication of the plurality of translated portions while displaying, via the display generation component, the first indication and the second indication; In response to receiving the request to select the respective indication, in accordance with determining that the request is a request to select the first indication, displaying, via the display generation component, a first translation user interface object that includes the first portion of the text and the translation of the first portion of the text without the translation of the second portion of the text; A method comprising:

120. 120. The method of claim 119, further comprising: in response to receiving the request to select the individual indication, in accordance with a determination that the request is a request to select the second indication, displaying, via the display generation component, a second translation user interface object that includes a second portion of text and the translation of the second portion of text without including a translation of the first portion of text.

121. 121. The method of claim 119 or 120, wherein the first translation user interface object includes a pronunciation option that, when enabled, causes the computer system to output an indication of how to pronounce the first portion of text, and a pronunciation option that, when enabled, causes the computer system to output an indication of how to pronounce the translation of the first portion of text.

122. 122. The method of any one of claims 119 to 121, wherein the representation of the field of view of the one or more cameras is a representation of previously captured media.

123. 123. A method according to any one of claims 119 to 122, wherein the representation of the field of view of the one or more cameras is a representation of the field of view of the one or more cameras as currently captured.

124. after displaying the first translation user interface object, receiving, via the one or more input devices, a request to share the first translation user interface object, the request including input detected while displaying the translation user interface object; In response to receiving the request to share the first translation user interface object, transmitting media corresponding to the first translation user interface object to one or more other computer systems; 124. The method of any one of claims 119 to 123, further comprising:

125. receiving, while displaying the first translation user interface object, a request via the one or more input devices to save the first translation user interface object; In response to receiving the request to store the first translation user interface object, storing media corresponding to the first translation user interface object in a library of translations accessible on the computer system; 125. The method of any one of claims 119 to 124, further comprising:

126. receiving, while displaying the representation of the field of view of the one or more cameras and the plurality of indications, a request to share the representation of the field of view of the one or more cameras via the one or more input devices; In response to receiving the request to share the representation of the field of view of the one or more cameras, transmitting media including at least a portion of the representation of the field of view of the one or more cameras and the plurality of indications; 126. The method of any one of claims 119 to 125, further comprising:

127. the computer system is in communication with a light source, and the method comprises: In response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, at a first location within the user interface via the display generation component, a first selectable user interface object that, when selected, changes an operational state of the light source in accordance with determining that the computer system is in a first active capture state; pursuant to determining that the computer system is not in the first active capture state, displaying, via the display generation component, at the first location within the user interface, a second selectable user interface object that, when selected, initiates a sharing process; 127. The method of any one of claims 119 to 126, further comprising:

128. 128. The method of any one of claims 119 to 127, wherein the first translation user interface object is displayed regardless of whether the computer system is in a second active capture state.

129. 129. The method of any one of claims 119 to 128, wherein a first portion of the representation of the field of view of the one or more cameras is displayed simultaneously with the first translation user interface object.

130. Displaying the representation of the field of view of the one or more cameras via the display generation component includes: in response to a change in the field of view of the one or more cameras; updating, via the display generation component, the representation of the field of view of the one or more cameras to reflect the change in the field of view of the one or more cameras in accordance with determining that the computer system is in a third active capture state; in response to determining that the computer system is not in the active capture state, forgoing, via the display generation component, updating the representation of the field of view of the one or more cameras to reflect the changes in the field of view of the one or more cameras; 130. The method of any one of claims 119 to 129, comprising:

131. 131. The method of claim 130, wherein the updated representation of the field of view of the one or more cameras is displayed simultaneously with the first translation user interface object.

132. a second portion of the representation of the field of view of the one or more cameras is displayed simultaneously with the first translation user interface object, the method comprising: receiving, via the one or more input devices, a request to cease displaying the first translation user interface object while simultaneously displaying, via the display generation component, the second portion of the representation of the field of view of the one or more cameras with the first translation user interface object; In response to receiving the request to discontinue displaying the first user interface object, discontinuing display of the first translation user interface object via the display generation component and displaying a portion of the representation not previously displayed while the first translation user interface object was displayed; 132. The method of any one of claims 119 to 131, further comprising:

133. 133. The method of any one of claims 119 to 132, wherein automatically displaying the plurality of indications of translated text via the display generation component comprises displaying the first indication of the translation of the first portion of text on top of the first portion of the text.

134. the first portion of text is displayed in a first color; the first indication of the translation is displayed in the first color; a second portion of the text is displayed in a second color different from the first color; 134. The method of any one of claims 119 to 133, wherein the second indication is displayed in the second color.

135. the first indication is displayed at a third location corresponding to the first portion of the text, and the method further comprises: receiving a request to display a second representation of the field of view of the one or more cameras while displaying the first indication at the third location and the representation of the field of view of the one or more cameras; In response to receiving the request to display a second representation of the field of view, in response to determining that the second representation includes the first portion of the text, displaying the second representation of the field of view; continuing to display the first indication at the third location; and 135. The method of any one of claims 119 to 134, further comprising:

136. The first translation user interface object is displayed at a third location, and the method further comprises: receiving, while displaying the first translation user interface object and the plurality of indications at the third location, a second request to select the individual indication via the one or more input devices; 136. The method of any one of claims 119 to 135, further comprising: in response to receiving the second request to select the individual indication, in accordance with a determination that the second request is a request to select the second indication, replacing the display of the first translation user interface object with a display of a third translation user interface object at the third location, wherein the third translation user interface object includes a second portion of the text and a translation of the second portion of the text without including a translation of the first portion of the text.

137. 137. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with a display generation component, the one or more programs including instructions for performing the method of any one of claims 119 to 136.

138. 1. A computer system configured in communication with one or more cameras, a display generation component, and one or more input devices, said computer system comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 137. A computer system comprising: said one or more programs comprising instructions for performing the method of any one of claims 119 to 136.

139. 1. A computer system configured in communication with one or more cameras, a display generation component, and one or more input devices, comprising:

137. A computer system comprising means for carrying out the method of any one of claims 119 to 136.

140. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system, the computer system being in communication with one or more cameras, a display generating component, and one or more input devices, the one or more programs comprising: receiving a request to display a representation of the field of view of the one or more cameras; In response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via the display generation component, the representation of the field of view of the one or more cameras, the representation including text that is within the field of view of the one or more cameras; automatically displaying, via the display generation component, a plurality of indications of the translated text, including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; receiving, via the one or more input devices, a request to select a respective indication of the plurality of translated portions while displaying the first indication and the second indication via the display generation component; a non-transitory computer-readable storage medium comprising instructions for, in response to receiving the request to select the individual indication, and in accordance with determining that the request is a request to select the first indication, displaying, via the display generation component, a first translation user interface object that includes the first portion of the text and the translation of the first portion of the text without including the translation of the second portion of the text.

141. 1. A computer system configured in communication with one or more cameras, a display generation component, and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 1. A computer system comprising: receiving a request to display a representation of the field of view of the one or more cameras; In response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via the display generation component, the representation of the field of view of the one or more cameras, the representation including text that is within the field of view of the one or more cameras; automatically displaying, via the display generation component, a plurality of indications of the translated text, including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; receiving, via the one or more input devices, a request to select a respective indication of the plurality of translated portions while displaying the first indication and the second indication via the display generation component; and, in response to receiving the request to select the individual indication, in accordance with a determination that the request is a request to select the first indication, displaying, via the display generation component, a first translation user interface object that includes the first portion of the text and the translation of the first portion of the text without including the translation of the second portion of the text.

142. 1. A computer system configured in communication with one or more cameras, a display generation component, and one or more input devices, comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; means for receiving a request to display a representation of the field of view of the one or more cameras; In response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via the display generation component, the representation of the field of view of the one or more cameras, the representation including text that is within the field of view of the one or more cameras; means for automatically displaying, via the display generation component, a plurality of indications of the translated text, including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; means for receiving, via the one or more input devices, a request to select a respective indication of the plurality of translated portions while displaying, via the display generation component, the first indication and the second indication; means for displaying, in response to receiving the request to select the respective indication, via the display generation component, a first translation user interface object including the first portion of the text and the translation of the first portion of the text without including the translation of the second portion of the text, in accordance with a determination that the request is a request to select the first indication; A computer system comprising:

143. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs comprising: displaying, via the display generation component, a camera user interface, including simultaneously displaying a representation of the media and a media capture affordance; While simultaneously displaying a representation of the media and the media capture affordance, displaying, via the display generation component, a first user interface object corresponding to one or more text management operations in accordance with a determination that a respective set of criteria is satisfied, the first set including criteria that are satisfied when a respective text is detected within the representation of the media; withholding display of the first user interface object in response to a determination that a respective set of criteria is not satisfied; Detecting a first input directed at the camera user interface while displaying the representation of the media; in response to detecting the first input directed at the camera user interface; Initiating capture of media to be added to a media library associated with the computer system in accordance with determining that the first input corresponds to a selection of the media capture affordance; and instructions for displaying, via the display generation component, a plurality of options for managing the individual pieces of text in accordance with determining that the first input corresponds to a selection of the first user interface object.

144. 20. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs including instructions for the method of any one of claims 1 to 19.

145. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs comprising: displaying, via the display generation component, a first representation of a previously captured media item; While displaying the first representation of the previously captured media item, detect an input via the one or more input devices corresponding to a request to display a second representation of the previously captured media item; displaying, via the display generation component, the second representation of the previously captured media item in response to detecting the input corresponding to a request to display the second representation of the previously captured media item; while displaying the second representation of the previously captured media item; pursuant to determining that a portion of text included in the second representation of the previously captured media item satisfies a respective set of criteria, displaying, via the display generation component, a visual indication corresponding to a portion of the text included in the second representation that was not displayed when the first representation of the previously captured media item was displayed. A computer program product containing instructions.

146. 50. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component and one or more input devices, the one or more programs including instructions for the method of any one of claims 26 to 49.

147. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, one or more input devices, and a display generation component, the one or more programs comprising: Displaying a first user interface including a text entry area; detecting a request to display a camera user interface while displaying the first user interface including the text entry area; In response to detecting the request to display the camera user interface, via the display generation component, a camera user interface, the camera user interface comprising: displaying a camera user interface including a representation of the field of view of the one or more cameras; displaying a selectable text insertion user interface object that inserts at least a portion of the detected text into the text entry area in accordance with a determination that the representation of the field of view of the one or more cameras includes detected text that meets one or more criteria; detecting, while simultaneously displaying the representation of the field of view and the text insertion user interface object, an input via the one or more input devices corresponding to a selection of the text insertion user interface object; a computer program product comprising instructions for, in response to detecting the input corresponding to a selection of the text insertion user interface object, inserting at least a portion of the detected text into the text entry area;

148. 81. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, one or more input devices, and a display generation component, the one or more programs including instructions for the method of any one of claims 56 to 80.

149. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs comprising: displaying, via the display generation component, a media user interface including a representation of the media; receiving, while displaying the media user interface including the representation of the media, a request to display additional information regarding a plurality of detected characteristics within the representation of the media; 1. A computer program product comprising: instructions for, in response to receiving the request to display additional information about the plurality of detected characteristics, displaying one or more indications of a plurality of detected characteristics in the media while displaying the media user interface including the representation of the media, the one or more indications of the plurality of detected characteristics including a first indication of a first detected characteristic displayed at a first location within the representation of the media, the first location corresponding to a location of the first detected characteristic within the representation of the media, and displaying the one or more indications includes: In response to determining that the first detected characteristic is a first type of characteristic, the first indication has a first appearance; In response to a determination that the first detected characteristic is a second type characteristic different from the first type characteristic, the first indication has a second appearance different from the first appearance. a computer program product,

150. 113. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, one or more input devices, and a display generation component, the one or more programs including instructions for the method of any one of claims 87 to 112.

151. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, a display generating component, and one or more input devices, the one or more programs comprising: receiving a request to display a representation of the field of view of the one or more cameras; In response to receiving the request to display the representation of the field of view of the one or more cameras, displaying, via the display generation component, the representation of the field of view of the one or more cameras, the representation including text that is within the field of view of the one or more cameras; automatically displaying, via the display generation component, a plurality of indications of the translated text, including a first indication of a translation of a first portion of the text and a second indication of a translation of a second portion of the text; receiving, via the one or more input devices, a request to select a respective indication of the plurality of translated portions while displaying the first indication and the second indication via the display generation component; and, in response to receiving the request to select the individual indication, in accordance with a determination that the request is a request to select the first indication, display, via the display generation component, a first translation user interface object that includes the first portion of the text and the translation of the first portion of the text without the translation of the second portion of the text.

152. 137. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with one or more cameras, a display generating component, and one or more input devices, the one or more programs including instructions for the method of any one of claims 119 to 136.

153. 1. A method comprising: a computer system in communication with a display generation component, While displaying a user interface that includes a representation of media, detecting a request to display additional information corresponding to the representation of the media; In response to detecting the request to display additional information corresponding to the representation of the media, displaying, via the display generation component, a first user interface object in accordance with determining that the detected text in the representation of media has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first action based on the detected text; displaying, via the display generation component, a second user interface object in accordance with determining that the detected text in the representation of media has a second set of properties that differs from the first set of properties, the second user interface object, when selected, causing the computer system to perform a second action based on the detected text, the second action being different from the first action; A method comprising:

154. 154. The method of claim 153, wherein the first user interface object is displayed simultaneously with the second user interface object.

155. 155. The method of claim 153 or 154, wherein the user interface includes a third user interface object, and wherein detecting the request to display additional information corresponding to the representation of the media includes detecting input directed at the third user interface object.

156. visually highlighting at least a first portion within the representation of the media relative to a second portion within the representation of the media in response to detecting the request to display additional information corresponding to the representation of the media; 156. The method of any one of claims 153 to 155, further comprising:

157. detecting, while displaying the first user interface object, a request to cease displaying additional information corresponding to the representation of the media; ceasing display of the first user interface object in response to detecting the request to cease displaying additional information corresponding to the representation of the media; 157. The method of any one of claims 153 to 156, further comprising:

158. 158. The method of any one of claims 153 to 157, wherein the first user interface object is a user interface object that copies a third portion of the representation of the media, and performing the first operation includes copying the third portion of the representation of the media.

159. 159. The method of any one of claims 153 to 158, wherein the first user interface object is a user interface object that initiates a communication session, and performing the first operation includes initiating a communication session with a second computer system associated with at least a first portion of the detected text.

160. 160. The method of any one of claims 153 to 159, wherein the first user interface object is a user interface object that converts a first value having a first unit of measurement to a second value having a second unit of measurement different from the first unit of measurement, and performing the first operation includes converting the first value having the first unit of measurement to the second value having the second unit of measurement.

161. The first user interface object is a user interface object that manages a first translation setting, and performing the first operation includes: configuring the computer system to operate in a translation mode in accordance with determining that the first translation setting is in a first state; configuring the computer system not to operate in the translation mode in accordance with determining that the first translation setting is in a second state different from the first state; 161. The method of any one of claims 153 to 160, comprising:

162. 162. The method of any one of claims 153 to 161, wherein the first user interface object is a user interface object that manages a second translation setting, and performing the first action includes ceasing to display a translated version of the second portion of the detected text.

163. 162. The method of any one of claims 153 to 161, wherein the first user interface object is a user interface object that manages a third translation setting, and performing the first operation includes displaying a translated version of a third portion of the detected text.

164. 164. The method of any one of claims 153 to 163, wherein the first user interface object is a user interface object that scans a fourth portion of the representation of the media, and wherein performing the first operation includes scanning the fourth portion of the representation of the media.

165. 165. The method of any one of claims 153 to 164, wherein the first user interface object is a user interface object that extracts one or more tables, a fifth portion of the media representation includes the first tables, and performing the first operation includes copying the first tables.

166. 166. The method of any one of claims 153 to 165, wherein the first user interface object is a user interface object that extracts information from one or more tables, a fifth portion of the media representation includes a second table, and performing the first operation includes displaying an indication that information in the second table has been selected.

167. 167. The method of any one of claims 153 to 166, wherein the first user interface object is a user interface object that manages one or more contacts, and performing the first operation includes adding a sixth portion of the representation of the media to a contact details form.

168. 168. The method of any one of claims 153 to 167, wherein the first user interface object is a user interface object for managing a shopping list, and performing the first operation includes adding a seventh portion of the representation of the media to the list.

169. the first user interface object is a user interface object for administering medication, and performing the first action includes: identifying medical information within the media representation; associating the medical information with a health-related application; 169. The method of any one of claims 153 to 168, comprising:

170. 170. The method of any one of claims 153 to 169, wherein the first user interface object is a user interface object for redeeming a gift card, and performing the first operation includes initiating a process for redeeming a gift card based on an eighth portion of the representation of the media, the eighth portion of the representation being identification information associated with the gift card.

171. 170. A method according to any one of claims 153 to 170, wherein the first user interface object is a user interface object for managing a barcode, and performing the first operation includes displaying first information about a product (and / or service) corresponding to the barcode, and the barcode is displayed within the representation of the media.

172. detecting an input directed at the barcode while displaying the representation of the media including the barcode; displaying second information about the product corresponding to the barcode in response to detecting the input directed at the barcode; and 172. The method of claim 171, further comprising:

173. the user interface includes a fourth user interface object displayed at a first location; detecting the request to display additional information corresponding to the media representation includes detecting input directed to the fourth user interface object; 170. The method of any one of claims 153 to 169, wherein, in response to detecting the request to display additional information corresponding to the representation of media, the first user interface object is displayed in the first location and the fourth user interface object is not displayed in the first location in accordance with a determination that the detected text in the representation of media has the first set of properties.

174. 174. The method of any one of claims 153 to 173, wherein the computer system is in communication with one or more cameras, and the representation of media is a representation of visual content being captured by the one or more cameras.

175. 175. The method of any one of claims 153 to 174, wherein the representation of the media is a representation of previously captured media.

176. 176. The method of any one of claims 153 to 175, wherein the media representation is a photograph or a video.

177. 177. The method of any one of claims 153 to 176, wherein the representation of the media is a screenshot.

178. 178. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs including instructions for performing the method of any one of claims 153 to 177.

179. 1. A computer system configured in communication with a display generation component, the computer system comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; 178. A computer system comprising: said one or more programs comprising instructions for performing the method of any one of claims 153 to 177.

180. 1. A computer system configured in communication with a display generation component, the computer system comprising:

178. A computer system comprising means for carrying out the method of any one of claims 153 to 177.

181. 178. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system configured to be in communication with a display generation component, the one or more programs including instructions for performing the method of any one of claims 153 to 177.

182. 1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of a computer system in communication with a display generation component, the one or more programs comprising: While displaying a user interface including a representation of media, detect a request to display additional information corresponding to the representation of the media, and in response to detecting the request to display additional information corresponding to the representation of the media, displaying, via the display generation component, a first user interface object in accordance with determining that the detected text in the representation of media has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first action based on the detected text; and displaying, via the display generation component, a second user interface object that, in accordance with determining that the detected text in the media representation has a second set of properties that is different from the first set of properties, when selected, causes the computer system to perform a second action based on the detected text, the second action being different from the first action. A non-transitory computer-readable storage medium containing instructions.

183. A computer system in communication with a display generation component, said computer system comprising: one or more processors; a memory storing one or more programs configured to be executed by the one or more processors; wherein the one or more programs: While displaying a user interface including a representation of media, detect a request to display additional information corresponding to the representation of the media; In response to detecting the request to display additional information corresponding to the representation of the media, displaying, via the display generation component, a first user interface object in accordance with determining that the detected text in the representation of media has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first action based on the detected text; and displaying, via the display generation component, a second user interface object that, in accordance with determining that the detected text in the media representation has a second set of properties that is different from the first set of properties, when selected, causes the computer system to perform a second action based on the detected text, the second action being different from the first action. A computer system including instructions.

184. a computer system in communication with a display generation component, means for detecting, while displaying a user interface including a representation of media, a request to display additional information corresponding to the representation of the media; In response to detecting the request to display additional information corresponding to the representation of the media, displaying, via the display generation component, a first user interface object in accordance with determining that the detected text in the representation of media has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first action based on the detected text; means for displaying, via the display generation component, a second user interface object in accordance with a determination that the detected text in the media representation has a second set of properties that differs from the first set of properties, the second user interface object, when selected, causing the computer system to perform a second action based on the detected text, the second action being different from the first action; A computer system comprising:

185. 1. A computer program product comprising one or more programs configured to be executed by one or more processors of a computer system in communication with a generation component, the one or more programs comprising: While displaying a user interface including a representation of media, detect a request to display additional information corresponding to the representation of the media, and in response to detecting the request to display additional information corresponding to the representation of the media, displaying, via the display generation component, a first user interface object in accordance with determining that the detected text in the representation of media has a first set of properties, the first user interface object, when selected, causing the computer system to perform a first action based on the detected text; and, in accordance with a determination that the detected text in the representation of the media has a second set of properties that differs from the first set of properties, display, via the display generation component, a second user interface object that, when selected, causes the computer system to perform a second action based on the detected text, the second action being different from the first action.

Citation Information

Patent Citations

  • Image processing device, image processing method, and map providing system

    JP2011137908A

  • Portable terminal device, program and display control method

    JP2013109687A

  • Plant monitoring device

    JP2015087778A

  • Methods of Displaying Information at Different Zoom Settings and Related Devices and Computer Program Products

    US20080252662A1

  • Method and portable electronic device for presenting text

    US20120089942A1