Display device and media content acquisition method

By displaying a graphic code on the user interface of the display device to obtain a link, using part of the image to generate media content in the target format and uploading it to the server, the problem of low efficiency in obtaining media content from the display device to the mobile terminal is solved, and efficient media content transmission and storage is achieved.

CN119356572BActive Publication Date: 2025-09-26HISENSE VISUAL TECH CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202411321901.0
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2024-09-20
Publication Date
2025-09-26
Estimated Expiration
2044-09-20

AI Technical Summary

Technical Problem

In the prior art, the acquisition efficiency of media content generated by a display device to a user's mobile terminal is low, requiring cumbersome device connections and a long data transmission time.

Method used

The display device displays a graphic code on the user interface, which includes a link to obtain the media content. It generates media content using part of the image in the target format and sends it to the server through a communication device to obtain the acquisition link. The user can scan and download the media content in the target format through a mobile terminal.

Benefits of technology

The data transmission volume is reduced, the integrity of the media content information is protected, and the display device is not limited by its own storage performance, thereby improving the efficiency of media content acquisition.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN119356572B_ABST
    Figure CN119356572B_ABST
Patent Text Reader

Abstract

The present application relates to a display device and a method for acquiring media content, and relates to the technical field of display devices. The display device includes: a display and a controller. The display is configured to display a user interface; the user interface includes at least a user interface of a target application; the controller is coupled to the display and is configured to: obtain an acquisition link for the media content in the target format when the generation of the media content in the video format is completed; in response to an acquisition instruction for the media content in the target format, display a graphic code in the user interface of the target application through the display; the graphic code includes an acquisition link for the media content in the target format; the target format is different from the video format; the media content in the target format includes some pictures from the media content in the video format; the part of the pictures is determined according to the text carrying form of the media content in the video format. The present application can reduce the amount of data transmitted when acquiring media content and protect the integrity of media content information.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the technical field of display devices, and in particular to a display device and a method for acquiring media content. Background Art

[0002] Display devices such as smart TVs can generate and play media content for users. The media content is presented in video format on the display device, and the generated media content can be stored on the display device.

[0003] In related technologies, if media content generated by a display device needs to be transferred to a user's mobile terminal or other device, cumbersome device connections or a long data transmission process are required, resulting in a technical problem of low efficiency in media content acquisition. Summary of the Invention

[0004] The present application provides a display device and a method for acquiring media content to solve the technical problem of low efficiency in acquiring media content.

[0005] In a first aspect, some embodiments provide a display device comprising: a display and a controller. The display is configured to display a user interface; the user interface includes at least a user interface of a target application; and the controller is coupled to the display and configured to:

[0006] When the media content in the video format is generated, obtaining an acquisition link for the media content in the target format;

[0007] In response to an instruction to obtain the media content in a target format, displaying a graphic code in a user interface of the target application via the display; the graphic code includes a link to obtain the media content in the target format;

[0008] The target format is different from the video format; the media content in the target format includes some pictures from the media content in the video format; and the some pictures are determined according to the text carrying form of the media content in the video format.

[0009] Technical effect: After the display device completes the generation of media content in video format and obtains media content in video format, it can convert the media content in video format into media content in target format, obtain an acquisition link for the media content in target format, and provide a graphic code containing the acquisition link for the media content in target format, so that users can scan and download the media content in target format through mobile terminals and other devices. The target format is different from the video format. The media content in the target format contains some pictures from the media content in video format. Therefore, the display device does not need to make all pictures of the media content into media content in target format, thereby reducing the amount of transmitted data, and the part of the pictures extracted by the display device is determined according to the text carrying form of the media content in video format to protect the integrity of the media content information.

[0010] In some embodiments of the present application, the display device further includes:

[0011] communication devices;

[0012] The controller is further configured to:

[0013] generating the media content in a target format according to the partial pictures of the media content in a video format;

[0014] sending the media content in the target format to a server via the communication device;

[0015] The acquisition link sent by the server is received through the communication device; the acquisition link is generated by the server according to the media content in the target format.

[0016] Technical effect: The display device can generate media content in a target format based on partial images of media content in a video format, thereby converting the media content in a video format into media content in a target format, and upload the media content in the target format to a server to obtain an acquisition link for the media content in the target format fed back by the server, so that the display device can display the acquisition link for the user to obtain the media content in the target format from the server based on the acquisition link through its mobile terminal or other device, so that the acquisition of the media content in the target format by the mobile terminal or other device does not need to be limited by the storage performance of the display device itself.

[0017] In some embodiments of the present application, the controller is further configured to: if the text carrying form is the first form, obtain the partial picture according to the target sampling period.

[0018] Technical effect: When the text is carried in the first form, the display device can extract images from media content in a video format through a certain sampling period, and form partial images efficiently and conveniently.

[0019] In some embodiments of the present application, the controller is further configured to: if the media content in video format does not have an audio track, perform text recognition on the image sequence of the media content in video format; if the text recognition result indicates that the media content in video format does not contain images with text, determine that the text carrying form is the first form.

[0020] Technical effect: It is possible to first determine whether the media content in video format has an audio track, and then perform text recognition on the image sequence of the media content in video format if there is no audio track. If no image with text is found in the media content, the text carrying form is identified as the first form, thereby accurately distinguishing between the two text carrying forms: the text having an independent audio track and the text having text but embedded in the video.

[0021] In some embodiments of the present application, the controller is further configured to:

[0022] If the text carrying form is the second form, the partial picture is obtained according to the picture similarity between two pictures of the media content in video format.

[0023] Technical effect: When the text carrying form is the second form, the display device can obtain a partial image based on the image similarity between two images, so that the partial image can be effectively extracted to form media content in the target format based on whether the two images are similar, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0024] In some embodiments of the present application, the controller is further configured to:

[0025] If the picture similarity between the first picture and the second picture of the two pictures is less than a threshold, the second picture is determined to be one of the partial pictures; and the first picture belongs to the partial pictures.

[0026] Technical effect: The display device can sequentially determine whether the image similarity of two images is less than a threshold value. If so, the second image can be determined as one of the partial images and the second image can be used as a basis to continue to determine until the last image in the image sequence of the media content. In this way, the extraction of partial images when the text carrying form is the second form can be accurately completed.

[0027] In some embodiments of the present application, the controller is further configured to:

[0028] If the media content in video format has an audio track, the text carrying format is determined to be the second format.

[0029] Technical effect: It is possible to first determine whether the media content in video format has an audio track. If it has an audio track, the text-carrying form can be directly identified as the second form, and quickly distinguished from the two text-carrying forms: one without text and the other with text embedded in the video.

[0030] In some embodiments of the present application, the controller is further configured to:

[0031] If the text carrying form is the third form, the partial picture is obtained based on the text consistency and picture similarity between two pictures of the media content in video format.

[0032] Technical effect: When the text carrying form is the third form, the display device can obtain a partial image based on the text consistency and image similarity between two images, so as to consider in turn whether the text between the two images is consistent and whether the two images are similar, accurately and effectively extract the partial image to form media content in the target format, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0033] In some embodiments of the present application, the controller is further configured to:

[0034] If the text consistency between the first picture and the second picture of the two pictures is inconsistent, the second picture is determined as one of the partial pictures; the first picture belongs to the partial pictures; if the text consistency is consistent, the partial picture is obtained according to the picture similarity between the first picture and the second picture.

[0035] Technical effect: The display device can perform a mixed verification of text and image similarity for two images. It can first determine whether the text of the first image is consistent with the second image. If not, the second image can also be determined as one of the partial images. If the text of the two is consistent, the partial image can be further obtained based on the similarity of the two images. In this way, when the text carrying form is the third form, it can be determined in turn based on the text consistency and image similarity whether to extract the image to form media content in the target format, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0036] In some embodiments of the present application, the controller is further configured to:

[0037] If the text consistency is consistent, and the image similarity between the first image and the second image is less than a threshold, the second image is determined as one of the partial images.

[0038] Technical effect: When the text carrying form is the third form, if the display device determines that the text of the first picture is consistent with that of the second picture, the display device can further determine the second picture as one of the partial pictures when the picture similarity between the first picture and the second picture is less than a threshold. In this way, even if two pictures correspond to the same text, they can be used to form media content in the target format when the pictures change to a certain extent, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0039] In some embodiments of the present application, the controller is further configured to:

[0040] If the media content in video format does not have an audio track, text recognition is performed on the image sequence of the media content in video format; if the text recognition result indicates that the media content in video format contains images with text, the text carrying form is determined to be the third form.

[0041] Technical effect: It is possible to first determine whether the media content in video format has an audio track, and then perform text recognition on the image sequence of the media content in video format if there is no audio track. At this time, as long as one of the images is recognized to have text, it can be determined that the text-carrying form is the third form, thereby being able to efficiently distinguish between the two text-carrying forms: the one without text and the one with text and an independent audio track.

[0042] In some embodiments of the present application, the target format includes a portable file format.

[0043] Technical effect: The display device can provide users with a link to obtain media content in a portable file format. The media content in a portable file format has strong compatibility and convenience, which facilitates the dissemination of media content.

[0044] In a second aspect, some embodiments further provide a method for acquiring media content, applied to a display device, wherein the display device includes a display and a controller; the controller is coupled to the display; the method includes:

[0045] When the generation of media content in a video format is completed, an acquisition link for the media content in a target format is obtained; in response to an acquisition instruction for the media content in the target format, a graphic code is displayed in the user interface of the target application via the display; the graphic code includes an acquisition link for the media content in the target format; wherein the target format is different from the video format; the media content in the target format includes partial images from the media content in the video format; the partial images are determined based on the text-carrying form of the media content in the video format.

[0046] Technical effect: After the media content in video format is generated, the media content in video format can be converted into media content in target format, and an acquisition link of the media content in target format can be obtained, and a graphic code containing the acquisition link of the media content in target format can be provided, so that users can scan and download the media content in target format through mobile terminals and other devices. The target format is different from the video format. The media content in the target format contains some pictures from the media content in video format. Therefore, the display device does not need to make all pictures of the media content into media content in target format, thereby reducing the amount of transmitted data, and the part of the pictures extracted by the display device is determined according to the text carrying form of the media content in video format to protect the integrity of the media content information. BRIEF DESCRIPTION OF THE DRAWINGS

[0047] In order to more clearly illustrate the technical solutions in the embodiments of the present application or related technologies, the following briefly introduces the drawings required for use in the embodiments of the present application or related technical descriptions. Obviously, the drawings described below are only some embodiments of the present application. For ordinary technicians in this field, other related drawings can be obtained based on these drawings without paying any creative work.

[0048] Figure 1 A schematic diagram of an operation scenario between a display device and a control device provided in some embodiments of the present application;

[0049] Figure 2 A schematic diagram of the hardware configuration of a display device provided in some embodiments of the present application;

[0050] Figure 3 A schematic diagram of the hardware configuration of a control device provided in some embodiments of the present application;

[0051] Figure 4 A schematic diagram of software configuration of a display device provided in some embodiments of the present application;

[0052] Figure 5 A flowchart of a method for acquiring media content provided in some embodiments of the present application;

[0053] Figure 6 A signaling interaction diagram of the picture book acquisition method provided in some embodiments of the present application. DETAILED DESCRIPTION

[0054] The following embodiments are described in detail, with examples illustrated in the accompanying drawings. When the following description refers to the drawings, identical numbers in different figures represent identical or similar elements unless otherwise indicated. The embodiments described in the following embodiments are not intended to represent all possible implementations consistent with the present application. They are merely examples of systems and methods consistent with certain aspects of the present application, as detailed in the claims.

[0055] It should be noted that the brief descriptions of terms in this application are only for the purpose of facilitating the understanding of the embodiments described below, and are not intended to limit the embodiments of this application. Unless otherwise specified, these terms should be understood according to their ordinary and usual meanings.

[0056] In the specification and claims of this application and the accompanying drawings, the terms "first," "second," "third," etc. are used to distinguish similar or similar objects or entities, and are not necessarily intended to limit a particular order or sequence, unless otherwise noted. It should be understood that the terms used in this manner are interchangeable under appropriate circumstances.

[0057] The terms "comprise," "include," and "have," and any variations thereof, are intended to cover but not exclude inclusion; for example, a product or device comprising a list of components is not necessarily limited to all the components expressly listed but may include other components not expressly listed or inherent to such product or device.

[0058] The term "module" refers to any known or later developed hardware, software, firmware, artificial intelligence, fuzzy logic, or combination of hardware and / or software code that is capable of performing the functionality associated with that element.

[0059] In the embodiments of the present application, the display device 200 generally refers to a device capable of displaying images and processing data. For example, the display device 200 includes but is not limited to a smart TV, a mobile terminal, a computer, a monitor, an advertising screen, a wearable device, a virtual reality device, an augmented reality device, etc.

[0060] Figure 1 This is a schematic diagram of an operation scenario between a display device and a control device provided in some embodiments of the present application. Figure 1 As shown in FIG, a user can operate the display device 200 through touch operation, the mobile terminal 300 and the control device 100. For example, the control device 100 can be a remote controller, a stylus pen, a handle, etc.

[0061] The mobile terminal 300 can function as a control device for performing human-computer interaction between a user and the display device 200. The mobile terminal 300 can also function as a communication device for establishing a communication connection with the display device 200 and exchanging data. In some embodiments, the mobile terminal 300 can install software applications with the display device 200, enabling connection and communication via a network communication protocol, enabling one-to-one control operations and data communication. Audio and video content displayed on the mobile terminal 300 can also be transmitted to the display device 200 for synchronized display.

[0062] like Figure 1As shown in FIG, the display device 200 also communicates data with the server 400 through various communication methods. The display device 200 may be allowed to communicate via a local area network (LAN), a wireless local area network (WLAN), and other networks.

[0063] The display device 200 may provide a broadcast receiving television function, and may also additionally provide an intelligent network television function with a computer support function, including but not limited to network television, smart TV, Internet Protocol television (IPTV), etc.

[0064] Figure 2 Some embodiments of this application provide Figure 1 2 is a block diagram of the hardware configuration of the display device 200.

[0065] In some embodiments, the display device 200 may include at least one of a tuner 210, a communication device 220, a detector 230, a device interface 240, a controller 250, a display 260, an audio output device 270, a memory, a power supply, and a user input interface.

[0066] In some embodiments, detector 230 is used to collect signals from the external environment or external interactions. For example, detector 230 may include a light receiver, such as a sensor for collecting ambient light intensity; or an image collector, such as a camera, for collecting external environmental scenes, user attributes, or user interaction gestures; or a sound collector, such as a microphone, for receiving external sounds.

[0067] In some embodiments, the display 260 includes a display component for presenting images and a driver component for driving image display. The display 260 is configured to receive image signals output from the controller 250 for display. For example, the display 260 can be used to display video content, image content, menu control interface components, and user control UI interfaces.

[0068] In some embodiments, the communication device 220 is a component used to communicate with an external device or server 400 according to various communication protocol types. The display device 200 can be provided with multiple communication devices 220 depending on the supported communication methods. For example, if the display device 200 supports wireless network communication, the display device 200 can be provided with a communication device 220 including WiFi functionality. If the display device 200 supports Bluetooth connection communication, the display device 200 needs to be provided with a communication device 220 including Bluetooth functionality.

[0069] The communication device 220 can establish a communication connection between the display device 200 and an external device or server 400 via a wireless or wired connection. A wired connection can connect the display device 200 to an external device via a data cable, an interface, or other components. A wireless connection can connect the display device 200 to an external device via a wireless signal or wireless network. The display device 200 can establish a connection with an external device directly or indirectly through a gateway, router, or connection device.

[0070] In some embodiments, the controller 250 may include at least one of a central processing unit (CPU), a video processor, an audio processor, a graphics processor, and a power processor, and first to nth interfaces for input / output. The controller 250 controls the operation of the display device and responds to user operations through various software control programs stored in a memory. The controller 250 controls the overall operation of the display device 200.

[0071] In some embodiments, the controller 250 and the tuner 210 may be located in different separate devices, that is, the tuner 210 may also be located in an external device of the main device where the controller 250 is located, such as an external set-top box.

[0072] In some embodiments, the user may input a user command through a graphical user interface (GUI) displayed on the display 260 , and the user input interface receives the user input command through the graphical user interface (GUI).

[0073] In some embodiments, the audio output device 270 may be a local speaker of the display device 200, or an external audio output device connected to the display device 200. For the external audio output device connected to the display device 200, the display device 200 may further be provided with an external audio output terminal, through which the audio output device may be connected to the display device 200 to output the sound of the display device 200.

[0074] In some embodiments, the user input interface 280 may be configured to receive instructions from a user.

[0075] Figure 3 Some embodiments of this application provide Figure 1 The hardware configuration diagram of the control device in the figure. Figure 3 As shown, the control device 100 may include: a controller 110, a communication interface 130, a user input / output interface, a memory, and a power supply.

[0076] The control device 100 is configured to control the display device 200 , and can receive user input operation instructions, and convert the operation instructions into instructions that the display device 200 can recognize and respond to, playing the role of an interactive intermediary between the user and the display device 200 .

[0077] In some embodiments, the control device 100 may be a smart device. For example, the control device 100 may be installed with various applications for controlling the display device 200 according to user needs.

[0078] In some embodiments, as Figure 1 As shown, the mobile terminal 300 or other intelligent electronic devices can play a similar function as the control device 100 after installing the application for controlling the display device 200 .

[0079] The controller 110 includes a processor 112, RAM 113, ROM 114, a communication interface 130, and a communication bus. The controller 110 is used to control the operation and operation of the control device 100, as well as the communication and cooperation between internal components and external and internal data processing functions.

[0080] Under the control of the controller 110, the communication interface 130 communicates control signals and data signals with the display device 200. The communication interface 130 may include at least one of a WiFi chip 131, a Bluetooth module 132, an NFC module 133, or other near field communication modules.

[0081] The user input / output interface 140 includes at least one of a microphone 141 , a touch panel 142 , a sensor 143 , a button 144 and other input interfaces.

[0082] In some embodiments, the control device 100 includes at least one of a communication interface 130 and an input / output interface 140. The control device 100 is configured with the communication interface 130, such as a WiFi, Bluetooth, or NFC module, to encode user input commands via the WiFi protocol, Bluetooth protocol, or NFC protocol and transmit them to the display device 200.

[0083] The memory 190 is used to store various operating programs, data and applications for driving and controlling the control device 100 under the control of the controller. The memory 190 can store various control signal instructions input by the user.

[0084] The power supply 180 is used to provide operating power support for each component of the control device 100 under the control of the controller.

[0085] To facilitate user interaction, in some embodiments, the display device 200 may run an operating system. The operating system is a computer program used to manage and control the hardware and software resources of the display device 200. The operating system may provide a user interface (to control the display device), allow the user to interact with the display device 200, and support the running of various application programs.

[0086] It should be noted that the operating system may be a native operating system based on a specific operating platform, or a third-party operating system deeply customized based on a specific operating platform, or an independent operating system specially developed for the display device.

[0087] The operating system can be divided into different modules or layers according to the functions implemented, e.g. Figure 4 As shown, in some embodiments, the system is divided into four layers, from top to bottom: the application layer (abbreviated as "application layer"), the application framework layer (abbreviated as "framework layer"), the system library layer and the kernel layer.

[0088] In some embodiments, the application layer provides services and interfaces for applications, enabling the display device 200 to run applications and interact with the user based on these applications. The application layer can host at least one application, which can include built-in window programs, system settings programs, clock programs, and other applications provided by the operating system, or applications developed by third-party developers. In specific implementations, the application packages in the application layer are not limited to the examples above.

[0089] The framework layer provides applications with an application programming interface (API) and programming framework. The application framework layer includes predefined functions. The application framework layer acts as a processing center, determining the actions taken by applications in the application layer. Through the API, applications can access system resources and services during execution.

[0090] like Figure 4As shown, in the embodiment of the present application, the application framework layer includes a view system, managers, content providers, etc., wherein the view system can design and implement the interface and interaction of the application, and the view system includes lists, grids, text boxes, buttons, etc. The manager includes at least one of the following modules: an activity manager for interacting with all activities running in the system; a location manager for providing system services or applications with access to the system location service; a package manager for retrieving various information related to the application packages currently installed on the device; a notification manager for controlling the display and clearing of notification messages; and a window manager for managing icons, windows, toolbars, wallpapers, and desktop widgets on the user interface.

[0091] In some embodiments, the activity manager is used to manage the lifecycle of each application and common navigation back functions, such as controlling application exit, opening, and back. The window manager is used to manage all window programs, such as obtaining the display screen size, determining whether there is a status bar, locking the screen, taking screenshots, and controlling changes in display windows, such as shrinking, shaking, or distorting the display window.

[0092] In some embodiments, the system runtime layer can provide support for the framework layer. When the framework layer is used, the operating system will run the instruction library contained in the system runtime layer, such as the C / C++ instruction library, to implement the functions to be implemented by the framework layer.

[0093] In some embodiments, the kernel layer is a functional layer between the hardware and software of the display device 200. The kernel layer can implement functions such as hardware abstraction, multitasking, and memory management. Figure 4 As shown, the kernel layer can be configured with hardware drivers, and the drivers included in the kernel layer can be at least one of the following drivers: audio driver, display driver, Bluetooth driver, camera driver, WIFI driver, USB driver, HDMI driver, sensor driver (such as fingerprint sensor, temperature sensor, pressure sensor, etc.), and power driver, etc.

[0094] It should be noted that the above example is only a simple division of the operating system functions and does not constitute a limitation on the specific operating system form of the display device 200 in the embodiment of the present application. Depending on factors such as the function of the display device and the type of operating system, the number of levels and specific level types contained in the operating system may be expressed in other forms.

[0095] In some embodiments, as Figure 1 and 2 As shown, the present application provides a display device 200, which may include a display 260 and a controller 250, and the controller 250 is coupled to the display 260; wherein the display 260 is configured to display a user interface; the user interface at least includes a user interface of a target application.

[0096] like Figure 5 As shown, the controller 250 is configured to:

[0097] Step 501: When the media content in the video format is generated, an acquisition link for the media content in the target format is acquired.

[0098] Step 502: In response to an instruction to obtain media content in a target format, a graphic code is displayed in a user interface of a target application via a display; the graphic code includes a link to obtain the media content in the target format; wherein the target format is different from the video format; the media content in the target format includes some pictures from the media content in the video format; and the some pictures are determined based on the text-carrying form of the media content in the video format.

[0099] The target application can be a picture book application, which can include multiple pre-set picture books and support user creation of new ones. Each picture book contains multiple picture images, with consistent character imagery across all images. Each picture image is generated based on a section of the story within the picture book, and each picture image is accompanied by an audio narration based on the story text associated with that picture image. Selecting a picture book in the picture book application and confirming playback will enable playback of the selected picture book in video mode. During video playback, the picture book images can be displayed statically or with a sliding effect. The story text associated with the page is also displayed over the picture image, and the duration of each picture image's display can correspond to the duration of the audio narration associated with that picture image. As an example, the switching between two adjacent picture images can be continuous, with a one-second interval between the audio narrations. Background music can be played during playback, matching the story content of the picture book.

[0100] The produced media content can be played through the display 260 of the display device 200, and the produced media content is presented in a video format on the display device 200. After the media content in the video format is produced, the controller 250 can obtain the media content in the video format and play it through the display 260.

[0101] The display device 200 can also provide a service for converting media content in a video format into a target format. The target format can be a format that allows users to view the media content on other devices (such as the mobile terminal 300), making it easier for users to view the media content on other devices. This target format is different from the video format. For example, it can be a Portable Document Format (PDF) or another format.

[0102] Once the video-formatted media content is generated, the controller 250 can obtain the video-formatted media content and a link to obtain the target-formatted media content. This link can be used by other devices (e.g., mobile terminal 300) to obtain the target-formatted media content. After obtaining the video-formatted media content, the controller 250 can also display an access point for obtaining the target-formatted media content via the user interface of the target application displayed on the display 260. This access point can be triggered by the user to provide the user with a service to obtain the target-formatted media content.

[0103] The user can press a button on the control device 100 to trigger the acquisition entry displayed on the user interface of the target application displayed on the display 260 to send an acquisition instruction for the media content in the target format to the controller 250 of the display device 200 .

[0104] In response to the instruction to retrieve the media content in the target format, the controller 250 of the display device 200 displays a graphical code containing a link to retrieve the media content in the target format in the user interface of the target application via the display 260. The graphical code may be a QR code, for example. As an example, the retrieval link may be displayed in the user interface of the target application on the display 260 in the form of a QR code, and the user may scan the retrieval link in the form of a QR code, for example, via the mobile terminal 300, to retrieve the media content in the target format. The media content in the target format includes some images from the media content in the video format. This means that it is not necessary to include all images from the media content in the video format in the media content in the target format, as this would result in an excessively large amount of data to be transmitted and would include a large number of duplicate images. For a picture book, the images may be video frames from the picture book in the video format. Thus, the images may be determined based on the text-carrying format of the media content in the video format. The text-carrying format refers to the text-carrying format of the media content in the video format, and may include, for example, whether the text is present and the form in which the text is present in the video format. For a picture book, the text may be subtitles, and the text-carrying format may be subtitles. Thus, the controller 250 can determine which pictures to extract and produce into the media content in the target format according to the text-carrying form of the media content in the video format, thereby achieving the purpose of reducing the amount of data to be transmitted and protecting the integrity of the media content information.

[0105] According to the technical solution of this embodiment, after the display device completes the generation of the media content in the video format and obtains the media content in the video format, it can convert the media content in the video format into the media content in the target format, obtain the acquisition link of the media content in the target format, and provide a graphic code containing the acquisition link of the media content in the target format, so that the user can scan and download the media content in the target format through a mobile terminal or other device. The target format is different from the video format. The media content in the target format contains some pictures from the media content in the video format. Therefore, the display device does not need to make all the pictures of the media content into media content in the target format, thereby reducing the amount of transmitted data, and the part of the pictures extracted by the display device is determined according to the text carrying form of the media content in the video format to protect the integrity of the media content information.

[0106] In some embodiments, as Figure 2 As shown, the display device 200 may further include: a communication device 220. The controller 250 is further configured to:

[0107] Generate media content in a target format based on partial images of media content in a video format; send the media content in the target format to a server via a communication device; receive an acquisition link sent by the server via the communication device; the acquisition link is generated by the server based on the media content in the target format.

[0108] In this embodiment, after determining a portion of the image of the media content in the video format, the controller 250 may generate the media content in the target format based on the portion of the image. The controller 250 may then send the media content in the target format to the server 400 via the communication device 220 for storage. After receiving the media content in the target format, the server 400 generates a corresponding acquisition link, which the server 400 returns to the display device 200. The communication device 220 of the display device 200 may receive the acquisition link sent by the server 400 and pass it to the controller 250. The controller 250 may then display the acquisition link in the form of a graphic code, such as a QR code, on the user interface of the target application on the display 260, allowing the user to scan the QR code through, for example, a mobile terminal 300. The server 400 then provides a download service for the media content in the target format, and the user downloads and obtains the media content in the target format through, for example, the mobile terminal 300.

[0109] According to the solution of this embodiment, the display device 200 can generate media content in a target format based on a partial image of media content in a video format, thereby converting the media content in a video format into media content in a target format, and uploading the media content in the target format to the server 400 to obtain an acquisition link for the media content in the target format fed back by the server 400, so that the display device 200 can display the acquisition link for the user to obtain the media content in the target format from the server 400 based on the acquisition link through its mobile terminal 300 or other devices, so that the acquisition of the media content in the target format by the mobile terminal 300 or other devices does not need to be limited by the storage performance of the display device 200 itself.

[0110] In some embodiments, the controller 250 is further configured to:

[0111] If the text carrying format is the first format, some images are acquired according to the target sampling period.

[0112] In this embodiment, the first form can be one of the multiple text carrying forms preset in the display device 200, or it can be other text carrying forms that are not preset in the display device 200. The first form may include a form without text. In this regard, the controller 250 of the display device 200 can extract the part of the picture from the media content in the video format according to the target sampling period. The target sampling period can be selected by the user or it can be a sampling period set by the display device 200. The controller 250 can extract the part of the pictures according to the target sampling period and insert them into the media content in the target format according to the order of the pictures in the media content.

[0113] In the solution of this embodiment, the display device 200 can extract pictures from media content in a video format through a certain sampling period when the text carrying form is the first form, and form a partial picture efficiently and conveniently.

[0114] In some embodiments, the controller 250 is further configured to:

[0115] If the media content in the video format does not have an audio track, text recognition is performed on the image sequence of the media content in the video format; if the text recognition result indicates that the media content in the video format does not contain images with text, the text carrying form is determined to be the first form.

[0116] The audio track may be an independent audio track in the media content in the video format. In this embodiment, the controller 250 may first determine whether the media content in the video format has an audio track. If the media content in the video format does not have an audio track, the controller 250 may perform text recognition on the image sequence of the media content in the video format, and may use OCR (Optical Character Recognition) to traverse each image in the image sequence of the media content in the video format. As an example, in the Android system, tools such as Firebase ML Kit or Tesseract OCR may be used to implement the OCR function. The controller 250 may thereby obtain a text recognition result. If the text recognition result indicates that the media content in the video format does not contain an image with text, the controller 250 may determine that the text carrying form of the media content in the video format is the first form, that is, the text carrying form of the media content in the video format may be a form without text.

[0117] The solution of this embodiment is: first, it is possible to determine whether the media content in video format has an audio track, and if there is no audio track, perform text recognition on the image sequence of the media content in video format. If no image with text is found in the media content, the text carrying form is identified as the first form, thereby accurately distinguishing between the two text carrying forms: the text having an independent audio track and the text having text but embedded in the video.

[0118] In some embodiments, the controller 250 is further configured to:

[0119] If the text carrying form is the second form, a partial picture is obtained based on the picture similarity between two pictures of the media content in the video format.

[0120] In this embodiment, the second form and the aforementioned first form are different text carrying forms. The second form can be one of the multiple text carrying forms preset in the display device 200. The second form can be a text carrying form that contains text and has an independent audio track. In this regard, the controller 250 of the display device 200 can obtain part of the image based on the image similarity between two images of the media content in the video format. As an example, the image similarity can be calculated using the mean squared error (MSE) or structural similarity index (SSIM) between the two images. In this way, the controller 250 can sequentially add images with changed image content to the media content in the target format based on the judgment of the image similarity between the two images, and can skip images with unchanged image content until the last frame is obtained to form the final media content in the target format.

[0121] In the solution of this embodiment, the display device 200 can obtain a partial image based on the image similarity between two images when the text carrying form is the second form, so that the partial image can be effectively extracted to form media content in the target format based on whether the two images are similar, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0122] In some embodiments, the controller 250 is further configured to:

[0123] If the image similarity between the first image and the second image in the two images is less than a threshold, the second image is determined to be one of the partial images; and the first image belongs to the partial image.

[0124] In this embodiment, the two images may include a first image and a second image, where the first image may be the preceding image and the second image may be the following image. The first image may be an image that has been determined to be included in the media content of the target format, i.e., the first image is a partial image. To this end, the image similarity between the first and second images may be determined. If the image similarity between the first and second images is less than a threshold, it indicates that the image content of the second image has changed relative to the first image. The second image can then be determined as one of the partial images. The second image can then be used as a basis for further determination until the last frame is retrieved to form the final media content of the target format. If the image similarity between the first and second images is greater than or equal to the threshold, it indicates that the image content of the second image may not have changed relative to the first image. In this case, the second image does not need to be considered a partial image. The second image can then be used as a basis for further determination until the last frame is retrieved to form the final media content of the target format.

[0125] In the solution of this embodiment, the display device 200 can sequentially determine whether the image similarity of two images is less than a threshold value. If so, the second image can be determined as one of the partial images and the second image can be used as a reference to continue the subsequent determination until the last image in the image sequence of the media content, thereby accurately completing the extraction of partial images when the text carrying form is the second form.

[0126] In some embodiments, the controller 250 is further configured to:

[0127] If the media content in the video format has an audio track, the text carrying format is determined to be the second format.

[0128] Wherein, the audio track can be an independent audio track in the media content of the video format. In the present embodiment, the controller 250 can first determine whether the media content of the video format has an audio track. If the media content of the video format has an audio track, the controller 250 can determine that the text carrying form of the media content of the video format is the second form.

[0129] In the solution of this embodiment, the controller 250 can first determine whether the media content in video format has an audio track. If it has an audio track, it can directly identify the text-carrying form as the second form, and quickly distinguish it from the two text-carrying forms: one without text and the other with text but embedded in the video.

[0130] In some embodiments, the controller 250 is further configured to:

[0131] If the text carrying form is the third form, partial images are obtained based on text consistency and image similarity between two images of the media content in the video format.

[0132] In this embodiment, the third form is a different text-carrying form from the first and second forms described above. The third form can be one of multiple text-carrying forms preset in the display device 200. The third form can also be a text-carrying form with text embedded in the video. To this end, the controller 250 of the display device 200 can obtain a portion of the image based on the text consistency and image similarity between two images of the media content in the video format. That is, unlike the second form, the third form also adds a judgment on text consistency, that is, a hybrid verification of text consistency and image similarity is used to obtain a portion of the image used to form the media content in the target format. Text consistency can be used to indicate whether the subtitles of the two images are consistent, and can include consistency and inconsistency. Therefore, in the third form, the controller 250 of the display device 200 can sequentially consider the text consistency and image similarity between the two images to determine the image to be added to the media content in the target format, thereby protecting the integrity of the media content information as much as possible.

[0133] In the solution of this embodiment, the display device 200 can obtain a partial image based on the text consistency and image similarity between two images when the text carrying form is the third form, so as to consider in turn whether the text between the two images is consistent and whether the two images are similar, and accurately and effectively extract the partial image to form media content in the target format, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0134] In some embodiments, the controller 250 is further configured to:

[0135] If the text consistency between the first and second images of the two images is inconsistent, the second image is determined as one of the partial images; the first image belongs to the partial image; if the text consistency is consistent, the partial image is obtained according to the image similarity between the first and second images.

[0136] In this embodiment, the two pictures may include a first picture and a second picture, the first picture may be the previous picture, and the second picture may be the next picture. The first picture may be a picture that has been determined to be added to the media content of the target format, that is, the first picture is a partial picture. In this regard, the controller 250 may first determine the text consistency of the first picture and the second picture. If the text consistency of the first picture and the second picture is inconsistent, the controller 250 needs to determine the second picture as one of the partial pictures to be added to the media content of the target format. If the text consistency of the first picture and the second picture is consistent, the controller 250 needs to continue to determine the image similarity of the first picture and the second picture, and determine whether to determine the second picture as one of the partial pictures based on the image similarity of the two pictures. Thus, the controller 250 can add the picture with changed image content to the media content of the target format based on the judgment of the image similarity of the two pictures, if the text consistency of the two pictures is consistent, and can skip the picture with unchanged image content. As an example, the image similarity may be calculated using a mean squared error (MSE) or a structural similarity index (SSIM) between two images.

[0137] In the solution of this embodiment, the display device 200 can perform a mixed verification of text and image similarity for two images. It can first determine whether the text of the first image is consistent with the second image. If not, the second image can also be determined as one of the partial images. If the text of the two images is consistent, the partial image can be further obtained based on the similarity of the two images. In this way, when the text carrying form is the third form, it can be determined in turn based on the text consistency and image similarity whether to extract the image to form the media content of the target format, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0138] In some embodiments, the controller 250 is further configured to:

[0139] If the text consistency is consistent and the image similarity between the first image and the second image is less than a threshold, the second image is determined as one of the partial images.

[0140] In this embodiment, when the text carrying form is the third form, if the controller 250 determines that the text consistency of the first picture and the second picture is consistent, it can continue to determine whether the image similarity between the first picture and the second picture is less than the threshold value. If the text consistency of the first picture and the second picture is consistent and the image similarity is less than the threshold value, the controller 250 can determine the second picture as one of the partial pictures. Among them, the first picture as mentioned above can be a picture of the media content that has been determined to be added to the target format, that is, the first picture belongs to the partial picture. Among them, when the text consistency of the first picture and the second picture is consistent, if the image similarity between the first picture and the second picture is less than the threshold value, it means that the image content of the second picture has changed relative to the first picture. The controller 250 can also determine the second picture as one of the partial pictures, and then continue to judge backward based on the second picture, and judge including the judgment of text consistency and image similarity until the last frame is taken to form the final media content of the target format. Among them, when the text consistency of the first picture and the second picture is consistent, if the image similarity between the first picture and the second picture is greater than or equal to the threshold, it means that the image content of the second picture may not have changed relative to the first picture, then there is no need to regard the second picture as one of the partial pictures, and then the backward judgment can still be continued based on the second picture, and the judgment also includes the judgment of text consistency and image similarity, until the last frame is taken to form the final media content in the target format.

[0141] In the solution of this embodiment, when the text carrying form is the third form, if the display device 200 determines that the text of the first picture is consistent with that of the second picture, the display device 200 can further determine the second picture as one of the partial pictures when the picture similarity between the first picture and the second picture is less than a threshold. In this way, even if two pictures correspond to the same text, they can be used to form media content in the target format when the pictures change to a certain extent, so as to protect the integrity of the media content information while reducing the amount of transmitted data.

[0142] In some embodiments, the controller 250 is further configured to:

[0143] If the media content in the video format does not have an audio track, text recognition is performed on the image sequence of the media content in the video format; if the text recognition result indicates that the media content in the video format contains images with text, the text carrying form is determined to be the third form.

[0144] The audio track can be an independent audio track in the video-formatted media content. In this embodiment, the controller 250 can first determine whether the video-formatted media content has an audio track. If the video-formatted media content does not have an audio track, the controller 250 can perform text recognition on the image sequence of the video-formatted media content. Optical Character Recognition (OCR) can be used to traverse each image in the image sequence of the video-formatted media content. As an example, in the Android system, tools such as Firebase ML Kit or Tesseract OCR can be used to implement OCR functionality. In this way, text recognition results can be continuously obtained each time a picture is identified. If the text recognition result indicates that the media content in the video format contains pictures with text, the controller 250 can determine that the text carrying form of the media content in the video format is the third form, that is, the text carrying form of the media content in the video format can be a text carrying form with text but the text is embedded in the video, that is, in the recognition process of traversing the picture sequence of the media content in the video format, as long as a certain picture is recognized as having text, the recognition process can be exited, and it can be determined that the text carrying form of the media content in the video format is the third form.

[0145] The solution of this embodiment can first determine whether the media content in video format has an audio track, and then perform text recognition on the image sequence of the media content in video format if there is no audio track. At this time, as long as one of the images is recognized to have text, it can be determined that the text carrying form is the third form, thereby being able to efficiently distinguish between the two text carrying forms: without text and with text and the text with an independent audio track.

[0146] In some embodiments, the target format comprises a portable document format.

[0147] In this embodiment, the controller 250 of the display device 200 can convert media content in video format into media content in portable file format, and then the controller 250 can send the media content in portable file format to the server 400 for storage. After receiving the media content in portable file format, the server 400 generates a corresponding acquisition link, and the server 400 returns the acquisition link to the display device 200. The controller 250 can display the acquisition link in the form of a QR code on the user interface of the target application of the display 260, so that the user can scan the QR code through its mobile terminal 300, for example. The server 400 provides a download service for the media content in portable file format, and the user downloads and obtains the media content in portable file format through its mobile terminal 300, for example.

[0148] In the solution of this embodiment, the display device 200 can provide users with a link to obtain media content in a portable file format. The media content in a portable file format has high compatibility and convenience, and is easy to spread.

[0149] In some embodiments, as Figure 5 As shown, a method for obtaining media content is provided, which can be applied to Figure 1 The display device 200 shown in FIG. Figure 2 As shown, the display device 200 may include a display 260 and a controller 250; the controller 250 is coupled to the display 260; the method may include the following steps:

[0150] Step 501: When the media content in the video format is generated, obtain a link for obtaining the media content in the target format.

[0151] Step 502: In response to an instruction to obtain the media content in the target format, a graphic code is displayed in the user interface of the target application via the display; the graphic code includes a link to obtain the media content in the target format; wherein the target format is different from the video format; the media content in the target format includes a partial picture from the media content in the video format; the partial picture is determined based on the text-carrying form of the media content in the video format.

[0152] According to the solution of this embodiment, after the media content in the video format is generated, the media content in the video format can be converted into media content in the target format, and an acquisition link for the media content in the target format can be obtained, and a graphic code containing the acquisition link for the media content in the target format can be provided, so that the user can scan and download the media content in the target format through a mobile terminal or other device. The target format is different from the video format. The media content in the target format contains some pictures from the media content in the video format. Therefore, the display device does not need to convert all pictures of the media content into media content in the target format, thereby reducing the amount of transmitted data. In addition, the part of the pictures extracted by the display device is determined according to the text carrying form of the media content in the video format to protect the integrity of the media content information.

[0153] In some embodiments, the above method may further include the following steps:

[0154] Generate the media content in a target format based on the partial picture of the media content in a video format; send the media content in the target format to a server via a communication device; receive the acquisition link sent by the server via the communication device; the acquisition link is generated by the server based on the media content in the target format.

[0155] In some embodiments, the above method may further include the following steps:

[0156] If the text carrying form is the first form, the partial picture is obtained according to the target sampling period.

[0157] In some embodiments, the above method may further include the following steps:

[0158] If the media content in video format does not have an audio track, text recognition is performed on the image sequence of the media content in video format; if the text recognition result indicates that the media content in video format does not contain images with text, it is determined that the text carrying form is the first form.

[0159] In some embodiments, the above method may further include the following steps:

[0160] If the text carrying form is the second form, the partial picture is obtained according to the picture similarity between two pictures of the media content in video format.

[0161] In some embodiments, the above method may further include the following steps:

[0162] If the picture similarity between the first picture and the second picture of the two pictures is less than a threshold, the second picture is determined to be one of the partial pictures; and the first picture belongs to the partial pictures.

[0163] In some embodiments, the above method may further include the following steps:

[0164] If the media content in video format has an audio track, the text carrying format is determined to be the second format.

[0165] In some embodiments, the above method may further include the following steps:

[0166] If the text carrying form is the third form, the partial picture is obtained based on the text consistency and picture similarity between two pictures of the media content in video format.

[0167] In some embodiments, the above method may further include the following steps:

[0168] If the text consistency between the first picture and the second picture of the two pictures is inconsistent, the second picture is determined as one of the partial pictures; the first picture belongs to the partial pictures; if the text consistency is consistent, the partial picture is obtained according to the picture similarity between the first picture and the second picture.

[0169] In some embodiments, the above method may further include the following steps:

[0170] If the text consistency is consistent, and the image similarity between the first image and the second image is less than a threshold, the second image is determined as one of the partial images.

[0171] In some embodiments, the above method may further include the following steps:

[0172] If the media content in video format does not have an audio track, text recognition is performed on the image sequence of the media content in video format; if the text recognition result indicates that the media content in video format contains images with text, the text carrying form is determined to be the third form.

[0173] In some embodiments, the target format comprises a portable document format.

[0174] In some embodiments, as Figure 6 As shown, a method for obtaining a picture book is provided.

[0175] In this embodiment, the target application can be a picture book application. Multiple picture books can be pre-set within the picture book application, and users can also create new picture books. Each picture book includes multiple picture book images, and the character images in multiple picture book images are uniform. Each picture book image is generated based on a section of the story text in the picture book story. Each picture book image is accompanied by an audio narration, which is derived from the story text corresponding to the picture book image. Select a picture book in the picture book application and confirm playback. The selected picture book can be played in video mode. During picture book video playback, the picture book images can be displayed statically or with a sliding effect. The story text corresponding to the page image is also displayed on the picture book image, and the display duration of each picture book image can correspond to the duration of the audio narration for that picture book image. As an example, the switching display of two adjacent picture book images can be continuous, and the audio narrations of the two adjacent picture book images can be played with an interval of, for example, 1 second. Background music can be played during picture book playback, and the background music matches the story content of the picture book.

[0176] The picture book acquisition method may include the following steps:

[0177] Step 601: When the picture book in video format is generated, obtain the picture book in video format.

[0178] Step 602: Send an instruction to obtain the picture book in a portable file format to the display device.

[0179] Step 603 : In response to the acquisition instruction, a portion of the video frames is determined according to the subtitle form of the picture book in the video format, and a picture book in a portable file format is generated according to the portion of the video frames.

[0180] Step 604: Send the picture book in a portable file format to the server.

[0181] Step 605: Send an acquisition link for the picture book in a portable file format to the display device.

[0182] Step 606: Display the acquisition link in the form of a QR code.

[0183] Step 607: Scan the QR code to obtain the link.

[0184] Step 608: Send an acquisition request to the server according to the acquisition link.

[0185] Step 609: In response to the acquisition request, the picture book in a portable file format is sent to the mobile terminal.

[0186] According to the technical solution of this embodiment, after the display device completes the generation of the picture book in video format and obtains the picture book in video format, it can convert the picture book in video format into a picture book in portable file format and upload the picture book in portable file format to the server to obtain the acquisition link for the picture book in portable file format fed back by the server, so that the display device can display the acquisition link in the form of a QR code for the user to scan through its mobile terminal to obtain the acquisition link, and obtain the picture book in portable file format from the server based on the acquisition link, wherein the picture book in portable file format contains some video frames from the picture book in video format, and the display device does not need to make all video frames into the picture book in portable file format, thereby reducing the amount of transmitted data, and the part of the video frames extracted by the display device can be determined according to the subtitle form of the picture book in video format to protect the integrity of the picture book information and facilitate the dissemination of the picture book.

[0187] It should be understood that, although the various steps in the flowcharts involved in the various embodiments described above are displayed in sequence according to the instructions of the arrows, these steps are not necessarily executed in sequence in the order indicated by the arrows. Unless otherwise specified herein, there is no strict order restriction on the execution of these steps, and these steps can be executed in other orders. Moreover, at least a portion of the steps in the flowcharts involved in the various embodiments described above can include multiple steps or multiple stages, and these steps or stages are not necessarily executed and completed at the same time, but can be executed at different times, and the execution order of these steps or stages is not necessarily to be carried out in sequence, but can be executed in turn or alternately with other steps or at least a portion of steps or stages in other steps.

[0188] Based on the same inventive concept, embodiments of the present application further provide a media content acquisition device for implementing the aforementioned media content acquisition method. The implementation solution provided by this device is similar to the implementation solution described in the aforementioned method. Therefore, the specific limitations of one or more media content acquisition device embodiments provided below can be found in the above-mentioned limitations of the media content acquisition method and will not be repeated here.

[0189] In some embodiments, a media content acquisition apparatus is provided, which is applied to a display device. The display device includes a display and a controller. The controller is coupled to the display. The media content acquisition apparatus may include:

[0190] An acquisition module, configured to acquire an acquisition link for the media content in a target format when the media content in the video format is generated;

[0191] a display module, configured to display a graphic code in a user interface of the target application via the display in response to an instruction to obtain the media content in the target format; the graphic code including a link for obtaining the media content in the target format;

[0192] The target format is different from the video format; the media content in the target format includes some pictures from the media content in the video format; and the some pictures are determined according to the text carrying form of the media content in the video format.

[0193] Each module in the above-mentioned media content acquisition device can be implemented in whole or in part through software, hardware, or a combination thereof. Each module can be embedded in or independent of the processor in the display device in hardware form, or can be stored in the display device in software form so that the processor can call and execute the corresponding operations of each module.

[0194] In some embodiments, a computer-readable storage medium is provided, on which a computer program is stored. When the computer program is executed by a processor, the steps in the above-mentioned method embodiments are implemented.

[0195] In some embodiments, a computer program product is provided, including a computer program, which implements the steps in the above method embodiments when executed by a processor.

[0196] It should be noted that the user information (including but not limited to user device information, user personal information, etc.) and data (including but not limited to data used for analysis, stored data, displayed data, etc.) involved in this application are all information and data authorized by the user or fully authorized by all parties, and the collection, use and processing of relevant data must comply with relevant regulations.

[0197] Those skilled in the art will understand that all or part of the processes in the above-mentioned embodiments can be implemented by instructing the relevant hardware through a computer program. The computer program can be stored in a non-volatile computer-readable storage medium. When the computer program is executed, it can include the processes of the embodiments of the above-mentioned methods. In particular, any reference to memory, database, or other media used in the embodiments provided in this application can include at least one of non-volatile memory and volatile memory. Non-volatile memory can include read-only memory (ROM), magnetic tape, floppy disk, flash memory, optical memory, high-density embedded non-volatile memory, resistive random access memory (ReRAM), magnetic random access memory (MRAM), ferroelectric random access memory (FRAM), phase change memory (PCM), graphene memory, etc. Volatile memory can include random access memory (RAM) or external cache memory, etc. By way of illustration and not limitation, RAM can take various forms, such as static random access memory (SRAM) or dynamic random access memory (DRAM). The databases involved in the various embodiments provided herein may include at least one of a relational database and a non-relational database. Non-relational databases may include, but are not limited to, blockchain-based distributed databases. The processors involved in the various embodiments provided herein may be, but are not limited to, general-purpose processors, central processing units (CPUs), graphics processing units (GPUs), digital signal processors (DSPs), programmable logic devices (PLDs), quantum computing-based data processing logic devices, artificial intelligence (AI) processors, and the like.

[0198] The technical features of the above embodiments can be combined arbitrarily. In order to make the description concise, not all possible combinations of the technical features in the above embodiments are described. However, as long as there is no contradiction in the combination of these technical features, they should be considered to be within the scope of this application.

[0199] The above-described embodiments merely represent several implementation methods of the present application. While the descriptions are relatively specific and detailed, they should not be construed as limiting the scope of the present application. It should be noted that a person of ordinary skill in the art may make various modifications and improvements without departing from the spirit of the present application, and these modifications and improvements fall within the scope of protection of the present application. Therefore, the scope of protection of the present application shall be determined by the appended claims.

Claims

1. A display device, characterized in that: include: A display configured to display a user interface; the user interface includes at least a user interface of a target application; a controller coupled to the display and configured to: When the media content in the video format is generated, obtaining an acquisition link for the media content in the target format; In response to an instruction to obtain the media content in a target format, displaying a graphic code in a user interface of the target application via the display; The graphic code includes an acquisition link for the media content in a target format; The target format is different from the video format; the media content in the target format includes some pictures from the media content in the video format; the some pictures are determined according to the text-carrying form of the media content in the video format; Wherein, the controller is further configured to: If the media content in video format does not have an audio track, text recognition is performed on the image sequence of the media content in video format; if the text recognition result indicates that the media content in video format does not contain images with text, the text carrying form is determined to be the first form; if the text recognition result indicates that the media content in video format contains images with text, the text carrying form is determined to be the third form; If the media content in video format has an audio track, the text carrying format is determined to be the second format.

2. The display device according to claim 1, wherein Also includes: communication devices; The controller is further configured to: generating the media content in a target format according to the partial pictures of the media content in a video format; sending the media content in the target format to a server via the communication device; Receiving the acquisition link sent by the server through the communication device; The acquisition link is generated by the server according to the media content in the target format.

3. The display device according to claim 1, wherein The controller is further configured to: If the text carrying form is the first form, the partial picture is obtained according to the target sampling period.

4. The display device according to claim 1, wherein The controller is further configured to: If the text carrying form is the second form, the partial picture is obtained according to the picture similarity between two pictures of the media content in video format.

5. The display device according to claim 4, wherein: The controller is further configured to: If the picture similarity between the first picture and the second picture of the two pictures is less than a threshold, the second picture is determined to be one of the partial pictures; and the first picture belongs to the partial pictures.

6. The display device according to claim 1, wherein The controller is further configured to: If the text carrying form is the third form, the partial picture is obtained based on the text consistency and picture similarity between two pictures of the media content in video format.

7. The display device according to claim 6, wherein: The controller is further configured to: If the text consistency between the first picture and the second picture of the two pictures is inconsistent, determining the second picture as one of the partial pictures; and the first picture belongs to the partial pictures; If the text consistency is consistent, the partial image is obtained according to the image similarity between the first image and the second image.

8. The display device according to claim 7, wherein: The controller is further configured to: If the text consistency is consistent, and the image similarity between the first image and the second image is less than a threshold, the second image is determined as one of the partial images.

9. The display device according to any one of claims 1 to 8, characterized in that The target format includes a portable document format.

10. A method for acquiring media content, characterized in that: Applicable to a display device, the display device comprising a display and a controller; The controller is coupled to the display; The method comprises: When the media content in the video format is generated, obtaining an acquisition link for the media content in the target format; In response to an instruction to obtain the media content in a target format, displaying a graphic code in a user interface of a target application via the display; the graphic code includes a link to obtain the media content in the target format; The target format is different from the video format; the media content in the target format includes some pictures from the media content in the video format; the some pictures are determined according to the text-carrying form of the media content in the video format; The method further comprises: If the media content in video format does not have an audio track, text recognition is performed on the image sequence of the media content in video format; if the text recognition result indicates that the media content in video format does not contain images with text, the text carrying form is determined to be the first form; if the text recognition result indicates that the media content in video format contains images with text, the text carrying form is determined to be the third form; If the media content in video format has an audio track, the text carrying format is determined to be the second format.

Citation Information

Patent Citations

  • Media data processing method, media data processing device and storage medium

    CN107257338A

  • Information card display method, device and equipment and storage medium

    CN111787391A