Interaction method, device, electronic device, storage medium, and computer program
The interaction method and apparatus enrich the form and efficiency of information acquisition from video and audio by displaying a text sentence list, allowing users to access information visually and auditorily, addressing the limitations of existing methods.
Patent Information
- Application Number
- JP2024571844
- Authority / Receiving Office
- JP · JP
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2022-08-15
- Filing Date
- 2023-08-14
- Publication Date
- 2025-07-08
- Estimated Expiration
- 2043-08-14
AI Technical Summary
Existing methods for obtaining valuable information from video and audio are inefficient and limited in form, leading to low acquisition efficiency.
An interaction method and apparatus that displays a text sentence list of video content in a predetermined area, allowing users to access information through both visual and auditory means, including features like text sentence display, interaction controls, and media content jump operations.
Enhances the form and efficiency of information acquisition by enabling users to obtain video content information through both visual and auditory means, improving user experience, especially in noisy environments or for those with hearing impairments.
Smart Images

Figure 2025521197000001_ABST
Abstract
Description
Technical Field
[0001] [Cross - Reference to Related Applications] This application claims the priority of a Chinese patent application with the application number 202210977564.5, which was filed with the China National Intellectual Property Administration on August 15, 2022, and all the contents of the said application are incorporated herein by reference.
[0002] This disclosure relates to the field of computer technology, for example, to an interaction method, apparatus, electronic device, and storage medium.
Background Art
[0003] Users can obtain valuable information in video and audio by listening to it. However, the methods for obtaining valuable information in video and audio are relatively single, and the efficiency of information acquisition is low.
Summary of the Invention
Problems to be Solved by the Invention
[0004] This disclosure provides an interaction method, apparatus, electronic device, and storage medium that enrich the acquisition form of valuable information in video and audio and improve the acquisition efficiency of valuable information in video and audio.
Means for Solving the Problems
[0005] Embodiments of this disclosure include receiving a text display operation on a first media content including video content, and in response to the text display operation, displaying a text sentence list of the first media content in a predetermined area, where the text sentence list includes at least two text sentences, and each of the at least two text sentences has a corresponding audio sentence in the target audio data of the first media content, and provides an interaction method.
[0006] Examples of the present disclosure An operation reception module configured to receive a text display operation on first media content including video content; A list display module configured to display a text sentence list of the first media content in a predetermined area in response to the text display operation, where the text sentence list includes at least two text sentences, and each of the at least two text sentences has a corresponding audio sentence in the target audio data of the first media content. An interaction device is further provided.
[0007] Examples of the present disclosure One or more processors; A memory configured to store one or more programs; and When the one or more programs are executed by the one or more processors, the one or more processors are caused to implement the interaction method according to any of the examples of the present disclosure. An electronic device is further provided.
[0008] Examples of the present disclosure further provide a computer-readable storage medium storing a computer program that, when executed by a processor, implements the interaction method according to the examples of the present disclosure.
Brief Description of the Drawings
[0009]
Figure 1
Figure 2
Figure 3
Figure 4
Figure 5
Figure 6
Figure 7
Figure 8
Figure 9
Figure 10
Embodiments for Carrying Out the Invention
[0010] Hereinafter, embodiments of the present disclosure will be described with reference to the drawings. Although some embodiments of the present disclosure are shown in the drawings, the present disclosure can be implemented in various forms and should not be construed as being limited to the embodiments described herein. On the contrary, these embodiments are provided to facilitate understanding of the present disclosure. The drawings and embodiments of the present disclosure are used only for illustration and not for limiting the protection scope of the present disclosure.
[0011] The multiple steps described in the method embodiments of the present disclosure may be executed in a different order and / or in parallel. Also, the method embodiments may include additional steps and / or omit the execution of the steps shown. The scope of the present disclosure is not limited thereto.
[0012] As used herein, the term "comprising" and its variations are to be construed in an open-ended manner, i.e., "including but not limited to". The term "based on" means "at least partially based on". The term "one embodiment" means "at least one embodiment", the term "another embodiment" means "at least one another embodiment", and the term "some embodiments" means "at least some embodiments". Related definitions of other terms are given in the following description.
[0013] The concepts such as "first", "second", etc. mentioned in the present disclosure are only for distinguishing different devices, modules or units, and are not for limiting the order or interdependence of the functions performed by these devices, modules or units.
[0014] The terms "one" and "a plurality" mentioned in the present disclosure are illustrative rather than restrictive, and should be understood as "one or more" unless otherwise clearly indicated in the context.
[0015] The names of the messages or information that interact between multiple devices in the embodiments of the present disclosure are only for the purpose of explanation and are not for limiting the scope of these messages or information.
[0016] Before using the technical solutions disclosed in the multiple embodiments of the present disclosure, the user should be informed of the type, scope of use, usage scenarios, etc. of the personal information related to the present disclosure in an appropriate form in accordance with the relevant laws and regulations, and the approval of the user should be obtained.
[0017] For example, when responding to an active request from a user, prompt information is sent to the user to explicitly prompt that the operation requesting execution needs to obtain and use the user's personal information. Thereby, the user can autonomously select whether to provide personal information to software or hardware such as an electronic device, an application, a server, or a storage medium that executes the operation of the technical solution of the present disclosure according to the prompt information.
[0018] As an optional but non-limiting implementation form, in response to receiving an active request from a user, the form of sending prompt information to the user may be, for example, in the form of a pop-up window, and the prompt information may be presented in the form of text in the pop-up window. Further, the pop-up window may be equipped with a selection control for the user to select "agree" or "disagree" to provide personal information to the electronic device.
[0019] The above-described notification and user approval acquisition procedures are merely exemplary and do not limit the implementation forms of the present disclosure. Other forms that comply with relevant laws and regulations can also be applied to the implementation forms of the present disclosure.
[0020] FIG. 1 is a schematic flowchart of an interaction method provided by an embodiment of the present disclosure. The method can be executed by an interaction device, wherein the device can be implemented by software and / or hardware and can be configured as an electronic device, such as a mobile phone or a tablet computer. The interaction method provided by the embodiment of the present disclosure is applicable to a scenario of checking a text list of a video, for example, a scenario of checking a text list of a video while watching the video. As shown in FIG. 1, the interaction method provided by the present embodiment includes the following.
[0021] S101: Receive a text display operation on a first media content including video content.
[0022] Among them, the text display operation may be a trigger operation for instructing the display of a text sentence list of media content, such as an operation for triggering a text display control of one piece of media content, an operation for triggering a media content jump control of another piece of media content (for example, a third piece of media content) generated based on the text sentence in one piece of media content, or a gesture operation for instructing the display of a text sentence list of another piece of media content. The first piece of media content may be media content whose display of the text sentence list of the first piece of media content is instructed by the text display operation. The media content may be, for example, video content, graphic text content, etc., and this embodiment does not limit its type. Optionally, the first piece of media content may be video content including audio data (for example, human voice) other than background music. Hereinafter, the case where the first piece of media content is video content will be described as an example.
[0023] Exemplarily, an operation for receiving a text display operation on the first piece of media content may be received. For example, when the first piece of media content is being displayed, an operation for the user to trigger the text display control of the first piece of media content is received, or when another piece of media content generated based on the text sentence in the first piece of media content is being displayed, a media content jump operation executed by the user is received. For example, an operation for the user to trigger the media content jump control corresponding to the third piece of media content is received.
[0024] S102: In response to the text display operation, display the text sentence list of the first piece of media content in a predetermined area. The text sentence list includes at least two text sentences, and the text sentence has a corresponding audio sentence in the target audio data of the first piece of media content.
[0025] Among them, the predetermined area may be an area for displaying a text sentence list of the first media content, and the position and size of the predetermined area may be flexibly set as necessary. The target audio data may be the audio data of the first media content. For example, among the audio data of the first media content, it may be voice data other than background music (e.g., valid voice data).
[0026] The text sentence list may be a list for displaying text sentences corresponding to at least one audio sentence in the target audio data of the first media content. The text sentence list may include text sentences corresponding to the audio sentences in the target audio data. For example, the text sentence list may include text sentences that have a one-to-one correspondence with the audio sentences in the target audio data. When there are at least two audio sentences in the target audio data, the text sentence list may include at least two text sentences. When there is only one audio sentence in the target audio data, the text sentence list may include only one text sentence. Each text sentence list may be arranged in the order before and after of the corresponding audio sentence in the target audio data.
[0027] The text sentences in the text sentence list may be recognized and input by the distributor of the first media content by listening to the target audio data. For example, they may be obtained by performing speech recognition on the target audio data. For example, after the distribution of the first media content is successful, the target audio data of the first media content may be subjected to speech recognition to obtain the text sentence list of the first media content. The text sentences in the text sentence list may be divided based on predetermined punctuation marks in the text sentence list. For example, the content between two adjacent predetermined punctuation marks in the text sentence list may be regarded as one text sentence. The predetermined punctuation marks may be set as needed and may include, for example, commas, semicolons, periods, and the like.
[0028] Whether or not to display the text sentence list has nothing to do with whether there are subtitles in the first media content. In other words, when there are subtitles in the first media content, the text sentence list of the first media content may be displayed. For example, when the first media content and the currently played subtitles of the first media content are displayed, the text sentence list of the first media content may also be displayed. Even when there are no subtitles in the first media content, the text sentence list of the first media content may be displayed. The text sentence list mentioned in this embodiment is different from subtitles. When subtitles are displayed, usually only the subtitles corresponding to the currently played audio sentence are displayed, and not all subtitle contents of the video content are displayed. However, the text sentence list displayed in this embodiment may include text sentences corresponding to at least one audio sentence in the target audio data. Therefore, through this text sentence list, the user can check not only the currently played text sentences but also the text sentences that have already been played and the text sentences that have not yet been played.
[0029] Exemplarily, when receiving a text display operation for the first media content, as shown in FIG. 2, a text sentence list 20 of the first media content may be displayed in a predetermined area. Exemplarily, simultaneously with playing the first media content, the text sentence list 20 of the first media content may be displayed. For example, the first media content is played in the media content playback area of the media content display page, and the text sentence list 20 of the first media content is displayed in a predetermined area of the media content display page. Here, there may or may not be an overlapping area between the media content playback area and the predetermined area, and this embodiment does not limit this.
[0030] Also, when displaying a text list of the first media content in a predetermined area, interaction controls preset for the first media content may be further displayed in the predetermined area. For example, in the predetermined area, a distributor identifier (such as an avatar, etc.) of the first media content, a like control for the first media content, a comment control, a favorite control, or a share control, etc. may be displayed. Thus, by triggering the distributor identifier of the first media content, the user can check the profile of the distributor of the first media content. By triggering the like control, the user can click the like button for the first media content. By triggering the comment control, the user can check the comment information of the first media content or instruct the current application to display a comment panel for the first media content to comment on the first media content. By triggering the favorite control, the user can add the first media content to the favorite list, or by triggering the share control, the user can instruct the current application to display a share panel for the first media content to share the first media content. When the user has no intention of checking the text list of the first media content, the user may instruct the current application to stop displaying the text list of the first media content by a corresponding trigger operation. The trigger operation may include, for example, triggering a close control displayed in the predetermined area, clicking on the media content playback area, and / or continuously swiping down when the first text sentence in the text area list is displayed in the predetermined area, etc.
[0031] In this embodiment, by displaying the text sentence list of the video content, in addition to the user being able to obtain the information in the video by watching the video screen or listening to the video audio, the user can also obtain the information in the video (for example, the video audio) by checking the text sentence list of the video content. Therefore, for example, when it is inconvenient to turn on the video audio, when in a noisy environment, or when there is a hearing impairment, etc., even in a scene where it is inconvenient for the user to listen to the video audio, the information contained in the video audio can still be quickly obtained. Compared with the prior art in which the user obtains the information contained in the video only by watching the video screen or listening to the video audio, the acquisition form of the information in the video (for example, the video audio) is enriched, the acquisition efficiency of the information in the video is improved, and the user experience can be improved.
[0032] In this embodiment, referring continuously to FIG. 2, when displaying the text sentence list 20 of the first media content, the current text sentence 21 (the sentence "These are what we really need to seriously consider" shown in FIG. 2) in the text sentence list 20 can be automatically moved and displayed within a predetermined area. The current text sentence 21 may have a different display mode from other text sentences in the text sentence list 20 other than the current text sentence 21 in the text sentence list 20, for example, having different fonts, font sizes, and / or colors, etc., in order for the user to clearly identify and easily check the current text sentence.
[0033] In this case, as an option, the at least two text sentences include the current text sentence and other text sentences other than the current text sentence that have a different display mode from the current text sentence, among which the current text sentence is the text sentence being currently played.
[0034] Among them, the current text sentence may be the text sentence corresponding to the currently played audio sentence in the target audio data of the first media content. The current text sentence may change as the first media content is played.
[0035] In this embodiment, when the search keyword to be waited for 22 is included, the search keyword to be waited for 22 and the content other than the search keyword to be waited for 22 in the text sentence list 20 may be displayed in the text sentence list 20 in different display modes. For example, the search keyword to be waited for 22 and the content other than the search keyword to be waited for 22 in the text sentence list 20 may be displayed in different fonts, font sizes, and / or colors, etc. For example, in order for the user to identify the search keyword to be waited for 22 in the text sentence list 20 and quickly search easily, a search tag (such as the magnifying glass mark displayed above the right of the search keyword "human capital" shown in FIG. 2) may be added to the search keyword to be waited for 22.
[0036] In this case, as an option, the search keyword to be waited for included in the at least two text sentences has a display mode different from the content other than the search keyword to be waited for in the at least two text sentences, and the search keyword to be waited for is used to trigger the display of search results that match the triggered search keyword to be waited for.
[0037] Among them, the search keyword to be waited for may be determined based on a preset determination rule. For example, candidate keywords may be manually marked in advance, and / or candidate keywords may be determined in advance according to the search frequency of a plurality of keywords in the current application or a candidate keyword determination model. The candidate keyword included in at least one text sentence of the text sentence list of the first media content may be used as the search keyword to be waited for in the text sentence list.
[0038] Therefore, when a trigger of one search waiting keyword in the text sentence list by the user is detected, search results matching the search waiting keyword may be obtained and displayed. For example, search results matching the search waiting keyword are obtained from the server, and the search results are displayed on the current page (for example, a predetermined area), or the current page is switched from the media content display page to the search result page, and the search results are displayed on the search result page.
[0039] In one embodiment, when the sentence being currently played is changed, the display position and / or display mode of the sentence being currently played within a predetermined area before and after the change may be automatically adjusted. For example, before the sentence being currently played is changed, the sentence being currently played before the change is displayed in a first display mode at the set position of the predetermined area, and the sentences not being currently played in the text sentence list may be displayed in a second display mode. When the sentence being currently played is changed, at least one text sentence (including the sentence being currently played before the change) in the text sentence list is controlled to move so that the sentence being currently played after the change moves to and is displayed at the set position of the predetermined area, and the sentence being currently played before the change is switched from the first display mode to the second display mode, and the sentence being currently played after the change may be switched from the second display mode to the first display mode. In this case, as an option, the current text sentence is displayed at the set position of the predetermined area, and the interaction method provided in this embodiment further includes moving and displaying the current text sentence after the change to the set position when the current text sentence is changed.
[0040] In one embodiment, after the text sentence list of the first media content is displayed in a predetermined area, in response to a phrase switching operation acting within the predetermined area, a position control for triggering switching the text sentence displayed in the predetermined area and moving and displaying the current text sentence to the set position of the predetermined area is further displayed.
[0041] Among them, the phrase switching operation may be a trigger operation for switching the text sentence, such as an operation for triggering a text sentence switch control or a slide operation acting within a predetermined area, for switching the text sentence displayed in the predetermined area. The position control may be a control for triggering the current text sentence to be moved to and displayed at the set position of the predetermined area. The position control may be displayed when receiving the phrase switching operation, or may be displayed when receiving the phrase switching operation and the currently played sentence is not displayed within the predetermined area, and may be set as required.
[0042] Exemplarily, when receiving the phrase switching operation, for example, when it is detected that the user is sliding vertically within the predetermined area, the text sentence displayed in the predetermined area may be switched. For example, in order to move and display a text sentence not displayed in the predetermined area within the predetermined area, at least one text sentence displayed in the predetermined area is controlled to move along with the user's sliding direction. And, as shown in FIG. 3, when receiving the phrase switching operation and / or the currently played sentence is not displayed within the predetermined area, the position control 30 may be displayed. Therefore, when the user tries to check the currently played sentence, the position control 30 may be triggered. Accordingly, when detecting that the user has triggered the position control 30, the current application program may control the plurality of text sentences displayed in the predetermined area to move so that the currently played sentence is moved to and displayed at the set position of the predetermined area.
[0043] In the above embodiment, after receiving the phrase switching operation, when the currently played sentence is changed, the changed currently played sentence may still be automatically moved to and displayed at the set position of the predetermined area.
[0044] Exemplarily, in order to meet the needs of a user who checks the text sentence displayed in a predetermined area after receiving a phrase switching operation, the currently playing sentence after the change may not be automatically moved to and displayed at the set position in the predetermined area. At this time, when it is detected that the user has triggered a position control, or when a trigger operation for switching the playback progress of the first media content is received (for example, when a playback operation for one text sentence in the text sentence list is received), the currently playing sentence after the change may be restored to be automatically moved to and displayed at the set position in the predetermined area again. That is, when the currently playing sentence is changed, the currently playing sentence after the change is automatically moved to and displayed at the set position in the predetermined area.
[0045] In one embodiment, it further includes playing the first media content before receiving a text display operation for the first media content, and displaying the text sentence list of the first media content in a predetermined area includes displaying the text sentence list of the first media content in the predetermined area and playing the first media content adjusted outside the predetermined area.
[0046] In the above embodiment, when viewing the first media content, the user may execute a text display operation on the first media content.
[0047] As shown in FIG. 4, the current application may play the first media content in the media content playback area of the media content display page. Therefore, when the user attempts to check the text sentence list 20 of the first media content, the user may perform a text display operation on the first media content. For example, the user may trigger the text display control 40 of the first media content. In response, when the current application program receives the text display operation on the first media content, as shown in FIG. 2, the text sentence list 20 of the first media content is displayed in a predetermined area of the media content display page, and the display range of the media content playback area is adjusted according to the display range of the predetermined area, and the first media content may be adjusted and played outside the predetermined area.
[0048] Among them, the display position of the text display control 40 can be set flexibly. For example, when playing the first media content, as shown in FIG. 4, the first media content may be displayed at a predetermined position on the media content display page, or the text display control 40 may be displayed on a predetermined panel of the first media content. Therefore, when the user attempts to perform a text display operation, the user may directly trigger the text display control 40 in the first media content, or perform a panel display operation to instruct the current application to display a predetermined panel of the first media content, and trigger the text display control 40 displayed on the predetermined panel.
[0049] In addition, the text display operation for the first media content may further include a mute playback operation for the first media content. For example, when it is detected that the user has adjusted the playback volume of the first media content to mute, it can be determined that a text display operation for the first media content has been received, and a text sentence list of the first media content is displayed in a predetermined area of the media content display page. In order for the user to easily obtain information in the target audio data of the first media content, the first media content is adjusted and played outside the predetermined area.
[0050] In another embodiment, the text display operation includes a media content jump operation for a third media content that is media content generated based on a second text sentence in the first media content, and further includes playing the third media content before receiving the text display operation for the first media content. Displaying the text sentence list of the first media content in a predetermined area includes playing the first media content outside the predetermined area and displaying the text sentence list of the first media content in the predetermined area.
[0051] In the above-described embodiment, when viewing the third media content generated based on the first media content, the user may execute a text display operation for the first media content.
[0052] Among them, the third media content may be media content generated based on the second text sentence in the first media content. For example, it may be media content generated by a media content generation operation for the second text sentence. The third media content may be video content, graphic text content, or the like. For example, it may be graphic text content. The second text sentence may include one or more text sentences in the text sentence list of the first media content, and may be the same as or different from the first text sentence. The third media content may be media content distributed by the current user or another user who is viewing the first media content. The media content jump operation may be a trigger operation for instructing a jump to the original media content (for example, the first media content) corresponding to the currently played media content (for example, the third media content). For example, when the currently played media content is media content generated based on a text sentence in another media content, the media content jump operation may be a trigger operation for jumping to the other media content. For example, it may be an operation for triggering the media content jump control of the currently played media content.
[0053] As shown in FIG. 5, the current application may display third media content generated based on a second text sentence in the first media content. Thus, when the user attempts to check the first media content and / or the text sentence list of the first media content, the user may perform a text display operation on the first media content. For example, the user may trigger a media content jump control 50 corresponding to the third media content. In response, when the current application receives a text display operation on the first media content, as shown in FIG. 2, the current application may play the first media content outside a predetermined area and display the text sentence list 20 of the first media content in the predetermined area. For example, the current page may be switched to the media content display page of the first media content, the first media content may be played outside the predetermined area of the media content display page, and the text sentence list 20 of the first media content may be displayed in the predetermined area of the media content display page.
[0054] In the above-described embodiment, the display timing of the media content jump control corresponding to the third media content may be set flexibly. For example, when the third media content is displayed, the media content jump control corresponding to the third media content may be displayed. Also, after receiving a control display operation for the third media content, the media content jump control corresponding to the third media content may be displayed. In this case, as an option, before receiving a text display operation for the first media content, in response to the control display operation for the third media content, the media content jump control for triggering the execution of the media content jump operation corresponding to the third media content may be further displayed. Among them, the control display operation for the third media content may be a trigger operation for instructing the display of the media content jump control corresponding to the third media content. For example, it may be a display pause operation for the third media content. In this case, illustratively, when receiving a display pause operation for the third media content, the display of the third media content may be paused, and the media content jump control corresponding to the third media content may be displayed.
[0055] In the above embodiment, playing the first media content outside a predetermined area may include playing the first media content outside the predetermined area with the start point of the first media content as the playback start point, or playing the first media content outside the predetermined area with the time node corresponding to the start point of the second text sentence in the first media content as the playback start point. That is, when playing the first media content outside a predetermined area in response to a media content jump operation for the third media content, the first media content may be played with the start point of the first media content as the playback start point, or the first media content may be played with the time node corresponding to the start point of the second text sentence in the first media content as the playback start point, which may be set as needed.
[0056] The interaction method provided by this embodiment receives a text display operation for the first media content including video content, and in response to the text display operation, displays a text sentence list of the first media content in a predetermined area. The text sentence list includes at least two text sentences, and there is a corresponding audio sentence in the target audio data of the first media content for the text sentence. By adopting the above technical solution, this embodiment can display the text sentence list of the video content, so that in addition to the user being able to obtain information in the video by watching the video screen or listening to the video audio, the user can also obtain information in the video (such as video audio) by checking the text sentence list of the video content, enriching the form of obtaining information in the video, improving the efficiency of obtaining information in the video, and improving the user experience.
[0057] FIG. 6 is a schematic flowchart of another interaction method provided by an embodiment of the present disclosure. The technical solution in this embodiment may be combined with one or more optional technical solutions in the above embodiments. Optionally, after displaying the text list of the first media content in a predetermined area, in response to a phrase selection operation acting in the predetermined area, identify the first text sentence selected by the phrase selection operation, and in response to a copy operation on the first text sentence, copy the first text sentence; in response to a sharing operation on the first text sentence, share the first text sentence; in response to a playback operation on the first text sentence, use the time node corresponding to the starting point of the first text sentence in the first media content as the playback start point to play the first media content; and in response to a media content generation operation on the first text sentence, generate a second media content including the first text sentence. The method further includes performing at least one of the above operations.
[0058] Optionally, after identifying the first text sentence selected by the phrase selection operation, the method further includes displaying the first text sentence in a selected state.
[0059] Accordingly, as shown in FIG. 6, the interaction method provided by this embodiment includes the following steps.
[0060] S201: Receive a text display operation on a first media content including video content.
[0061] S202: In response to the text display operation, display the text list of the first media content in a predetermined area. The text list includes at least two text sentences, and there is a corresponding audio sentence in the target audio data of the first media content for each text sentence.
[0062] S203: In response to a phrase selection operation acting within the predetermined area, identify the first text sentence selected by the phrase selection operation, display the first text sentence in a selected state, and execute at least one of S204 to S207.
[0063] Among them, the phrase selection operation may be an operation of selecting one or more text sentences in the text sentence list of the first media content. For example, it may be a click operation acting within the predetermined area. The first text sentence may be the text sentence selected by the phrase selection operation or may include one or more text sentences.
[0064] Exemplarily, when receiving a phrase selection operation acting within the predetermined area, the first text sentence selected by the phrase selection operation can be identified, the first text sentence is displayed in a selected state, and as shown in FIG. 7, a copy control 70, a sharing control (not shown in FIG. 7), a playback control 71, and / or a media content generation control 72 corresponding to the first text sentence may be displayed.
[0065] Exemplarily, when receiving a click operation acting within the predetermined area, the text sentence displayed at the trigger position of the click operation is regarded as the first text sentence corresponding to the click operation, the first text sentence is displayed in a selected state, and a copy control, a sharing control, a playback control, and / or a media content generation control corresponding to the first text sentence may be displayed. Further, the user can adjust the selected range by adjusting the position of the start point marker and / or the end point marker of the selected first text sentence, thereby adjusting the content included in the first text sentence.
[0066] S204: In response to a copy operation on the first text sentence, copy the first text sentence.
[0067] Exemplarily, upon receiving a copy operation for the first text sentence, the first text sentence may be copied, or prompt information may be displayed, and the success of copying the first text sentence may be presented to the user by the prompt information. Thus, thereafter, the user can input the first text sentence to other pages of the current application or other applications in a pasting form. Among them, the copy operation for the first text sentence may be a trigger operation for instructing copying of the first text sentence, for example, an operation for triggering a copy control 70 (shown in FIG. 7) corresponding to the first text sentence.
[0068] S205: In response to a sharing operation for the first text sentence, share the first text sentence.
[0069] Exemplarily, upon receiving a sharing operation for the first text sentence, the first text sentence may be shared. For example, in order for the current user to select other users who are sharing partners of the first text sentence, a sharing panel may be displayed, and after the selection of the current user is completed, the first text sentence may be shared with other users selected by the current user, or the first text sentence may be directly shared with target users pre-associated with the current user. Among them, the copy operation for the first text sentence may be a trigger operation for instructing display of the sharing panel for sharing the first text sentence, or may be a trigger operation for instructing sharing of the first text sentence, for example, an operation for triggering a sharing control corresponding to the first text sentence.
[0070] S206: In response to a playback operation for the first text sentence, using the time node corresponding to the start point of the first text sentence in the first media content as the playback start point, play the first media content.
[0071] Exemplarily, when receiving a playback operation for the first text sentence, the playback progress of the first media content may be adjusted to the playback progress indicated by the time node corresponding to the start point of the first text sentence in the first media content. That is, in order for the user to easily listen to the audio sentence corresponding to the first text sentence, the first media content may be played with the time node corresponding to the start point of the first text sentence in the first media content as the playback start point. Among them, the playback operation for the first text sentence may be a trigger operation for instructing the playback of the audio sentence corresponding to the first text sentence. For example, it may be an operation for triggering the playback control 71 (shown in FIG. 7) corresponding to the first text sentence.
[0072] S207: In response to the media content generation operation for the first text sentence, generate a second media content including the first text sentence.
[0073] Among them, the media content generation operation for the first text sentence may be a trigger operation for instructing to generate new media content using the first text sentence. For example, it may be an operation for triggering the media content generation control 72 (shown in FIG. 7) corresponding to the first text sentence. The second media content may be media content generated using the first text sentence. The second media content may be video content or graphic text content, etc. For example, it may be graphic text content. The graphic text content may be understood as media content with pictures as the content and characters as the introduction information of the media content.
[0074] Exemplarily, when receiving a media content generation operation for the first text sentence, a second media content including the first text sentence may be generated.
[0075] Taking as an example that the second media content is graphic text content, when generating the second media content including the first text sentence, a card including the second text sentence may be generated, and the card may be used as the content of the media content, and the music corresponding to the second text sentence and / or the card may be used as the background music of the media content to generate the second media content. In this case, optionally, generating the second media content including the first text sentence includes generating a target card including the first text sentence, using the target card as the content, using the music corresponding to the first text sentence and / or the target card as the background music, and generating the second media content.
[0076] The target card may be a picture including the first text sentence. The music corresponding to the first text sentence / target card may be music associated with the first text sentence / target card. The music associated with multiple text sentences / cards may be determined by a model obtained through pre-training or may be preset. The target card may be determined by being randomly selected or selected from a plurality of preset cards based on the first text sentence. Different cards may have different display modes, and this embodiment does not limit this.
[0077] In this embodiment, regardless of the length of the first text sentence, when receiving a media content generation operation for the first text sentence, second media content including the first text sentence may be generated. Also, the length of the first text sentence may be considered. Only when the length of the first text sentence is within a predetermined length range (for example, less than 150 characters), when receiving a media content generation operation for the first text sentence, second media content including the first text sentence is generated. However, when the length of the first text sentence is not within the predetermined length range, second media content may not be generated based on the first text sentence. For example, when the length of the first text sentence is not within the predetermined length range, a media content generation control corresponding to the first text sentence may not be displayed and may be flexibly set as needed.
[0078] In this embodiment, secondary creation may be performed based on the text sentences in the text sentence list of the first media content to generate new media content. Therefore, a new method for creating media content can be provided, reducing the difficulty of creating media content and improving the user's creation experience.
[0079] In one embodiment, in response to a media content generation operation for the first text sentence, second media content including the first text sentence may be generated and distributed. For example, the media content generation operation may be a trigger operation for instructing the generation and distribution of the second media content. For example, it may include a distribution operation for the second media content including the first text sentence. Therefore, when receiving a media content generation operation for the first text sentence, second media content including the first text sentence may be generated and the second media content may be distributed. In this case, as an option, generating the second media content including the first text sentence includes generating the second media content including the first text sentence and distributing the second media content.
[0080] In another embodiment, when receiving a media content generation operation for the first text sentence, only generate a second media content including the first text sentence, and after receiving a distribution operation for the second media content, distribute the second media content. For example, when receiving a media content generation operation for the first text sentence, generate a second media content including the first text sentence, and as shown in FIG. 8, in order for a user to edit the second media content, an editing page of the second media content may be displayed, and after receiving a distribution operation acting on the editing page of the second media content (for example, an operation of triggering a distribution control 80 on the editing page of the second media content) or a distribution operation acting within a distribution page of the second media content, distribute the second media content. In this case, optionally, after generating the second media content including the first text sentence, further include distributing the second media content in response to a distribution operation for the second media content.
[0081] The display of the text list of the first media content and the secondary creation based on the text in the text list of the first media content are both carried out on the premise of obtaining the permission of the distributor of the first media content. For example, the distributor of the first media content may turn on or off the permission for displaying the text list or secondary creation for all media content (including the first media content) distributed by himself on the setting page, or may turn on or off the permission for displaying the text list or secondary creation for the first media content when distributing the first media content or on the setting panel of the first media content. Also, when the distributor of the first media content does not permit the display of the text list of the first media content, the text list of the first media content shall not be displayed in response to the user's text display operation. When the distributor of the first media content does not permit the secondary creation based on the text in the text list of the first media content, no response shall be made to the media content generation operation by the user for the text in the text list of the first media content.
[0082] The interaction method provided by this embodiment can support the user to copy, share and / or play the text in the text list, or create new media content based on the text in the text list, can meet the different needs of the user, reduce the difficulty of media content creation, and improve the user experience.
[0083] FIG. 9 is a block diagram of an interaction device provided by an embodiment of the present disclosure. The device can be implemented by software and / or hardware, and may be configured as an electronic device, for example, configured as a mobile phone or a tablet computer, and can display a text list of a video by executing an interaction method. As shown in FIG. 9, the interaction device provided by this embodiment may include an operation reception module 901 and a list display module 902.
[0084] The operation reception module 901 is configured to receive a text display operation on a first media content including video content.
[0085] The list display module 902 is configured to display a text list of the first media content in a predetermined area in response to the text display operation. The text list includes at least two text sentences, and there is a corresponding audio sentence in the target audio data of the first media content for the text sentence.
[0086] The interaction device provided by this embodiment receives a text display operation on a first media content including video content by the operation reception module 901, and in response to the text display operation, the list display module 902 displays a text list of the first media content in a predetermined area. The text list includes at least two text sentences, and there is a corresponding audio sentence in the target audio data of the first media content for the text sentence. By adopting the above technical solution, this embodiment can display the text list of the video content, and in addition to obtaining the information in the video by watching the video screen or listening to the video audio, the information in the video (especially the video audio) can also be obtained by checking the text list of the video content, enriching the acquisition form of the information in the video, improving the acquisition efficiency of the information in the video, and improving the user experience.
[0087] In the above technical solution, the at least two text sentences may include the current text sentence and other text sentences other than the current text sentence that have a display mode different from that of the current text sentence, where the current text sentence may be the text sentence being currently played.
[0088] In the above technical solution, the current text sentence may be displayed at the set position in the predetermined area, and the interaction device provided in this embodiment may further include a sentence movement module configured to move and display the changed current text sentence at the set position when the current text sentence is changed.
[0089] Exemplarily, after the interaction device provided in this embodiment displays the text sentence list of the first media content in a predetermined area, in response to a sentence switching operation acting in the predetermined area, the interaction device may further include a sentence switching module configured to display a position control for triggering switching the text sentence displayed in the predetermined area and moving and displaying the current text sentence at the set position in the predetermined area.
[0090] Exemplarily, after the interaction device provided in this embodiment displays the text sentence list of the first media content in a predetermined area, in response to a word selection operation acting within the predetermined area, it identifies the first text sentence selected by the word selection operation, and then, in response to a copy operation on the first text sentence, it is configured with a word copy module for copying the first text sentence, a word sharing module configured to share the first text sentence in response to a sharing operation on the first text sentence, a word playback module configured to play the first media content with the time node corresponding to the start point of the first text sentence in the first media content as the playback start point in response to a playback operation on the first text sentence, and a media content generation module configured to generate a second media content including the first text sentence in response to a media content generation operation on the first text sentence, and may further include a word selection module configured to call at least one of them.
[0091] In the above technical solution, the media content generation module may include a card generation unit configured to generate a target card including the first text sentence, and a media content generation unit configured to use the target card as content and the music corresponding to the first text sentence and / or the target card as background music to generate a second media content.
[0092] In the above technical solution, the media content generation module may be configured to generate a second media content including the first text sentence and deliver the second media content, or the interaction device provided in this embodiment may further include a media content delivery module configured to deliver the second media content in response to a delivery operation on the second media content after generating the second media content including the first text sentence.
[0093] In the above technical solution, after identifying the first text sentence selected by the sentence selection operation, the sentence selection module may be configured to display the first text sentence in a selected state.
[0094] In the above technical solution, the search keyword waiting to be searched included in the at least two text sentences may have a display mode different from other contents in the at least two text sentences other than the search keyword waiting to be searched, and the search keyword waiting to be searched may be used to trigger the display of search results matching the triggered search keyword waiting to be searched.
[0095] Exemplarily, the interaction device provided in this embodiment may further include a first playback module configured to play the first media content before receiving a text display operation on the first media content, and the list display module 902 may display a text sentence list of the first media content in a predetermined area, and may be configured to adjust and play the first media content outside the predetermined area.
[0096] In the above technical solution, the text display operation may include a media content jump operation on the third media content, and the interaction device provided in this embodiment may further include a second playback module configured to play the third media content, which is the media content generated based on the second text sentence in the first media content, before receiving a text display operation on the first media content, and the list display module 902 may be configured to play the first media content outside a predetermined area and display a text sentence list of the first media content in the predetermined area.
[0097] Exemplarily, before receiving a text display operation on the first media content, the interaction device provided in this embodiment may further include a control display module configured to display a jump control for triggering execution of the media content jump operation corresponding to the third media content in response to a control display operation on the third media content.
[0098] In the above technical solution, the list display module 902 is configured to play the first media content outside the predetermined area with the start point of the first media content as the playback start point, or in the first media content, with the time node corresponding to the start point of the second text sentence as the playback start point, and may be provided to play the first media content outside the predetermined area.
[0099] The interaction device provided in the embodiments of the present disclosure can execute the interaction method provided in any embodiment of the present disclosure, and includes functional modules and effects corresponding to the execution of the interaction method. For technical details not described in detail in this embodiment, reference may be made to the interaction method provided in any embodiment of the present disclosure.
[0100] Referring to FIG. 10 below, FIG. 10 shows a schematic configuration diagram of an electronic device (for example, a terminal device) 1000 suitable for implementing the embodiments of the present disclosure. The terminal device in the embodiments of the present disclosure may include, for example, mobile terminals such as mobile phones, notebook computers, digital broadcast receivers, personal digital assistants (PDAs), tablets (Portable Android Devices, PADs), portable multimedia players (PMPs), in-vehicle terminals (for example, car navigation terminals), etc., and fixed terminals such as digital TVs, desktop computers, etc. The electronic device shown in FIG. 10 is only an example and should not limit the functions and usage ranges of the embodiments of the present disclosure.
[0101] As shown in FIG. 10, the electronic device 1000 may include a processing device (e.g., a central processing unit, a graphic text processor, etc.) 1001 that can execute various appropriate operations and processes according to a program stored in a read-only memory (ROM) 1002 or a program loaded from a storage device 1008 into a random access memory (RAM) 1003. Various programs and data necessary for the operation of the electronic device 1000 are also stored in the RAM 1003. The processing device 1001, the ROM 1002, and the RAM 1003 are connected to each other via a bus 1004. An input / output (I / O) interface 1005 is also connected to the bus 1004.
[0102] An input device 1006 including, for example, a touch screen, a touch pad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc., an output device 1007 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc., a storage device 1008 including, for example, a magnetic tape, a hard disk, etc., and a communication device 1009 may be connected to the I / O interface 1005. The communication device 1009 may allow the electronic device 1000 to communicate with other devices wirelessly or wiredly to exchange data. FIG. 10 shows an electronic device 1000 having various devices, but it is not necessary to implement or include all the shown devices. More devices or fewer devices may be implemented or included instead.
[0103] According to an embodiment of the present disclosure, the above-described process described with reference to the flowchart may be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product including a computer program mounted on a non-transitory computer-readable medium including program code for executing the method shown in the flowchart. In such an embodiment, the computer program may be downloaded and installed from a network by the communication device 1009, may be installed from the storage device 1008, or may be installed from the ROM 1002. When the computer program is executed by the processing device 1001, the above-described functions limited in the method of the embodiment of the present disclosure are executed.
[0104] The above-mentioned computer-readable medium in the present disclosure may be a computer-readable signal medium or a computer-readable storage medium, or any combination of the above two. The computer-readable storage medium may be, for example, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination of the above. The computer-readable storage medium may include an electrical connection having one or more wires, a portable computer disk, a hard disk, a RAM, a ROM, an Erasable Programmable Read Only Memory (EPROM), a flash memory, an optical fiber, a portable Compact Disc Read Only Memory (CD-ROM), an optical memory device, a magnetic memory device, or any suitable combination of the above. In the present disclosure, the computer-readable storage medium may be any tangible medium that includes or stores a program that can be used by or in combination with an instruction execution system, apparatus, or device. On the other hand, in the present disclosure, the computer-readable signal medium may include a data signal propagated in a baseband or as part of a carrier wave, on which a computer-readable program code is carried. Such a propagated data signal may adopt various forms including an electromagnetic signal, an optical signal, or any suitable combination of the above. Further, the computer-readable signal medium may be any computer-readable medium other than a computer-readable storage medium that can transmit, propagate, or transmit a program for use by or in combination with an instruction execution system, apparatus, or device. The program code included in the computer-readable medium may be transmitted through any suitable medium including a wire, an optical cable, a Radio Frequency (RF), etc., or any suitable combination of the above.
[0105] In some embodiments, the client and the server can communicate using any network protocol known currently or to be developed in the future, such as, for example, the Hyper Text Transfer Protocol (HTTP), and can be connected to digital data communication of any form or medium (e.g., a communication network). Examples of communication networks include Local Area Networks (LANs), Wide Area Networks (WANs), extranets (e.g., the Internet), and end-to-end networks (e.g., ad hoc end-to-end networks), and any network known currently or to be developed in the future.
[0106] The computer-readable medium described above may be included in the electronic device described above or may exist separately without being incorporated into the electronic device.
[0107] The computer-readable medium described above stores one or more programs. When the one or more programs described above are executed by the electronic device, the electronic device receives a text display operation for first media content including video content, and in response to the text display operation, displays a text sentence list of the first media content in a predetermined area. The text sentence list includes at least two text sentences, and for each text sentence, there is a corresponding audio sentence in the target audio data of the first media content.
[0108] The computer program code for performing the operations of the present disclosure may be written in one or more programming languages, or combinations thereof, including object-oriented programming languages such as Java, Smalltalk, C++, etc., and also including conventional procedural programming languages such as the "C" language or similar programming languages. The program code may be executed entirely on a user computer, partially executed on a user computer, executed as an independent software package, partially executed on a user computer and partially executed on a remote computer, or executed entirely on a remote computer or server. When a remote computer is involved, the remote computer may be connected to the user computer via any type of network including a LAN or WAN, or may be connected to an external computer (e.g., connected using an Internet service provider via the Internet).
[0109] Flowcharts and block diagrams in the drawings illustrate the architecture, functionality, and operations that can be implemented according to systems, methods, and computer program products in various embodiments of the present disclosure. In this regard, each block in a flowchart or block diagram may represent one module, program segment, or portion of code that includes one or more executable instructions for implementing a given logical function. In some alternative implementations, the functions represented in the blocks may occur in an order different from that shown in the drawings. For example, two blocks shown in succession may actually be executed substantially in parallel, or may sometimes be executed in the reverse order depending on the functions involved. Each block in the block diagrams and / or flowcharts, as well as combinations of blocks in the block diagrams and / or flowcharts, may be implemented by a dedicated hardware-based system for performing a given function or operation, or may be implemented by a combination of dedicated hardware and computer instructions.
[0110] The units according to the embodiments of the present disclosure described may be implemented in software or in hardware. Among them, the name of the unit may not constitute a limitation on the unit itself in some cases.
[0111] The functions described above in this specification may be executed, at least in part, by one or more hardware logic components. For example, without limitation, exemplary types of hardware logic components that may be used include Field Programmable Gate Array (FPGA), Application Specific IntegraTEd Circuit (ASIC), Application Specific Standard Parts (ASSP), System on Chip (SOC), Complex Programmable Logic Device (CPLD), and the like.
[0112] In the context of the present disclosure, a machine-readable medium may be a tangible medium that includes or stores a program used by or in combination with an instruction execution system, apparatus, or device. A machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium may include electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatus, or devices, or any suitable combination of the foregoing. A machine-readable storage medium may include one or more wire-based electrical connections, portable computer disks, hard disks, RAM, ROM, EPROM, flash memory, optical fibers, portable CD-ROMs, optical storage devices, magnetic storage devices, or any suitable combination of the foregoing. The storage medium may be a non-transitory storage medium.
[0113] According to one or more embodiments of the present disclosure, Example 1 includes receiving a text display operation on first media content including video content, and in response to the text display operation, displaying a list of text sentences of the first media content in a predetermined area, and the list of text sentences includes at least two text sentences, and the text sentences provide an interaction method in which corresponding audio sentences exist in the target audio data of the first media content.
[0114] According to one or more embodiments of the present disclosure, Example 2 is the method described in Example 1, wherein the at least two text sentences include the current text sentence and another text sentence other than the current text sentence having a display mode different from that of the current text sentence, and the current text sentence is the text sentence being currently played.
[0115] According to one or more embodiments of the present disclosure, Example 3 is the method described in Example 2, wherein the current text sentence is displayed at a set position in the predetermined area, and the method further includes: when the current text sentence is changed, moving and displaying the changed current text sentence at the set position.
[0116] According to one or more embodiments of the present disclosure, Example 4 is the method described in Example 2, and after displaying the list of text sentences of the first media content in a predetermined area, in response to a phrase switching operation acting within the predetermined area, switching the text sentence displayed in the predetermined area, and further including displaying a position control for triggering moving and displaying the current text sentence to a set position in the predetermined area.
[0117] According to one or more embodiments of the present disclosure, Example 5 is the method described in Example 1, and after displaying the list of text sentences of the first media content in a predetermined area, Identifying a first text sentence selected by the sentence selection operation in response to the sentence selection operation acting within the predetermined area; Further including performing at least one of: copying the first text sentence in response to a copy operation on the first text sentence; sharing the first text sentence in response to a sharing operation on the first text sentence; playing the first media content with a playback start point being a time node corresponding to a start point of the first text sentence in the first media content in response to a playback operation on the first text sentence; and generating a second media content including the first text sentence in response to a media content generation operation on the first text sentence.
[0118] According to one or more embodiments of the present disclosure, Example 6 is the method described in Example 5, wherein generating a second media content including the first text sentence includes generating a target card including the first text sentence; and generating a second media content with the target card as content and music corresponding to the first text sentence and / or the target card as background music.
[0119] According to one or more embodiments of the present disclosure, Example 7 is the method described in Example 5, wherein generating a second media content including the first text sentence includes generating a second media content including the first text sentence and delivering the second media content, or further includes delivering the second media content in response to a delivery operation on the second media content after generating the second media content including the first text sentence.
[0120] According to one or more embodiments of the present disclosure, Example 8 is the method described in Example 5, wherein after identifying the first text sentence selected by the sentence selection operation, Further including displaying the first text sentence in a selected state.
[0121] According to one or more embodiments of the present disclosure, Example 9 is the method described in Example 1, wherein the search waiting keywords included in the at least two text sentences have a display mode different from other contents other than the search waiting keywords in the at least two text sentences, and the search waiting keywords are used to trigger the display of search results that match the triggered search waiting keywords.
[0122] According to one or more embodiments of the present disclosure, Example 10 is the method described in any one of Examples 1 to 9, further including playing the first media content before displaying the text sentence list of the first media content in a predetermined area, includes displaying the text sentence list of the first media content in a predetermined area and adjusting and playing the first media content outside the predetermined area.
[0123] According to one or more embodiments of the present disclosure, Example 11 is the method described in any one of Examples 1 to 9, wherein the text display operation includes a media content jump operation for a third media content that is media content generated based on a second text sentence in the first media content. Before receiving the text display operation for the first media content, further including playing the third media content, displaying the text sentence list of the first media content in a predetermined area, includes playing the first media content outside the predetermined area and displaying the text sentence list of the first media content in the predetermined area.
[0124] According to one or more embodiments of the present disclosure, Example 12 is the method described in any of Examples 1-11, and before receiving a text display operation for the first media content, further comprising, in response to a control display operation for the third media content, displaying a jump control for triggering execution of the media content jump operation corresponding to the third media content.
[0125] According to one or more embodiments of the present disclosure, Example 13 is the method described in Example 11, and playing the first media content outside a predetermined region comprises the step of playing the first media content outside the predetermined region with the start point of the first media content as the playback start point, or comprises the step of playing the first media content outside the predetermined region with the time node corresponding to the start point of the second text sentence in the first media content as the playback start point.
[0126] According to one or more embodiments of the present disclosure, Example 14 is an operation reception module configured to receive a text display operation for first media content including video content, and a list display module configured to display a text sentence list of the first media content in a predetermined region in response to the text display operation, wherein the text sentence list includes at least two text sentences, and the text sentences provide an interaction device where corresponding audio sentences exist in the target audio data of the first media content.
[0127] According to one or more embodiments of the present disclosure, Example 15 is one or more processors, and a memory configured to store one or more programs, When the one or more programs are executed by the one or more processors, an electronic device is provided that causes the one or more processors to implement the interaction method described in any of Examples 1 to 13.
[0128] According to one or more embodiments of the present disclosure, Example 16 provides a computer-readable storage medium storing a computer program that, when executed by a processor, implements the interaction method described in any of Examples 1 to 13.
[0129] Also, although a plurality of operations are depicted in a particular order, it should not be understood that these operations are to be performed in the particular order shown or that performance in order is required. In certain environments, multitasking and parallel processing may be advantageous. Similarly, although the above discussion contains multiple implementation details, these should not be construed as limiting the scope of the present disclosure. Some features described in the context of individual embodiments may be implemented in combination in a single embodiment. Conversely, various features described in the context of a single embodiment may also be implemented in multiple embodiments individually or in any suitable sub-combination.
[0130] Although the subject matter has been described in language specific to structural features and / or methodological logic operations, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or operations described above. Conversely, the specific features and operations described above are merely exemplary forms for implementing the claims.
Claims
1. Receiving a text display operation for first media content including video content; In response to the text display operation, displaying a list of text sentences of the first media content in a predetermined area; The list of text sentences includes at least two text sentences, and for each of the at least two text sentences, there is a corresponding audio sentence in the target audio data of the first media content; An interaction method.
2. The at least two text sentences include a current text sentence and another text sentence other than the current text sentence having a display mode different from that of the current text sentence, and the current text sentence is the text sentence currently being played; The interaction method according to Claim 1.
3. The current text sentence is displayed at a set position in the predetermined area, and the interaction method further includes: When the current text sentence is changed, moving and displaying the changed current text sentence at the set position; The interaction method according to Claim 2.
4. After displaying the list of text sentences of the first media content in a predetermined area, In response to a phrase switching operation acting within the predetermined area, switching the text sentence displayed in the predetermined area and displaying a position control for triggering moving and displaying the current text sentence to the set position in the predetermined area; The interaction method according to Claim 2.
5. After displaying the list of text sentences of the first media content in a predetermined area, In response to a phrase selection operation acting within the predetermined area, identifying a first text sentence selected by the phrase selection operation; In response to a copy operation on the first text sentence, copying the first text sentence; in response to a sharing operation on the first text sentence, sharing the first text sentence; in response to a playback operation on the first text sentence, using a time node corresponding to a start point of the first text sentence in the first media content as a playback start point to play the first media content; in response to a media content generation operation on the first text sentence, generating a second media content including the first text sentence; and further including performing at least one of the above. The interaction method according to claim 1.
6. Generating the second media content including the first text sentence includes generating a target card including the first text sentence; and using the target card as content and using music corresponding to at least one of the first text sentence and the target card as background music to generate the second media content. The interaction method according to claim 5.
7. Generating the second media content including the first text sentence includes generating the second media content including the first text sentence and delivering the second media content, or after generating the second media content including the first text sentence, further including delivering the second media content in response to a delivery operation on the second media content. The interaction method according to claim 5.
8. After identifying the first text sentence selected by the sentence selection operation, further including displaying the first text sentence in a selected state. The interaction method according to claim 5.
9. The search waiting keywords included in the at least two text sentences have a display mode different from other content in the at least two text sentences other than the search waiting keywords, and the search waiting keywords are used to trigger the display of search results matching the triggered search waiting keywords. The interaction method according to claim 1.
10. Before receiving a text display operation on the first media content, further comprising playing the first media content, displaying a text sentence list of the first media content in a predetermined area, displaying the text sentence list of the first media content in the predetermined area, and playing the first media content adjusted outside the predetermined area, The interaction method according to any one of claims 1 to 9.
11. The text display operation includes a media content jump operation for a third media content that is media content generated based on a second text sentence in the first media content, and before receiving the text display operation for the first media content, further comprising playing the third media content, displaying a text sentence list of the first media content in a predetermined area, playing the first media content outside the predetermined area, and displaying the text sentence list of the first media content in the predetermined area, The interaction method according to any one of claims 1 to 9.
12. before receiving the text display operation for the first media content, further comprising displaying a jump control for triggering execution of the media content jump operation corresponding to the third media content in response to a control display operation for the third media content, The interaction method according to claim 11.
13. Playing the first media content outside the predetermined area includes: playing the first media content outside the predetermined area with the start point of the first media content as the playback start point, or playing the first media content outside the predetermined area with the time node corresponding to the start point of the second text sentence in the first media content as the playback start point, The interaction method according to claim 11.
14. An operation reception module configured to receive a text display operation for first media content including video content, and a list display module configured to display a text sentence list of the first media content in a predetermined area in response to the text display operation. The text sentence list includes at least two text sentences, and for each of the at least two text sentences, there is a corresponding audio sentence in the target audio data of the first media content. Interaction device.
15. At least one processor, A memory communicatively connected to the at least one processor, and The memory stores a computer program executable by the at least one processor, and when the computer program is executed by the at least one processor, the at least one processor is caused to execute the interaction method according to any one of Claims 1 to 13. Electronic device.
16. A computer-readable storage medium storing computer instructions for realizing the interaction method according to any one of Claims 1 to 13 when executed by a processor. Computer-readable storage medium.
Citation Information
Patent Citations
Manuscript display control method and device, electronic equipment and storage medium
CN111970257A
Caption display device and method
JP2003018491A
Video display device
JP2009152753A
Content display, content display method, program and recording medium
JP2009218741A
MOBILE TERMINAL VIDEO RECORDING METHOD AND DEVICE - Patent application
JP2019512174A