Display method, device and electronic device
By obtaining and displaying subtitle information and its annotation information in multimedia conferences, the problem of difficulty in understanding for conference participants is solved, and the accuracy and efficiency of conference interaction are improved.
Patent Information
- Application Number
- CN202210495727.6
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2022-04-29
- Publication Date
- 2025-09-26
- Estimated Expiration
- 2042-04-29
AI Technical Summary
In multimedia conferences, participants may have difficulty understanding voice information, which may lead to interruptions or misunderstandings, affecting the accuracy and efficiency of interaction.
By obtaining the subtitle information of the multimedia conference, determining the content to be annotated, and obtaining the corresponding annotation information, the subtitle information and its corresponding annotation information are displayed so that the participants can better understand the conference content.
It improves the interactive accuracy and efficiency of multimedia conferences and avoids conference interruptions and misunderstandings due to difficulty in understanding.
Smart Images

Figure CN117014660B_ABST
Abstract
Description
Technical Field
[0001] The present disclosure relates to the field of computer technology, and in particular to a display method, device, and electronic device. Background Art
[0002] With the development of the internet, users are increasingly utilizing the functions of their devices, making work and life more convenient. For example, users can use their devices to initiate multimedia conferences with other users online. Online multimedia conferences allow users to interact remotely and can also facilitate meetings without physically gathering in one place. Multimedia conferences largely avoid the location and venue restrictions of traditional face-to-face meetings. Summary of the Invention
[0003] This disclosure section is provided to briefly introduce concepts that will be described in detail in the detailed description section below. This disclosure section is not intended to identify key features or essential features of the claimed technical solution, nor is it intended to limit the scope of the claimed technical solution.
[0004] In a first aspect, an embodiment of the present disclosure provides a display method, which includes: obtaining subtitle information of a multimedia conference, wherein the subtitle information is converted based on voice information in the multimedia conference; determining content to be annotated in the subtitle information; obtaining annotation information of the content to be annotated; displaying the subtitle information, and displaying annotation information corresponding to the content to be annotated.
[0005] In a second aspect, an embodiment of the present disclosure provides a display device, comprising: a first acquisition unit, for acquiring subtitle information of a multimedia conference, wherein the subtitle information is converted based on voice information in the multimedia conference; a determination unit, for determining the content to be annotated in the subtitle information; a second acquisition unit, for acquiring annotation information of the content to be annotated; and a display unit, for displaying the subtitle information, and displaying the annotation information corresponding to the content to be annotated.
[0006] In a third aspect, an embodiment of the present disclosure provides an electronic device comprising: one or more processors; a storage device for storing one or more programs, wherein when the one or more programs are executed by the one or more processors, the one or more processors implement the display method as described in the first aspect.
[0007] In a fourth aspect, an embodiment of the present disclosure provides a computer-readable medium having a computer program stored thereon, which, when executed by a processor, implements the steps of the display method described in the first aspect.
[0008] The display method, device and electronic device provided by the embodiments of the present disclosure obtain subtitle information of a multimedia conference, and this subtitle information can be converted according to the voice information in the multimedia conference; the words can be determined from the subtitle information to obtain the content to be annotated; then, the annotation information of the content to be annotated can be obtained; then, the subtitle information can be displayed, and the annotation information corresponding to the annotation words in the subtitle information can be displayed; thereby, a new display method can be obtained, through which the annotation information can be displayed to the participants of the multimedia conference. Through these annotation information, the participants can improve the speed of understanding the multimedia conference, avoid interrupting the conference process or misunderstanding due to not understanding the meaning of other participating users, and thus improve the interaction accuracy and interaction efficiency of the multimedia conference. BRIEF DESCRIPTION OF THE DRAWINGS
[0009] The above and other features, advantages, and aspects of the various embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. Throughout the drawings, the same or similar reference numerals represent the same or similar elements. It should be understood that the drawings are schematic and that the originals and elements are not necessarily drawn to scale.
[0010] Figure 1 is a flow chart of one embodiment of a display method according to the present disclosure;
[0011] Figure 2 、 Figure 3 and Figure 4 is a schematic diagram of an application scenario of the display method according to the present disclosure;
[0012] Figure 5 is a schematic structural diagram of an embodiment of a display device according to the present disclosure;
[0013] Figure 6 is an exemplary system architecture in which the demonstration method of one embodiment of the present disclosure may be applied;
[0014] Figure 7 It is a schematic diagram of the basic structure of an electronic device provided according to an embodiment of the present disclosure. DETAILED DESCRIPTION
[0015] The following describes embodiments of the present disclosure in more detail with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments described herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.
[0016] It should be understood that the various steps described in the method embodiments of the present disclosure may be performed in different orders and / or in parallel. In addition, the method embodiments may include additional steps and / or omit the steps shown. The scope of the present disclosure is not limited in this respect.
[0017] As used herein, the term "including" and its variations are open-ended, i.e., "including but not limited to." The term "based on" means "based, at least in part, on." The term "one embodiment" means "at least one embodiment," the term "another embodiment" means "at least one additional embodiment," and the term "some embodiments" means "at least some embodiments." Other terms are defined in the following description.
[0018] It should be noted that the concepts of "first" and "second" mentioned in this disclosure are only used to distinguish different devices, modules or units, and are not used to limit the order or interdependence of the functions performed by these devices, modules or units.
[0019] It should be noted that the modifications of "one" and "multiple" mentioned in the present disclosure are illustrative rather than restrictive, and those skilled in the art should understand that unless otherwise clearly indicated in the context, they should be understood as "one or more".
[0020] The names of the messages or information exchanged between multiple devices in the embodiments of the present disclosure are only used for illustrative purposes and are not used to limit the scope of these messages or information.
[0021] Please refer to Figure 1 , which shows the process of an embodiment of the display method according to the present disclosure. Figure 1 The display method shown includes the following steps:
[0022] Step 101: Acquire subtitle information of a multimedia conference.
[0023] Here, the subtitle information is converted according to the voice information in the multimedia conference.
[0024] In this embodiment, the multimedia conference may be an online conference conducted using multimedia. The multimedia may include but is not limited to at least one of audio and video. The multimedia conference interface may be a related interface of the multimedia conference.
[0025] In this embodiment, the application for initiating the multimedia conference can be any type of application and is not limited here. For example, the application can be an instant video conferencing application, a communication application, a video playback application, an email application, etc.
[0026] In this embodiment, the voice information of participants in a multimedia conference can be converted into subtitle information. For example, the participants in a multimedia conference may include user A, user B, and user C. When a participant speaks, the participant's speech can be converted into subtitle information.
[0027] In this embodiment, subtitle information can be displayed in real time or with a time lag relative to the audio. Optionally, subtitle information can be displayed on a multimedia stream display interface, i.e., the participant's real-time multimedia stream. Optionally, the interface displaying subtitle information can also be displayed alongside the multimedia stream display interface.
[0028] In some embodiments, the multimedia conference is a real-time multimedia conference in progress. During the real-time multimedia conference, subtitle information and annotation information are displayed.
[0029] In some embodiments, the multimedia conference is a multimedia conference that has ended. Optionally, the subtitle information can also be displayed after the multimedia conference ends.
[0030] Step 102: Determine the content to be annotated in the subtitle information.
[0031] Here, a predefined determination basis is used to determine the content to be annotated from the subtitle information.
[0032] In this embodiment, the content to be annotated may be content to which annotations are to be added. Optionally, the content to be annotated may be a word, that is, the content to be annotated may also be referred to as a word to be annotated.
[0033] Here, the content to be annotated can be words that are difficult for the participants to understand. The content to be annotated can be determined by pre-set conditions. The pre-set conditions can be set according to actual conditions and are not limited here.
[0034] Step 103: Obtain annotation information of the content to be annotated.
[0035] In this embodiment, annotation information of the content to be annotated may be obtained. The annotation information may be information explaining the content to be annotated.
[0036] Alternatively, the annotation information may be obtained from a preset database. Alternatively, the annotation information may be obtained from the Internet.
[0037] Step 104: display the subtitle information and the annotation information corresponding to the content to be annotated.
[0038] In this embodiment, subtitle information can be displayed, and the displayed subtitle information can include the content to be annotated. That is, when the subtitle information is displayed, the content to be annotated can also be displayed; the annotation information of the content to be annotated can also be displayed.
[0039] Optionally, annotation information may be displayed corresponding to the displayed content to be annotated; for example, the annotation information is adjacent to the display position of the content to be annotated.
[0040] It should be noted that the display method provided in this embodiment obtains the subtitle information of the multimedia conference, and this subtitle information can be converted according to the voice information in the multimedia conference; the words can be determined from the subtitle information to obtain the content to be annotated; then, the annotation information of the content to be annotated can be obtained; then, the subtitle information can be displayed, and the annotation information corresponding to the annotation words in the subtitle information can be displayed. Thus, a new display method can be obtained. Through this display method, the annotation information can be displayed to the participants of the multimedia conference. Through these annotation information, the participants can improve the speed of understanding the multimedia conference, avoid interrupting the conference process or misunderstanding due to not understanding the meaning of other participating users, thereby improving the interaction accuracy and interaction efficiency of the multimedia conference.
[0041] In some embodiments, step 102 may include: determining a user group vocabulary corresponding to the user group to which the participating users of the multimedia conference belong; and determining the content to be annotated in the subtitle information based on the user group vocabulary.
[0042] Here, the user group vocabulary includes entries and entry explanations. The setting method of the entries in the user group vocabulary is not limited here.
[0043] Optionally, a user group vocabulary corresponding to a user group may be determined based on the user group to which all participants in the multimedia conference belong. For example, if participants A, B, and C belong to the same user group, the user group vocabulary of this user group may be used to determine the content to be annotated.
[0044] Optionally, a user group vocabulary corresponding to a user group may be determined based on the user groups to which some of the participants in the multimedia conference belong. For example, participants A and B belong to a first user group, and participant C belongs to a second user group. Optionally, the user group vocabulary corresponding to the first user group may be used to determine the content to be annotated. Optionally, the user group vocabulary corresponding to the first user group and the user group vocabulary corresponding to the second user group may be used to determine the content to be annotated. Optionally, the user group vocabulary corresponding to the second user group may be used to determine the content to be annotated.
[0045] Optionally, for the content to be annotated determined using the user group vocabulary corresponding to the first user group, the annotation information of the content to be annotated can be displayed to the participants belonging to the first user group, or to the participants belonging to the second user group. For the content to be annotated determined using the user group vocabulary corresponding to the first user group and the user group vocabulary corresponding to the second user group, the annotation information of the content to be annotated can be displayed to the participants belonging to the first user group and to the participants belonging to the second user group. For the content to be annotated determined using the user group vocabulary corresponding to the second user group, the annotation information of the content to be annotated can be displayed to the participants belonging to the second user group, or to the participants belonging to the second user group.
[0046] It should be noted that by determining the user group to which the participating users belong, determining the user group vocabulary, and determining the content to be annotated based on the user group vocabulary, the participating users can be shown entries that may have specific meanings within the enterprise, helping users to accurately understand the meaning expressed by other participating users and improving the interaction efficiency of multimedia conferences.
[0047] As an example, you can refer to Figure 2 , Figure 2 The scene of displaying annotation information for the enterprise entry in the subtitle information is shown in FIG. Figure 2 In the multimedia stream display area 201, the multimedia stream of the multimedia conference (such as real-time videos of participating users, shared content, etc.) can be displayed. The subtitle display area 202 can display the subtitle corresponding to user A's voice, "The Doudou value in the calculation result is relatively reasonable." The "Doudou value" can be an entry in the user group vocabulary corresponding to the user group, and the annotation information "growth rate" can be displayed for the corresponding "Doudou value."
[0048] Optionally, determining the content to be annotated in the subtitle information based on the user group vocabulary may include: selecting, from the words in the subtitle information, words that can match entries in the user group vocabulary as the content to be annotated.
[0049] In some embodiments, determining the content to be annotated in the subtitle information based on the user group vocabulary may include: selecting words from the words in the subtitle information that hit entries with predefined features in the user group vocabulary as the content to be annotated.
[0050] Here, the entries in the user group vocabulary may have some attributes, which may include but are not limited to at least one of the following: entry length, entry language, whether the entry is an abbreviated entry, the number of times the entry appears in the preset document set, the frequency of occurrence of the entry in the preset document set, and whether the entry appears in the preset dictionary set.
[0051] Here, the predefined feature may indicate a predefined feature used to indicate words that may be difficult for the participants to understand.
[0052] As an example, the predefined features may include but are not limited to at least one of the following: the length of the term is greater than a preset length threshold, the language of the term is a preset language, the term is an abbreviated term, the number of times the term appears in a preset document set is not greater than a preset number threshold, the frequency of appearance of the term in a preset document set is not greater than a preset frequency threshold, and the term does not appear in a preset dictionary set.
[0053] It should be noted that the entries with predefined characteristics in the user group vocabulary can be used as the basis for determining the content to be annotated, which can effectively narrow the scope of the content to be annotated and avoid the situation where a large number of words are annotated in the subtitle information and the focus is blurred.
[0054] In some embodiments, the predefined feature includes that the term is an abbreviation term, and the annotation information includes the full name of the term. The step 104 includes: displaying the full name of the term corresponding to the content to be annotated.
[0055] Here, an entry is an abbreviation, which indicates that the entry is an abbreviation of an entry with the same meaning. Generally, in communication, people may use simpler words to represent more complex words to improve communication efficiency. However, for those who do not master simpler words (such as abbreviations), the meaning of the simpler words may be confusing.
[0056] For example, the abbreviation for Internet Data Center can be IDC. If "IDC" appears in the subtitle information, and the user group vocabulary includes "IDC" (which is an abbreviation) in the predefined characteristics, then "IDC" in the subtitle information can be determined as the term to be annotated, and the full name of the term to be annotated can include Internet Data Center and / or Internet Data Center.
[0057] As an example, you can refer to Figure 3 , Figure 3 The following shows the scene where annotation information is displayed for the abbreviation terms in the enterprise terms in the subtitle information. Figure 3 In the multimedia stream display area 301, the multimedia stream of the multimedia conference (such as real-time videos of participating users, shared content, etc.) can be displayed. The subtitle display area 302 can display the subtitles corresponding to user B's voice, "IDCs are generally established in remote areas." Where "IDC" can be an abbreviation in the user group vocabulary corresponding to the user group, and the annotation information "Internet Data Center" can be displayed for "IDC".
[0058] It should be noted that using the abbreviated terms in the user group's vocabulary as the basis for determining the content to be annotated can improve interaction efficiency. Specifically, for participants who use abbreviated terms, when they want to express themselves using abbreviated terms, they don't have to worry about other participants not understanding and having to change their expression (changing expression requires additional time and thinking); for participants who receive abbreviated terms but don't know the meaning, they can refer to the annotation information to quickly understand the meaning of the abbreviated terms, avoiding misinterpretation due to misunderstanding, thereby improving the interaction efficiency of multimedia conferences.
[0059] In some embodiments, step 102 may include: determining a first language of the multimedia conference; and determining content to be annotated from words in a second language in the subtitle information.
[0060] Here, the first language may be a basic language of the multimedia conference, and the basic language may be a language mainly used in the multimedia conference.
[0061] In some embodiments, the first language may be determined based on at least one of the following of the multimedia conference, but not limited to: the region where the participants of the multimedia conference are located, the native language of the participants of the multimedia conference, and the preset language of the multimedia conference.
[0062] In some embodiments, the first language is set by the participant.
[0063] In some embodiments, the first language can be determined by recognizing the voice of the conference participant.
[0064] In some embodiments, the first language may be determined based on the participants, for example, based on the region where the participants are located, or based on the language used by the participants in the application.
[0065] Here, the second language is a language different from the first language. Optionally, the second language is a language other than the first language. Optionally, the second language is a designated language different from the first language.
[0066] As an example, the first language of the multimedia conference can be determined to be Chinese, and a language other than Chinese can be determined to be a second language, such as English. If Chinese and English appear in the multimedia conference, the content to be annotated can be selected from the English words.
[0067] It should be noted that in a multimedia conference, if languages other than the first language appear, it may cause participants to have difficulty understanding the multimedia conference. Setting the words in the second language as content to be annotated can improve the participants' full understanding of the meaning expressed by the words in the second language and improve the interactive efficiency of the multimedia conference.
[0068] In some embodiments, determining the content to be annotated from the second language words in the subtitle information includes: selecting second language words that match a preset second language vocabulary from the second language words in the subtitle information as the content to be annotated.
[0069] Here, the second language vocabulary includes second language entries and corresponding first language definitions.
[0070] Here, the method of presetting the second language vocabulary can be set according to actual conditions and is not limited here.
[0071] In some embodiments, the preset second language vocabulary may include entries whose usage frequency is less than a preset usage frequency threshold.
[0072] It should be noted that some commonly used words in the second language may not present a comprehension barrier for conference users. Selecting a few words from the second language as a pre-defined second language vocabulary and using the words that match this vocabulary as the content to be annotated can reduce the number of times second language words in subtitle information are identified as content to be annotated. This allows for more accurate selection of words that are difficult for conference users to understand, avoids extensive annotation of second language words in subtitle information, and reduces disruption to conference users.
[0073] In some embodiments, the first language interpretation includes at least two sub-interpretations; and step 104 may include: selecting a sub-interpretation associated with the context information from the first language interpretation corresponding to the content to be annotated based on context information in the subtitle information of the content to be annotated; and displaying the selected sub-interpretation corresponding to the content to be annotated.
[0074] For example, the second language includes English. The content to be annotated (English word) selected from the subtitle information may have a Chinese interpretation, and the Chinese interpretation may include at least two sub-interpretations. In this case, based on the contextual information in the subtitle information containing the content to be annotated (English word), the sub-interpretation associated with the contextual information can be selected from the Chinese interpretation of the content to be annotated.
[0075] In some embodiments, the context may include a first language word and / or a second language word, and sub-definition related to the context may be selected based on the meaning expressed by the first language word and / or the second language word.
[0076] It should be noted that by identifying and processing contextual information, the selected sub-definitions can be tailored to the context, improving the accuracy of the displayed sub-definitions. This can improve participants' understanding of second language terms, thereby improving their understanding of multimedia conferences and increasing the accuracy and efficiency of multimedia conference information exchange.
[0077] In some embodiments, step 102 may include: selecting words from the words in the subtitle information that match a preset rare word database as the content to be annotated.
[0078] In some embodiments, the language to which the generated characters in the rare character database belong may include words in any language, for example, words in the first language and words in a non-first language.
[0079] It should be noted that by setting up a database of uncommon characters, selecting content to be annotated using the database, and displaying annotation information for uncommon characters selected from subtitle information, participants in a multimedia conference can accurately understand the content of other participants using uncommon characters through the annotation information of uncommon characters, thereby improving their understanding of the multimedia conference and further enhancing the interaction efficiency of the multimedia conference.
[0080] In some embodiments, the above step 104 may include: displaying the content to be annotated in the displayed subtitle information in association with the corresponding annotation information.
[0081] Here, the specific implementation method of the associated display can be set according to the actual application scenario and is not limited here.
[0082] As an example, the annotation information may be displayed above or below the content to be annotated, thereby realizing associated display.
[0083] As an example, the annotation information may be displayed in a bubble, with the bubble pointing to the content to be annotated, thereby achieving associated display.
[0084] It should be noted that, through the associated display, it is possible to clearly indicate the content to be annotated that the annotation information is directed to, thereby improving the accuracy of the provided annotation information.
[0085] In some embodiments, the display of the content to be annotated in the displayed subtitle information in association with the corresponding annotation information includes at least one of the following but is not limited to: there is at least one difference in display style between the content to be annotated and the other content in the displayed subtitles except the content to be annotated; the annotation information display area does not overlap with the subtitle display area; the annotation information display area is embedded in the subtitle display area; in response to a triggering operation on the content to be annotated, the annotation information corresponding to the content to be annotated is displayed; the annotation information is displayed in association with the corresponding content to be annotated in the form of a floating window; in a real-time multimedia conference, the first display duration of the content to be annotated in the subtitle information is not greater than the second display duration of the corresponding annotation information.
[0086] Here, the content to be annotated and the other content in the displayed subtitles except the content to be annotated have at least one different display style. As an example, Figure 2 The content to be annotated (Bean Value) can be displayed in a different display style from other content (calculation results, relatively reasonable). For example, the content to be annotated can be displayed in bold or italics (the difference in display is not shown in the figure).
[0087] Here, the annotation information display area does not overlap with the subtitle display area. In other words, the annotation information and the subtitle information can be displayed in different areas.
[0088] Here, the annotation information display area is embedded in the subtitle display area; in other words, the annotation information is displayed in the subtitle display area, for example Figure 2 shown.
[0089] Here, in response to a triggering operation on the content to be annotated, the annotation information corresponding to the content to be annotated can be displayed; in other words, when the annotation information is displayed in response to a user trigger, the interference of the annotation information on the participants viewing the subtitle information can be reduced.
[0090] Optionally, the above step 102 may include at least one of the following but is not limited to: determining the content to be annotated according to a user operation; and the electronic device automatically determining the content to be annotated.
[0091] Here, the annotation information is displayed in a floating window in association with the corresponding content to be annotated; thus, the annotation information can be displayed in association with the content to be annotated without affecting the display of the subtitle information.
[0092] In a real-time multimedia conference, the first display duration of the annotated content in the subtitle information is no longer than the second display duration of the corresponding annotation information. In other words, the duration that the annotation information remains on the interface is greater than or equal to the duration that the annotated content remains on the subtitle information terminal. This ensures that the annotation information is fully displayed, helping users understand the annotated content.
[0093] In some embodiments, the method further includes: for the same content to be annotated that appears at least twice in the multimedia conference, determining the interval between the non-first occurrence of the content to be annotated and the previous corresponding display of the content to be annotated with annotation information; in response to the interval satisfying a preset interval condition, displaying annotation information corresponding to the non-first occurrence of the content to be annotated.
[0094] Here, for the first occurrence of the content to be annotated, annotation information can be correspondingly displayed.
[0095] As an example, the same content to be annotated may appear multiple times during the multimedia conference. For the same word that appears multiple times, the frequency of its appearance can be controlled. For example, for the same word, the time interval between two displays of annotation information is greater than one minute. For example, for the same word, the position interval between two displays of annotation information is greater than 5 lines. <000^212>
[0096] As an example, please refer to Figure 4 , Figure 4 shows an exemplary scenario for controlling the annotation frequency of the content to be annotated. In Figure 4 , in the subtitle information display area 201, the subtitle information corresponding to the voices of the participating users can be displayed. In the subtitle information corresponding to the voices of User A, User B, User C, and User D, the English word "cute" corresponding to the Chinese word "可爱" appears, that is, "cute" appears multiple times during the multimedia conference. The "cute" said by User A is the first occurrence, and the "cute" said by User B, User C, and User D is the non-first occurrence of the content to be annotated.
[0097] Figure 4 , when the subtitle information of the voice information of User B includes the content to be annotated ("cute"), it can be determined that the interval between the "cute" said by User B and the previous corresponding display of "cute" with annotation information is 0 lines; if the interval condition is not less than 1 line, it can be determined that the "cute" said by User B does not meet the interval condition, and the "cute" said by User B does not correspondingly display annotation information (i.e., 可爱).
[0098] Figure 4 , when the subtitle information of the voice information of User C includes the content to be annotated ("cute"), it can be determined that the interval between the "cute" said by User C and the previous corresponding display of "cute" with annotation information is 1 line; if the interval condition is not less than 1 line, it can be determined that the "cute" said by User C meets the interval condition, and the "cute" said by User C correspondingly displays annotation information (i.e., 可爱).
[0099] It's important to note that by controlling the frequency of displaying annotations for the same content, we can reduce the disruption caused by frequent display of annotations for users who may have temporarily understood the content. We can also help participants understand the content if they haven't seen the annotations for a long time. This balances reducing user disruption and providing timely reminders, improving the efficiency of multimedia conference interactions.
[0100] In some embodiments, the interval includes a time interval, and the time interval indicates the interval between the voice information corresponding to the content to be annotated.
[0101] In some embodiments, the method further includes: if the time interval between the non-first appearance of the to-be-annotated content and the last corresponding to-be-annotated content with annotation information displayed is greater than a preset time threshold, determining that the time interval meets a preset interval condition.
[0102] For example, the same content to be annotated may appear multiple times during a multimedia conference. For multiple occurrences of the same word, the frequency of its appearance can be controlled. For example, the interval between two presentations of annotation information for the same word can be greater than one minute.
[0103] It should be noted that by displaying the time interval of annotation information for the same content to be annotated, we can effectively adapt to the human forgetfulness law and provide reminders. For example, if the human short memory time is 5 minutes, the display interval of the same annotation information can be controlled to be no less than 5 minutes. In this way, when the user at the meeting may still remember the annotation information for the content to be annotated for 5 minutes, the annotation information will not be displayed; when the user at the meeting may have forgotten the annotation information for 5 minutes, the annotation information will be displayed again.
[0104] In some embodiments, the interval includes a text interval, and the text interval is used to represent the position interval of the same to-be-annotated content in the subtitle information.
[0105] In some embodiments, the method further includes: if the text interval between the non-first-appearing content to be annotated and the previous corresponding content to be annotated with annotation information displayed is greater than a preset text length threshold, determining that the text interval meets a preset interval condition.
[0106] For example, the same content to be annotated may appear multiple times during a multimedia conference. For multiple occurrences of the same word, the frequency of its appearance can be controlled. For example, the interval between the two locations where annotation information for the same word is displayed is greater than 5 lines.
[0107] Optionally, text spacing can indicate the spacing between texts. Text spacing can include, but is not limited to, at least one of the following: the size of the text displayed on the interface, or the spacing between sentences to which the text belongs. The size of the text displayed on the interface can be expressed in absolute terms or as the line difference between the text lines in the subtitle information.
[0108] In some embodiments, the text length threshold includes a line difference between text lines in the subtitle information being greater than a preset line difference.
[0109] It should be noted that by displaying the text interval of the annotation information of the same content to be annotated, it is possible to effectively prompt the user in accordance with the display method of the subtitle information and the viewing method of the user. Figure 2 The subtitle information display area can display 5 lines of subtitle information at a time, and the display interval of the same annotation information can be controlled to be no less than 5 lines; thus, the participating users can see 5 lines of annotation information of the content to be annotated, and the annotation information is not displayed; when the participating users may have forgotten the annotation information and there is no prompt in the current page (that is, after a 5-line interval), the annotation information will be displayed again.
[0110] Further references Figure 5 As an implementation of the methods shown in the above figures, the present disclosure provides an embodiment of a display device, which is similar to Figure 1 Corresponding to the method embodiment shown, the device can be specifically applied to various electronic devices.
[0111] like Figure 5 As shown, the display device of this embodiment includes: a first acquisition unit 501, a determination unit 502, a second acquisition unit 503, and a display unit 504. The first acquisition unit is used to acquire subtitle information of a multimedia conference, wherein the subtitle information is converted from voice information in the multimedia conference; the determination unit is used to determine the content to be annotated in the subtitle information; the second acquisition unit is used to acquire annotation information for the content to be annotated; and the display unit is used to display the subtitle information and the annotation information corresponding to the content to be annotated.
[0112] In this embodiment, the specific processing of the first acquisition unit 501, the determination unit 502, the second acquisition unit 503 and the display unit 504 of the display device and the technical effects thereof can be referred to respectively. Figure 1 The relevant descriptions of step 101, step 102, step 103 and step 104 in the corresponding embodiment are not repeated here.
[0113] In some embodiments, determining the content to be annotated in the subtitle information includes: determining a user group vocabulary corresponding to the user group to which the participating users of the multimedia conference belong, wherein the user group vocabulary includes entries and entry explanations; and determining the content to be annotated in the subtitle information based on the user group vocabulary.
[0114] In some embodiments, determining the content to be annotated in the subtitle information based on the user group vocabulary includes: selecting, from the words in the subtitle information, words that hit entries with predefined features in the user group vocabulary as the content to be annotated.
[0115] In some embodiments, the predefined features include the entry being an abbreviated entry, the annotation information including the full name of the entry; and the display of subtitle information, and the display of annotation information corresponding to the content to be annotated, including: displaying the full name of the entry corresponding to the content to be annotated.
[0116] In some embodiments, determining the content to be annotated in the subtitle information includes: determining a first language of the multimedia conference, wherein the first language is determined based on at least one of the following items of the multimedia conference: the region where the participants of the multimedia conference are located, the native language of the participants of the multimedia conference, and the preset language of the multimedia conference; determining the content to be annotated from the second language words in the subtitle information, wherein the second language is a language other than the first language.
[0117] In some embodiments, determining the content to be annotated from the second language words in the subtitle information includes: selecting, from the second language words in the subtitle information, second language words that hit a preset second language vocabulary as the content to be annotated, wherein the second language vocabulary includes second language entries and corresponding first language definitions.
[0118] In some embodiments, the first language interpretation includes at least two sub-interpretations; and the displaying of subtitle information and the displaying of annotation information corresponding to the content to be annotated include:
[0119] According to the context information in the subtitle information of the content to be annotated, a sub-definition associated with the context information is selected from the first language interpretation corresponding to the content to be annotated; and the selected sub-definition is displayed corresponding to the content to be annotated.
[0120] In some embodiments, determining the content to be annotated in the subtitle information includes: selecting words that match a preset rare word database from the words in the subtitle information as the content to be annotated.
[0121] In some embodiments, the displaying of subtitle information and the displaying of annotation information corresponding to the content to be annotated include: displaying the content to be annotated in the displayed subtitle information in association with the corresponding annotation information.
[0122] In some embodiments, the display of the content to be annotated in the displayed subtitle information in association with the corresponding annotation information includes at least one of the following: there is at least one difference in display style between the content to be annotated and other content in the displayed subtitles except the content to be annotated; the annotation information display area does not overlap with the subtitle display area; the annotation information display area is embedded in the subtitle display area; in response to a triggering operation on the content to be annotated, the annotation information corresponding to the content to be annotated is displayed; the annotation information is displayed in association with the corresponding content to be annotated in the form of a floating window; in a real-time multimedia conference, the first display duration of the content to be annotated in the subtitle information is not greater than the second display duration of the corresponding annotation information.
[0123] In some embodiments, the device is also used to: for the same content to be annotated that appears at least twice in the multimedia conference, determine the interval between the content to be annotated that does not appear for the first time and the previous corresponding content to be annotated with annotation information displayed; in response to the interval satisfying a preset interval condition, display annotation information corresponding to the content to be annotated that does not appear for the first time.
[0124] In some embodiments, the interval includes a time interval, which indicates the interval between the voice information corresponding to the content to be annotated; and the device is also used to: if the time interval between the content to be annotated that does not appear for the first time and the last corresponding content to be annotated with annotation information displayed is greater than a preset time threshold, then determine that the time interval meets the preset interval condition.
[0125] In some embodiments, the interval includes a text interval, which is used to represent the position interval of the same content to be annotated that appears adjacently in the order of appearance in the subtitle information; and the device is also used to: if the text interval between the content to be annotated that does not appear for the first time and the previous corresponding content to be annotated with annotation information displayed is greater than a preset text length threshold, then determine that the text interval meets the preset interval condition.
[0126] In some embodiments, the text length threshold includes a line difference between text lines in the subtitle information being greater than a preset line difference.
[0127] In some embodiments, the multimedia conference is a real-time multimedia conference in progress.
[0128] Please refer to Figure 6 , Figure 6 An exemplary system architecture is shown in which the presentation method of one embodiment of the present disclosure may be applied.
[0129] like Figure 6 As shown, the system architecture may include terminal devices 601, 602, 603, a network 604, and a server 605. The network 604 is used to provide a medium for communication links between the terminal devices 601, 602, 603 and the server 605. The network 604 may include various connection types, such as wired or wireless communication links or fiber optic cables.
[0130] Terminal devices 601, 602, and 603 can interact with server 605 via network 604 to receive or send messages, etc. Various client applications can be installed on terminal devices 601, 602, and 603, such as web browser applications, search applications, and news and information applications. The client applications in terminal devices 601, 602, and 603 can receive user instructions and perform corresponding functions based on the user instructions, such as adding corresponding information to the message based on the user's instructions.
[0131] Terminal devices 601, 602, and 603 can be hardware or software. When terminal devices 601, 602, and 603 are hardware, they can be various electronic devices with display screens and support web browsing, including but not limited to smart phones, tablet computers, e-book readers, MP3 players (Moving Picture Experts Group Audio Layer III, Moving Picture Experts Group Audio Layer 3), MP4 (Moving Picture Experts Group Audio Layer IV, Moving Picture Experts Group Audio Layer 4) players, laptop computers, and desktop computers, etc. When terminal devices 601, 602, and 603 are software, they can be installed in the electronic devices listed above. It can be implemented as multiple software or software modules (for example, software or software modules used to provide distributed services), or it can be implemented as a single software or software module. No specific limitation is made here.
[0132] The server 605 may be a server that provides various services, such as receiving information acquisition requests sent by the terminal devices 601, 602, and 603, acquiring display information corresponding to the information acquisition requests through various means according to the information acquisition requests, and sending relevant data of the display information to the terminal devices 601, 602, and 603.
[0133] It should be noted that the display method provided in the embodiment of the present disclosure can be executed by a terminal device, and accordingly, the display device can be set in the terminal devices 601, 602, and 603. In addition, the display method provided in the embodiment of the present disclosure can also be executed by a server 605, and accordingly, the display device can be set in the server 605.
[0134] It should be understood that Figure 6 The number of terminal devices, networks and servers in the embodiment is merely illustrative. Any number of terminal devices, networks and servers may be provided as required.
[0135] Reference below Figure 7 , which shows an electronic device (eg Figure 6 The terminal device in the embodiments of the present disclosure may include, but is not limited to, mobile terminals such as mobile phones, laptop computers, digital broadcast receivers, PDAs (personal digital assistants), PADs (tablet computers), PMPs (portable multimedia players), in-vehicle terminals (such as in-vehicle navigation terminals), and fixed terminals such as digital TVs and desktop computers. Figure 7 The electronic device shown is only an example and should not limit the functions and scope of use of the embodiments of the present disclosure.
[0136] like Figure 7 As shown, the electronic device may include a processing device (e.g., a central processing unit, a graphics processing unit, etc.) 701, which can perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 702 or a program loaded from a storage device 708 into a random access memory (RAM) 703. Various programs and data required for the operation of the electronic device 700 are also stored in the RAM 703. The processing device 701, the ROM 702, and the RAM 703 are connected to each other via a bus 704. An input / output (I / O) interface 705 is also connected to the bus 704.
[0137] Typically, the following devices may be connected to the I / O interface 705: an input device 706 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 707 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 708 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 709. The communication device 709 may allow the electronic device to communicate with other devices wirelessly or by wire to exchange data. Although Figure 7 The electronic device is shown with various devices, but it should be understood that it is not required to implement or possess all of the devices shown. More or fewer devices may be implemented or possessed instead.
[0138] In particular, according to an embodiment of the present disclosure, the process described above with reference to the flowchart can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes a program code for executing the method shown in the flowchart. In such an embodiment, the computer program can be downloaded and installed from the network through the communication device 709, or installed from the storage device 708, or installed from the ROM 702. When the computer program is executed by the processing device 701, the above-mentioned functions defined in the method of the embodiment of the present disclosure are performed.
[0139] It should be noted that the computer-readable medium mentioned above in the present disclosure may be a computer-readable signal medium or a computer-readable storage medium, or any combination of the two. A computer-readable storage medium may be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of computer-readable storage media may include, but are not limited to, an electrical connection having one or more conductors, a portable computer disk, a hard disk, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination thereof. In the present disclosure, a computer-readable storage medium may be any tangible medium containing or storing a program that can be used by or in conjunction with an instruction execution system, apparatus, or device. In the present disclosure, a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, carrying computer-readable program code. Such a propagated data signal may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination thereof. A computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium that can transmit, propagate, or transport a program for use by or in conjunction with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium may be transmitted using any suitable medium, including but not limited to wires, optical cables, RF (radio frequency), etc., or any suitable combination thereof.
[0140] In some embodiments, the client and server can communicate using any currently known or future developed network protocol, such as HTTP (HyperText Transfer Protocol), and can be interconnected with any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network ("LAN"), a wide area network ("WAN"), an internet (e.g., the Internet), and a peer-to-peer network (e.g., an ad hoc peer-to-peer network), as well as any currently known or future developed network.
[0141] The computer-readable medium may be included in the electronic device, or may exist independently without being incorporated into the electronic device.
[0142] The above-mentioned computer-readable medium carries one or more programs. When the above-mentioned one or more programs are executed by the electronic device, the electronic device is enabled to: obtain subtitle information of the multimedia conference, wherein the subtitle information is converted based on the voice information in the multimedia conference; determine the content to be annotated in the subtitle information; obtain annotation information of the content to be annotated; display the subtitle information, and display the annotation information corresponding to the content to be annotated.
[0143] Computer program code for performing the operations of the present disclosure may be written in one or more programming languages, or a combination thereof, including, but not limited to, object-oriented programming languages such as Java, Smalltalk, C++, and conventional procedural programming languages such as "C" or similar programming languages. The program code may be executed entirely on the user's computer, partially on the user's computer, as a stand-alone software package, partially on the user's computer and partially on a remote computer, or entirely on the remote computer or server. In cases involving a remote computer, the remote computer may be connected to the user's computer through any type of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (e.g., through the Internet using an Internet service provider).
[0144] The flowcharts and block diagrams in the accompanying drawings illustrate the possible implementation architecture, functions and operations of the systems, methods and computer program products according to various embodiments of the present disclosure. In this regard, each box in the flowchart or block diagram can represent a module, program segment, or a part of code, and the module, program segment, or a part of code contains one or more executable instructions for realizing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the box can also occur in a different order than that marked in the accompanying drawings. For example, two boxes represented in succession can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flowchart, and the combination of the boxes in the block diagram and / or flowchart, can be implemented with a dedicated hardware-based system that performs the specified function or operation, or can be implemented with a combination of dedicated hardware and computer instructions.
[0145] The units involved in the embodiments described in this disclosure may be implemented by software or hardware. In some cases, the name of a unit does not limit the unit itself. For example, the first acquisition unit may also be described as a "unit for acquiring subtitle information."
[0146] The functions described above herein may be performed, at least in part, by one or more hardware logic components. For example, and without limitation, exemplary types of hardware logic components that may be used include: field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chip (SOCs), complex programmable logic devices (CPLDs), and the like.
[0147] In the context of the present disclosure, a machine-readable medium may be a tangible medium that may contain or store a program for use by or in conjunction with an instruction execution system, apparatus, or device. A machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium may include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any suitable combination of the foregoing. More specific examples of machine-readable storage media would include an electrical connection based on one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.
[0148] The above description is merely a preferred embodiment of the present disclosure and an illustration of the technical principles employed. Those skilled in the art should understand that the scope of disclosure involved in the present disclosure is not limited to the technical solutions formed by the specific combination of the above-mentioned technical features, but also includes other technical solutions formed by any combination of the above-mentioned technical features or their equivalents without departing from the above-mentioned disclosed concepts. For example, a technical solution formed by replacing the above-mentioned features with (but not limited to) technical features with similar functions disclosed in this disclosure.
[0149] In addition, although each operation is described in a specific order, this should not be understood as requiring these operations to be performed in the specific order shown or in a sequential order. Under certain circumstances, multitasking and parallel processing may be advantageous. Similarly, although some specific implementation details have been included in the above discussion, these should not be interpreted as limiting the scope of the present disclosure. Some features described in the context of a separate embodiment can also be implemented in a single embodiment in combination. On the contrary, the various features described in the context of a single embodiment can also be implemented in multiple embodiments individually or in any suitable sub-combination mode.
[0150] Although the subject matter has been described in language specific to structural features and / or methodological logical acts, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or acts described above. Rather, the specific features and acts described above are merely example forms of implementing the claims.
Claims
1. A display method, characterized in that: include: Acquiring subtitle information of a multimedia conference, wherein the subtitle information is converted based on voice information in the multimedia conference; Determining the content to be annotated in the subtitle information; Obtaining annotation information of the content to be annotated; Display subtitle information and annotation information corresponding to the content to be annotated; The method further includes: for the same content to be annotated that appears at least twice in the multimedia conference, determining an interval between a non-first occurrence of the content to be annotated and a previous occurrence of the content to be annotated for which annotation information is displayed; and displaying annotation information corresponding to the non-first occurrence of the content to be annotated in response to the interval satisfying a preset interval condition; The interval includes a time interval, which indicates the interval between the voice information corresponding to the content to be annotated; and the method also includes: if the time interval between the content to be annotated that does not appear for the first time and the previous corresponding content to be annotated with annotation information displayed is greater than a preset time threshold, then determining that the time interval meets the preset interval condition.
2. The method according to claim 1, characterized in that The determining of the content to be annotated in the subtitle information includes: Determining a user group vocabulary corresponding to the user group according to the user group to which the participating users of the multimedia conference belong, wherein the user group vocabulary includes entries and entry explanations; Based on the user group vocabulary, content to be annotated in the subtitle information is determined.
3. The method according to claim 2, characterized in that The determining, based on the user group vocabulary, the content to be annotated in the subtitle information includes: From the words in the subtitle information, words that match entries with predefined features in the user group vocabulary are selected as the content to be annotated.
4. The method according to claim 3, characterized in that The predefined features include that the term is an abbreviated term, and the annotation information includes the full name of the term; and The display of subtitle information and the display of annotation information corresponding to the content to be annotated include: Display the full name of the entry corresponding to the content to be annotated.
5. The method according to claim 1, wherein The determining of the content to be annotated in the subtitle information includes: Determining a first language for the multimedia conference, wherein the first language is determined based on at least one of the following: a region where participants of the multimedia conference are located, a native language of the participants of the multimedia conference, and a preset language for the multimedia conference; Determine the content to be annotated from the second language words in the subtitle information, wherein the second language is a language other than the first language.
6. The method according to claim 5, characterized in that The step of determining the content to be annotated from the second language words in the subtitle information includes: From the second language words in the subtitle information, second language words that match a preset second language vocabulary are selected as the content to be annotated, wherein the second language vocabulary includes second language entries and corresponding first language definitions.
7. The method according to claim 6, characterized in that The first language definition includes at least two sub-definitions; as well as The display of subtitle information and the display of annotation information corresponding to the content to be annotated include: Selecting, according to context information in the subtitle information of the content to be annotated, a sub-definition associated with the context information from the first language interpretation corresponding to the content to be annotated; The selected sub-definition is displayed corresponding to the content to be annotated.
8. The method according to claim 1, characterized in that The determining of the content to be annotated in the subtitle information includes: From the words in the subtitle information, words that match a preset rare word database are selected as the content to be annotated.
9. The method according to claim 1, characterized in that The display of subtitle information and the display of annotation information corresponding to the content to be annotated include: The content to be annotated in the displayed subtitle information is associated with the corresponding annotation information and displayed.
10. The method according to claim 9, characterized in that The displaying of the to-be-annotated content in the displayed subtitle information in association with the corresponding annotation information includes at least one of the following: There is at least one difference in display styles between the content to be annotated and the other content in the displayed subtitles except the content to be annotated; The annotation information display area does not overlap with the subtitle display area; The annotation information display area is embedded in the subtitle display area; In response to a triggering operation on the content to be annotated, displaying annotation information corresponding to the content to be annotated; Display the annotation information in a floating window in association with the corresponding content to be annotated; In a real-time multimedia conference, a first display duration of content to be annotated in subtitle information is not greater than a second display duration of corresponding annotation information.
11. The method according to claim 1, wherein The interval includes a text interval, and the text interval is used to represent the position interval of the same to-be-annotated content in the subtitle information; as well as The method further comprises: If the text interval between the non-first-appearing content to be annotated and the previous corresponding content to be annotated with annotation information displayed is greater than a preset text length threshold, it is determined that the text interval meets the preset interval condition.
12. The method according to claim 11, characterized in that The text length threshold includes a line difference between text lines in the subtitle information being greater than a preset line difference.
13. The method according to claim 1, wherein The multimedia conference is a real-time multimedia conference in progress.
14. A display device, characterized in that: include: A first acquiring unit is configured to acquire subtitle information of a multimedia conference, wherein the subtitle information is converted from voice information in the multimedia conference; a determining unit, configured to determine content to be annotated in the subtitle information; A second acquiring unit, configured to acquire annotation information of the content to be annotated; A display unit, used to display subtitle information and annotation information corresponding to the content to be annotated; The apparatus is further configured to: for the same content to be annotated that appears at least twice in the multimedia conference, determine an interval between a non-first occurrence of the content to be annotated and a previous occurrence of the content to be annotated for which annotation information is displayed; and in response to the interval satisfying a preset interval condition, display annotation information corresponding to the non-first occurrence of the content to be annotated; The interval includes a time interval, which indicates the interval between the voice information corresponding to the content to be annotated; and the device is also used to: if the time interval between the content to be annotated that does not appear for the first time and the previous corresponding content to be annotated with annotation information displayed is greater than a preset time threshold, then determine that the time interval meets the preset interval condition.
15. An electronic device, characterized in that: include: one or more processors; a storage device for storing one or more programs, When the one or more programs are executed by the one or more processors, the one or more processors implement the method according to any one of claims 1 to 13.
16. A computer-readable medium having a computer program stored thereon, characterized in that: When the program is executed by a processor, the method according to any one of claims 1 to 13 is implemented.
Citation Information
Patent Citations
Method and device for editing caption of video
CN108650543A
Data processing method and device, electronic equipment and storage medium
CN111161737A