Comment information publishing method and device, equipment and storage medium
By allowing direct voice or text comment input on multimedia playback pages through specific operations, the method addresses the inefficiency of traditional comment publishing, enhancing interaction efficiency and user convenience.
Patent Information
- Application Number
- CN202510422169.4
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-04-03
- Publication Date
- 2025-07-15
- Estimated Expiration
- 2045-04-03
AI Technical Summary
In the prior art, users need to open the comment panel and perform multiple steps when posting comment information, resulting in inefficient human-computer interaction.
Set up a comment input area on the multimedia playback page. By displaying voice input controls or voice to text controls through preset operations on this area, users can directly enter voice or text comment information without opening the comment panel.
It simplifies the process of publishing comment information, improves human-computer interaction efficiency, provides diversified input selection, meets the comment habits of different users, and optimizes the operation experience.
Smart Images

Figure CN120321444A_ABST
Abstract
Description
Technical Field
[0001] The present disclosure relates to the field of computer technologies, and particularly to a method, apparatus, device, and storage medium for publishing comment information. Background Art
[0002] With the development of computer technologies, users can watch works (videos, pictures, articles, etc.) they are interested in through applications installed on terminals. During the process of watching a work, users can publish comment information on the work. Among them, the comment information can be voice, text, pictures, etc. Currently, when users publish comment information, they need to first open the comment panel, then input the comment information, and finally send the comment information. However, the operation of sending comment information in the above manner is relatively cumbersome, and the interaction link is relatively long, resulting in low human-computer interaction efficiency. Summary of the Invention
[0003] The present disclosure provides a method, apparatus, device, and storage medium for publishing comment information. Compared with the existing methods for publishing comment information, this solution does not require opening the comment panel, and can quickly input voice information and publish comment information through a control displayed on the multimedia playback page, with simple and convenient operation, shortening the interaction link of the process of publishing comment information, and improving the human-computer interaction efficiency.
[0004] According to one aspect of the embodiments of the present disclosure, a method for publishing comment information is provided, the method including:
[0005] Displaying a multimedia playback page, the multimedia playback page including a comment input area for inputting comment information to be published;
[0006] In response to a preset operation on the comment input area, displaying at least one of a voice input control or a voice-to-text control as a comment input control;
[0007] In response to a trigger operation on any one of the comment input controls, publishing comment information in a form corresponding to the comment input control.
[0008] According to another aspect of the embodiments of the present disclosure, a device for publishing comment information is provided, the device including:
[0009] A display unit configured to display a multimedia playback page, the multimedia playback page including a comment input area for inputting comment information to be published;
[0010] The display unit is further configured to, in response to a preset operation on the comment input area, display at least one of a voice input control or a voice-to-text control as a comment input control;
[0011] A publishing unit, configured to publish comment information in a corresponding form to the comment input control in response to a triggering operation on any of the comment input controls.
[0012] In some embodiments, the display unit is configured to, in response to a preset operation on the comment input area, display the voice input control and the voice-to-text control on the multimedia playback page when the multimedia playback page supports publishing voice comment information; and display the voice-to-text control on the multimedia playback page when the multimedia playback page does not support publishing voice comment information.
[0013] In some embodiments, the preset operation is a long-press operation or a preset gesture operation.
[0014] In some embodiments, the preset operation is the long-press operation;
[0015] The display unit is configured to, in response to a long-press operation on the comment input area, if the current user is a first type of user, display the voice input control in a triggered state and the voice-to-text control in an untriggered state on the multimedia playback page, where the first type of user is a user who prefers to publish voice comment information; if the current user is a second type of user, display the voice input control in an untriggered state and the voice-to-text control in a triggered state on the multimedia playback page, where the second type of user is a user who prefers to publish text comment information.
[0016] In some embodiments, the display unit is further configured to, in response to the position of the long-press operation moving to a position in a first area, set the voice input control to a triggered state and the voice-to-text control to an untriggered state, where the first area is the triggering area of the voice input control; and in response to the position of the long-press operation moving to a position in a second area, set the voice input control to an untriggered state and the voice-to-text control to a triggered state, where the second area is the triggering area of the voice-to-text control.
[0017] In some embodiments, the preset operation is the preset gesture operation;
[0018] The display unit is configured to, in response to a preset gesture operation on the comment input area, display at least one of the voice input control and the voice-to-text control at a preset position when the preset gesture operation is input completely.
[0019] In some embodiments, the display unit is further configured to display a first prompt message on the multimedia playback page, where the first prompt message is used to prompt the triggering of the preset operation; if the preset operation is not detected after the first prompt message is displayed for a preset duration, the display of the first prompt message is cancelled.
[0020] In some embodiments, the publishing unit is further configured to cancel the publishing of the voice comment information in response to a first cancel publishing operation based on the voice input control; and cancel the publishing of the text comment information in response to a second cancel publishing operation based on the voice-to-text control.
[0021] In some embodiments, the publishing unit is further configured to, in response to a swiping-up operation on the voice-to-text control, cancel the speech recognition of the input voice signal, display a second prompt message, where the second prompt message is used to prompt to cancel the publishing of the text comment information corresponding to the voice signal after the swiping-up operation ends; and cancel the publishing of the text comment information in response to the end of the swiping-up operation.
[0022] In some embodiments, the publishing unit is configured to, in response to a triggering operation on the voice input control, display a plurality of tone options, where each tone option corresponds to a voice tone; and in response to a selection operation on any one of the tone options, publish the voice comment information with the voice tone corresponding to the tone option.
[0023] In some embodiments, the publishing unit is configured to, in response to a triggering operation on the voice-to-text control, display the text corresponding to the voice during the process of inputting the voice signal; and publish the text comment information in response to the end of the voice input.
[0024] According to another aspect of the embodiments of the present disclosure, there is provided an electronic device, which includes:
[0025] One or more processors;
[0026] A memory for storing program code executable by the processor;
[0027] Wherein, the processor is configured to execute the program code to implement the above-mentioned comment information publishing method.
[0028] According to another aspect of the embodiments of the present disclosure, there is provided a computer-readable storage medium, when the instructions in the computer-readable storage medium are executed by the processor of the electronic device, enabling the electronic device to execute the above-mentioned comment information publishing method.
[0029] According to another aspect of the embodiments of the present disclosure, there is provided a computer program product, including a computer program, where the computer program implements the above-mentioned comment information publishing method when executed by a processor.
[0030] Embodiments of the present disclosure provide a comment information publishing solution. By setting a comment input area on the multimedia playback page, it is convenient for users to input the comment information to be published. When a user performs a preset operation on the comment input area, at least one of a voice input control or a voice-to-text control can be displayed. This display method provides users with diversified input options. That is, users can evoke the voice input control or the voice-to-text control in the multimedia playback page through the preset operation on the comment input area without opening the comment panel, realizing the dual use of one box (that is, the conventional operation of the comment input box realizes text input, and the preset operation realizes voice input). This solution is simple and convenient to operate, shortens the interaction link of the process of publishing comment information, and improves the human-computer interaction efficiency.
[0031] It should be understood that the above general description and the following detailed description are only exemplary and explanatory, and cannot limit the present disclosure.
[0032] By setting a comment input area on the multimedia playback page and displaying comment input controls such as voice input or voice-to-text according to the preset operation on this area, users can directly perform various ways of comment input through the comment input box. That is, users can evoke the voice input control or the voice-to-text control through the preset operation on the comment input area without opening the comment panel, realizing the dual use of one box (that is, the conventional operation of the comment input box realizes text input, and the preset operation realizes voice input), thus realizing the rapid input of voice information and the publishing of comment information. This solution is simple and convenient to operate, shortens the interaction link of the process of publishing comment information, and improves the human-computer interaction efficiency. BRIEF DESCRIPTION OF THE DRAWINGS
[0033] The drawings herein are incorporated into the specification and form a part of the specification, showing embodiments consistent with the present disclosure, and are used together with the specification to explain the principles of the present disclosure and do not constitute an improper limitation to the present disclosure.
[0034] Figure 1 is a schematic diagram of the implementation environment of a comment information publishing method shown according to an exemplary embodiment.
[0035] Figure 2 is a flowchart of a comment information publishing method shown according to an exemplary embodiment.
[0036] Figure 3 is a flowchart of another comment information publishing method shown according to an exemplary embodiment.
[0037] Figure 4 is a schematic diagram of a multimedia playback page provided according to an exemplary embodiment.
[0038] Figure 5 It is a schematic diagram of another multimedia playback page according to an exemplary embodiment.
[0039] Figure 6 It is a schematic diagram of another multimedia playback page provided according to an exemplary embodiment.
[0040] Figure 7 It is a schematic diagram of another multimedia playback page provided according to an exemplary embodiment.
[0041] Figure 8 It is a flowchart of a user publishing comment information according to an exemplary embodiment.
[0042] Figure 9 It is a block diagram of a comment information publishing device shown according to an exemplary embodiment.
[0043] Figure 10 It is a block diagram of an electronic device shown according to an exemplary embodiment. Detailed implementation manners
[0044] In order to enable those of ordinary skill in the art to better understand the technical solutions of the present disclosure, the technical solutions in the embodiments of the present disclosure will be clearly and completely described below with reference to the accompanying drawings.
[0045] It should be noted that the terms "first", "second", etc. in the specification and claims of the present disclosure and the above-mentioned drawings are used to distinguish similar objects, and do not necessarily need to describe a specific order or sequence. It should be understood that such data used can be interchanged under appropriate circumstances so that the embodiments of the present disclosure described herein can be implemented in an order other than those illustrated or described herein. The implementation manners described in the following exemplary embodiments do not represent all implementation manners consistent with the present disclosure. On the contrary, they are merely examples of devices and methods consistent with some aspects of the present disclosure as detailed in the appended claims.
[0046] It should be noted that the information (including but not limited to user device information, user personal information, etc.), data (including but not limited to data for analysis, stored data, displayed data, etc.) and signals involved in the present disclosure are all authorized by the user or fully authorized by all parties, and the collection, use and processing of relevant data need to comply with relevant laws, regulations and standards of relevant countries and regions. For example, the voice information involved in the present disclosure is obtained under full authorization.
[0047] Figure 1 It is a schematic diagram of the implementation environment of a comment information publishing method shown according to an exemplary embodiment. See Figure 1, the implementation environment specifically includes: a terminal 101 and a server 102. The terminal 101 can be connected to the server 102 through a wireless network or a wired network.
[0048] The terminal 101 can be at least one of devices such as a smart phone, a smart watch, a desktop computer, a laptop computer, an MP3 player (Moving Picture Experts Group Audio Layer III), an MP4 (Moving Picture Experts Group Audio Layer IV) player, and a laptop portable computer. An application program can be installed and run on the terminal 101. The application program is used to view multimedia resources such as videos, live broadcasts, etc. The application program is associated with the server 102, and the server 102 provides background services to the terminal 101.
[0049] The terminal 101 can generally refer to one of multiple terminals. In this embodiment, the terminal 101 is used as an example for illustration. Those skilled in the art can know that the number of the above terminals can be more or less. For example, the above terminals can be several, or the above terminals are dozens or hundreds, or a larger number. The embodiments of the present disclosure do not limit the number and device type of the terminals.
[0050] The server 102 is at least one of a server, multiple servers, a cloud computing platform, and a virtualization center. Optionally, the number of the above servers can be more or less, and the embodiments of the present disclosure do not limit this. Of course, the server 102 can also include other functional servers to provide more comprehensive and diverse services. In some embodiments, the server 102 undertakes the main computing work, and the terminal 101 undertakes the secondary computing work; or, the server 102 undertakes the secondary computing work, and the terminal 101 undertakes the main computing work; or, the server 102 and the terminal 101 adopt a distributed computing architecture for collaborative computing. The server 102 can be connected to the terminal 101 and other terminals through a wireless network or a wired network. Optionally, the number of the above servers can be more or less, and the embodiments of the present disclosure do not limit this.
[0051] Figure 2 is a flowchart of a method for publishing comment information shown according to an exemplary embodiment, as Figure 2 shown, the method is executed by an electronic device and includes the following steps:
[0052] In step S201, a multimedia playback page is displayed. The multimedia playback page includes a comment input area for inputting comment information to be published.
[0053] In an embodiment of the present disclosure, a user can browse multimedia resources through a multimedia playback page. When the user wants to post comment information about the multimedia resource, the comment information to be posted can be input through a comment input area displayed on the multimedia playback page. The comment input area can be used to input text comment information and can also be used to input voice comment information.
[0054] In step S202, in response to a preset operation on the comment input area, at least one of a voice input control or a voice-to-text control is displayed as a comment input control.
[0055] In an embodiment of the present disclosure, a user can input text by clicking on the comment input area and can also input a voice signal by performing a preset operation on the comment input area. The voice signal can be directly posted as voice comment information or can be posted as text comment information after voice-to-text conversion. Correspondingly, in response to the user's preset operation on the comment input area, at least one of the two comment input controls can be displayed, that is, the voice input control is displayed, or the voice-to-text control is displayed, or both the voice input control and the voice-to-text control are displayed.
[0056] In step S203, in response to a trigger operation on any one of the comment input controls, comment information in a form corresponding to the comment input control is posted.
[0057] In an embodiment of the present disclosure, a user can input voice information during the process of triggering the voice input control and then post the voice information as voice comment information on the multimedia playback page. The user can also input voice information during the process of triggering the voice-to-text control. The computer device can perform speech recognition on the input voice signal to obtain the text corresponding to the voice signal, and then post the text as text comment information on the multimedia playback page.
[0058] The embodiment of the present disclosure provides a comment information posting solution. By setting a comment input area on the multimedia playback page, it is convenient for users to input comment information to be posted. When the user performs a preset operation on the comment input area, at least one of a voice input control or a voice-to-text control is displayed as a comment input control. This display method provides users with diversified input options. That is, the user does not need to open the comment panel and can evoke the voice input control or the voice-to-text control on the multimedia playback page by performing a preset operation on the comment input area, realizing the dual use of one box (that is, the comment input box realizes text input through regular operations and voice input through preset operations). The solution is simple and convenient to operate, shortens the interaction link of the process of posting comment information, and improves the human-computer interaction efficiency.
[0059] In some embodiments, in response to a preset operation on a comment input area, at least one of a voice input control or a voice-to-text control is displayed. The comment input control includes:
[0060] In response to a preset operation on the comment input area, when the multimedia playback page supports publishing voice comment information, a voice input control and a voice-to-text control are displayed on the multimedia playback page;
[0061] When the multimedia playback page does not support publishing voice comment information, a voice-to-text control is displayed on the multimedia playback page.
[0062] In the embodiments of the present disclosure, by adapting the multimedia playback page according to different comment functions of the virtual live room, when voice comments are supported, a voice input control and a voice-to-text control are displayed, facilitating users to select voice input and convert it to text as needed; when voice comments are not supported, only the voice-to-text control is retained, which not only conforms to the actual functions of the live room but also provides users with flexible and diverse interaction controls that match the functions, improving the operation convenience and participation of users in the virtual live room.
[0063] In some embodiments, the preset operation is a long-press operation or a preset gesture operation.
[0064] In the embodiments of the present disclosure, the long-press operation on the comment input area can reduce the probability of misoperation, provide a thinking buffer time, and enhance the confirmability of the operation. The preset operation on the comment input area can increase the operation interest and personalization, improve the operation efficiency of proficient users, and enrich the interaction methods. The above two methods optimize the experience of comment input on the multimedia page from different perspectives, meet diverse needs, and improve the human-computer interaction efficiency.
[0065] In some embodiments, the preset operation is a long-press operation;
[0066] In response to a preset operation on a comment input area, at least one of a voice input control or a voice-to-text control is displayed. The comment input control includes:
[0067] In response to a long-press operation on the comment input area, if the current user is a first-type user, a voice input control in a triggered state and a voice-to-text control in an untriggered state are displayed on the multimedia playback page. The first-type user is a user who prefers to publish voice comment information;
[0068] If the current user is a second-type user, a voice input control in an untriggered state and a voice-to-text control in a triggered state are displayed on the multimedia playback page. The second-type user is a user who prefers to publish text comment information.
[0069] In the disclosed embodiment, for the first type of users who prefer to post voice comments, a long press will display a triggered voice input control and an untriggered voice-to-text control, allowing them to directly input comments by voice; for the second type of users who prefer text comments, a long press will present an untriggered voice input control and a triggered voice-to-text control, allowing them to quickly transcribe voice into text for commenting. This method not only fits the comment input habits of different users, reduces the number of operating steps, and improves input efficiency, but also optimizes the interactive experience of the multimedia playback page, meets the diverse needs of users, and improves the efficiency of human-computer interaction.
[0070] In some embodiments, the method further comprises:
[0071] In response to the position of the long press operation moving to a position in the first area, the voice input control is set to a triggered state, and the voice-to-text control is set to a non-triggered state, and the first area is a trigger area of the voice input control;
[0072] In response to the position of the long press operation moving to a position in the second area, the voice input control is set to an untriggered state, and the voice-to-text control is set to a triggered state, and the second area is a trigger area of the voice-to-text control.
[0073] In the embodiments of the present disclosure, by associating the position that responds to the long press operation with the triggering state of the voice input control and the voice-to-text control, that is, when the long press position is moved to the first area serving as the triggering area of the voice input control, the voice input control is automatically set to the triggered state and the voice-to-text control is automatically set to the untriggered state, and vice versa when moved to the second area serving as the triggering area of the voice-to-text control. This can more accurately adapt to the user's operating intentions, improve the efficiency of switching controls, improve the convenience of comment input, and improve the efficiency of human-computer interaction.
[0074] In some embodiments, the preset operation is a preset gesture operation;
[0075] Displaying at least one of a voice input control and a voice-to-text control on a multimedia playback page includes:
[0076] In response to a preset gesture operation on the comment input area, when the preset gesture operation is input, at least one comment input control of a voice input control and a voice-to-text control is displayed at a preset position.
[0077] In the disclosed embodiments, comment input is triggered by a preset gesture operation, which provides users with a novel and convenient way to start the operation, and gets rid of the reliance on traditional button clicks; and at the key node when the preset gesture operation is input, a voice input or text conversion control is presented on demand at a preset position, which not only meets the continuity of the operation, allowing users to choose the comment input form they like, but also saves operation steps and improves the efficiency of human-computer interaction.
[0078] In some embodiments, the method further includes:
[0079] Displaying a first prompt message on the multimedia playback page, the first prompt message being used to prompt the triggering of a preset operation;
[0080] If the preset operation is not detected after the first prompt message is displayed for a preset duration, cancel the display of the first prompt message.
[0081] In the embodiments of the present disclosure, displaying the first prompt message for prompting the triggering of the preset operation on the multimedia playback page can timely guide the user to understand how to enable comment input and lower the operation threshold. If the first prompt message is not used after being displayed for the preset duration, it will be automatically cancelled, avoiding information redundancy, ensuring that users who are new to this can easily get started, maintaining the page simplicity, and improving the human-computer interaction efficiency.
[0082] In some embodiments, the method further includes:
[0083] In response to a first cancel publishing operation based on the voice input control, cancel the publishing of the voice comment information;
[0084] In response to a second cancel publishing operation based on the voice-to-text control, cancel the publishing of the text comment information.
[0085] In the embodiments of the present disclosure, by setting the first cancel publishing operation based on the voice input control and the second cancel publishing operation based on the voice-to-text control, users can cancel the publishing immediately and conveniently during the input of voice comment information or text comment information, effectively avoiding problems caused by mispublishing comments and improving the human-computer interaction efficiency.
[0086] In some embodiments, canceling the publishing of the text comment information in response to the second cancel publishing operation based on the voice-to-text control includes:
[0087] In response to a swiping-up operation on the voice-to-text control, cancel the speech recognition of the input voice signal, and display a second prompt message, the second prompt message being used to prompt to cancel the publishing of the text comment information corresponding to the voice signal after the swiping-up operation ends;
[0088] In response to the end of the swiping-up operation, cancel the publishing of the text comment information.
[0089] In the embodiments of the present disclosure, by setting a swiping-up operation and the associated prompt message on the voice-to-text control, giving the user real-time and intuitive operation feedback, enabling the user to conveniently interrupt the speech recognition process during the conversion of voice to text comment, avoiding unnecessary text generation, and improving the human-computer interaction efficiency.
[0090] In some embodiments, in response to a triggering operation on any type of comment input control, comment information in a corresponding form to the comment input control is published, including:
[0091] In response to a triggering operation on the voice input control, multiple voice color options are displayed, and each voice color option corresponds to a voice timbre.
[0092] In response to a selection operation on any voice color option, voice comment information with the voice timbre corresponding to the voice color option is published.
[0093] In the embodiments of the present disclosure, by providing multiple voice color options after the user triggers the voice input control, the user can select their favorite voice timbre to publish voice comment information, which increases the personalization and interest of voice comments, enriches the diversity of user expressions, significantly improves the user's sense of participation and experience, and improves the efficiency of human-computer interaction.
[0094] In some embodiments, in response to a triggering operation on any type of comment input control, comment information in a corresponding form to the comment input control is published, including:
[0095] In response to a triggering operation on the voice-to-text control, the text corresponding to the voice is displayed during the input of the voice signal.
[0096] In response to the end of voice input, text comment information is published.
[0097] In the embodiments of the present disclosure, by displaying the corresponding text in real time when the user triggers the voice-to-text control to input voice, the text comment can be published immediately after the input ends, which greatly improves the accuracy and convenience of voice-to-text comments, reduces the cost of the user manually inputting text, and improves the efficiency of human-computer interaction.
[0098] The above Figure 2 shows a flowchart of a method for publishing comment information according to the present disclosure. The comment information publishing solution provided by the present disclosure will be further elaborated below. Figure 3 is a flowchart of another method for publishing comment information shown according to an exemplary embodiment. Refer to Figure 3 , this method is executed by an electronic device and includes the following steps:
[0099] In step S301, a multimedia playback page is displayed, and the multimedia playback page displays multimedia resources and a comment input area.
[0100] In the embodiments of the present disclosure, the multimedia playback page is a page for displaying multimedia resources. Multimedia resources include various forms, commonly live streams (such as virtual live rooms), videos (such as movies, TV series, short videos, etc.), audio (such as music, audiobooks, radio programs, etc.), and can also include picture sets, animations, and other different types of audio-visual materials.
[0101] For example, a video resource is played in a video playback area in the middle of a multimedia playback page. An audio resource can be displayed together with corresponding audio visualization elements (such as an audio waveform diagram, etc.) and playback control buttons, etc., so that users can intuitively see or hear the corresponding multimedia resource.
[0102] A comment input area is an interactive component. Users can express their opinions, feelings, suggestions, etc. about the multimedia resource being played through this comment input area. Optionally, the comment input area will have a specific visual presentation form on the multimedia playback page, such as a button or input box with the word "Comment". In response to a click operation on the comment input area, a text input cursor is displayed, and users can input text through a virtual keyboard or a physical keyboard, and then publish text comment information. In response to a preset operation on the comment input area, at least one comment input control can be displayed, such as a voice input control or a speech-to-text control. Users can input voice through the comment input control and then publish voice comment information or text comment information.
[0103] In some embodiments, the multimedia playback page is the live page of a virtual live room. Some virtual live rooms can publish voice comment information, while some virtual live rooms cannot publish voice comment information but can send text comment information. Correspondingly, in response to a preset operation on the comment input area, when the multimedia playback page supports publishing voice comment information, a voice input control and a speech-to-text control are displayed on the multimedia playback page. When the multimedia playback page does not support publishing voice comment information, a speech-to-text control is displayed on the multimedia playback page. By adapting the multimedia playback page according to different comment function settings for virtual live rooms, when voice comments are supported, a voice input control and a speech-to-text control are displayed to facilitate users to choose voice input and convert text as needed; when voice comments are not supported, only the speech-to-text control is retained, which not only conforms to the actual functions of the live room but also provides users with flexible and diverse interaction controls that match the functions, improving the operation convenience and sense of participation of users in the virtual live room.
[0104] For example, see Figure 4 as shown Figure 4 is a schematic diagram of a multimedia playback page provided according to an exemplary embodiment. As Figure 4 shown in (a) of, when the multimedia playback page supports publishing voice comment information, in response to a preset operation on the comment input area, such as a long press operation, a voice input control 401 and a speech-to-text control 402 are displayed. As Figure 4 shown in (b) of, when the multimedia playback page does not support publishing voice comment information, a speech-to-text control 402 is displayed.
[0105] Optionally, the voice input control and the voice-to-text control can also be displayed according to the user's usage habits. Accordingly, if it is determined based on the user's historical behavior information that the user is accustomed to sending voice comment information, the voice input control is displayed on the multimedia playback page in response to the preset operation of the comment input area. If it is determined based on the user's historical behavior information that the user is accustomed to sending text comment information, the voice-to-text control is displayed on the multimedia playback page in response to the preset operation of the comment input area. By judging the user's habits based on the user's historical behavior information, the voice input control or the voice-to-text control is displayed on the multimedia playback page in a targeted manner, so that the operation interface can dynamically adapt to the preferences of individual users, which not only meets the convenient input needs of users who are accustomed to voice communication, but also takes into account the needs of users who prefer text expression to quickly convert voice content, thereby improving the efficiency of human-computer interaction.
[0106] In some embodiments, when a user visits a multimedia playback page for the first time, or pushes the function provided by the present disclosure to the user for the first time, a prompt message can be used to prompt the user how to perform a preset operation on the comment input area. Accordingly, a first prompt message is displayed on the multimedia playback page, and the first prompt message is used to prompt the triggering of the preset operation. If the preset operation is still not detected after the first prompt message is displayed for a preset time, the first prompt message is canceled. That is, a first prompt message will be displayed on the multimedia playback page, and the function of this prompt message is to tell the user how to trigger the preset operation. If the user has not triggered the preset operation after the prompt message is displayed for a preset time, the system will cancel this prompt message and no longer display it. It should be noted that the first prompt message can be a streamer effect or other effect displayed on the comment input area to prompt the user to trigger the preset operation through the comment input area. Alternatively, the first prompt message can also be in the form of a small bubble or a bullet screen, prompting the user how to trigger the comment preset operation through text. Alternatively, a flashing exclamation mark icon is displayed on the multimedia playback page, and a text prompt message is displayed after the mouse hovers over it.
[0107] The following example uses the multimedia playback page to support the release of voice comment information, and the preset operation is a long press operation. According to the type of user, the multimedia playback page displays different states of voice input controls and voice-to-text controls to further shorten the interactive link of posting comment information, thereby improving the efficiency of human-computer interaction.
[0108] In step S302, in response to a long press operation on the comment input area, if the current user is a first type of user, a voice input control in a triggered state and a voice-to-text control in a non-triggered state are displayed on the multimedia playback page.
[0109] In the embodiments of the present disclosure, the user inputs text by clicking on the comment input area and inputs voice by long-pressing on the comment input area, and then at least one comment input control is displayed. Before displaying the comment input control, the type of the user can be determined first. Illustratively, the users are divided into two types. The first type of users are those who prefer to post voice comment information, and the second type of users are those who prefer to post text comment information. Optionally, in the case where the user authorizes the collection of the user's historical behavior data, the user can be classified based on the collected historical behavior data to determine the type to which the user belongs.
[0110] If the current user is a first-type user, that is, a user who prefers to post voice comment information, the computer device can display two comment input controls on the multimedia playback page, namely, a voice input control and a voice-to-text control. And since the first-type user prefers to post voice comment information, the computer device can set the voice input control to the triggered state. That is, when the user long-presses on the comment input area, it is equivalent to directly triggering the voice input control, and the user can directly input voice information without having to perform the operation of selecting the voice input control, shortening the interaction link. Optionally, the voice input control in the triggered state is highlighted, and the voice-to-text control in the un-triggered state is presented in a gray and non-clickable style.
[0111] For example, refer to Figure 5 shown in Figure 5 is a schematic diagram of another multimedia playback page according to an exemplary embodiment. In response to a long-press operation on the comment input area, a voice input control and a voice-to-text control are displayed. As shown in Figure 5 (a) therein, a voice input control 501 in the triggered state and a voice-to-text control 502 in the un-triggered state are displayed.
[0112] In some embodiments, when the user posts voice comment information through the voice input control, voice variation can also be performed. Correspondingly, in response to a triggering operation on the voice input control, a plurality of timbre options are displayed, and each timbre option corresponds to a voice timbre. In response to a selection operation on any one of the timbre options, a voice comment information with the voice timbre corresponding to the timbre option is posted. By providing a plurality of timbre options after the user triggers the voice input control, the user can select their favorite voice timbre to post voice comment information, increasing the personalization and interest of the voice comment, enriching the diversity of user expression, significantly enhancing the user's sense of participation and experience, and improving the human-computer interaction efficiency.
[0113] In step S303, in response to the position of the long-press operation moving to a position in the second area, the voice input control is set to the un-triggered state, and the voice-to-text control is set to the triggered state.
[0114] In an embodiment of the present disclosure, the multimedia playback page includes a first area and a second area. Among them, the first area is the triggering area of the voice input control, and the second area is the triggering area of the voice-to-text control. Since the user does not end the pressing after a long press operation, for example, the operation medium such as a finger (when using a touch screen device) or a mouse pointer (when operating on a computer) does not release, but remains in a long press state. Therefore, by moving the finger or the mouse pointer, the user can move the position where the long press operation is located from the original place to the second area, that is, the triggering area of the voice-to-text control.
[0115] When the position of the long press operation moves into the second area, it means that the voice-to-text control is triggered. That is, the voice input control, which was originally in a triggered state where it could receive the user's voice input and be used at any time, changes to an untriggered state because the long press operation enters the second area. Correspondingly, the voice-to-text control, which was previously in an untriggered state, is activated after the long press operation enters its dedicated triggering area (i.e., the second area) and then switches to a triggered state. At this time, the user can use the voice-to-text function to publish text comment information.
[0116] For example, continue to refer to Figure 5 , in response to a long press operation on the comment input area, the voice input control and the voice-to-text control are displayed. As shown in (b) of Figure 5 , the voice input control 501 in an untriggered state and the voice-to-text control 502 in a triggered state are displayed.
[0117] In step S304, in response to a second cancel publishing operation based on the voice-to-text control, the text comment information is cancelled from being published.
[0118] In an embodiment of the present disclosure, during the process of the user triggering the voice-to-text control to perform voice input and convert it to text, the user can also cancel the publication of the text comment information through a cancel publishing operation. For the sake of convenience in description, this cancel publishing operation is called the second cancel publishing operation. Optionally, the second cancel publishing operation can be a swiping up operation on the voice-to-text control, or it can be sliding the position of the long press operation to the cancel control and ending the pressing, such as letting go. The embodiment of the present disclosure does not limit this.
[0119] In some embodiments, the second cancellation and publication operation may cancel speech recognition without sending text comment information, or the second cancellation and publication operation may cancel sending the recognized text comment information. Correspondingly, in response to a swiping-up operation on the speech-to-text control, the speech recognition of the input speech signal is cancelled, and a second prompt message is displayed, which is used to prompt that the text comment information corresponding to the speech signal is cancelled after the swiping-up operation ends. In response to the end of the swiping-up operation, the text comment information is cancelled and published. By setting a swiping-up operation and the associated prompt message on the speech-to-text control, real-time and intuitive operation feedback is given to the user, enabling the user to conveniently interrupt the speech recognition process during the conversion of speech to text comments, avoiding unnecessary text generation, and improving the human-computer interaction efficiency.
[0120] For example, refer to Figure 6 as shown Figure 6 is a schematic diagram of another multimedia playback page provided according to an exemplary embodiment. As Figure 6 shown, a second prompt message 601 is displayed.
[0121] In step S305, in response to a long-press operation on the comment input area, if the current user is a second-type user, a voice input control in an untriggered state and a speech-to-text control in a triggered state are displayed on the multimedia playback page.
[0122] In the embodiments of the present disclosure, similar to the above step S302, if the current user is a second-type user, that is, the current user is a user who prefers to publish text comment information, voice input controls and speech-to-text controls in different states from those of the first-type users are displayed.
[0123] Since the current user is a second-type user, that is, a user who prefers to publish text comment information, the speech-to-text control can be set to the triggered state. That is, when the user long-presses the comment input area, it is equivalent to directly triggering the speech-to-text control. The user can directly input voice information, and correspondingly, the voice information is recognized in real time and the recognized text information is displayed, and the user does not need to perform the operation of selecting the speech-to-text control again, shortening the interaction link. Optionally, the speech-to-text control in the triggered state is highlighted, and the voice input control in the untriggered state is presented in a gray and non-clickable style.
[0124] For example, continue to refer to Figure 5 , as Figure 5 shown in (b) of, a voice input control 501 in an untriggered state and a speech-to-text control 502 in a triggered state are displayed.
[0125] In some embodiments, the user can preview the text information obtained by voice recognition in real time. Accordingly, in response to a trigger operation on the voice-to-text control, the text corresponding to the voice is displayed during the input of the voice signal. In response to the end of the voice input, the text comment information is published. By displaying the corresponding text in real time when the user triggers the voice-to-text control to input voice, the text comment can be published immediately after the input ends, which greatly improves the accuracy and convenience of voice-to-text comments, reduces the cost of the user manually inputting text, and improves the human-computer interaction efficiency.
[0126] Optionally, after the voice input ends, the user can also modify the text obtained by voice recognition. After the modification is completed, the text comment information is published through a publish operation. Optionally, in response to the end of the voice input, a modification prompt message is displayed. The modification prompt message is used to prompt whether to modify the text information obtained by voice-to-text. If not modified, it is directly published. If modified, in response to the publish operation, the modified text is published as the text comment information.
[0127] In step S306, in response to the position of the long-press operation moving to a position in the first area, the voice input control is set to the trigger state, and the voice-to-text control is set to the untriggered state.
[0128] In the embodiments of the present disclosure, similarly to the above step s303, this first area is the trigger area of the voice input control. Since the user triggers the voice-to-text control after a long-press operation and does not release it, but keeps the long-press state, the user can move the position where the long-press operation is located from the original place to the first area by moving the finger or the mouse pointer, that is, the trigger area of the voice input control.
[0129] When the position of the long-press operation moves to the first area, it means that the voice input control is triggered. That is, the voice-to-text control that was originally in a state where it could receive the user's voice input and perform voice recognition and text display will change because the long-press operation enters the first area, that is, it will become the untriggered state. Correspondingly, the voice input control that was previously in the untriggered state is activated after the long-press operation enters its exclusive trigger area (i.e., the first area). At this time, the user can use the voice input function to publish voice comment information.
[0130] For example, as shown in (a) of Figure 5 there is a voice input control 501 in the trigger state and a voice-to-text control 502 in the untriggered state displayed.
[0131] In step S307, in response to the first cancel publish operation based on the voice input control, the voice comment information is cancelled from being published.
[0132] In the embodiments of the present disclosure, similar to the above step S304, during the process of the user triggering the voice input control for voice input, the user can also cancel the publication of the voice comment information through a cancellation operation. For the convenience of description, this cancellation operation is referred to as the first cancellation operation. Optionally, the first cancellation operation can be a swiping-up operation on the voice input control, or can be sliding the position of the long-press operation to the cancellation control and then releasing the hand. The embodiments of the present disclosure do not limit this.
[0133] In some embodiments, the above step is described by taking the preset operation as a long-press operation as an example. Optionally, the preset operation can also be a preset gesture operation. Correspondingly, in response to the preset gesture operation on the comment input area, when the preset gesture operation is input completely, at least one of the voice input control and the voice-to-text control, i.e., the comment input control, is displayed at a preset position. Among them, the preset gesture operation can be an interactive gesture operation such as single-click / double-click / swiping-up / swiping-down / swiping-left / swiping-right / double-finger zooming / in two-finger long-press, etc. The embodiments of the present disclosure do not limit this. The preset position can be the end position of the preset gesture operation, or can be a preset fixed position on the page. The embodiments of the present disclosure do not limit this. It should be noted that the embodiments of the present disclosure do not limit the relative positions of the voice input control and the voice-to-text control. For example, the voice input control and the voice-to-text control can be arranged horizontally, or can be arranged vertically. Triggering the comment input through the preset gesture operation provides a novel and convenient operation start method for the user, getting rid of the dependence on the traditional button click; and at the key node when the preset gesture operation is input completely, presenting the voice input or text conversion control as needed at the preset position, which not only conforms to the operation coherence, enabling the user to conveniently select the desired comment input form, but also saves operation steps and improves the human-computer interaction efficiency.
[0134] For example, referring to Figure 7 as shown in Figure 7 is a schematic diagram of another multimedia playback page provided according to an exemplary embodiment. As Figure 7 (a) shows, the voice input control and the voice-to-text control are arranged horizontally. As Figure 7 (b) shows, the voice input control and the voice-to-text control are arranged vertically.
[0135] It should be noted that, in order to make the comment information publishing solution provided by the embodiments of the present disclosure easier to understand, referring to Figure 8 as shown in Figure 8 is a flowchart of a user publishing comment information provided according to an exemplary embodiment. As Figure 8As shown, it includes the following steps: 801. The user enters the virtual live room. 802. The user long-presses the comment input area. 803. Determine whether the virtual live room supports publishing voice comment information. If it supports, execute 804; if not, execute 805. 804. Determine whether the user prefers to use the voice-to-text function. If the user prefers it, execute 805; if the user does not prefer it, execute 806. 806. Trigger voice input while the user keeps long-pressing. 807. When the user stops long-pressing, publish the voice comment information. 808. When the user swipes up, cancel the publishing of the voice comment information. 805. Trigger voice-to-text while the user keeps long-pressing. 809. When the user stops long-pressing, end the voice-to-text. 810. Click to send to publish the text comment information. Optionally, if clicking to send voice, publish the voice comment information. 811. When the user swipes up, cancel the publishing of the text comment information.
[0136] Embodiments of the present disclosure provide a comment information publishing solution. By setting a comment input area on the multimedia playback page, it is convenient for users to input the comment information to be published. When the user performs a preset operation on the comment input area, at least one of a voice input control or a voice-to-text control can be displayed as a comment input control. This display method provides users with diversified input options. That is, users do not need to open the comment panel and can, through the preset operation on the comment input area, evoke the voice input control or the voice-to-text control in the multimedia playback page, realizing the dual use of one box (i.e., the conventional operation of the comment input box is used to input text, and the preset operation is used to input voice). This solution is simple and convenient to operate, shortens the interaction link of the process of publishing comment information, and improves the human-computer interaction efficiency.
[0137] Figure 9 is a block diagram of a comment information publishing device shown according to an exemplary embodiment. As Figure 9 shown, the device includes: a display unit 901 and a publishing unit 902.
[0138] The display unit 901 is configured to display a multimedia playback page, and the multimedia playback page includes a comment input area for inputting the comment information to be published;
[0139] The publishing unit 902 is configured to, in response to a preset operation on the comment input area, display at least one of a voice input control or a voice-to-text control as a comment input control;
[0140] The publishing unit 902 is further configured to, in response to a triggering operation on any one of the comment input controls, publish the comment information in the corresponding form of the comment input control.
[0141] In some embodiments, the display unit 901 is configured to, in response to a preset operation on the comment input area, when the multimedia playback page supports publishing voice comment information, display a voice input control and a voice-to-text control on the multimedia playback page; when the multimedia playback page does not support publishing voice comment information, display a voice-to-text control on the multimedia playback page.
[0142] In some embodiments, the preset operation is a long-press operation or a preset gesture operation.
[0143] In some embodiments, the preset operation is a long-press operation;
[0144] The display unit 901 is configured to, in response to a long-press operation on the comment input area, if the current user is a first type of user, display a voice input control in a triggered state and a voice-to-text control in an untriggered state on the multimedia playback page, where the first type of user is a user who prefers to publish voice comment information; if the current user is a second type of user, display a voice input control in an untriggered state and a voice-to-text control in a triggered state on the multimedia playback page, where the second type of user is a user who prefers to publish text comment information.
[0145] In some embodiments, the display unit 901 is further configured to, in response to the position of the long-press operation moving to a position in the first area, set the voice input control to the triggered state and set the voice-to-text control to the untriggered state, where the first area is the trigger area of the voice input control; in response to the position of the long-press operation moving to a position in the second area, set the voice input control to the untriggered state and set the voice-to-text control to the triggered state, where the second area is the trigger area of the voice-to-text control.
[0146] In some embodiments, the preset operation is a preset gesture operation;
[0147] The display unit 901 is configured to, in response to a preset gesture operation on the comment input area, when the preset gesture operation is input completely, display at least one of a voice input control and a voice-to-text control as a comment input control at a preset position.
[0148] In some embodiments, the display unit 901 is further configured to display a first prompt message on the multimedia playback page, where the first prompt message is used to prompt triggering the preset operation; if the preset operation is not detected after the first prompt message is displayed for a preset duration, cancel the display of the first prompt message.
[0149] In some embodiments, the publishing unit 902 is further configured to, in response to a first cancel publishing operation based on the voice input control, cancel the publishing of voice comment information; in response to a second cancel publishing operation based on the voice-to-text control, cancel the publishing of text comment information.
[0150] In some embodiments, the publishing unit 902 is further configured to cancel speech recognition of the input speech signal and display a second prompt message in response to a swiping-up operation on the speech-to-text control. The second prompt message is used to prompt canceling the publishing of the text comment information corresponding to the speech signal after the swiping-up operation ends. In response to the end of the swiping-up operation, the text comment information is not published.
[0151] In some embodiments, the publishing unit is configured to display multiple timbre options in response to a triggering operation on the voice input control, where each timbre option corresponds to a voice timbre; and in response to a selection operation on any one of the timbre options, publish a voice comment information with the voice timbre corresponding to the timbre option.
[0152] In some embodiments, the publishing unit is configured to display the text corresponding to the voice during the process of inputting the speech signal in response to a triggering operation on the speech-to-text control; and in response to the end of the speech input, publish the text comment information.
[0153] The embodiments of the present disclosure provide a comment information publishing device. By setting a comment input area on the multimedia playback page, it is convenient for users to input comment information to be published. When a user performs a preset operation on the comment input area, at least one of a voice input control or a speech-to-text control can be displayed. This display method provides users with diversified input options. That is, users do not need to open the comment panel and can evoke the voice input control or the speech-to-text control in the multimedia playback page through the preset operation on the comment input area, realizing the dual use of one box (i.e., the comment input box can be used to input text through regular operations and input voice through preset operations). This solution is simple and convenient to operate, shortens the interaction link of the process of publishing comment information, and improves the human-computer interaction efficiency.
[0154] It should be noted that for the comment information publishing device provided in the above embodiments, only the division of the above-mentioned functional units is used for illustration. In practical applications, the above functions can be allocated to different functional units according to needs, that is, the internal structure of the electronic device can be divided into different functional units to complete all or part of the functions described above. In addition, the comment information publishing device provided in the above embodiments and the embodiments of the comment information publishing method belong to the same concept, and the specific implementation process can be seen in the method embodiments, which will not be elaborated here.
[0155] Regarding the comment information publishing device in the above embodiments, the specific manners in which each module performs operations have been described in detail in the embodiments related to the method, and will not be elaborated here.
[0156] In the embodiments of the present disclosure, the electronic device may be a terminal or a server. When the electronic device is a terminal, the terminal serves as the execution entity to implement the technical solutions provided by the embodiments of the present disclosure; when the electronic device is a server, the server serves as the execution entity to implement the technical solutions provided by the embodiments of the present disclosure; or, the technical solutions provided by the present disclosure are implemented through the interaction between the terminal and the server. The embodiments of the present disclosure do not limit this.
[0157] Figure 10 It is a block diagram of an electronic device shown according to an exemplary embodiment. Generally, the electronic device 1000 includes: a processor 1001 and a memory 1002.
[0158] The processor 1001 may include one or more processing cores, such as a 4-core processor, an 8-core processor, etc. The processor 1001 may be implemented in at least one hardware form of DSP (Digital Signal Processing), FPGA (Field-Programmable Gate Array), and PLA (Programmable Logic Array). The processor 1001 may also include a main processor and a coprocessor. The main processor is a processor for processing data in the wake state, also known as the CPU (Central Processing Unit); the coprocessor is a low-power processor for processing data in the standby state. In some embodiments, the processor 1001 may be integrated with a GPU (Graphics Processing Unit), and the GPU is responsible for rendering and drawing the content to be displayed on the display screen. In some embodiments, the processor 1001 may further include an AI (Artificial Intelligence) processor, and the AI processor is used to process computational operations related to machine learning.
[0159] The memory 1002 may include one or more computer-readable storage media, and the computer-readable storage media may be non-transitory. The memory 1002 may further include high-speed random access memory and non-volatile memory, such as one or more disk storage devices and flash storage devices. In some embodiments, the non-transitory computer-readable storage media in the memory 1002 is used to store at least one program code, and the at least one program code is used to be executed by the processor 1001 to implement the comment information publishing method provided by the method embodiments in the present disclosure.
[0160] In some embodiments, the electronic device 1000 may further optionally include: a peripheral device interface 1003 and at least one peripheral device. The processor 1001, the memory 1002, and the peripheral device interface 1003 may be connected through a bus or signal lines. Each peripheral device may be connected to the peripheral device interface 1003 through a bus, signal lines, or a circuit board. Specifically, the peripheral device includes at least one of a radio frequency circuit 1004, a display screen 1005, a camera assembly 1006, an audio circuit 1007, and a power supply 1008.
[0161] The peripheral device interface 1003 can be used to connect at least one peripheral device related to I / O (Input / Output) to the processor 1001 and the memory 1002. In some embodiments, the processor 1001, the memory 1002, and the peripheral device interface 1003 are integrated on the same chip or circuit board; in some other embodiments, any one or two of the processor 1001, the memory 1002, and the peripheral device interface 1003 can be implemented on a separate chip or circuit board, and this embodiment does not limit this.
[0162] The radio frequency circuit 1004 is used to receive and transmit RF (Radio Frequency) signals, also known as electromagnetic signals. The radio frequency circuit 1004 communicates with a communication network and other communication devices through electromagnetic signals. The radio frequency circuit 1004 converts an electrical signal into an electromagnetic signal for transmission, or converts the received electromagnetic signal into an electrical signal. Optionally, the radio frequency circuit 1004 includes: an antenna system, an RF transceiver, one or more amplifiers, a tuner, an oscillator, a digital signal processor, a codec chipset, a subscriber identity module card, and so on. The radio frequency circuit 1004 can communicate with other electronic devices through at least one wireless communication protocol. The wireless communication protocol includes but is not limited to: a metropolitan area network, each generation of mobile communication networks (2G, 3G, 4G, and 5G), a wireless local area network, and / or a WiFi (Wireless Fidelity) network. In some embodiments, the radio frequency circuit 1004 may further include a circuit related to NFC (Near Field Communication), and this disclosure does not limit this.
[0163] The display screen 1005 is used to display the UI (User Interface). The UI may include graphics, text, icons, videos, and any combination thereof. When the display screen 1005 is a touch display screen, the display screen 1005 also has the ability to collect touch signals on or above the surface of the display screen 1005. The touch signals can be input to the processor 1001 as control signals for processing. At this time, the display screen 1005 can also be used to provide virtual buttons and / or virtual keyboards, also known as soft buttons and / or soft keyboards. In some embodiments, there may be one display screen 1005, which is provided on the front panel of the electronic device 1000; in other embodiments, there may be at least two display screens 1005, which are respectively provided on different surfaces of the electronic device 1000 or in a foldable design; in still other embodiments, the display screen 1005 may be a flexible display screen, which is provided on the curved surface or foldable surface of the electronic device 1000. Even, the display screen 1005 can also be set to an irregular non-rectangular shape, that is, a special-shaped screen. The display screen 1005 can be prepared from materials such as LCD (Liquid Crystal Display) and OLED (Organic Light-Emitting Diode).
[0164] The camera module 1006 is used to capture images or videos. Optionally, the camera module 1006 includes a front camera and a rear camera. Generally, the front camera is provided on the front panel of the electronic device, and the rear camera is provided on the back of the electronic device. In some embodiments, there are at least two rear cameras, which are any one of a main camera, a depth camera, a wide-angle camera, and a telephoto camera respectively, so as to realize the function of background blurring by fusing the main camera and the depth camera, panoramic shooting and VR (Virtual Reality) shooting functions or other fused shooting functions by fusing the main camera and the wide-angle camera. In some embodiments, the camera module 1006 may also include a flash. The flash can be a single-color temperature flash or a dual-color temperature flash. The dual-color temperature flash refers to the combination of a warm light flash and a cold light flash, which can be used for light compensation under different color temperatures.
[0165] The audio circuit 1007 may include a microphone and a speaker. The microphone is used to collect sound waves of the user and the environment, and convert the sound waves into electrical signals for input to the processor 1001 for processing, or input to the radio frequency circuit 1004 to enable voice communication. For the purpose of stereo collection or noise reduction, there may be multiple microphones, which are respectively arranged at different parts of the electronic device 1000. The microphone may also be an array microphone or an omnidirectional collection microphone. The speaker is used to convert the electrical signal from the processor 1001 or the radio frequency circuit 1004 into sound waves. The speaker may be a traditional thin film speaker or a piezoelectric ceramic speaker. When the speaker is a piezoelectric ceramic speaker, it can not only convert the electrical signal into sound waves audible to humans, but also convert the electrical signal into sound waves inaudible to humans for uses such as ranging. In some embodiments, the audio circuit 1007 may further include a headphone jack.
[0166] The power supply 1008 is used to supply power to each component in the electronic device 1000. The power supply 1008 may be alternating current, direct current, a primary battery or a rechargeable battery. When the power supply 1008 includes a rechargeable battery, the rechargeable battery may support wired charging or wireless charging. The rechargeable battery may also be used to support fast charging technology.
[0167] Those skilled in the art can understand that Figure 10 the structure shown in does not constitute a limitation on the electronic device 1000, and may include more or fewer components than shown in the figure, or combine certain components, or adopt different component arrangements.
[0168] In an exemplary embodiment, there is also provided a computer-readable storage medium including instructions, such as the memory 1002 including instructions, and the above instructions can be executed by the processor 1001 of the electronic device 1000 to complete the above comment information publishing method. Optionally, the computer-readable storage medium may be a ROM, a random access memory (RAM), a CD-ROM, a magnetic tape, a floppy disk, an optical data storage device, etc.
[0169] A computer program product includes a computer program, and when the computer program is executed by a processor, it implements the above comment information publishing method.
[0170] After considering the specification and practicing the invention disclosed herein, those skilled in the art will readily conceive of other embodiments of the present disclosure. The present disclosure is intended to cover any variations, uses, or adaptations of the present disclosure, which follow the general principles of the present disclosure and include known common knowledge or conventional technical means in the technical field not disclosed in the present disclosure. The specification and the embodiments are only regarded as exemplary, and the true scope and spirit of the present disclosure are pointed out by the following claims.
[0171] It should be understood that the present disclosure is not limited to the exact structures described above and shown in the drawings, and various modifications and changes can be made without departing from its scope. The scope of the present disclosure is only limited by the appended claims.
Claims
1. A method for publishing comment information, characterized in that, The method includes: Displaying a multimedia playback page, where the multimedia playback page includes a comment input area for inputting comment information to be published; In response to a preset operation on the comment input area, displaying at least one of a voice input control or a voice-to-text control as a comment input control; In response to a trigger operation on any one of the comment input controls, publishing comment information in a corresponding form to the comment input control.
2. The comment information publishing method according to claim 1, wherein The step of, in response to a preset operation on the comment input area, displaying at least one of a voice input control or a voice-to-text control as a comment input control includes: In response to a preset operation on the comment input area, when the multimedia playback page supports publishing voice comment information, displaying the voice input control and the voice-to-text control on the multimedia playback page; When the multimedia playback page does not support publishing voice comment information, displaying the voice-to-text control on the multimedia playback page.
3. The comment information publishing method according to claim 1, wherein The preset operation is a long-press operation or a preset gesture operation.
4. The comment information publishing method according to claim 3, wherein The preset operation is the long-press operation; The step of, in response to a preset operation on the comment input area, displaying at least one of a voice input control or a voice-to-text control as a comment input control includes: In response to a long-press operation on the comment input area, if the current user is a first type of user, displaying the voice input control in a triggered state and the voice-to-text control in an untriggered state on the multimedia playback page, where the first type of user is a user who prefers to publish voice comment information; If the current user is a second type of user, displaying the voice input control in an untriggered state and the voice-to-text control in a triggered state on the multimedia playback page, where the second type of user is a user who prefers to publish text comment information.
5. The comment information publishing method according to claim 4, wherein The method further includes: In response to the position of the long-press operation moving to a position in the first area, setting the voice input control to a triggered state and setting the voice-to-text control to an untriggered state, where the first area is the trigger area of the voice input control; In response to the position of the long-press operation moving to a position in the second area, setting the voice input control to an untriggered state and setting the voice-to-text control to a triggered state, where the second area is the trigger area of the voice-to-text control.
6. The comment information publishing method according to claim 3, wherein The preset operation is the preset gesture operation; The step of, in response to a preset operation on the comment input area, displaying at least one of a voice input control or a voice-to-text control as a comment input control includes: In response to a preset gesture operation on the comment input area, when the preset gesture operation is input completely, displaying at least one of the voice input control and the voice-to-text control at a preset position.
7. The comment information publishing method according to claim 1, wherein The method further includes: Displaying a first prompt message on the multimedia playback page, where the first prompt message is used to prompt triggering the preset operation; If the preset operation is not detected after the first prompt message is displayed for a preset duration, canceling the display of the first prompt message.
8. The comment information publishing method according to any one of claims 1-7, characterized in that, The method further includes: Upon receiving a first unpublishing operation based on the voice input control, unpublish the voice comment information; Upon receiving a second unpublishing operation based on the voice-to-text control, unpublish the text comment information.
9. The comment information publishing method according to claim 8, wherein The step of, upon receiving a second unpublishing operation based on the voice-to-text control, unpublishing the text comment information includes: Upon receiving a swiping-up operation on the voice-to-text control, cancel the voice recognition of the input voice signal, and display a second prompt message for prompting to unpublish the text comment information corresponding to the voice signal after the swiping-up operation ends; Upon the swiping-up operation ending, unpublish the text comment information.
10. The method for publishing comment information according to any one of claims 1-7, characterized in that The step of, upon receiving a triggering operation on any comment input control, publishing comment information in a form corresponding to the comment input control includes: Upon receiving a triggering operation on the voice input control, display a plurality of timbre options, each timbre option corresponding to a voice timbre; Upon receiving a selection operation on any timbre option, publish voice comment information with the voice timbre corresponding to the timbre option.
11. The method for publishing comment information according to any one of claims 1-7, characterized in that The step of, upon receiving a triggering operation on any comment input control, publishing comment information in a form corresponding to the comment input control includes: Upon receiving a triggering operation on the voice-to-text control, display the text corresponding to the voice during the process of inputting the voice signal; Upon the voice input ending, publish the text comment information.
12. A comment information publishing device, characterized in that The device includes: A display unit configured to display a multimedia playback page, the multimedia playback page including a comment input area for inputting comment information to be published; The display unit is further configured to, upon receiving a preset operation on the comment input area, display at least one of a voice input control or a voice-to-text control as a comment input control; A publishing unit configured to, upon receiving a triggering operation on any comment input control, publish comment information in a form corresponding to the comment input control.
13. An electronic device, characterized in that, The electronic device includes: One or more processors; A memory for storing program code executable by the processor; Wherein, the processor is configured to execute the program code to implement the comment information publishing method according to any one of claims 1 to 11.
14. A computer-readable storage medium, characterized in that, When the instructions in the computer-readable storage medium are executed by the processor of the electronic device, the electronic device is enabled to execute the comment information publishing method according to any one of claims 1 to 11.
15. A computer program product, including a computer program which, when executed by a processor, implements the comment information publishing method according to any one of claims 1 to 11.
Citation Information
Patent Citations
Live client voice input method and terminal device
CN106648535A
Voice comment display method and system, medium and electronic equipment
CN110377842A
Comment method and device, terminal equipment and computer storage medium
CN111128204A
Comment information display method and device, comment information release method and device, computer equipment and medium
CN117111788A
Information input method and device, electronic equipment and storage medium
CN118018794A