Information processing method, device, electronic device and storage medium
By combining user-provided pictures with personalized content in foreign language learning, the problem of boring text learning is solved, and more efficient learning results and increased user interest are achieved.
Patent Information
- Application Number
- CN202410528288.3
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2024-04-28
- Publication Date
- 2025-09-19
- Estimated Expiration
- 2044-04-28
AI Technical Summary
Simple text learning is abstract and boring in foreign language learning, which is not conducive to memory and application.
By combining user-provided images to generate personalized content, and utilizing server and client information processing methods to obtain and display descriptive information associated with the images, the personalization and diversity of the content can be improved.
The generated content is closer to users' lives, improves learning effects and interests, and optimizes users' experience.
Smart Images

Figure CN118535793B_ABST
Abstract
Description
Technical Field
[0001] The present disclosure relates to the field of computer technology, in particular to the field of artificial intelligence technologies such as large models and computer vision, and specifically to an information processing method, device, electronic device and storage medium. Background Art
[0002] When learning a foreign language, simply studying text is abstract and boring, making it difficult to memorize. Using images can more intuitively demonstrate the meaning and usage of words, establish connections between words and specific things, and improve learning efficiency. Summary of the Invention
[0003] The present disclosure aims to solve one of the technical problems in the related art at least to a certain extent.
[0004] To this end, the purpose of the present disclosure is to propose an information processing method, device, electronic device and storage medium that can generate personalized content based on pictures provided by users themselves, thereby improving the personalization and diversity of the generated content, making the content closer to users' lives, and providing conditions for improving learning effects and interests.
[0005] According to a first aspect of the present disclosure, there is provided an information processing method, which is executed by a server and includes:
[0006] Receive a content loading request sent by a client, wherein the loading request includes an identifier of the client;
[0007] Acquiring target content from a database associated with the identifier of the client, wherein the database includes a picture set associated with the client and picture description information generated based on the picture set;
[0008] The target content is sent to the client.
[0009] According to a second aspect of the present disclosure, there is provided an information processing method, which is executed by a client and includes:
[0010] In the case where it is detected that any description information in the display interface is selected or the first control in the display interface is triggered, determining that a content loading instruction is received;
[0011] Sending a content loading request to a server, wherein the loading request includes an identifier of the client;
[0012] receiving target content returned by the server, wherein the target content includes a picture associated with the client and picture description information generated by the server based on the picture;
[0013] The target content is displayed.
[0014] According to a third aspect of the present disclosure, there is provided an information processing device, configured in a server, comprising:
[0015] A first receiving module is configured to receive a content loading request sent by a client, wherein the loading request includes an identifier of the client;
[0016] an acquisition module, configured to acquire target content from a database associated with the identifier of the client, wherein the database includes a picture set associated with the client and picture description information generated based on the picture set;
[0017] The first sending module is configured to send the target content to the client.
[0018] According to a fourth aspect of the present disclosure, there is provided an information processing device, configured in a client, comprising:
[0019] A second receiving module is configured to determine that a content loading instruction has been received when detecting that any description information in the display interface is selected or the first control in the display interface is triggered;
[0020] A second sending module is configured to send a content loading request to a server, wherein the loading request includes an identifier of the client;
[0021] a third receiving module, configured to receive target content returned by the server, wherein the target content includes a picture associated with the client and picture description information generated by the server based on the picture;
[0022] A display module is configured to display the target content.
[0023] According to a fifth aspect of the present disclosure, there is provided an electronic device, including:
[0024] at least one processor; and
[0025] a memory communicatively connected to the at least one processor; wherein,
[0026] The memory stores instructions that can be executed by the at least one processor, and the instructions are executed by the at least one processor to enable the at least one processor to perform the information processing method as described in the first aspect or the third aspect.
[0027] According to a sixth aspect of the present disclosure, a non-transitory computer-readable storage medium storing computer instructions is provided, wherein the computer instructions are used to enable the computer to execute the information processing method as described in the first aspect or the third aspect.
[0028] According to a seventh aspect of the present disclosure, a computer program product is provided, comprising computer instructions, which, when executed by a processor, implement the steps of the information processing method as described in the first aspect or the third aspect.
[0029] The information processing method, device, electronic device, and storage medium provided by the present disclosure have the following beneficial effects:
[0030] In this disclosure, a content loading request is first received from a client. Then, based on the client identifier included in the loading request, the target content is retrieved from a database associated with the client identifier, and then the target content is sent to the client. Thus, by retrieving personalized content associated with the client on the server and sending it to the client for processing, the personalization and diversity of the content can be improved, providing conditions for improving effectiveness and interest, and optimizing the user experience.
[0031] It should be understood that the contents described in this section are not intended to identify the key or important features of the embodiments of the present disclosure, nor are they intended to limit the scope of the present disclosure. Other features of the present disclosure will become readily understood through the following description. BRIEF DESCRIPTION OF THE DRAWINGS
[0032] The above and / or additional aspects and advantages of the present disclosure will become apparent and readily understood from the following description of the embodiments in conjunction with the accompanying drawings, which are provided for a better understanding of the present solution and do not constitute a limitation of the present disclosure.
[0033] Figure 1a is a flowchart of an information processing method proposed according to an embodiment of the present disclosure;
[0034] Figure 1b A schematic diagram of a display interface of a client in the information processing method provided by an embodiment of the present disclosure;
[0035] Figure 1c This is a schematic diagram of another display interface of the client in the information processing method provided by an embodiment of the present disclosure;
[0036] Figure 2 is a flowchart of an information processing method according to another embodiment of the present disclosure;
[0037] Figure 3 is a flowchart of an information processing method proposed according to another embodiment of the present disclosure;
[0038] Figure 4 is a flowchart of an information processing method proposed according to another embodiment of the present disclosure;
[0039] Figure 5 is a flowchart of an information processing method proposed according to another embodiment of the present disclosure;
[0040] Figure 6 is a structural diagram of an information processing device proposed according to an embodiment of the present disclosure;
[0041] Figure 7 is a structural diagram of an information processing device proposed according to another embodiment of the present disclosure;
[0042] Figure 8 A block diagram of an exemplary electronic device suitable for implementing embodiments of the present disclosure is shown. DETAILED DESCRIPTION
[0043] The following description of exemplary embodiments of the present disclosure is made in conjunction with the accompanying drawings, including various details of the embodiments of the present disclosure to facilitate understanding. These details should be considered as merely exemplary. Therefore, those skilled in the art will recognize that various changes and modifications may be made to the embodiments described herein without departing from the scope and spirit of the present disclosure. Similarly, for the sake of clarity and conciseness, descriptions of well-known functions and structures are omitted in the following description.
[0044] The embodiments of the present disclosure relate to artificial intelligence technology fields such as large models and computer vision.
[0045] Artificial Intelligence (AI) is a new technical science that studies and develops theories, methods, technologies, and application systems for simulating, extending, and expanding human intelligence.
[0046] A large model refers to a machine learning model with a large number of parameters and complex structure. It can process massive amounts of data and complete various complex tasks, such as natural language processing, computer vision, and speech recognition.
[0047] Computer vision refers to the use of cameras and computers to replace the human eye to identify, track, and measure targets, and further perform graphic processing so that the computer processing becomes an image that is more suitable for human eye observation or transmission to instrument detection.
[0048] In the technical solutions disclosed herein, the collection, storage, use, processing, transmission, provision and disclosure of user personal information involved comply with the provisions of relevant laws and regulations and do not violate public order and good morals.
[0049] The following describes an information processing method, apparatus, electronic device, and storage medium according to embodiments of the present disclosure with reference to the accompanying drawings.
[0050] It should be noted that the executor of the information processing method of this embodiment is an information processing device, which can be implemented by software and / or hardware. The device can be configured in an electronic device, which may include but is not limited to a terminal, a server, etc.
[0051] This disclosure addresses the problem in foreign language learning or other application scenarios where textual content is abstract and boring, making it difficult to memorize and apply. By combining user images, this method generates personalized content for users. This generated content is more relevant to users' lives and more practical, thereby improving user interest and efficiency.
[0052] The information processing method, device, electronic device, and storage medium provided by the present disclosure are described in detail below with reference to the following contents and accompanying drawings.
[0053] Figure 1a It is a flowchart of an information processing method proposed according to an embodiment of the present disclosure.
[0054] like Figure 1a As shown, the information processing method, executed by the server, may include:
[0055] S101: Receive a content loading request sent by a client.
[0056] The loading request includes the client's identifier.
[0057] It should be noted that the identifier of the client may be an identity document (ID) of a logged-in user, or may be a device ID, etc., which is not limited in the present disclosure.
[0058] In the present disclosure, when a client has a learning need or other usage need, it can send a content loading request to the server to obtain the content it needs.
[0059] S102: Acquire target content from a database associated with the client's identifier.
[0060] The target content includes the picture associated with the client and the picture description information generated based on the picture.
[0061] Accordingly, the database may include a collection of images associated with the client and image description information generated based on the collection of images. The image description information may include characters and / or sentences describing entities in the image, the entity's font, and other information. The collection of images associated with the client may be images captured and stored by the user to whom the client belongs, or images selected and uploaded to the server by the user to whom the client belongs, though this disclosure does not limit this.
[0062] Optionally, the database may only include picture description information generated based on a certain type of language, or may also include picture description information generated based on several types of languages, which is not limited in the present disclosure.
[0063] It should be noted that when a client connects to the server for the first time or when a user registers an account, the server may first create a database associated with the client's ID. Based on the client's ID, the server then retrieves the associated image collection and identifies the images in the collection to obtain description information for each image. The image and its corresponding description information are then stored in the associated database.
[0064] It should be noted that, in the present disclosure, each content in the database is data corresponding to a picture and picture description information. Each picture in the picture set may correspond to multiple picture description information, and each picture description information may include word description information and sentence description information. Word description information may refer to words that describe entities or entity states contained in the picture. Sentence description information may be a sentence used to describe the content of the picture, or it may be a sentence corresponding to the word description information of the entity in the picture, used to explain the meaning and / or usage of the word. The present disclosure does not limit this.
[0065] In an embodiment of the present disclosure, the picture set associated with the client can be obtained from a storage system associated with the client after user authorization, for example, from a network disk, cloud disk or other storage system associated with the client; or, it can be obtained from a terminal where the client is located (such as a mobile phone, computer, etc.) after user authorization, and the present disclosure does not limit this.
[0066] It is understandable that since the picture sets associated with different clients may be different, after receiving a content loading request sent by any client, the database associated with the client identifier can be queried in the server based on the client identifier contained in the loading request, and then part of the content can be obtained from the associated database as the target content.
[0067] In some possible embodiments, the loading request may also include the client's indication of the content to be obtained, for example, the loading request may include an identifier of the target content. Therefore, when the server receives different data contained in the loading request, the target content obtained is also different.
[0068] Optionally, when the loading request includes an identifier of the target content, the server may obtain the target content corresponding to the identifier of the target content from the database.
[0069] The identifier of the target content may be an ID of description information, an identifier of a corresponding image, or any other identifier that can uniquely identify a content in the database, which is not limited in the present disclosure.
[0070] In some possible embodiments, the content loading request received by the server may include an image. The server may then traverse the database based on the image to retrieve the target content. It should be noted that if the server fails to retrieve the corresponding description information for the image after traversing the database, it may identify the image to generate its corresponding description information, including its display location, and use this as the target content. The detailed process for the server generating the image description information can be found in the detailed description of the following embodiments and is not further elaborated here.
[0071] In some possible embodiments, the content loading request received by the server may include the identifier of a certain image and / or the identifier of a certain descriptive information in the image. The server can then traverse the database to obtain the target content based on the received identifier of the image and / or the identifier of a certain descriptive information in the image.
[0072] Figure 1b and 1c A schematic diagram of the display interface of the client in the information processing method provided in an embodiment of the present disclosure.
[0073] For example, if the client triggers a content loading request when the display interface is shown in 1b, and the content currently displayed by the client is identified as "Picture #11", where "Picture #1" is the identifier of the currently displayed picture, and "Picture #11" indicates that the currently displayed description information is description information #1 associated with "Picture #1". Then the loading request received by the server may include "Picture #12", so the server can query the database based on "Picture #12" to obtain other description information associated with the picture, such as Figure 1c The target content shown includes descriptive information and picture #1 for explaining the grammar or usage of "grassland", etc., which is not limited in this disclosure.
[0074] It should be noted that the above description of the form and meaning of the content identification is only an illustrative description and cannot be used as a restrictive description of the solution provided by this disclosure.
[0075] In an embodiment of the present disclosure, if the loading request received by the server contains an identifier of the target content to define the current target learning content, the server can search the database for content that matches the identifier based on the identifier of the target content as the target content required by the client sending the request.
[0076] Alternatively, when the loading request does not include the identifier of the target content and the client's history record is not obtained, the server may determine any content in the database as the target content.
[0077] In the embodiment of the present disclosure, when the client requests to load content for the first time, the server does not have the client's historical record, so it is impossible to select the target content to be returned based on the historical record, and the loading request does not contain the identifier of the target content, so the server can obtain any content in the database and determine it as the target content.
[0078] Alternatively, when the loading request does not include the identifier of the target content, the server may obtain the target content from the database based on the client's historical records.
[0079] That is to say, when the client's history record can be obtained in the server based on the identifier of the client sending the request, and when the identifier of the target content to be obtained is not specifically indicated in the loading request, the server can obtain the content whose learning process is interrupted or has not yet been learned from the database as the target content based on the client's history record.
[0080] In the embodiment of the present disclosure, the target content in the database is obtained according to different situations based on whether the loading request contains the identifier of the target content and whether there is a historical record of the client requesting to load the content in the server, thereby improving the diversity and reliability of content acquisition and optimizing the user experience.
[0081] Optionally, the target content may include a picture, description information of the picture, and one or more position information. The position information is used to indicate the relative display position between the description information and the picture.
[0082] The location information may be the coordinates of the description information in the image coordinate system.
[0083] It is understood that the location information contained in the target content and the image description information can be in a one-to-one correspondence. In other words, each description information can be used to determine its relative display position with respect to the image through the corresponding location information. For example, whether the description information is displayed within the image or outside of the image, etc. This makes the client display of the target content clearer and easier to read, providing conditions for improving user learning efficiency.
[0084] S103: Send the target content to the client.
[0085] It should be noted that the target content retrieved by the server from the database may include one or more descriptive information. Therefore, if the target content includes multiple descriptive information, the multiple descriptive information can be sent to the client via a queue. The order of the multiple descriptive information in the queue can be randomly generated or arranged according to a specific rule, such as alphabetical order, difficulty level, etc., which is not limited in this disclosure.
[0086] In this embodiment, a content loading request is first received from a client. Based on the client identifier included in the loading request, the target content is retrieved from a database associated with the client identifier, and then the target content is sent to the client. Thus, by retrieving personalized content corresponding to the image associated with the client on the server and sending it to the client for display, the personalization and diversity of the content can be improved, the practicality of the content is enhanced, conditions are provided for improving learning effectiveness and efficiency, and the user experience is optimized.
[0087] Figure 2 It is a flowchart of an information processing method proposed in another embodiment of the present disclosure.
[0088] like Figure 2 As shown, the information processing method, executed by the server, may include:
[0089] S201: Receive a content loading request sent by a client.
[0090] The description of S201 can be found in the above embodiment and will not be repeated here.
[0091] S202: When the loading request does not include the identifier of the target content, determine the display count of each content in the database associated with the identifier of the client.
[0092] The number of times displayed refers to the number of times the corresponding content has been displayed historically on the display interface of the client, which can be stored in association with the content in the database as attribute information.
[0093] In an embodiment of the present disclosure, when the client's history record can be obtained in the server based on the identifier of the client sending the request, and when the identifier of the target content to be obtained is not specifically indicated in the loading request, the number of times each content in the database has been displayed can be determined first, and then the target content can be filtered and obtained based on the number of times it has been displayed.
[0094] S203: According to the historical records of the client, content corresponding to any historical record whose display times are less than a threshold is obtained from the database as target content.
[0095] The number threshold is a value used to determine whether any content needs to be displayed again, and can be customized according to user needs. For example, the number threshold can be 3 or 5, etc., and this disclosure does not limit this.
[0096] It should be noted that in learning scenarios, when users have low content retention rates, the number of times a piece of content can be displayed can be increased. When the number of times a piece of content has been displayed exceeds the number of times threshold, the client is considered to have completed learning the content and will not be displayed again.
[0097] In the embodiment of the present disclosure, the server can maintain and record the number of times all content has been displayed based on all historical records of the client. When receiving a content loading request sent by the client, the server can compare the relationship between the number of times each content has been displayed and the number threshold. If the number of times the content corresponding to any historical record has been displayed is less than the number threshold, it can be determined that the client has browsed the content and has not completed the learning of the content. Therefore, the content can be prioritized as the target content.
[0098] It should be noted that if the display counts for all content in the history records meet the display count threshold, no content can be retrieved from the history records as the target content. Instead, any content in the database with a display count of zero can be retrieved as the target content. In other words, any content that has not been viewed by the client can be used as the target content.
[0099] In the disclosed embodiment, by obtaining content that has been displayed less than a threshold number of times in the client's historical records as the target content to be waited for by the client, the continuity of the learning process, as well as the integrity and orderliness of the learning are ensured, further optimizing the user experience.
[0100] S204: Add 1 to the display count of the target content to obtain an updated display count corresponding to the target content.
[0101] It is understandable that when any content is determined as target content, it will be sent to the client for display, so its corresponding display count can be increased by 1, so the display count of the target content can be increased by 1, and the display count corresponding to the target content can be updated in the database.
[0102] S205: Send the target content and the number of times it has been displayed to the client.
[0103] In the embodiment of the present disclosure, the target content may be associated with its corresponding number of times displayed and sent to the client.
[0104] In the embodiment of the present disclosure, in a learning scenario, by updating the number of times the target content has been displayed in real time and sending the number of times displayed together with the target content to the client, the client can clearly and in real time obtain the learning progress of each content, which helps to enhance the user's learning motivation, improve learning efficiency, and optimize the user experience.
[0105] In this embodiment, by selecting content from the client's history records that has been displayed fewer than a threshold number of times and using it as the target content, the continuity of the learning process, as well as the integrity and coherence of learning, is ensured, further optimizing the user experience. Furthermore, by updating the display count corresponding to the target content and sending this count along with the target content to the client, the client can clearly and in real time access its learning progress, helping to enhance the user's learning motivation, improve learning efficiency, and optimize the user experience.
[0106] Figure 3 It is a flowchart of an information processing method proposed in another embodiment of the present disclosure.
[0107] like Figure 3 As shown, the information processing method, executed by the server, may include:
[0108] S301: Receive a content loading request sent by a client.
[0109] The description of S301 can be found in the above embodiment and will not be repeated here.
[0110] S302: When the loading request does not include the identifier of the target content, determine the first record in the history records that has the smallest difference with the current time.
[0111] In an embodiment of the present disclosure, in a learning scenario, the client's historical record can be obtained in the server based on the identifier of the client sending the request, and when the identifier of the target content to be obtained is not specifically indicated in the loading request, the last generated record can be obtained in the historical record to determine the target content based on the last generated record, thereby ensuring the correlation between this learning and the last learning, and improving the continuity and efficiency of learning.
[0112] It should be noted that each record in the historical record may correspond to a record generation time, so the record with the smallest difference between the record generation time corresponding to all records and the current time may be the last generated record, ie, the first record.
[0113] S303: Determine the first content corresponding to the first record.
[0114] In the embodiment of the present disclosure, the first content corresponding to the first record may be determined according to an identifier of the content included in the first record.
[0115] S304: Acquire a content belonging to the same content group as the first content from the database as the target content.
[0116] It should be noted that a content group can contain multiple contents. Multiple contents in the same content group can have a high degree of similarity or correlation. Alternatively, the server can randomly divide the images in the image collection and the generated image description information into different groups to maximize the diversity and flexibility of the content in each content group.
[0117] In the disclosed embodiment, all stored content can be grouped in advance in the database. The identifier of the content group to which each content belongs can then be associated and stored as the content's attribute information. After determining the first content corresponding to the first record, the identifier of the content group to which the first content belongs is used to search for another content associated with the content group identifier, which is then used as the target content.
[0118] In the embodiment of the present disclosure, by obtaining the content corresponding to the last record generated in the historical record and other content belonging to the same content group as the target content, a strong correlation between the current content and the previous content can be guaranteed, and the coherence between the displayed content can be improved, which provides conditions for improving the user's learning efficiency and interest and further optimizes the user experience.
[0119] It should be noted that whether two contents belong to the same content group can be determined based on the similarity between the images and / or description information contained in the two contents.
[0120] Optionally, a first similarity between a first picture in the first content and a second picture in the second content, and a second similarity between first description information of the first picture and second description information of the second picture may be determined first.
[0121] In the embodiment of the present disclosure, there are multiple ways to determine the first similarity and the second similarity. For example, the first similarity between the first image and the second image can be determined based on the type of entities contained in the first image and the second image (such as people, animals, objects, etc.) or the scenes shown in the images (such as schools, forests, etc.). In addition, the second similarity between the first description information and the second description information can be determined based on the semantics corresponding to the first description information and the second description information, and at least one of the verbs, nouns, etc. contained therein, etc., and the present disclosure does not limit this.
[0122] Then, when at least one of the first similarity and the second similarity is greater than the similarity threshold, it can be determined that the first content and the second content belong to the same content group.
[0123] The similarity threshold can be determined based on the accuracy requirements in actual applications, and this disclosure does not limit this.
[0124] That is, when the first similarity is greater than the similarity threshold and the second similarity is less than or equal to the similarity threshold, it can be determined that the first content and the second content belong to the same content group. Alternatively, when the second similarity is greater than the similarity threshold and the first similarity is less than or equal to the similarity threshold, it can also be determined that the first content and the second content belong to the same content group.
[0125] In the embodiment of the present disclosure, by grouping content containing similar or close images, or grouping content containing similar or close image description information, conditions are provided for increasing the display frequency of similar content within a certain period of time, thereby providing conditions for improving the user's learning efficiency.
[0126] In some embodiments, when determining the target content required by the client, the server may cyclically determine multiple content items within a content group as target content and return them to the client for display until all content items within the content group have been displayed a certain number of times, and then determine new target content from the new content group for display. This cyclical recycling of displayed content makes the user's learning process more efficient, maintaining the user's learning interest while improving learning efficiency.
[0127] It should be noted that in order to avoid a content group containing too much content, which would result in a long time interval between two displays of the same content, or repeated display of content with high similarity, which would affect the diversity of displayed content, the similarity threshold can be appropriately raised in the embodiment of the present disclosure, or the maximum number of contents contained in a content group can be set.
[0128] In the disclosed embodiment, upon receiving a content load request, the server identifies content from the same content group as the most recently displayed content as the target content and returns it to the client. This prevents repeated display of the same content while ensuring the relevance of consecutively displayed content, thus improving learning continuity and efficiency.
[0129] S305: Send the target content to the client.
[0130] The description of S305 can be found in the above embodiment and will not be repeated here.
[0131] In this embodiment, if the load request does not include the identifier of the target content, the system first determines the first record in the historical record with the smallest difference from the current time. It then determines the first content corresponding to the first record. Finally, it retrieves a piece of content from the database that belongs to the same content group as the first content and uses it as the target content. This ensures a strong correlation between the currently returned target content and the previously displayed first content, improving learning consistency and efficiency and further optimizing the user experience.
[0132] Figure 4 It is a flowchart of an information processing method proposed in another embodiment of the present disclosure.
[0133] like Figure 4 As shown, the information processing method, executed by the server, may include:
[0134] S401: Acquire a picture set associated with the client's identifier.
[0135] In the disclosed embodiment, before obtaining target content from a database associated with a client's identifier, the server must first generate content associated with the client based on the client's associated picture collection and store it in the database. The client's associated picture collection may be pictures taken and stored by the client's user, or pictures selected and uploaded to the server by the client's user, though this disclosure does not limit this.
[0136] It should be noted that the method for obtaining image sets is not fixed. Different image sets associated with the client's identifier can be obtained through different channels to increase the diversity of generated content. However, regardless of the acquisition method, it is carried out after confirming the user's consent, in compliance with relevant laws and regulations, and does not violate public order and good morals.
[0137] Optionally, the picture set may be obtained from a storage system associated with the client's identifier.
[0138] The storage system may refer to any network disk or cloud where images are stored.
[0139] In the embodiment of the present disclosure, the storage system may be bound to the client first, and then the associated storage system may be determined according to the identifier of the client, and the picture set may be obtained from the storage system.
[0140] Alternatively, you can also obtain the picture collection stored in the terminal where the client is located.
[0141] That is, a plurality of pictures stored in an album of a terminal (such as a smart phone, a tablet computer, etc.) where the client is located may also be obtained to determine a picture set associated with the identifier of the client.
[0142] In the disclosed embodiments, by acquiring a picture collection associated with a client identifier through multiple methods, the diversity of ways to acquire the picture collection can be increased, thereby providing conditions for increasing the diversity of generated content. Furthermore, the pictures in the acquired picture collection are not familiar to the user, thereby ensuring the user's interest in the content generated based on the picture collection, providing conditions for improving the user's learning interest and efficiency.
[0143] S402: Identify the pictures in the picture set to determine the entities contained in the pictures, the locations of the entities in the pictures, and the status information of the entities.
[0144] Among them, the status information may include the expression or action of the person, the shape of the object (such as color, size, etc.) and the weather condition (such as sunny, rainy, etc.), etc., and the present disclosure does not limit this.
[0145] In the disclosed embodiment, a large model capable of identifying and analyzing image content may be used to identify each image in the image set, determine the entities contained in the image, the location of the entity in the image, and the status information of the entity.
[0146] S403: Generate description information of the image based on the target language currently corresponding to the client and the entity and its status information.
[0147] The target language may be any language such as Chinese, English, Japanese, etc., which is not limited in this disclosure. The target language currently corresponding to the client may be a default language or one or more languages selected by the user as needed.
[0148] In the disclosed embodiment, a large model capable of translating and generating sentences can be used to generate words and / or sentences corresponding to the target language according to the current target language corresponding to the client, as well as entities and entity status information to obtain description information of the image.
[0149] It should be noted that in the present disclosure, the server may periodically obtain the picture set associated with the client and update the content in the database based on the newly obtained picture set. Alternatively, the server may monitor the picture set associated with the client and, upon determining that the picture set associated with the client has been updated, synchronously update the content in the database. This disclosure is not limited to this.
[0150] It should be noted that, in order to further improve the diversity and personalization of the content, the generated description information may also have different language styles to increase the interest of the content.
[0151] Alternatively, the client can determine the target style currently being used, and then generate image description information based on the target style, entities, and entity status information. For example, after determining the target style, a prompt corresponding to the target style can be determined. The determined prompt, entity, and entity status information can then be input into the macro model to generate the image description information.
[0152] The server can provide users with a variety of styles to choose from as needed, such as default style, humorous style, colloquial style, literary style, celebrity speaking style, etc. Users can modify the target style corresponding to the client through the client.
[0153] It should be noted that different styles result in different prompt information input into the macro model when generating description information. The style can be set by the user or automatically determined based on the type of images in the image collection. For example, if the image type is comics, the target style can be determined to be humorous, etc. This disclosure does not limit this.
[0154] In the disclosed embodiment, different styles are used to generate description information of different styles, which can further improve the diversity, personalization and fun of the content and optimize the experience.
[0155] It should be noted that in some embodiments, there may be a situation where the number of entities contained in an image is particularly large. Therefore, after the image is recognized, in order to avoid the situation where too much description information blocks each other when the content is displayed, the present disclosure can also limit the number of description information associated with each image.
[0156] Optionally, candidate description information that matches the entity and / or entity status information can be obtained first, and then, when the number of candidate description information of any picture is greater than a quantity threshold, the target description information of any picture can be determined from the candidate description information based on the difficulty level corresponding to each candidate description information of any picture and the frequency of appearance in the generated description information.
[0157] Among them, the quantity threshold can be pre-set, or can be determined based on the length of the candidate description information, the density requirements of the interface display, and the actual capability requirements, etc. For example, the quantity threshold can be 6, 10, etc., and this disclosure does not limit this.
[0158] The difficulty level can be divided into simple, complex, basic, advanced, etc., which is not limited in this disclosure. In other words, in the embodiment of the present disclosure, the server can generate description information of different difficulty levels based on the same image to meet the learning and usage needs of different users.
[0159] In the embodiment of the present disclosure, when the number of candidate description information of any picture is greater than a quantity threshold, it is necessary to filter description information corresponding to the quantity threshold from the candidate description information as target description information of the any picture.
[0160] It should be noted that each candidate description information can be divided into different difficulty levels according to the difficulty of the vocabulary it contains and / or the difficulty of the grammar used. For example, when the language corresponding to the candidate description information is English, it can be divided into different difficulty levels according to which test level the vocabulary it contains belongs to (such as level 4, level 6, etc.). In addition, the frequency of any candidate description information appearing in the description information of other generated pictures can be counted to determine the description information with an appearance frequency less than a certain frequency threshold as the target description information, thereby avoiding the repeated appearance of the same description information that the user is already familiar with, and providing conditions for saving user time.
[0161] In an embodiment of the present disclosure, after determining the difficulty level corresponding to each candidate description information of any picture and the frequency of occurrence in the generated description information, the description information corresponding to the quantity threshold can be filtered according to the preset filtering rules as the target description information of any picture.
[0162] It should be noted that the preset screening rules can be adjusted according to the user's capabilities, etc., and this disclosure does not limit this.
[0163] For example, when the quantity threshold is 6, the preset filtering rule may be 2 descriptions of simple difficulty and 4 descriptions of complex difficulty, and the descriptions must appear no more than 2 times. Then, from all candidate descriptions for any image, 2 descriptions of simple difficulty and 4 descriptions of complex difficulty, with an appearance frequency no more than 2, can be selected as the target description for that image.
[0164] In the disclosed embodiment, by screening all candidate description information of the image and selecting an appropriate number of description information as the target description information of the image, the amount of data displayed can be reduced, the visibility and simplicity of the content can be improved, and mutual occlusion can be avoided, thereby providing conditions for improving learning efficiency.
[0165] Optionally, after generating the description information of the image, the historical records of the client may be obtained, and the current capability level of the user to which the client belongs may be determined based on the historical records.
[0166] The level of user ability may correspond to the difficulty level of the description information, and may be divided into basic, advanced, and the like.
[0167] In the embodiment of the present disclosure, the current ability level of the user to which the client belongs may be determined based on the difficulty level corresponding to the description information of the completed learning included in the historical records.
[0168] Then, in the case that the user's current ability level does not match the difficulty level of the description information of any picture, the description information of the any picture can be updated to obtain updated description information.
[0169] It is understandable that if the user's current ability level is advanced, but the difficulty level of any image description is normal, it can be determined that the two do not match. In this case, the description of the image can be updated by replacing high-difficulty vocabulary or high-difficulty grammar, etc., to obtain an updated description. This makes the updated description more in line with the user's needs, avoids returning content that does not meet the user's ability, and avoids wasting the user's time.
[0170] It should be noted that the server can update the description information in real time, or it can collect statistics on the client's historical records after a certain period to determine the user's latest ability level, and then update the description information in the database before loading content next time.
[0171] Afterwards, the content corresponding to any image in the database may be updated based on the updated description information.
[0172] It should be noted that after the description information is updated, the corresponding number of times it has been displayed before the update needs to be reset.
[0173] In the embodiment of the present disclosure, whether to update the description information of any picture is determined by judging whether the user's current ability level matches the difficulty level of the description information. In this way, it can be ensured that the current target content matches the user's ability, avoiding the generation of meaningless content, improving content acquisition efficiency, and further optimizing the user experience.
[0174] Optionally, the server can also initiate an update of the description information corresponding to the image in any content in the database when it determines that the number of times any content in the database has been displayed has reached a threshold. In other words, when the number of times any content in the database has reached the threshold, the description information in the content is updated. By updating the content in the database in real time, the diversity and flexibility of the content in the database are improved.
[0175] S404: Based on the location of the entity in the picture, the description information is associated with the picture and stored in a database.
[0176] In the embodiment of the present disclosure, the display position of the description information corresponding to the entity in the picture can be determined according to the position of the entity in the picture, and the description information can be associated with the picture and stored in the database.
[0177] In this embodiment, a collection of images associated with the client's identifier is first obtained. The images in the collection are then identified to determine the entities contained in the images, their locations within the images, and their status information. Based on the client's current target language, descriptions of the images are generated based on the entities and their status information. The descriptions are then associated with the images and stored in a database based on the locations of the entities within the images. This allows for personalized content acquisition, improved practicality, and greater diversity in the content obtained, helping to increase the efficiency of content acquisition and optimizing the user experience.
[0178] Figure 5 It is a flowchart of an information processing method proposed in another embodiment of the present disclosure.
[0179] like Figure 5 As shown, the information processing method, executed by the client, may include:
[0180] S501: When it is detected that any description information in the display interface is selected or the first control in the display interface is triggered, it is determined that a content loading instruction is received.
[0181] Among them, the first control is a functional control for triggering the acquisition of new content, and can be a control named "Next" or "Start" in the display interface. The present disclosure does not limit the name of the first control.
[0182] The following combination Figure 1b Further explanation is given on how the client determines that it has received the content loading instruction. Figure 1b As shown, the display interface may include a picture and five description information corresponding to the picture, and the display position of each description information corresponds to the position of the corresponding entity in the picture.
[0183] In the embodiment of the present disclosure, if the user Figure 1b On the display interface shown, click any descriptive information on the picture or click the first control named "Start Disk" below the display interface, and the client can monitor that any descriptive information in the display interface is selected or the first control in the display interface is triggered, and then determine that the content loading instruction is received.
[0184] It should be noted that when a user logs into the client for the first time, a pop-up window can be displayed at the corresponding position of the first control to introduce the function of the first control, so as to optimize the user experience. Figure 1b In the process, a message "Click the word or button to start" can be popped up above the first control price "Start" to help users master the various functions of the client.
[0185] Optionally, when it is detected that any description information in the display interface is selected, the identifier of the target content to be loaded can be determined according to the selected description information and the identifier of the picture currently displayed in the display interface.
[0186] It should be noted that the target content identifier may be composed of the identifier of the selected description information and the identifier of the image, or may be determined in other ways, and the present disclosure does not limit this.
[0187] Then, a content loading request is sent to the server, wherein the content loading request may include an identifier of the target content.
[0188] In the embodiment of the present disclosure, by obtaining the identifier of the content corresponding to any descriptive information when any descriptive information is selected, and then sending a content loading request to the server based on the identifier of the content, the accuracy and efficiency of content acquisition can be improved, so that the target content requested to be loaded is more in line with user needs, and the user experience is further optimized.
[0189] Optionally, the client can also determine the identifier of the target content to be loaded based on the content when detecting that the first control is triggered and the content currently displayed on the display interface is any content. Then, it sends a content loading request to the server. The process and logic of the client determining the target content to be loaded based on the currently displayed content can refer to the above Figure 1b and Figure 1c The relevant description will not be repeated here.
[0190] It should be noted that when there is no content displayed in the current display interface of the client and the first control is detected to be triggered, a loading request can be directly sent to the server, so that the server can select a content from the database according to the preset logic and return it to the client.
[0191] In the embodiment of the present disclosure, when the content currently displayed on the display interface is any content, the first control price is triggered, indicating that the client needs to obtain and load the next content. The image identifier included in the identifier of the target content to be loaded may be the same as the image identifier corresponding to the any content, or may be different.
[0192] Therefore, when the content currently displayed on the client display interface is not the last content associated with its corresponding picture, the identifier of the target content to be loaded can be composed of the identifier of the currently displayed picture and the identifier of another associated description information. For example, the content identifiers associated with picture #1 are: picture #11, picture #12, picture #13... picture #17. Then when the content identifier currently displayed on the client display interface is "picture #13", if the user triggers the first control (for example, Figure 1c ), the client can determine that the target content is identified as "Picture #14".
[0193] That is, when the user Figure 1c When the first control is triggered in the interface shown, the identifier of the target content to be loaded can be determined for any content according to the content currently displayed on the display interface, and then a loading request is sent to the server.
[0194] It should be noted that if Figure 1c As shown in the figure, the number of times the description information has been displayed can also be displayed next to the word-type description information corresponding to the entity in the picture. Figure 1c The number threshold is 3 as an example. Each time the description information is displayed, a five-pointed star icon will light up. When all the five-pointed star icons are lit, that is, the number of times the description information has been displayed meets the number threshold, the description information will no longer be displayed.
[0195] It should be noted that in Figure 1c As can be seen in the figure, the display interface may also include controls for voice playback and adjustment controls for controlling whether the translation is displayed in the display interface, etc., which are not limited in this disclosure. In addition, in this disclosure, different voice styles can be selected for voice broadcasting of content, such as crisp female voice, young male voice, etc.
[0196] The voice style used for the broadcast content may be automatically determined by the system based on the attribute information of the user to which the client belongs, or automatically determined based on the content to be broadcast, or may be selected by the user, and this disclosure does not limit this.
[0197] In the embodiments of the present disclosure, the identifier of the target content may also be the identifier of any currently displayed content plus a special character to instruct the server to load content corresponding to other images associated with the content. For example, if the identifier of the currently displayed content is "Image #13" and the user triggers the first control, if the first control is a "next word" control, then the identifier of the target content determined by the client may be "Image #13+"; if the first control is a "previous word" control, then the identifier of the target content determined by the client may be "Image #13-", and so on. This disclosure does not limit this.
[0198] It should be noted that the above-mentioned method of determining the identifier of the target content by the client is only an illustrative description and cannot be used as a restrictive interpretation of the technical solution of the present disclosure.
[0199] In an embodiment of the present disclosure, by determining the identifier of the target content based on any content when the first control is triggered and the content currently displayed on the display interface is any content, and sending a content loading request to the server, the accuracy and efficiency of obtaining content can be improved, so that the target content requested to be loaded meets user needs, and the user experience is further optimized.
[0200] S502: Send a content loading request to the server.
[0201] The loading request includes the client's identifier.
[0202] S503: Receive the target content returned by the server.
[0203] The target content includes the pictures associated with the client and the picture description information generated by the server based on the pictures.
[0204] S504: Display target content.
[0205] In the embodiment of the present disclosure, the received target content can be Figure 1b or Figure 1c That is, the pictures associated with the client and the picture description information generated by the server based on the pictures are displayed in different positions in the display interface.
[0206] Optionally, when the description information is the name of the first entity in the picture and the target content includes the location information of the first entity, the name of the first entity may be displayed at the location of the first entity in the picture.
[0207] In an embodiment of the present disclosure, when the description information included in the target content is the name of the first entity in the picture and includes the location information of the first entity, the name of the first entity can be displayed at the corresponding position in the picture according to the location information of the first entity.
[0208] It should be noted that the name of one entity can be displayed in one picture, or the names of multiple first entities can be displayed at the same time, so that users can grasp the target learning content as a whole, better understand the internal connection between the descriptive information, and improve learning efficiency.
[0209] like Figure 1bAs shown, when the description information is the names of the first entities in the picture, such as "sky", "forest", "hike", "grassland", and "trail", according to the position information corresponding to each name of the first entity in the target content, the description information "sky" can be displayed at the position of "sky" in the picture, the description information "forest" can be displayed at the position of "forest" in the picture, and so on. Then, the client interface display diagram as shown in Figure 1b can be obtained.
[0210] Optionally, in order to improve the readability of the displayed content, the entities associated with the description information in the picture can also be marked. As shown in Figure 1b , an identifier (such as the circle shown in Figure 1b ) is added at the entity "sky" associated with "sky".
[0211] Optionally, when the description information includes the name of the second entity in the picture and the description statement composed of the name of the second entity, and the target content contains the position information of the second entity, the name of the second entity is displayed at the position where the second entity is located in the picture, and the description statement is displayed in a preset area of the display interface.
[0212] It should be noted that the preset area for displaying the description statement can be a fixed area in the display interface (such as below the picture), or it can be determined according to the content of the description statement to be actually displayed, etc. The present disclosure does not limit this.
[0213] In the embodiments of the present disclosure, by displaying the name of the second entity on the picture in the display interface and displaying the description statement composed of the name of the second entity at a preset position in the display interface, users can obtain more content at the same time, improving the richness of the content and the organization of the learning content display, and optimizing the user learning experience.
[0214] As shown in Figure 1c , when the description information includes the name "grassland" of the second entity "grassland" in the picture and the description statement composed of "grassland", the description information "grassland" can be displayed at the position of "grassland" in the picture, and the description statement can be displayed below the picture. Then, the client interface display diagram as shown in Figure 1c can be obtained.
[0215] In this embodiment, upon detecting that any descriptive information in the display interface has been selected or the first control in the display interface has been triggered, the system determines that a content loading instruction has been received, then sends a content loading request to the server. After receiving the target content generated based on the client-associated image from the server, the target content is then displayed. This allows for personalized content to be generated for the client's learning or use based on the client's associated image, improving the personalization and practicality of the content obtained, ensuring the user's interest in the content, and thus providing conditions for improving the user's learning efficiency and further optimizing the user experience.
[0216] Figure 6 It is a structural diagram of an information processing device proposed in one embodiment of the present disclosure.
[0217] like Figure 6 As shown, the information processing device 600, configured in a server, may include:
[0218] A first receiving module 601 is configured to receive a content loading request sent by a client, wherein the loading request includes an identifier of the client;
[0219] An acquisition module 602 is configured to acquire target content from a database associated with the client identifier, wherein the database includes a picture set associated with the client and picture description information generated based on the picture set;
[0220] The first sending module 603 is configured to send the target content to the client.
[0221] Optionally, the acquisition module 602 may be configured to:
[0222] When the loading request includes the identifier of the target content, the target content corresponding to the identifier of the target content is obtained from the database; or
[0223] If the loading request does not include the identifier of the target content and the client's history record is not obtained, any content in the database is determined as the target content; or
[0224] When the load request does not include the identifier of the target content, the target content is obtained from the database according to the client's history.
[0225] Optionally, the acquisition module 602 may be configured to:
[0226] Determine the number of times each content in the database has been displayed;
[0227] According to the historical records, the content corresponding to any historical record whose display times are less than the threshold is obtained from the database as the target content.
[0228] Optionally, the acquisition module 602 may be configured to:
[0229] Add 1 to the target content's display count to obtain the updated display count corresponding to the target content;
[0230] Send the target content and the number of times it has been displayed to the client.
[0231] Optionally, the acquisition module 602 may be configured to:
[0232] Determine the first record in the history that has the smallest difference with the current time;
[0233] Determining first content corresponding to the first record;
[0234] A content belonging to the same content group as the first content is obtained from the database as the target content.
[0235] Optionally, the acquisition module 602 may also be used to:
[0236] Determining a first similarity between a first picture in the first content and a second picture in the second content, and a second similarity between first picture description information of the first picture and second picture description information of the second picture;
[0237] When at least one of the first similarity and the second similarity is greater than a similarity threshold, it is determined that the first content and the second content belong to the same content group.
[0238] Optionally, the acquisition module 602 may also be used to:
[0239] Get the image set associated with the client's logo;
[0240] Identify images in the image collection to determine the entities contained in the images, the locations of the entities in the images, and the status information of the entities;
[0241] Generate image descriptions based on the target language currently supported by the client and the entity and its status information.
[0242] Based on the location of the entity in the image, the description information is associated with the image and stored in the database.
[0243] Optionally, the acquisition module 602 may be configured to:
[0244] Get the image set from the storage system associated with the client's ID; or,
[0245] Get the image collection stored in the terminal where the client is located.
[0246] Optionally, the acquisition module 602 may be configured to:
[0247] Determine the target style currently corresponding to the client;
[0248] Generate image description information based on the target style, entity and entity status information.
[0249] Optionally, the acquisition module 602 may be configured to:
[0250] Obtain candidate description information matching the entity and / or the entity's state information;
[0251] When the number of candidate description information of any picture is greater than the quantity threshold, the target description information of any picture is determined from the candidate description information according to the difficulty level corresponding to each candidate description information of any picture and the frequency of occurrence in the generated description information.
[0252] Optionally, the acquisition module 602 may also be used to:
[0253] Get the client's history;
[0254] Determine the current capability level of the client user based on historical records;
[0255] When the user's current ability level does not match the difficulty level of the description information of any picture, the description information of the picture is updated to obtain updated description information;
[0256] The content corresponding to any image in the database is updated based on the updated description information.
[0257] It should be noted that the above explanation of the information processing method is also applicable to the information processing device of this embodiment and will not be repeated here.
[0258] In this embodiment, a content loading request is first received from a client. Then, based on the client identifier included in the loading request, the target content is retrieved from a database associated with the client identifier, and the target content is then sent to the client. Thus, by retrieving personalized content associated with the client on the server and sending it to the client for processing, the personalization and diversity of the content can be enhanced, which provides conditions for improving effectiveness and interest, and optimizes the user experience.
[0259] Figure 7 It is a structural diagram of an information processing device proposed in another embodiment of the present disclosure.
[0260] like Figure 7 As shown, the information processing device 700, configured in the client, may include:
[0261] The second receiving module 701 is configured to determine that a content loading instruction has been received when detecting that any description information in the display interface is selected or the first control in the display interface is triggered;
[0262] The second sending module 702 is configured to send a content loading request to the server, wherein the loading request includes an identifier of the client;
[0263] The third receiving module 703 is configured to receive target content returned by the server, wherein the target content includes a picture associated with the client and picture description information generated by the server based on the picture;
[0264] The display module 704 is configured to display target content.
[0265] Optionally, the second sending module 702 may also be configured to:
[0266] When it is detected that any description information in the display interface is selected, the identifier of the target content to be loaded is determined according to the selected description information and the identifier of the image currently displayed in the display interface;
[0267] A content loading request is sent to the server, wherein the content loading request includes an identifier of the target content.
[0268] Optionally, the second sending module 702 may also be configured to:
[0269] When it is detected that the first control is triggered and the content currently displayed on the display interface is any content, determining the identifier of the target content to be loaded according to the any content;
[0270] A content loading request is sent to the server, wherein the content loading request includes an identifier of the target content.
[0271] It should be noted that the above explanation of the information processing method is also applicable to the information processing device of this embodiment and will not be repeated here.
[0272] In this embodiment, upon detecting that any descriptive information in the display interface has been selected or the first control in the display interface has been triggered, it is determined that a content loading instruction has been received, and then a content loading request is sent to the server. The target content is then received and displayed. Thus, by selecting descriptive information or triggering a control, a content loading request is sent to the server, and the target content is retrieved and displayed. This increases the diversity of content loading request transmission, improves information processing efficiency, and further optimizes the user experience.
[0273] According to an embodiment of the present disclosure, the present disclosure also provides an electronic device, a readable storage medium, and a computer program product.
[0274] Figure 8 A schematic block diagram of an example electronic device 800 that can be used to implement embodiments of the present disclosure is shown. The electronic device is intended to represent various forms of digital computers, such as laptop computers, desktop computers, workstations, personal digital assistants, servers, blade servers, mainframe computers, and other suitable computers. The electronic device can also represent various forms of mobile devices, such as personal digital assistants, cellular phones, smartphones, wearable devices, and other similar computing devices. The components shown herein, their connections and relationships, and their functions are provided as examples only and are not intended to limit the implementation of the present disclosure described and / or claimed herein.
[0275] like Figure 8 As shown, the device 800 includes a computing unit 801, which can perform various appropriate actions and processes according to a computer program stored in a read-only memory (ROM) 802 or a computer program loaded from a storage unit 808 into a random access memory (RAM) 803. Various programs and data required for the operation of the device 800 can also be stored in the RAM 803. The computing unit 801, the ROM 802, and the RAM 803 are connected to each other via a bus 804. An input / output (I / O) interface 805 is also connected to the bus 804.
[0276] Various components in device 800 are connected to I / O interface 805, including an input unit 806, such as a keyboard, mouse, etc.; an output unit 807, such as various types of displays, speakers, etc.; a storage unit 808, such as a magnetic disk, optical disk, etc.; and a communication unit 809, such as a network card, modem, wireless communication transceiver, etc. The communication unit 809 allows device 800 to exchange information / data with other devices via a computer network such as the Internet and / or various telecommunication networks.
[0277] The computing unit 801 can be a variety of general and / or special processing components with processing and computing capabilities. Some examples of the computing unit 801 include, but are not limited to, a central processing unit (CPU), a graphics processing unit (GPU), various dedicated artificial intelligence (AI) computing chips, various computing units that run machine model algorithms, digital signal processors (DSPs), and any appropriate processors, controllers, microcontrollers, etc. The computing unit 801 performs the various methods and processes described above, such as the information processing method. For example, in some embodiments, the information processing method can be implemented as a computer software program that is tangibly contained in a machine-readable medium, such as the storage unit 808. In some embodiments, part or all of the computer program can be loaded and / or installed on the device 800 via the ROM 802 and / or the communication unit 809. When the computer program is loaded into the RAM 803 and executed by the computing unit 801, one or more steps of the information processing method described above can be performed. Alternatively, in other embodiments, the computing unit 801 can be configured to perform the information processing method by any other appropriate means (e.g., by means of firmware).
[0278] Various embodiments of the systems and techniques described herein can be implemented in digital electronic circuit systems, integrated circuit systems, field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), system-on-chip systems (SOCs), programmable logic devices (CPLDs), computer hardware, firmware, software, and / or combinations thereof. These various embodiments can include being implemented in one or more computer programs that are executable and / or interpreted on a programmable system that includes at least one programmable processor, which can be a special purpose or general purpose programmable processor that can receive data and instructions from a storage system, at least one input device, and at least one output device, and transmit data and instructions to the storage system, the at least one input device, and the at least one output device.
[0279] The program code for implementing the method of the present disclosure can be written in any combination of one or more programming languages. These program codes can be provided to a processor or controller of a general-purpose computer, a special-purpose computer, or other programmable data processing device so that when the program code is executed by the processor or controller, the functions / operations specified in the flow chart and / or block diagram are implemented. The program code can be executed entirely on the machine, partially on the machine, as a stand-alone software package, partially on the machine and partially on a remote machine, or entirely on a remote machine or server.
[0280] In the context of the present disclosure, a machine-readable medium can be a tangible medium that can contain or store a program for use by or in conjunction with an instruction execution system, device or equipment. A machine-readable medium can be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium can include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, device or equipment, or any suitable combination of the foregoing. A more specific example of a machine-readable storage medium can include an electrical connection based on one or more lines, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.
[0281] To provide interaction with a user, the systems and techniques described herein can be implemented on a computer having: a display device (e.g., a CRT (cathode ray tube) or LCD (liquid crystal display) monitor) for displaying information to the user; and a keyboard and pointing device (e.g., a mouse or trackball) through which the user can provide input to the computer. Other types of devices can also be used to provide interaction with the user; for example, the feedback provided to the user can be any form of sensory feedback (e.g., visual feedback, auditory feedback, or tactile feedback); and input from the user can be received in any form (including acoustic input, voice input, or tactile input).
[0282] The systems and techniques described herein can be implemented in a computing system that includes back-end components (e.g., as a data server), or a computing system that includes middleware components (e.g., an application server), or a computing system that includes front-end components (e.g., a user computer with a graphical user interface or a web browser through which a user can interact with implementations of the systems and techniques described herein), or a computing system that includes any combination of such back-end components, middleware components, or front-end components. The components of the system can be interconnected by any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include: a local area network (LAN), a wide area network (WAN), the Internet, and a blockchain network.
[0283] A computer system may include a client and a server. The client and server are generally remote from each other and typically interact via a communication network. This client-server relationship is established by computer programs running on the respective computers, establishing a client-server relationship. The server may be a cloud server, also known as a cloud computing server or cloud host, a host product within the cloud computing service ecosystem that addresses the management difficulties and limited scalability of traditional physical hosts and VPS services ("Virtual Private Servers" or simply "VPS"). The server may also be a server in a distributed system or a server integrated with blockchain.
[0284] It should be understood that the various forms of the processes shown above can be used to reorder, add, or delete steps. For example, the steps described in this disclosure can be performed in parallel, sequentially, or in a different order, as long as the desired results of the technical solutions disclosed in this disclosure can be achieved. This is not a limitation herein.
[0285] In addition, the terms "first" and "second" are used for descriptive purposes only and should not be understood as indicating or implying relative importance or implicitly indicating the number of the indicated technical features. Thus, a feature defined as "first" or "second" may explicitly or implicitly include at least one such feature. In the description of the present disclosure, the meaning of "plurality" is at least two, such as two, three, etc., unless otherwise clearly and specifically defined. In the description of the present disclosure, the words "if" and "if" used can be interpreted as "at the time of" or "when" or "in response to a determination" or "under the circumstances of".
[0286] The above specific embodiments do not constitute a limitation on the scope of protection of this disclosure. Those skilled in the art will appreciate that various modifications, combinations, sub-combinations, and substitutions may be made based on design requirements and other factors. Any modifications, equivalent substitutions, and improvements made within the spirit and principles of this disclosure shall be included within the scope of protection of this disclosure.
Claims
1. An information processing method, characterized in that: The method is executed by a server, and includes: Receive a content loading request sent by a client, wherein the loading request includes an identifier of the client; Acquiring target content from a database associated with the identifier of the client, wherein the database includes a picture set associated with the client and picture description information generated based on the picture set; Sending the target content to the client; Before acquiring the target content from the database associated with the identifier of the client, the method further includes: Obtaining a picture set associated with the client's identifier; Identify the pictures in the picture set to determine entities contained in the pictures, positions of the entities in the pictures, and status information of the entities; Generate description information of the image based on the target language currently corresponding to the client and the entity and the state information of the entity; Based on the position of the entity in the picture, the description information is associated with the picture and stored in the database; After generating the description information of the picture according to the entity and the state information of the entity, the method further includes: Obtaining the client's history; Determining the current capability level of the user to which the client belongs based on the historical records; When the user's current ability level does not match the difficulty level of the description information of any picture, updating the description information of the any picture to obtain updated description information; The content corresponding to any picture in the database is updated based on the updated description information.
2. The method according to claim 1, wherein The acquiring target content from a database associated with the identifier of the client includes: In a case where the loading request includes an identifier of the target content, acquiring the target content corresponding to the identifier of the target content from the database; or In a case where the loading request does not include an identifier of the target content and the historical record of the client is not obtained, any content in the database is determined as the target content; or In a case where the loading request does not include an identifier of the target content, the target content is acquired from the database according to the historical records of the client.
3. The method according to claim 2, wherein: The acquiring target content from the database according to the historical records of the client includes: Determine the number of times each content in the database has been displayed; According to the historical records, content corresponding to any historical record whose display times are less than a threshold is obtained from the database as target content.
4. The method according to claim 3, wherein: The sending the target content to the client includes: Add 1 to the number of times the target content has been displayed to obtain an updated number of times the target content has been displayed; The target content and the number of times it has been displayed are sent to the client.
5. The method according to claim 2, wherein: The acquiring target content from the database according to the historical records of the client includes: Determine the first record in the historical records that has the smallest difference with the current time; determining first content corresponding to the first record; A content belonging to the same content group as the first content is acquired from the database as target content.
6. The method according to claim 5, wherein: Before acquiring a content belonging to the same content group as the first content from the database, the method further includes: Determining a first similarity between a first picture in the first content and a second picture in the second content, and a second similarity between first picture description information of the first picture and second picture description information of the second picture; When at least one of the first similarity and the second similarity is greater than a similarity threshold, it is determined that the first content and the second content belong to the same content group.
7. The method of claim 1, wherein: The method further comprises: Acquire the picture set from the storage system associated with the client's identifier; or, Obtain a picture set stored in the terminal where the client is located.
8. The method according to any one of claims 1 to 7, wherein: The target content includes a picture, description information of the picture, and one or more position information, where the position information is used to indicate a relative display position between the description information and the picture.
9. The method of claim 1, wherein: Generating description information of the image according to the entity and the state information of the entity includes: Determine the target style currently corresponding to the client; Generate description information of the image according to the target style, the entity, and the status information of the entity.
10. The method of claim 1, wherein: Generating description information of the image according to the entity and the state information of the entity includes: Acquire candidate description information matching the entity and / or the state information of the entity; When the number of candidate description information of any picture is greater than the quantity threshold, the target description information of any picture is determined from the candidate description information based on the difficulty level corresponding to each candidate description information of the picture and the frequency of appearance in the generated description information.
11. An information processing method, characterized in that: The method is executed by a client, and includes: In the case where it is detected that any description information in the display interface is selected or the first control in the display interface is triggered, determining that a content loading instruction is received; Sending a content loading request to a server, wherein the loading request includes an identifier of the client; Receiving target content returned by the server according to the information processing method according to any one of claims 1 to 10, wherein the target content includes a picture associated with the client and picture description information generated by the server based on the picture; The target content is displayed.
12. The method of claim 11, wherein: The sending of the content loading request to the server includes: When it is detected that any description information in the display interface is selected, the identifier of the target content to be loaded is determined according to the selected description information and the identifier of the picture currently displayed in the display interface; A content loading request is sent to the server, wherein the content loading request includes an identifier of the target content.
13. The method of claim 11, wherein: The sending of the content loading request to the server includes: When it is detected that the first control is triggered and the content currently displayed on the display interface is any content, determining an identifier of the target content to be loaded according to the any content; A content loading request is sent to the server, wherein the content loading request includes an identifier of the target content.
14. The method according to any one of claims 11 to 13, wherein: The displaying of the target content includes: When the description information is the name of the first entity in the picture and the target content includes the location information of the first entity, the name of the first entity is displayed at the location of the first entity in the picture.
15. The method according to any one of claims 11 to 13, wherein: The displaying of the target content includes: When the description information includes the name of the second entity in the picture and a description statement consisting of the name of the second entity, and the target content contains the location information of the second entity, the name of the second entity is displayed at the location of the second entity in the picture, and the description statement is displayed in a preset area of the display interface.
16. An information processing device, characterized in that: The device is configured on a server and includes: A first receiving module is configured to receive a content loading request sent by a client, wherein the loading request includes an identifier of the client; an acquisition module, configured to acquire target content from a database associated with the identifier of the client, wherein the database includes a picture set associated with the client and picture description information generated based on the picture set; A first sending module, configured to send the target content to the client; The acquisition module is further used to: Obtaining a picture set associated with the client's identifier; Identify the pictures in the picture set to determine entities contained in the pictures, positions of the entities in the pictures, and status information of the entities; Generate description information of the image based on the target language currently corresponding to the client and the entity and the state information of the entity; Based on the position of the entity in the picture, the description information is associated with the picture and stored in the database; Wherein, the acquisition module is further used to: Obtaining the client's history; Determining the current capability level of the user to which the client belongs based on the historical records; When the user's current ability level does not match the difficulty level of the description information of any picture, updating the description information of the any picture to obtain updated description information; The content corresponding to any picture in the database is updated based on the updated description information.
17. The apparatus of claim 16, wherein: The acquisition module is specifically used to: In a case where the loading request includes an identifier of the target content, acquiring the target content corresponding to the identifier of the target content from the database; or In a case where the loading request does not include an identifier of the target content and the historical record of the client is not obtained, any content in the database is determined as the target content; or In a case where the loading request does not include an identifier of the target content, the target content is acquired from the database according to the historical records of the client.
18. The apparatus of claim 17, wherein: The acquisition module is specifically used to: Determine the number of times each content in the database has been displayed; According to the historical records, content corresponding to any historical record whose display times are less than a threshold is obtained from the database as target content.
19. The apparatus of claim 18, wherein: The acquisition module is further used to: Add 1 to the number of times the target content has been displayed to obtain an updated number of times the target content has been displayed; The target content and the number of times it has been displayed are sent to the client.
20. The apparatus of claim 18, wherein The acquisition module is further used to: Determine the first record in the historical records that has the smallest difference with the current time; determining first content corresponding to the first record; A content belonging to the same content group as the first content is acquired from the database as target content.
21. The apparatus of claim 20, wherein: The acquisition module is further used to: Determining a first similarity between a first picture in the first content and a second picture in the second content, and a second similarity between first picture description information of the first picture and second picture description information of the second picture; When at least one of the first similarity and the second similarity is greater than a similarity threshold, it is determined that the first content and the second content belong to the same content group.
22. The apparatus of claim 16, wherein: The acquisition module is further used to: Acquire the picture set from the storage system associated with the client's identifier; or, Obtain a picture set stored in the terminal where the client is located.
23. The device according to any one of claims 16 to 22, wherein: The target content includes a picture, description information of the picture, and one or more position information, where the position information is used to indicate a relative display position between the description information and the picture.
24. The apparatus of claim 16, wherein: The acquisition module is further used to: Determine the target style currently corresponding to the client; Generate description information of the image according to the target style, the entity, and the status information of the entity.
25. The apparatus of claim 16, wherein: The acquisition module is further used to: Acquire candidate description information matching the entity and / or the state information of the entity; When the number of candidate description information of any picture is greater than the quantity threshold, the target description information of any picture is determined from the candidate description information based on the difficulty level corresponding to each candidate description information of the picture and the frequency of appearance in the generated description information.
26. An information processing device, characterized in that The device is configured on the client side and includes: A second receiving module is configured to determine that a content loading instruction has been received when detecting that any description information in the display interface is selected or the first control in the display interface is triggered; A second sending module is configured to send a content loading request to a server, wherein the loading request includes an identifier of the client; a third receiving module, configured to receive target content returned by the server according to the information processing method according to any one of claims 1 to 10, wherein the target content includes a picture associated with the client and picture description information generated by the server based on the picture; A display module is configured to display the target content.
27. The apparatus of claim 26, wherein: The second sending module is further configured to: When it is detected that any description information in the display interface is selected, the identifier of the target content to be loaded is determined according to the selected description information and the identifier of the picture currently displayed in the display interface; A content loading request is sent to the server, wherein the content loading request includes an identifier of the target content.
28. The apparatus of claim 26, wherein: The second sending module is further configured to: When it is detected that the first control is triggered and the content currently displayed on the display interface is any content, determining an identifier of the target content to be loaded according to the any content; A content loading request is sent to the server, wherein the content loading request includes an identifier of the target content.
29. The device according to any one of claims 26 to 28, wherein: The display module is further used for: When the description information is the name of the first entity in the picture and the target content includes the location information of the first entity, the name of the first entity is displayed at the location of the first entity in the picture.
30. The device according to any one of claims 26 to 28, wherein: The display module is further used for: When the description information includes the name of the second entity in the picture and a description statement consisting of the name of the second entity, and the target content contains the location information of the second entity, the name of the second entity is displayed at the location of the second entity in the picture, and the description statement is displayed in a preset area of the display interface.
31. An electronic device comprising: at least one processor; as well as a memory communicatively connected to the at least one processor; wherein, The memory stores instructions that can be executed by the at least one processor. The instructions are executed by the at least one processor to enable the at least one processor to perform the information processing method according to any one of claims 1 to 15.
32. A non-transitory computer-readable storage medium storing computer instructions, characterized in that: in, The computer instructions are used to enable the computer to execute the information processing method according to any one of claims 1 to 15.
33. A computer program product, characterized in that The invention comprises a computer program, which implements the steps of the information processing method according to any one of claims 1 to 15 when being executed by a processor.
Citation Information
Patent Citations
Picture translation method and system
CN104090871A