Information processing method and device, equipment, medium and product

By receiving the user's initial question information, the system uses NLP models and multimodal large models to retrieve matching target images from the image library and extracts answer information based on the question-and-answer intent. This solves the problem of low information acquisition efficiency in existing technologies and achieves intuitive information acquisition and efficient operation convenience.

CN121327166APending Publication Date: 2026-01-13VIVO MOBILE COMM CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202511394463.5
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-09-26
Publication Date
2026-01-13

AI Technical Summary

Technical Problem

In existing technologies, the large number of images leads to low information retrieval efficiency.

Method used

By receiving the user's initial question information, the system uses NLP models and multimodal large models to retrieve matching target images from the image library, extracts answer information based on the question-and-answer intent, and displays the images and answer information.

Benefits of technology

It enables users to intuitively receive answers corresponding to the questions, improves information acquisition efficiency, reduces information conversion costs, supports the combined use of scenario-based functions, and enhances ease of operation and practicality.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN121327166A_ABST
    Figure CN121327166A_ABST
Patent Text Reader

Abstract

The invention discloses an information processing method and device, equipment, a medium and a product, and relates to the field of information processing.The method comprises the steps that first input is received; the first input is used for inputting first question information; in response to the first input, displaying first answer information and at least one picture; wherein the first answer information is determined from at least one picture according to the first question information.
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] The present application belongs to the technical field of information processing, and particularly relates to an information processing method, device, equipment, medium and product. BACKGROUND

[0002] With the development of digital image technology, pictures have become an important carrier for users to store and record information, and picture search function has become one of the core functions of various terminal devices and applications.

[0003] At present, when a user wants to obtain some information, the user can obtain the information by searching for related pictures. However, due to a large number of pictures, the information obtaining efficiency is low. SUMMARY

[0004] Embodiments of the present application provide an information processing method, device, equipment, medium and product, and improve information obtaining efficiency.

[0005] In a first aspect, an information processing method is provided, comprising:

[0006] receiving a first input; the first input is used to input first question information;

[0007] in response to the first input, displaying first answer information and at least one picture; wherein the first answer information is determined from the at least one picture according to the first question information.

[0008] In a second aspect, an information processing device is provided, comprising:

[0009] a receiving module configured to receive a first input; the first input is used to input first question information;

[0010] a display module configured to, in response to the first input, display first answer information and at least one picture; wherein the first answer information is determined from the at least one picture according to the first question information.

[0011] In a third aspect, an electronic device is provided, which comprises a processor and a memory. The memory stores programs or instructions executable on the processor. When the programs or instructions are executed by the processor, the steps of the method according to the first aspect are implemented.

[0012] In a fourth aspect, a readable storage medium is provided. The readable storage medium stores programs or instructions. When the programs or instructions are executed by a processor, the steps of the method according to the first aspect are implemented.

[0013] In a fifth aspect, an embodiment of the present application provides a computer program product stored in a storage medium, which is executed by at least one processor to implement the steps of the method according to the first aspect.

[0014] In the embodiment of the present application, after receiving the first input of the first question information, the first answer information corresponding to the first question information determined from the picture can be displayed in addition to the picture. In this way, the user can be intuitively provided with the answer corresponding to the first question information, and the information acquisition efficiency is improved. BRIEF DESCRIPTION OF DRAWINGS

[0015] Figure 1 A flowchart of the information processing method provided by the embodiment of the present application is provided.

[0016] Figure 2 A first interface schematic diagram in response to the first input provided by the embodiment of the present application is provided.

[0017] Figure 3 A second interface schematic diagram in response to the first input provided by the embodiment of the present application is provided.

[0018] Figure 4 A third interface schematic diagram in response to the first input provided by the embodiment of the present application is provided.

[0019] Figure 5 A fourth interface schematic diagram in response to the first input provided by the embodiment of the present application is provided.

[0020] Figure 6 A structural schematic diagram of the information processing apparatus provided by the embodiment of the present application is provided.

[0021] Figure 7 A structural schematic diagram of the electronic device provided by the embodiment of the present application is provided.

[0022] Figure 8 A hardware structural schematic diagram of the electronic device provided by the embodiment of the present application is provided. DETAILED DESCRIPTION

[0023] The technical solutions in the embodiments of the present application will be described clearly and completely below with reference to the accompanying drawings in the embodiments of the present application. Obviously, the described embodiments are only some of the embodiments of the present application, but not all the embodiments of the present application. Based on the embodiments in the present application, all other embodiments obtained by those of ordinary skill in the art belong to the scope of protection of the present application.

[0024] The terms "first", "second", etc. in the specification of this application are used to distinguish similar objects, rather than to describe a specific order or sequence. It should be understood that the data used in this way can be interchanged under appropriate circumstances, so that the embodiments of this application can be implemented in an order other than those illustrated or described herein, and the objects distinguished by "first", "second", etc. are usually of the same category, and the number of objects is not limited. For example, the first object can be one or more. In addition, "and / or" in the specification means at least one of the connected objects, and the character "and / or" generally indicates that the associated objects before and after are in an "or" relationship.

[0025] The following will combine the accompanying drawings and, through specific embodiments and their application scenarios, provide a detailed description of the information processing method, device, equipment, medium, and product provided by the embodiments of this application.

[0026] Figure 1 FIG. is a flowchart of an information processing method provided by an embodiment of this application. This information processing method can be applied to electronic devices such as mobile phones, computers, and tablets.

[0027] As Figure 1 shown, this information processing method may include the following steps:

[0028] S110. Receive a first input; the first input is used to input first question information;

[0029] S120. In response to the first input, display first answer information and at least one picture; wherein, the first answer information is determined from at least one picture according to the first question information.

[0030] In the embodiments of this application, after receiving the first input for inputting the first question information, in addition to displaying pictures, the first answer information corresponding to the first question information determined from the pictures can also be displayed. In this way, the answer corresponding to the first question information can be intuitively provided to the user, improving the information acquisition efficiency.

[0031] The above steps will be described in detail as follows:

[0032] In S110, the execution subject may be an electronic device. Taking the electronic device as a mobile phone and searching for information on the mobile phone as an example, when the user needs some information, the user can input the first question information on the mobile phone, and the mobile phone determines the first question information input by the user in response to the input of the user's first question information.

[0033] For example, the first question can be a simple image search question, such as "baby's passport"; or it can be a question that extracts information from an image, such as "how long until my baby's passport expires" or "what was the name of the tea restaurant I ate at in Hong Kong last year?". Of course, the first question can be text or voice information.

[0034] After receiving the first question information input by the user, the electronic device further analyzes the first question information to determine the image search conditions included in the first question information and whether there is a question-and-answer intent.

[0035] Image search criteria can be obtained by extracting keywords from the image search elements in the first question information.

[0036] For example, the keywords of image retrieval elements can be identified by an NLP model to identify entity words, attribute words and semantic relationships in the first question information. Entity words include objects, time and location, and semantic relationships include location relationships and association relationships.

[0037] For example, the keywords for image retrieval elements extracted from the first question information can also be extracted using a word segmentation tool.

[0038] For example, for the question "How long until my baby's passport expires?", the image search keywords could be "baby", "passport", and "expiration time". These keywords would constitute the image search criteria.

[0039] For example, for the question "What was the name of the tea restaurant I ate at in Hong Kong last year?", the image search keywords could be "last year", "Hong Kong", and "tea restaurant". These keywords would constitute the image search criteria.

[0040] For example, for "cat on the floor", the keywords could be "floor" and "cat", with the semantic relationship being "cat on the floor".

[0041] After analyzing the information of the first question, when determining whether there is any question-and-answer intent in the information of the first question, question-and-answer intent identification can be performed on the information of the first question.

[0042] For example, the question-and-answer intent recognition of the first question information can be based on the analysis of the text features of the first question information through an intent recognition model. If the first question information contains information query terms or a clear information acquisition need, it is determined to contain question-and-answer intent.

[0043] For example, by performing question-and-answer intent recognition on the first question information, semantic recognition can be used to determine whether there is question-and-answer intent in the first question information. For instance, if the search term is "How long until my passport expires", semantic recognition can determine that the above search term contains question-and-answer intent.

[0044] For example, semantic recognition of the search term "a little girl running on the beach" indicates that it does not have a question-and-answer intent. However, if the search term is "the clothes a little girl wearing while running on the beach," semantic recognition indicates that it does have a question-and-answer intent.

[0045] For example, by performing question-and-answer intent recognition on the first question information, it can also be determined whether there are interrogative words in the first question information; if interrogative words are present, it can be determined that there is question-and-answer intent in the first question information. Interrogative words can be "what", "is", etc. For example, if the search term is "What was the name of the restaurant I ate at last year?", since there is also the interrogative word "what", it can be concluded that the above search term has question-and-answer intent.

[0046] In fact, even if there are no interrogative words in the search terms, the first question information may still contain the intent to ask and answer questions. For example, the search term is "the validity period of a baby's passport". Although there are no interrogative words, the above search term still contains the intent to ask and answer questions.

[0047] By extracting keywords from search elements, we can accurately match corresponding resources in the image library, avoiding false positives or false negatives caused by semantic ambiguity. Through question-and-answer intent recognition, we can accurately determine whether a user only needs to obtain images or needs to obtain specific information based on images. For requests without question-and-answer intent, the process can be simplified to directly return images; for requests with question-and-answer intent, a deep information extraction process is triggered to achieve a precise "on-demand service" response.

[0048] In this embodiment of the application, when analyzing the first question information, a pre-retrieval judgment step may be included, in which the first question information is judged by a binary classification search model. If the first question information is a single-element short text, keyword retrieval is used; if the first question information is a multi-element long text and contains semantic relationships, semantic retrieval is triggered.

[0049] It should be noted that, in order to improve the accuracy of the search, semantic search can be invoked when analyzing the first question information. However, not all question information requires semantic search. Therefore, when analyzing the first question information, it is necessary to first determine the text length and search elements within the first question information.

[0050] For example, semantic search is not needed when the text length is greater than the preset text length and / or the number of search elements is not greater than the preset number of elements. For instance, for short text search terms such as "Nanjing", "cat", and "hamburger", calling semantic search may actually retrieve more messy results and reduce search accuracy.

[0051] For example, when the text length exceeds the preset text length and / or the number of search elements exceeds the preset number of elements, the image search method is determined to be semantic search. For instance, for search terms with multiple elements and long texts such as "cats photographed in Nanjing last year" or "hamburgers eaten in Hong Kong," semantic search must be invoked to achieve accurate retrieval.

[0052] For example, the preset text length and / or the number of search elements can be the number of characters. For instance, the preset number of elements is 4 characters. When the text length is greater than the preset text length and / or the number of search elements is greater than 4 characters, semantic search is invoked to achieve accurate retrieval.

[0053] Based on the above image retrieval criteria, at least one image matching the image retrieval request can be retrieved from the image library and used as the target image.

[0054] In some embodiments, retrieving a target image from an image library that matches the first question information may include:

[0055] The target image matching the information in the first question is retrieved from the image library using semantic retrieval.

[0056] For example, such as Figure 2 As shown, for the question "How long until my baby's passport expires?", based on the image search criteria, two images matching the question "How long until my baby's passport expires?" can be found in the image library through semantic search.

[0057] For example, semantic retrieval methods can improve matching accuracy by understanding contextual relationships through pre-trained language models (such as GPT).

[0058] By using deep semantic matching, the problems of missed or false detections caused by missing or ambiguous keywords in traditional retrieval are avoided. This achieves an upgrade from "fuzzy matching" to "precise positioning", breaks the traditional keyword search method, and reduces search costs.

[0059] Once the existence of a question-and-answer intent is confirmed, the image information of the target image is analyzed based on the question-and-answer intent to obtain the answer information of the question-and-answer intent. Alternatively, the image information of the target image can be analyzed based on a multimodal large model to obtain the answer information of the question-and-answer intent.

[0060] For example, by analyzing the image information of a target image based on a multimodal large model, text information in the image can be extracted using an OCR model, and image content description can be generated using an image description model. The text information, content description, and retrieval request are then input into the multimodal large model, and structured information corresponding to the question-answering intent is output.

[0061] For example, such as Figure 2 As shown, for questions with the intent to answer, such as "How long until my baby's passport expires?", the multimodal model can directly identify the validity period text and calculate the remaining time from the passport image. For example, the answer could be, "My baby's passport expires in October 2026, so there are currently 1 year and 2 months left."

[0062] For example, such as Figure 3 As shown, for the question "What was the name of the tea restaurant you ate at in Hong Kong last year?", the multimodal model can automatically extract the shop name, photo date, and photo location from the image. Based on the extracted information, it can find images that meet the requirements. The model can also determine whether the existing information can answer the user's question based on the image information. If so, it provides the answer to the user, for example, the answer could be "The tea restaurant you ate at in Hong Kong last year was called XXX".

[0063] Of course, the large model can also provide more detailed answer information based on the image information. For example, the answer information could be "The tea restaurant you ate at in Hong Kong last year was called XXX, and the breakfast included bread, coffee, and sandwiches."

[0064] By extracting information from images using a multimodal large model, users are no longer required to manually identify, calculate, or analyze the information, thus reducing the cost of information extraction and solving the problem of high information conversion costs in existing technologies.

[0065] In this embodiment of the application, user dialogue information can also be received;

[0066] Generate response information corresponding to the dialogue information based on the answer information and / or online search information;

[0067] Output the reply message.

[0068] For example, after a user receives the question and answer results, the user can ask further questions based on those results.

[0069] like Figure 4 As shown, for example:

[0070] User: How long until my baby's passport expires?

[0071] Search results: 1 year and 2 months remaining.

[0072] User: When did it take effect?

[0073] Search results: From March 18, 2015.

[0074] For example, the answer information could be information extracted from the target image or data obtained through online searches.

[0075] for example:

[0076] User: What is the name of the grassland I went to wearing a red windbreaker?

[0077] Search answer: Guizhou's Asilixi.

[0078] User: What kind of terrain is that area?

[0079] Search answer: The topography of the Ashilixi Grassland is very diverse, including both karst and basalt landforms.

[0080] ...

[0081] By allowing users to ask follow-up questions about the results, the limitations of a single search are broken, and users' progressive needs for information can be met.

[0082] In S120, after obtaining at least one image and answer information, the electronic device can further display at least one image and first answer information, which is determined by the electronic device from at least one image based on the first question information.

[0083] For example, the output of the first answer information can be in the form of audio or it can be displayed directly.

[0084] In some embodiments, displaying a first answer and at least one image in response to a first input may include:

[0085] In response to the first input, display the first answer information, at least one image, and information processing controls;

[0086] Receive a second input to the information processing control;

[0087] In response to the second input, the first answer information is processed according to the information processing strategy corresponding to the information processing control.

[0088] In this embodiment, the electronic device not only displays the first answer information and at least one image, but also displays information processing controls. Users can perform preset operations on the first answer information through the information processing controls and enter preset function pages. This makes the first answer information no longer statically displayed text, but a directly operable function entry point, realizing scenario-based function linkage and significantly improving the practicality of the function.

[0089] In this embodiment of the application, the information processing control can be a control for copying operations, and in response to a second input to the information processing control, the first answer information can be copied.

[0090] For example, for at least one image and the first answer information, corresponding backward services can be provided based on the category of key information. One such backward service is copying, where the answer text can be copied, making it easy for users to copy it to other scenarios for saving, searching, forwarding, and other operations. This realizes a service loop from "information acquisition" to "information utilization," significantly enhancing the practical value of the image retrieval function while improving operational convenience.

[0091] In some embodiments, after the electronic device displays the first answer information and at least one image, the electronic device may also:

[0092] Receive a third input for the first image, which is at least one image;

[0093] In response to the third input, perform the target operation on the first image;

[0094] The target operation includes at least one of the following:

[0095] Share the first image, create an album, add to an album, copy or cut the first image.

[0096] In this embodiment, after the electronic device displays the first answer information and at least one image, it can share the first image, for example, by sharing it to a social media application; it can copy or cut the first image, for example, by adjusting the image size; it can also create an album based on the first image or add the first image to an album. In this way, the first image is no longer a statically displayed image, but a directly operable function entry point, realizing contextualized function integration and significantly improving the practicality of the function.

[0097] In some embodiments, after the electronic device displays the first answer information and at least one image, the electronic device may also:

[0098] Receive the fourth input; the fourth input is used to overlay the information from the second question.

[0099] In response to the fourth input, if the first answer information corresponding to the first question information and at least one image are displayed, the second answer information corresponding to the second question information and at least one image are displayed.

[0100] For example, such as Figure 5As shown, after the electronic device displays the first answer information and at least one image, the user can trigger the "+" button below the first question input box. The electronic device then displays a second question input box, allowing the user to enter the second question information. Responding to this input, the electronic device displays the second answer information and at least one image corresponding to the first question information, while also displaying the first answer information and at least one image corresponding to the first question information. For example, if the user's first question is "What was the name of the tea restaurant I ate at in Hong Kong last year?", the electronic device displays the first answer as "The tea restaurant you ate at in Hong Kong last year was called XXX, and breakfast included bread, coffee, and a sandwich," and displays images of the tea restaurant and the bread, coffee, and sandwich. Similarly, if the user triggers the "+" button below the first question input box and enters the second question information in the second question input box, and the user's first question is "What was the name of the restaurant I ate at in Macau the year before last?", the electronic device displays the second answer as "The restaurant you ate at in Macau the year before last was called XX, and breakfast included buttered toast and coffee," and displays images of the restaurant and the buttered toast and coffee.

[0101] It should be noted that, in order to ensure that the images corresponding to the first problem information and the second problem information are displayed completely, in this embodiment of the application, the electronic device can determine the display parameters of at least one image corresponding to the first problem information and the display parameters of at least one image corresponding to the second problem information based on the size of the image display area corresponding to the first problem information.

[0102] For example, the display parameters can be the length and width data of the image. The display size of the image corresponding to the first problem information and the display size of the image corresponding to the second problem information are determined according to the size of the image display area, so that at least one image corresponding to the first problem information and the second problem information can be fully displayed in the image display area.

[0103] In some embodiments, the electronic device may also:

[0104] Receive the seventh input for the image display area corresponding to the first question information or the image display area corresponding to the second question information;

[0105] In response to the seventh input, update the size of the image display area corresponding to the first question information or the size of the image display area corresponding to the second question information.

[0106] In practical applications, the display area for the images corresponding to the first or second problem information can be adjusted. For example, if the number of images corresponding to the first problem information is greater than the number of images corresponding to the second problem information, the size of the display area for the images corresponding to the first problem information can be increased; similarly, if the number of images corresponding to the second problem information is greater than the number of images corresponding to the first problem information, the size of the display area for the images corresponding to the second problem information can be increased, so that the images corresponding to the first and second problem information are displayed completely and reasonably.

[0107] In some embodiments, after displaying the first answer information and at least one image corresponding to the first question information, and the second answer information and at least one image corresponding to the second question information, the electronic device may further:

[0108] Receive a fifth input for the second image from at least one image corresponding to the first question information and at least one image corresponding to the second question information;

[0109] In response to the fifth input, perform the target operation on the second image;

[0110] The target operation includes at least one of the following:

[0111] Share a second image, create an album, add it to an album, or copy or cut the second image.

[0112] In this embodiment of the application, after the electronic device displays the first answer information and at least one image corresponding to the first question information, and the second answer information and at least one image corresponding to the second question information, it can also share the second image among the at least one image, such as sharing the second image to a social application; it can copy or cut the second image, such as adjusting the size of the image; it can also create an album based on the second image or add the first image to the album.

[0113] In this embodiment of the application, the first answer information includes preset type information.

[0114] In some embodiments, the electronic device may also receive a sixth input of information of a preset type;

[0115] In response to the sixth input, display an image corresponding to the information of the preset type.

[0116] For example, if the preset type information is a merchant identifier, the electronic device can jump to the target application to view the merchant information corresponding to the merchant identifier;

[0117] When the preset type information is date, the electronic device can search for the first image set corresponding to the date in the image library and display the first image set;

[0118] When the preset type information is location, the electronic device can search for a second set of images corresponding to the location in the image library and display the second set of images.

[0119] For example, given at least one image and answer information, a targeted back-end service can also be triggered when the answer text contains key information of a special category such as store name, date, or city name.

[0120] For example, if the preset type of information is the store name, you can be redirected to the target application by triggering an action on the store name to view the relevant information of the store.

[0121] For example, if the preset type information is date, users can quickly locate all pictures in the album corresponding to that date by using the date, helping them recall pictures from that day.

[0122] For example, the preset type information is the city name. By using the city name, users can quickly locate all the pictures in the album that correspond to that city, helping them recall pictures related to that city.

[0123] Targeted backward services based on merchant identifiers, dates, and locations can establish a direct link between search results and contextualized services. Through precise contextual association, efficient operation simplification, and deep emotional connection, it solves the problems of "fragmented information, fragmented services, and difficulty in memory recall" in traditional image retrieval, achieving a dual enhancement of technological value.

[0124] In this embodiment, after receiving the first input of the first question information, in addition to displaying an image, the system can also display the first answer information corresponding to the first question information determined from the image. This allows for a more intuitive presentation of the answer to the first question information to the user, improving information retrieval efficiency.

[0125] It should be noted that the information processing method provided in this application embodiment can be executed by an information processing device or a processing module within that information processing device for executing the information processing method. This application embodiment uses an information processing device executing the information processing method as an example to illustrate the information processing device provided in this application embodiment.

[0126] Figure 6 This is a schematic diagram of the structure of an information processing device provided in an embodiment of this application.

[0127] like Figure 6 As shown, the information processing device 600 may include:

[0128] The receiving module 601 is used to receive the first input; the first input is used to input the first question information;

[0129] Display module 602 is configured to display first answer information and at least one image in response to a first input; wherein the first answer information is determined from at least one image based on the first question information.

[0130] In this embodiment, after receiving the first input of the first question information, in addition to displaying an image, the system can also display the first answer information corresponding to the first question information determined from the image. This allows for a more intuitive presentation of the answer to the first question information to the user, improving information retrieval efficiency.

[0131] In some possible implementations of the embodiments of this application, the display module 602 is further configured to:

[0132] In response to the first input, display the first answer information, at least one image, and information processing controls;

[0133] Receive a second input to the information processing control;

[0134] In response to the second input, the first answer information is processed according to the information processing strategy corresponding to the information processing control.

[0135] In some possible implementations of the embodiments of this application, the receiving module 601 may also be used for:

[0136] Receive a third input for the first image, which is at least one image;

[0137] In response to the third input, perform the target operation on the first image;

[0138] The target operation includes at least one of the following:

[0139] Share the first image, create an album, add to an album, copy or cut the first image.

[0140] In some possible implementations of the embodiments of this application, the receiving module 601 may also be used for:

[0141] Receive the fourth input; the fourth input is used to overlay the information from the second question.

[0142] In response to the fourth input, if the first answer information corresponding to the first question information and at least one image are displayed, the second answer information corresponding to the second question information and at least one image are displayed.

[0143] In some possible implementations of the embodiments of this application, the information processing device 600 may further include: a determining module:

[0144] The determination module is used to determine the display parameters of at least one image corresponding to the first problem information and the display parameters of at least one image corresponding to the second problem information based on the size of the image display area corresponding to the first problem information.

[0145] In some possible implementations of the embodiments of this application, the receiving module 601 may also be used for:

[0146] Receive a fifth input for the second image from at least one image corresponding to the first question information and at least one image corresponding to the second question information;

[0147] In response to the fifth input, perform the target operation on the second image;

[0148] The target operation includes at least one of the following:

[0149] Share a second image, create an album, add it to an album, or copy or cut the second image.

[0150] In some possible implementations of the embodiments of this application, the receiving module 601 may also be used for:

[0151] Receive the seventh input for the image display area corresponding to the first question information or the image display area corresponding to the second question information;

[0152] In response to the seventh input, update the size of the image display area corresponding to the first question information or the size of the image display area corresponding to the second question information.

[0153] In some possible implementations of the embodiments of this application, the first answer information includes information of a preset type;

[0154] The receiving module 601 can also be used for:

[0155] Receive the sixth input of information of the preset type;

[0156] In response to the sixth input, display an image corresponding to the information of the preset type.

[0157] In this embodiment, after receiving the first input of the first question information, in addition to displaying an image, the system can also display the first answer information corresponding to the first question information determined from the image. This allows for a more intuitive presentation of the answer to the first question information to the user, improving information retrieval efficiency.

[0158] The information processing device in this application embodiment can be a device or a component in an electronic device, such as an integrated circuit or a chip. For example, the electronic device can be a mobile phone, tablet computer, laptop computer, PDA, in-vehicle electronic device, mobile internet device (MID), augmented reality (AR) / virtual reality (VR) device, robot, wearable device, ultra-mobile personal computer (UMPC), netbook, or personal digital assistant (PDA), etc. It can also be a server, network attached storage (NAS), personal computer (PC), television (TV), ATM, or self-service machine, etc. This application embodiment does not specifically limit the device.

[0159] The electronic device in this application embodiment can be a terminal with an operating system. This operating system can be Android, iOS, or other possible operating systems; this application embodiment does not specifically limit the specific operating system used.

[0160] The information processing device provided in this application embodiment can achieve... Figures 1 to 5 The various processes in the information processing method embodiments can achieve the same technical effect, and will not be described again here to avoid repetition.

[0161] like Figure 7 As shown, this application embodiment also provides an electronic device 700, including a processor 701 and a memory 702. The memory 702 stores programs or instructions that can run on the processor 701. When the program or instructions are executed by the processor 701, they implement the various steps of the above-described information processing method embodiment and can achieve the same technical effect. To avoid repetition, they will not be described again here.

[0162] It should be noted that the electronic devices in the embodiments of this application include the mobile terminals and non-mobile terminals mentioned above.

[0163] Figure 8 This is a schematic diagram of the hardware structure of an electronic device provided in an embodiment of this application.

[0164] The electronic device 800 includes, but is not limited to, components such as: radio frequency unit 801, network module 802, audio output unit 803, input unit 804, sensor 805, display unit 806, user input unit 807, interface unit 808, memory 809, and processor 810.

[0165] Those skilled in the art will understand that the electronic device 800 may also include a power supply (such as a battery) for supplying power to various components. The power supply may be logically connected to the processor 810 through a power management system, thereby enabling functions such as managing charging, discharging, and power consumption through the power management system. Figure 8 The structure of the electronic device 800 shown does not constitute a limitation on the electronic device 800. The electronic device 800 may include more or fewer components than shown, or combine certain components, or have different component arrangements, which will not be described in detail here.

[0166] The user input unit 807 is used to receive a first input; the first input is used to input first question information.

[0167] Display unit 806 is configured to display first answer information and at least one image in response to a first input; wherein the first answer information is determined from at least one image based on first question information.

[0168] In this embodiment, after receiving the first input of the first question information, in addition to displaying an image, the system can also display the first answer information corresponding to the first question information determined from the image. This allows for a more intuitive presentation of the answer to the first question information to the user, improving information retrieval efficiency.

[0169] In some possible implementations of the embodiments of this application, the display unit 806 is further configured to:

[0170] In response to the first input, display the first answer information, at least one image, and information processing controls;

[0171] Receive a second input to the information processing control;

[0172] In response to the second input, the first answer information is processed according to the information processing strategy corresponding to the information processing control.

[0173] In some possible implementations of the embodiments of this application, the user input unit 807 is further configured to:

[0174] Receive a third input for the first image, which is at least one image;

[0175] In response to the third input, perform the target operation on the first image;

[0176] The target operation includes at least one of the following:

[0177] Share the first image, create an album, add to an album, copy or cut the first image.

[0178] In some possible implementations of the embodiments of this application, the user input unit 807 is further configured to:

[0179] Receive the fourth input; the fourth input is used to overlay the information from the second question.

[0180] In response to the fourth input, if the first answer information corresponding to the first question information and at least one image are displayed, the second answer information corresponding to the second question information and at least one image are displayed.

[0181] In some possible implementations of the embodiments of this application, the processor 810 may also be used for:

[0182] Based on the size of the image display area corresponding to the first question information, determine the display parameters of at least one image corresponding to the first question information, and the display parameters of at least one image corresponding to the second question information.

[0183] In some possible implementations of the embodiments of this application, the user input unit 807 is further configured to:

[0184] Receive a fifth input for the second image from at least one image corresponding to the first question information and at least one image corresponding to the second question information;

[0185] In response to the fifth input, perform the target operation on the second image;

[0186] The target operation includes at least one of the following:

[0187] Share a second image, create an album, add it to an album, or copy or cut the second image.

[0188] In some possible implementations of the embodiments of this application, the user input unit 807 is further configured to:

[0189] Receive the seventh input for the image display area corresponding to the first question information or the image display area corresponding to the second question information;

[0190] In response to the seventh input, update the size of the image display area corresponding to the first question information or the size of the image display area corresponding to the second question information.

[0191] In some possible implementations of the embodiments of this application, the first answer information includes information of a preset type;

[0192] User input unit 807 is also used for:

[0193] Receive the sixth input of information of the preset type;

[0194] In response to the sixth input, display an image corresponding to the information of the preset type.

[0195] In this embodiment, after receiving the first input of the first question information, in addition to displaying an image, the system can also display the first answer information corresponding to the first question information determined from the image. This allows for a more intuitive presentation of the answer to the first question information to the user, improving information retrieval efficiency.

[0196] It should be understood that, in this embodiment, the input unit 804 may include a graphics processing unit (GPU) 8041 and a microphone 8042. The GPU 8041 processes image data of still images or videos obtained by an image capture device (such as a camera) in video capture mode or image capture mode. The display unit 806 may include a display panel 8061, which may be configured in the form of a liquid crystal display, an organic light-emitting diode, or the like. The user input unit 807 includes at least one of a touch panel 8071 and other input devices 8072. The touch panel 8071 is also called a touch screen. The touch panel 8071 may include a touch detection device and a touch controller. Other input devices 8072 may include, but are not limited to, physical keyboards, function keys (such as volume control buttons, power buttons, etc.), trackballs, mice, and joysticks, which will not be described in detail here.

[0197] The memory 809 can be used to store software programs and various data. The memory 809 may primarily include a first storage area for storing programs or instructions and a second storage area for storing data. The first storage area may store the operating system, application programs or instructions required for at least one function (such as sound playback function, image playback function, etc.). Furthermore, the memory 809 may include volatile memory or non-volatile memory, or both. The non-volatile memory may be read-only memory (ROM), programmable read-only memory (PROM), erasable programmable read-only memory (EPROM), electrically erasable programmable read-only memory (EEPROM), or flash memory. Volatile memory can be random access memory (RAM), static random access memory (SRAM), dynamic random access memory (DRAM), synchronous dynamic random access memory (SDRAM), double data rate synchronous dynamic random access memory (DDRSDRAM), enhanced synchronous dynamic random access memory (ESDRAM), synchronous link dynamic random access memory (SLDRAM), and direct memory bus RAM (DRRAM). The memory 709 in the embodiments of this application includes, but is not limited to, these and any other suitable types of memory.

[0198] Processor 810 may include one or more processing units; optionally, processor 810 integrates an application processor and a modem processor, wherein the application processor mainly handles operations involving the operating system, user interface, and applications, and the modem processor mainly handles wireless communication signals, such as a baseband processor. It is understood that the aforementioned modem processor may also not be integrated into processor 810.

[0199] This application also provides a readable storage medium storing a program or instructions. When the program or instructions are executed by a processor, they implement the various processes of the above-described information processing method embodiments and achieve the same technical effects. To avoid repetition, they will not be described again here.

[0200] The processor is the processor in the electronic device described in the above embodiments. The readable storage medium includes computer-readable storage media, such as computer read-only memory (ROM), random access memory (RAM), magnetic disk, or optical disk.

[0201] This application also provides a chip, which includes a processor and a communication interface. The communication interface and the processor are coupled. The processor is used to run programs or instructions to implement the various processes of the above-described information processing method embodiments and can achieve the same technical effect. To avoid repetition, it will not be described again here.

[0202] It should be understood that the chip mentioned in the embodiments of this application may also be referred to as a system-on-a-chip, system chip, chip system, or system-on-a-chip, etc.

[0203] This application provides a computer program product, which is stored in a storage medium and executed by at least one processor to implement the various processes of the information processing method embodiments described above, and can achieve the same technical effect. To avoid repetition, it will not be described again here.

[0204] It should be noted that, in this document, the terms "comprising," "including," or any other variations thereof are intended to cover non-exclusive inclusion, such that a process, method, article, or apparatus that comprises a list of elements includes not only those elements but also other elements not expressly listed, or elements inherent to such a process, method, article, or apparatus. Without further limitations, an element defined by the phrase "comprising one..." does not exclude the presence of other identical elements in the process, method, article, or apparatus that includes that element. Furthermore, it should be noted that the scope of the methods and apparatuses in the embodiments of this application is not limited to performing functions in the order shown or discussed, but may also include performing functions substantially simultaneously or in the reverse order, depending on the functions involved. For example, the described methods may be performed in a different order than described, and various steps may be added, omitted, or combined. Additionally, features described with reference to certain examples may be combined in other examples.

[0205] Through the above description of the embodiments, those skilled in the art can clearly understand that the methods of the above embodiments can be implemented by means of software plus necessary general-purpose hardware platforms. Of course, they can also be implemented by hardware, but in many cases the former is a better implementation method. Based on this understanding, the technical solution of this application, in essence, or the part that contributes to the related technology, can be embodied in the form of a computer software product. This computer software product is stored in a storage medium (such as ROM / RAM, magnetic disk, optical disk) and includes several instructions to cause a terminal (which may be a mobile phone, computer, server, or network device, etc.) to execute the methods described in the various embodiments of this application.

[0206] The embodiments of this application have been described above with reference to the accompanying drawings. However, this application is not limited to the specific embodiments described above. The specific embodiments described above are merely illustrative and not restrictive. Those skilled in the art can make many other forms under the guidance of this application without departing from the spirit and scope of the claims, and all of these forms are within the protection scope of this application.

Claims

1. An information processing method, characterized in that, include: Receive the first input; The first input is used to input the first question information; In response to the first input, first answer information and at least one image are displayed; wherein the first answer information is determined from the at least one image based on the first question information.

2. The method according to claim 1, characterized in that, The response to the first input, displaying the first answer information and at least one image, includes: In response to the first input, display the first answer information, at least one image, and information processing controls; Receive a second input to the information processing control; In response to the second input, the first answer information is processed according to the information processing strategy corresponding to the information processing control.

3. The method according to claim 1, characterized in that, The method further includes: Receive a third input for the first image among the at least one image; In response to the third input, the target operation is performed on the first image; The target operation includes at least one of the following: Share the first image, create an album, add to an album, copy or cut the first image.

4. The method according to claim 1, characterized in that, The method further includes: Receive a fourth input; the fourth input is used to overlay the second problem information; In response to the fourth input, if the first answer information corresponding to the first question information and at least one image are displayed, the second answer information corresponding to the second question information and at least one image are displayed.

5. The method according to claim 4, characterized in that, The method further includes: Based on the size of the image display area corresponding to the first problem information, determine the display parameters of at least one image corresponding to the first problem information, and the display parameters of at least one image corresponding to the second problem information.

6. The method according to claim 5, characterized in that, The method further includes: Receive a fifth input for the second image in at least one image corresponding to the first problem information and at least one image corresponding to the second problem information; In response to the fifth input, the target operation is performed on the second image; The target operation includes at least one of the following: Share the second image, create an album, add it to an album, or copy or cut the second image.

7. The method according to claim 5, characterized in that, The method further includes: Receive a seventh input for the image display area corresponding to the first problem information or the image display area corresponding to the second problem information; In response to the seventh input, update the size of the image display area corresponding to the first problem information or the size of the image display area corresponding to the second problem information.

8. The method according to claim 1, characterized in that, The first answer information includes information of a preset type; The method further includes: Receive a sixth input for information of the preset type; In response to the sixth input, an image corresponding to the information of the preset type is displayed.

9. An electronic device, characterized in that, The electronic device includes a processor and a memory, the memory storing programs or instructions that can run on the processor, the programs or instructions being executed by the processor to implement the steps of the method as described in any one of claims 1 to 8.

10. A readable storage medium, characterized in that, The readable storage medium stores a program or instructions that, when executed by a processor, implement the steps of the method as described in any one of claims 1 to 8.