A method and apparatus for a dialogue based on user information for a vehicle
By identifying users and generating personalized responses, the potential dangers of commands that are not suitable for children or others to operate in vehicle voice systems are solved, thus improving the intelligence and safety of vehicle dialogue systems.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- CHINA FAW CO LTD
- Filing Date
- 2023-03-27
- Publication Date
- 2026-07-21
AI Technical Summary
Existing vehicle voice systems cannot intelligently identify the user's identity, which may lead to potential dangers if voice commands are not suitable for children or others to operate them, and the responses lack personalization.
By acquiring users' voice and video information, the system identifies users and generates personalized dialogue information. This includes obtaining the similarity between the user's facial information and preset user facial information, determining the user's identity, and generating corresponding dialogue information, which is divided into positive and negative dialogue sets. Different responses are generated based on the user's identity and instructions.
It enables the generation of personalized dialogue responses based on user identity and instructions, preventing the execution of inappropriate actions and improving the intelligence and security of the vehicle dialogue system.
Smart Images

Figure CN116343821B_ABST
Abstract
Description
Technical Field
[0001] This application relates to the field of automotive voice control technology, specifically to a method and apparatus for vehicle-based dialogue based on user information. Background Technology
[0002] In existing voice systems, most execute user commands based on semantic parsing. For example, when a user controls the system by giving the voice command "open the car window," the vehicle's infotainment system will recognize the command, execute the operation of opening the window, and generate a response, such as: "Okay, opening the window for you."
[0003] In existing technologies, vehicles cannot intelligently recognize whether a user's command is valid or potentially dangerous. For example, if a child commands the car to open a window, there is a potential hazard if the child actually opens the window.
[0004] Therefore, there is a need for a technical solution to address or at least mitigate the aforementioned shortcomings of existing technologies. Summary of the Invention
[0005] The purpose of this invention is to provide a method for vehicle-based dialogue based on user information to at least solve one of the above-mentioned technical problems.
[0006] One aspect of the present invention provides a method for vehicle-based dialogue based on user information, the method comprising:
[0007] Obtain user command information;
[0008] Obtain user identity information;
[0009] Script information is generated based on user identity information and user instruction information.
[0010] Optionally, obtaining user instruction information includes:
[0011] Obtain the user's voice information;
[0012] The user's voice information is recognized by the speech semantic recognition engine, thereby obtaining user command information.
[0013] Optionally, obtaining user identity information includes:
[0014] During the process of acquiring users' voice information, video information of people inside the vehicle is acquired through in-vehicle cameras;
[0015] Facial information of occupants whose lips are moving, obtained from video footage;
[0016] Obtain a user database, which includes at least one preset user identity information and preset user facial information corresponding to each preset user identity information;
[0017] Determine whether the similarity between the acquired user facial information and the preset user facial information exceeds a preset threshold. If so, then...
[0018] Obtain user identity information corresponding to the facial information of a preset user whose similarity exceeds a preset threshold.
[0019] Optionally, generating the script information based on user identity information and user instruction information includes:
[0020] Obtain a randomly constructed dialogue script database, which includes response tendencies and at least one dialogue script database. One response tendency corresponds to one dialogue script database. The dialogue script database includes multiple prefix dialogue scripts and multiple suffix dialogue scripts. One response tendency corresponds to at least one preset user instruction information.
[0021] Obtain basic style information based on feature information;
[0022] Based on the basic style information and the preset user instruction information corresponding to the user instruction information, obtain the response tendency corresponding to the preset user instruction information;
[0023] Based on the basic style information and response tendency, a prefix dialogue and a suffix dialogue are obtained from the dialogue database corresponding to the response tendency, wherein the prefix dialogue and the suffix dialogue constitute the dialogue information.
[0024] Optionally, obtaining the characteristic information of the person in the vehicle who issued the user command information based on the acquired video information of the person in the vehicle includes:
[0025] Extract multiple frames of images of people inside the vehicle who have issued user commands from the video information;
[0026] Extract feature information from multiple frames of images.
[0027] Optionally, obtaining basic style information based on feature information includes:
[0028] The image features are input into the basic style classifier to obtain basic style information.
[0029] Optionally, the vehicle-based dialogue method based on user information further includes:
[0030] Obtain a positive set of dialogue and a negative set of dialogue. The positive set of dialogue includes at least one dialogue information, and the negative set of dialogue includes at least one dialogue information. Any dialogue information in the positive set is different from any dialogue information in the negative set.
[0031] The system determines whether the acquired script information belongs to a positive script set or a negative script set. If it belongs to a positive script set, the system generates a control command based on the user instruction information, sends the control command to the corresponding control mechanism in the vehicle, and broadcasts the script information via voice.
[0032] Optionally, it can be determined whether the obtained script information belongs to a positive script set or a negative script set. If it belongs to a negative script set, the generated script will be broadcast through voice broadcast.
[0033] This application also provides a vehicle-mounted device for dialogue based on user information, the vehicle-mounted device for dialogue based on user information comprising:
[0034] User instruction information acquisition module, the user instruction information acquisition module is used to acquire user instruction information;
[0035] User identity information acquisition module, the user identity information acquisition module is used to acquire user identity information;
[0036] The script information generation module is used to generate script information based on user identity information and user instruction information.
[0037] Beneficial effects
[0038] The method for vehicle-based dialogue based on user information in this application generates corresponding dialogue information according to the user's identity and instructions, thereby preventing the user's requests from being followed in all circumstances and the generation of the same dialogue regardless of the user's identity. Compared with the prior art, it can adjust the dialogue information more intelligently. Attached Figure Description
[0039] Figure 1 This is a flowchart illustrating a method for vehicle-based dialogue based on user information, provided in an embodiment of this application.
[0040] Figure 2 It is used to implement Figure 1 The diagram shows an electronic device for a vehicle that uses a user information-based dialogue method. Detailed Implementation
[0041] To make the objectives, technical solutions, and advantages of this application clearer, the technical solutions in the embodiments of this application will be described in more detail below with reference to the accompanying drawings. In the drawings, the same or similar reference numerals denote the same or similar elements or elements having the same or similar functions throughout. The described embodiments are some, but not all, embodiments of this application. The embodiments described below with reference to the accompanying drawings are exemplary and intended to explain this application, and should not be construed as limiting this application. All other embodiments obtained by those skilled in the art based on the embodiments of this application without creative effort are within the scope of protection of this application. The embodiments of this application will be described in detail below with reference to the accompanying drawings.
[0042] Figure 1 This is a flowchart illustrating a method for vehicle-based dialogue based on user information, provided in an embodiment of this application.
[0043] like Figure 1 The methods for vehicle-based dialogue based on user information shown include:
[0044] Obtain user command information;
[0045] Obtain user identity information;
[0046] Script information is generated based on user identity information and user instruction information.
[0047] The method for vehicle-based dialogue based on user information in this application generates corresponding dialogue information according to the user's identity and instructions, thereby preventing the user's requests from being followed in all circumstances and the generation of the same dialogue regardless of the user's identity. Compared with the prior art, it can adjust the dialogue information more intelligently.
[0048] In this embodiment, obtaining user instruction information includes:
[0049] Obtain the user's voice information;
[0050] The user's voice information is recognized by the speech semantic recognition engine, thereby obtaining user command information.
[0051] Specifically, user command information can be obtained using the following methods:
[0052] The speech to be recognized is acquired and input into the recognition decoder to obtain sound features, which include the recognized text, voiceprint confidence, audio features and acoustic model features;
[0053] Semantic analysis is performed based on voiceprint confidence and recognized text to obtain business domain information;
[0054] The confidence level of the recognition result is obtained based on business domain information, audio features, and the recognized text.
[0055] The confidence level of the instruction word is obtained based on the recognition result, the confidence level of the recognition result, the acoustic model features, and the business domain information. Among them, the instruction word with the highest confidence level is the user instruction information.
[0056] In this embodiment, obtaining user identity information includes:
[0057] During the process of acquiring users' voice information, video information of people inside the vehicle is acquired through in-vehicle cameras;
[0058] Facial information of occupants whose lips are moving, obtained from video footage;
[0059] Obtain a user database, which includes at least one preset user identity information and preset user facial information corresponding to each preset user identity information;
[0060] Determine whether the similarity between the acquired user facial information and the preset user facial information exceeds a preset threshold. If so, then...
[0061] Obtain user identity information corresponding to the facial information of a preset user whose similarity exceeds a preset threshold.
[0062] In this embodiment, when identifying user identity information, it is first necessary to determine which user in the vehicle is speaking. Therefore, the user who is speaking is first identified by recognizing lip movements. For example, if user A is sitting in the driver's seat and user B is sitting in the passenger seat, and user B is found to be speaking by recognizing lip movements, then it is determined that the user's instruction information comes from user B.
[0063] In this embodiment, after determining that the user command information comes from user B, the facial information of user B is obtained, which can be obtained through the vehicle-mounted camera device.
[0064] The facial information of user B is compared with the facial information of a preset user. If the similarity exceeds a preset threshold, the identity information of the user exceeding the preset threshold is obtained.
[0065] In this embodiment, generating the script information based on the user's identity information and user instruction information includes:
[0066] Obtain a preset dialogue script database, which includes at least one preset dialogue script group and at least one preset user identity information. One preset user identity information corresponds to one preset dialogue script group. Each preset dialogue script group includes at least one preset instruction and a preset dialogue script corresponding to each preset instruction.
[0067] Obtain the preset dialogue group corresponding to the preset user identity information that is the same as the user identity information;
[0068] Retrieve the preset script corresponding to the preset command that is the same as the user's command information.
[0069] In this embodiment, it is first necessary to obtain the preset dialogue group corresponding to the preset user identity information that is the same as the user identity information. For example, assuming that the person issuing the instruction is the car owner, the car owner has a preset dialogue group. For some specific instructions, the operation is only allowed when the person is the car owner. For example, the instruction is "navigate to the company". This instruction is only allowed to be executed when the car owner says it. At this time, in the car owner's preset dialogue group, there will be the dialogue "start navigation to the company".
[0070] However, if the person issuing the command is not the car owner, and the command is "navigate to the company", even if the person issuing the command has preset user identity information, there will not be a good "start navigation to the company" message in their preset message group. It may be other messages, such as "Sorry, you do not have permission to navigate to the company".
[0071] In this embodiment, the method for vehicle-based dialogue based on user information further includes:
[0072] Determine whether the similarity between the acquired user facial information and preset user facial information in the user database exceeds a preset threshold; if not, then...
[0073] Based on the video information of the people inside the vehicle, obtain the characteristic information of the people inside the vehicle who issued the user's instructions;
[0074] Obtain a randomly constructed dialogue script database, which includes response tendencies and at least one dialogue script database. One response tendency corresponds to one dialogue script database. The dialogue script database includes multiple prefix dialogue scripts and multiple suffix dialogue scripts. One response tendency corresponds to at least one preset user instruction information.
[0075] Obtain basic style information based on feature information;
[0076] Based on the basic style information and the preset user instruction information corresponding to the user instruction information, obtain the response tendency corresponding to the preset user instruction information;
[0077] Based on the basic style information and response tendency, a prefix dialogue and a suffix dialogue are obtained from the dialogue database corresponding to the response tendency, wherein the prefix dialogue and the suffix dialogue constitute the dialogue information.
[0078] In this embodiment, obtaining the characteristic information of the person in the vehicle who issued the user command information based on the acquired video information of the person in the vehicle includes:
[0079] Extract multiple frames of images of people inside the vehicle who have issued user commands from the video information;
[0080] Extract feature information from multiple frames of images.
[0081] In this embodiment, obtaining basic style information based on feature information includes:
[0082] The image features are input into the basic style classifier to obtain basic style information.
[0083] In some cases, the speaker may not be someone from the user database. In such cases, the script can be randomly generated, making the method in this application more intelligent.
[0084] For example, if the speaker is not someone in the user database, then the camera device can obtain the speaker's image information, which may include various information such as facial information, clothing information, and hair color information.
[0085] By identifying the various pieces of information mentioned above, basic style information can be obtained. For example, basic style information may include the following styles:
[0086] Male punk style (identified as male by facial image, or by dyed hair, denim clothing, etc.).
[0087] Male child style (gender and approximate age can be identified through facial images, thus determining the style as male child).
[0088] Female child style (gender and approximate age can be identified through facial images, thus determining the style as female child).
[0089] Through the above identification, basic style information is obtained.
[0090] For example, suppose image recognition determines that user C's style is that of a male child. In this case, user C's user instruction is to open a window.
[0091] At this point, the response tendency corresponding to the preset user instruction information is obtained based on the basic style information and the user instruction information. Specifically, since the user is a child, the response tendency is negative, meaning that the command is not allowed to be executed through this instruction.
[0092] By identifying response tendencies, a database of corresponding response phrases is obtained. In this database, at least one phrase in each prefix phrase explicitly expresses a negative meaning, and at least one phrase in each suffix phrase explicitly expresses a negative meaning. When selecting prefix and suffix phrases, regardless of the combination used, the selected phrases must explicitly express a negative meaning.
[0093] For example, prefix phrases may include the following:
[0094] Little friend, you can't!
[0095] Hey handsome young man, you can't just open the window without permission.
[0096] For example, suffix phrases may include the following:
[0097] You must not open the window without permission; please ask your parents to open it for you.
[0098] Please ask your parents to open the window for you.
[0099] Using this method, the resulting script information can be as follows:
[0100] Young man, you can't open the window by yourself. Please ask your parents to open the window for you.
[0101] This approach allows users to be informed of specific response scripts, while also providing a fresh experience through different combinations, preventing them from becoming bored with unchanging human-computer interaction language.
[0102] In this embodiment, the method for vehicle-based dialogue based on user information further includes:
[0103] Obtain a positive set of dialogue and a negative set of dialogue. The positive set of dialogue includes at least one dialogue information, and the negative set of dialogue includes at least one dialogue information. Any dialogue information in the positive set is different from any dialogue information in the negative set.
[0104] The system determines whether the acquired script information belongs to a positive script set or a negative script set. If it belongs to a positive script set, the system generates a control command based on the user instruction information, sends the control command to the corresponding control mechanism in the vehicle, and broadcasts the script information via voice.
[0105] In this embodiment, it is determined whether the acquired script information belongs to the positive script set or the negative script set. If it belongs to the negative script set, the generated script will be broadcast through voice broadcasting.
[0106] In this embodiment, it is necessary not only to respond to the person sending the user's instruction information via a script, but also to determine whether to execute the corresponding operation based on the specific script. In this case, the script information is divided into a positive script set and a negative script set. When the script information belongs to the positive script set, not only is the script output, but the corresponding action instruction is also given. For example, if the car owner gives the instruction: "Navigate to the company," and parsing reveals that the script information corresponding to this instruction belongs to the positive script set, then the script might be: "Okay, start navigating to the company," and the navigation operation is executed.
[0107] If the child's style is as described above (male child), then the verbal message is identified as belonging to the negative verbal message set. In this case, only the following voice message will be broadcast: "Handsome boy, you can't open the window by yourself. Please ask your parents to open the window for you." No actual action will be taken.
[0108] This application also provides a vehicle-mounted device for dialogue based on user information, the device comprising a user instruction information acquisition module, a user identity information acquisition module, and a dialogue script information generation module, wherein...
[0109] The user instruction information acquisition module is used to acquire user instruction information;
[0110] The user identity information acquisition module is used to acquire user identity information;
[0111] The script information generation module is used to generate script information based on user identity information and user instruction information.
[0112] It should be noted that the foregoing explanation of the method embodiments also applies to the apparatus of this embodiment, and will not be repeated here.
[0113] This application also provides an electronic device, including a memory, a processor, and a computer program stored in the memory and capable of running on the processor, wherein the processor executes the computer program to implement the above-described method for vehicle-based dialogue based on user information.
[0114] This application also provides a computer-readable storage medium storing a computer program that, when executed by a processor, enables the above-described vehicle-based dialogue method based on user information.
[0115] Figure 2 This is an exemplary structural diagram of an electronic device capable of implementing a method for dialogue based on user information in a vehicle according to an embodiment of this application.
[0116] like Figure 2As shown, the electronic device includes an input device 501, an input interface 502, a central processing unit 503, a memory 504, an output interface 505, and an output device 506. The input interface 502, central processing unit 503, memory 504, and output interface 505 are interconnected via a bus 507. The input device 501 and output device 506 are connected to the bus 507 via the input interface 502 and output interface 505, respectively, and thus connected to other components of the electronic device. Specifically, the input device 504 receives input information from the outside and transmits it to the central processing unit 503 via the input interface 502. The central processing unit 503 processes the input information based on computer-executable instructions stored in the memory 504 to generate output information, temporarily or permanently storing the output information in the memory 504, and then transmitting the output information to the output device 506 via the output interface 505. The output device 506 outputs the output information to the outside of the electronic device for user use.
[0117] In other words, Figure 2 The illustrated electronic device may also be implemented as including: a memory storing computer-executable instructions; and one or more processors, which can be coupled when executing the computer-executable instructions. Figure 1 The method described describes a car-based dialogue system that uses user information.
[0118] In one embodiment, Figure 2 The illustrated electronic device can be implemented to include: a memory 504 configured to store executable program code; and one or more processors 503 configured to run the executable program code stored in the memory 504 to perform the vehicle-based user information-based dialogue method in the above embodiments.
[0119] In a typical configuration, a computing device includes one or more processors (CPU), input / output interfaces, network interfaces, and memory.
[0120] Memory may include non-persistent storage in computer-readable media, such as random access memory (RAM) and / or non-volatile memory, such as read-only memory (ROM) or flash RAM. Memory is an example of computer-readable media.
[0121] Computer-readable media include both permanent and non-permanent, removable and non-removable media, and information storage can be achieved by any method or technology. Information can be computer-readable instructions, data structures, program modules, or other data. Examples of computer storage media include, but are not limited to, phase-change memory (PRAM), static random access memory (SRAM), dynamic random access memory (DRAM), other types of random access memory (RAM), read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory or other memory technologies, CD-ROM, DVD or other optical storage, magnetic tape, magnetic disk storage or other magnetic storage devices, or any other non-transfer medium that can be used to store information accessible by a computing device.
[0122] Those skilled in the art will understand that embodiments of this application can be provided as methods, systems, or computer program products. Therefore, this application can take the form of a completely hardware embodiment, a completely software embodiment, or an embodiment combining software and hardware aspects. Furthermore, this application can take the form of a computer program product embodied on one or more computer-usable storage media (including, but not limited to, disk storage, CD-ROM, optical storage, etc.) containing computer-usable program code.
[0123] Furthermore, it is clear that the word "comprising" does not exclude other units or steps. Multiple units, modules, or devices recited in the apparatus claims may also be implemented by a single unit or overall apparatus via software or hardware.
[0124] The flowcharts and block diagrams in the accompanying drawings illustrate the architecture, functionality, and operation of possible implementations of systems, methods, and computer program products according to various embodiments of this application. In this regard, each block in a flowchart or block diagram may represent a module, segment, or portion of code, which includes one or more executable instructions for implementing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the blocks may occur in a different order than those marked in the drawings. For example, two consecutively marked blocks may actually be executed substantially in parallel, and they may sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each block in a block diagram and / or flowchart, and combinations of blocks in block diagrams and / or the overall flowchart, can be implemented using a dedicated hardware-based system that performs the specified function or operation, or using a combination of dedicated hardware and computer instructions.
[0125] In this embodiment, the processor may be a Central Processing Unit (CPU), or other general-purpose processors, digital signal processors (DSPs), application-specific integrated circuits (ASICs), field-programmable gate arrays (FPGAs), or other programmable logic devices, discrete gate or transistor logic devices, discrete hardware components, etc. A general-purpose processor may be a microprocessor or any conventional processor.
[0126] Memory can be used to store computer programs and / or modules. The processor implements various functions of the device / terminal equipment by running or executing the computer programs and / or modules stored in the memory, and by accessing data stored in the memory. Memory can mainly include a program storage area and a data storage area. The program storage area can store the operating system, at least one application program required for a function (such as sound playback, image playback, etc.), etc.; the data storage area can store data created based on the use of the mobile phone (such as audio data, phonebook, etc.). Furthermore, memory can include high-speed random access memory, and can also include non-volatile memory, such as hard disks, RAM, plug-in hard disks, SmartMediaCards (SMC), Secure Digital (SD) cards, FlashCards, at least one disk storage device, flash memory device, or other volatile solid-state storage devices.
[0127] In this embodiment, if the modules / units integrated into the device / terminal equipment are implemented as software functional units and sold or used as independent products, they can be stored in a computer-readable storage medium. Based on this understanding, all or part of the processes in the methods of the above embodiments of the present invention can also be implemented by a computer program instructing related hardware. The computer program can be stored in a computer-readable storage medium, and when the computer program is executed by a processor, it can implement the steps of the various method embodiments described above. The computer program includes computer program code, which can be in the form of source code, object code, executable files, or certain intermediate forms. The computer-readable medium can include: any entity or device capable of carrying computer program code, recording media, USB flash drives, portable hard drives, magnetic disks, optical disks, computer memory, read-only memory (ROM), random access memory (RAM), electrical carrier signals, telecommunication signals, and software distribution media, etc. It should be noted that the content contained in the computer-readable medium can be appropriately added or removed according to the requirements of legislation and patent practice in the jurisdiction. Although this application discloses preferred embodiments as described above, it is not intended to limit this application. Any person skilled in the art can make possible changes and modifications without departing from the spirit and scope of this application. Therefore, the scope of protection of this application should be determined by the scope defined in the claims of this application.
[0128] Those skilled in the art will understand that embodiments of this application can be provided as methods, systems, or computer program products. Therefore, this application can take the form of a completely hardware embodiment, a completely software embodiment, or an embodiment combining software and hardware aspects. Furthermore, this application can take the form of a computer program product embodied on one or more computer-usable storage media (including, but not limited to, disk storage, CD-ROM, optical storage, etc.) containing computer-usable program code.
[0129] Furthermore, it is clear that the word "comprising" does not exclude other units or steps. Multiple units, modules, or devices recited in the apparatus claims may also be implemented by a single unit or overall apparatus via software or hardware.
[0130] Although the present invention has been described in detail above with general descriptions and specific embodiments, modifications or improvements can be made to it, which will be obvious to those skilled in the art. Therefore, all such modifications or improvements made without departing from the spirit of the present invention fall within the scope of protection claimed by the present invention.
Claims
1. A method for vehicle-mounted dialogue based on user information, characterized in that, The vehicle-based dialogue method based on user information includes: Obtain user command information; Obtain user identity information; Generate script information based on user identity information and user instruction information; The acquisition of user instruction information includes: Obtain the user's voice information; The user's voice information is recognized by the speech semantic recognition engine, thereby obtaining the user's command information; The acquisition of user identity information includes: During the process of acquiring users' voice information, video information of people inside the vehicle is acquired through in-vehicle cameras; Facial information of occupants whose lips are moving, obtained from video footage; Obtain a user database, which includes at least one preset user identity information and preset user facial information corresponding to each preset user identity information; Determine whether the similarity between the acquired user facial information and the preset user facial information exceeds a preset threshold. If so, then... Obtain user identity information corresponding to the facial information of a preset user whose similarity exceeds a preset threshold; The process of generating script information based on user identity information and user instruction information includes: Obtain a preset dialogue script database, which includes at least one preset dialogue script group and at least one preset user identity information. One preset user identity information corresponds to one preset dialogue script group. Each preset dialogue script group includes at least one preset instruction and a preset dialogue script corresponding to each preset instruction. Obtain the preset dialogue group corresponding to the preset user identity information that is the same as the user identity information; Retrieve the preset script corresponding to the preset command that is the same as the user's command information.
2. The method for vehicle-based dialogue based on user information as described in claim 1, characterized in that, The vehicle-based dialogue method based on user information further includes: Determine whether the similarity between the acquired user facial information and preset user facial information in the user database exceeds a preset threshold; if not, then... Based on the video information of the people inside the vehicle, obtain the characteristic information of the people inside the vehicle who issued the user's instructions; Obtain a randomly constructed dialogue script database, which includes response tendencies and at least one dialogue script database. One response tendency corresponds to one dialogue script database. The dialogue script database includes multiple prefix dialogue scripts and multiple suffix dialogue scripts. One response tendency corresponds to at least one preset user instruction information. Obtain basic style information based on feature information; Based on the basic style information and the preset user instruction information corresponding to the user instruction information, obtain the response tendency corresponding to the preset user instruction information; Based on the basic style information and response tendency, a prefix dialogue and a suffix dialogue are obtained from the dialogue database corresponding to the response tendency, wherein the prefix dialogue and the suffix dialogue constitute the dialogue information.
3. The method for vehicle-based dialogue based on user information as described in claim 2, characterized in that, The step of obtaining the characteristic information of the person in the vehicle who issued the user command information based on the acquired video information of the person in the vehicle includes: Extract multiple frames of images of people inside the vehicle who have issued user commands from the video information; Extract feature information from multiple frames of images.
4. The method for vehicle-based dialogue based on user information as described in claim 3, characterized in that, The step of obtaining basic style information based on feature information includes: The image features are input into a basic style classifier to obtain basic style information.
5. The method for vehicle-based dialogue based on user information as described in claim 4, characterized in that, The vehicle-based dialogue method based on user information further includes: Obtain a positive set of dialogue and a negative set of dialogue. The positive set of dialogue includes at least one dialogue information, and the negative set of dialogue includes at least one dialogue information. Any dialogue information in the positive set is different from any dialogue information in the negative set. The system determines whether the acquired script information belongs to a positive script set or a negative script set. If it belongs to a positive script set, the system generates a control command based on the user instruction information, sends the control command to the corresponding control mechanism in the vehicle, and broadcasts the script information via voice.
6. The method for vehicle-based dialogue based on user information as described in claim 5, characterized in that, Determine whether the acquired script information belongs to the positive script set or the negative script set. If it belongs to the negative script set, the generated script will be broadcast via voice.
7. A vehicle-mounted device for dialogue based on user information, used in the vehicle-mounted method for dialogue based on user information as described in any one of claims 1 to 6, characterized in that, The vehicle-mounted device for dialogue based on user information includes: User instruction information acquisition module, the user instruction information acquisition module is used to acquire user instruction information; User identity information acquisition module, the user identity information acquisition module is used to acquire user identity information; The script information generation module is used to generate script information based on user identity information and user instruction information.