Vehicle Control Data Learning for Preference-Based Drive Tuning

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing vehicle control systems require extensive manual effort from skilled workers to set up and adapt operation amounts of electronic devices based on vehicle states, leading to inefficiencies and increased man-hours.

Innovation Solution

A vehicle control data generation method using reinforcement learning to adjust the relationship between vehicle states and action variables, providing rewards based on user preferences for elements like acceleration response, noise, and energy efficiency, thereby automating the optimization process and reducing manual intervention.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Manufacturing precision

If manual adaptation of filter and operation amounts is performed by skilled workers, then appropriate operation amounts are set, but a great number of man-hours are required

Engineering Contradiction:
Improveappropriateness of operation amountsVSAvoidman-hours for adaptation
Core Design Contradiction:
Manufacturing precisionVSLoss of time

Solution Approach 1:

The system performs self-adaptation through reinforcement learning, where the vehicle controller automatically learns optimal operation amounts for electronic devices based on vehicle states and user preferences, eliminating the need for manual adjustment by skilled workers

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical/manual adjustment process with an automated computational system using reinforcement learning algorithms that process sensor data, preference variables, and reward signals to determine optimal control parameters

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Extent of automation

If reinforcement learning is used to automatically set operation amounts, then manual intervention is reduced, but the system complexity increases

Engineering Contradiction:
Improveautomatic operation amount settingVSAvoidsystem complexity
Core Design Contradiction:
Extent of automationVSDevice complexity

Solution Approach 1:

The vehicle controller performs multiple functions including normal vehicle control, sensor data processing, preference variable interpretation, reward calculation, and reinforcement learning-based adaptation, consolidating these diverse functions into a single integrated system

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system dynamically adjusts operation amounts by changing control parameters based on learned relationships between vehicle states, user preferences, and reward signals, allowing flexible adaptation without hardware modifications

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS11840245B2Vehicle control data generation method, vehicle controller, vehicle control system, vehicle learning device, vehicle control data generation device, and memory medium
Publication Date: 2023.12.12 TOYOTA JIDOSHA KK
  • US11840245B2 patent drawing
  • US11840245B2 patent drawing
  • US11840245B2 patent drawing

AI summary

A vehicle control data generation method is provided. A preference variable indicates a relative preference of a user for two or more requested elements that include at least two of three requested elements including a requested element indicating a high acceleration response of a vehicle, a requested element indicating at least one of vibration and noise of the vehicle is small, and a requested element indicating a high energy use efficiency. The reward calculating process includes a changing process that changes a reward provided when a characteristic of the vehicle is a predetermined characteristic in a case where a value of the preference variable is a second value such that the changed reward differs from the reward provided when the characteristic is the predetermined characteristic in a case where the value of the preference variable is a first value.