Electronic device, authentication method and authentication program

The electronic device uses speech recognition to authenticate users by comparing user-uttered words with reference values, addressing cumbersome authentication issues in conventional methods and ensuring secure access.

JP7729118B2Active Publication Date: 2025-08-26CASIO COMPUTER CO LTD
View PDF 7 Cites 0 Cited by

Patent Information

Application Number
JP2021140269
Authority / Receiving Office
JP · JP
Patent Type
Patents
Current Assignee / Owner
Filing Date
2021-08-30
Publication Date
2025-08-26
Estimated Expiration
2041-08-30

AI Technical Summary

Technical Problem

Conventional electronic devices require cumbersome and strict conditions for personal authentication, such as registered PINs, fingerprints, or facial images, which can be unreliable due to environmental factors or user conditions, leading to security challenges.

Method used

An electronic device utilizing speech recognition to authenticate users by comparing user-uttered words with a reference value, allowing easy and secure personal authentication through voice-based evaluation values.

Benefits of technology

Enables easy and secure personal authentication by ensuring that user voices match predefined criteria, facilitating secure access to device functions.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 0007729118000001
    Figure 0007729118000001
  • Figure 0007729118000002
    Figure 0007729118000002
  • Figure 0007729118000003
    Figure 0007729118000003
Patent Text Reader

Abstract

To provide an electronic apparatus capable of simply performing personal authentication while ensuring the security.SOLUTION: An electronic apparatus according to an embodiment of the present invention includes a control unit. When the control unit has acquired voice of at least one word uttered by a user, the control unit derives a first evaluation value on the basis of comparison with a previously set reference value, and registers the first evaluation value as authentication determination information for determining execution propriety of at least one content process. When the control unit has accepted an execution request of the at least one content process, the control unit presents the at least one word as information for allowing the user to utter, and when the control unit has acquired voice of the at least one word uttered by the user on the basis of the presented information, the control unit derives a second evaluation value on the basis of a comparison with the reference value. When it is determined that a result of comparison between the first evaluation value registered as the authentication determination information and the second evaluation value satisfies a previously determined condition, the control unit executes the at least one content process.SELECTED DRAWING: Figure 1
Need to check novelty before this filing date? Find Prior Art

Description

[Technical Field]

[0001] The present invention relates to an electronic device, an authentication method, and an authentication program. [Background technology]

[0002] In recent years, electronic devices have become commonplace for storing various user-specific information and personal data, and so they are equipped with security features that use personal authentication. For example, personal computers, tablet devices, smartphones, and other devices are equipped with security features such as PIN authentication, fingerprint authentication, and face authentication. Some electronic dictionaries also have a password setting function.

[0003] Security functions provided in conventional electronic devices require users to register personal authentication data in advance. For example, for PIN authentication, it is necessary to register multi-digit alphanumeric characters, for fingerprint authentication, it is necessary to register fingerprint patterns of multiple fingers (e.g., three fingers), and for facial authentication, it is necessary to register a facial image pattern.

[0004] Furthermore, in conventional electronic devices, it has been considered that a special device is attached to the user and personal authentication is performed based on data input from this device (see, for example, Patent Document 1). [Prior art documents] [Patent documents]

[0005] [Patent Document 1] Japanese Patent Application Publication No. 2019-62377 Summary of the Invention [Problem to be solved by the invention]

[0006] As described above, in conventional electronic devices, when security functions are installed, personal authentication data must be registered in advance, and to ensure security, the pre-registered authentication data cannot be easily changed.

[0007] Therefore, with PIN authentication, if you forget your password, you will not be able to use the electronic device. Also, with fingerprint authentication, if your fingertips are dirty or scratched, your fingerprint will not be read correctly, and similarly, you will not be able to use the electronic device. Also, with face authentication, capturing a face image requires a suitable shooting environment, such as pointing the camera at your face under lighting appropriate for capturing the image, and if your face image cannot be captured correctly, you will similarly not be able to use the electronic device.

[0008] As described above, with the security functions installed in conventional electronic devices, the conditions for ensuring security are strict, and the work required for personal authentication can be cumbersome.

[0009] The present invention has been made in consideration of the above-mentioned problems, and aims to provide an electronic device, an authentication method, and an authentication program that can easily perform personal authentication while ensuring security. [Means for solving the problem]

[0010] In order to solve the above problems, the present invention provides The first aspect The electronic device is an electronic device including a control unit, and the control unit Small At least 1 one word The user uttered audio The first voice is If you get The first voice and Compared with the preset reference value Pertaining to pronunciation Based on comparison It is a numerical value a first evaluation value is derived, the first evaluation value is registered as authentication determination information for determining whether or not at least one content process can be executed, and when a request to execute the at least one content process is received, NotationThe system presents a word as information for the user to utter, and then performs a speech recognition based on the presented information. and converting the word into a second voice uttered by the user. If you get audio, The second voice and The reference value Pertaining to pronunciation Based on comparison It is a numerical value A second evaluation value is derived, and when it is determined that the comparison result of the first evaluation value registered as the authentication judgment information and the second evaluation value satisfies a predetermined condition, the at least one content processing is executed.

[0011] In order to solve the above problems, the present invention provides a The first aspect The authentication method is an authentication method for an electronic device including a control unit, the control unit comprising: Small At least 1 one word The user uttered audio The first voice is If you get The first voice and Compared with the preset reference value Pertaining to pronunciation Based on comparison It is a numerical value a first evaluation value is derived, the first evaluation value is registered as authentication determination information for determining whether or not at least one content process can be executed, and when a request to execute the at least one content process is received, Notation The system presents a word as information for the user to utter, and then performs a speech recognition based on the presented information. and converting the word into a second voice uttered by the user. If you get audio, The second voice and The reference value Pertaining to pronunciation Based on comparison It is a numerical value A second evaluation value is derived, and when it is determined that the comparison result of the first evaluation value registered as the authentication judgment information and the second evaluation value satisfies a predetermined condition, the at least one content processing is executed.

[0012] In order to solve the above problems, the present invention provides a The first aspect The authentication program allows the computer to: Small At least 1 one word The user uttered audio The first voice is If you get The first voice and Compared with the preset reference value Pertaining to pronunciation Based on comparison It is a numerical value a first evaluation value is derived, the first evaluation value is registered as authentication determination information for determining whether or not at least one content process can be executed, and when a request to execute the at least one content process is received, Notation The system presents a word as information for the user to utter, and then performs a speech recognition based on the presented information. and converting the word into a second voice uttered by the user. If you get audio, The second voice and The reference value Pertaining to pronunciation Based on comparison It is a numerical value A second evaluation value is derived, and when it is determined that the comparison result of comparing the first evaluation value registered as the authentication judgment information with the second evaluation value satisfies a predetermined condition, the function is to execute the at least one content processing. In addition, in order to solve the above problem, an electronic device of a second aspect of an embodiment of the present invention is an electronic device that has a control unit, wherein when a first voice, which is a voice uttered by a user of at least one word, is acquired, the control unit derives a first evaluation value, which is a value based on a comparison of the first voice with a predetermined reference value, and registers the first evaluation value as authentication judgment information for determining whether at least one content process can be executed, and when a request to execute the at least one content process is received, the control unit presents the word as information for the user to speak, and when a second voice, which is a voice uttered by the user of the word based on the presented information, the control unit derives a second evaluation value, which is a value based on a comparison of the second voice with the reference value, and when a sum of the differences between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple second evaluation values ​​corresponding to each of them is within a predetermined range, the control unit determines that a predetermined condition is met and executes the at least one content process. In addition, in order to solve the above-mentioned problem, an authentication method of a second aspect of an embodiment of the present invention is an authentication method in an electronic device having a control unit, wherein when the control unit acquires a first voice, which is a voice uttered by a user of at least one word, it derives a first evaluation value, which is a value based on a comparison of the first voice with a predetermined reference value, and registers the first evaluation value as authentication judgment information for determining whether at least one content process can be executed, and when a request to execute the at least one content process is accepted, it presents the word as information for the user to speak, and when the control unit acquires a second voice, which is a voice uttered by the user of the word based on the presented information, it derives a second evaluation value, which is a value based on a comparison of the second voice with the reference value, and when the sum of the differences between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple second evaluation values ​​corresponding to each of them is within a predetermined range, it determines that a predetermined condition is met and executes the at least one content process. In addition, in order to solve the above-mentioned problem, an authentication program of a second aspect of an embodiment of the present invention causes a computer to function as follows: when a first voice, which is a voice in which a user speaks at least one word, is acquired, derive a first evaluation value, which is a value based on a comparison of the first voice with a predetermined reference value; register the first evaluation value as authentication judgment information for determining whether at least one content processing can be executed; when a request to execute the at least one content processing is received, present the word as information for the user to speak; when a second voice, which is a voice in which the user speaks the word based on the presented information, derive a second evaluation value, which is a value based on a comparison of the second voice with the reference value; and when a sum of the differences between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple second evaluation values ​​corresponding to each of them is within a predetermined range, determine that a predetermined condition is met and execute the at least one content processing. In addition, in order to solve the above problem, an electronic device of a third aspect of an embodiment of the present invention is an electronic device that has a control unit, wherein when a first voice, which is a voice uttered by a user of at least one word, is acquired, the control unit derives a first evaluation value, which is a value based on a comparison of the first voice with a predetermined reference value, and registers the first evaluation value as authentication judgment information for determining whether at least one content process can be executed, and when a request to execute the at least one content process is received, the control unit presents the word as information for the user to speak, and when a second voice, which is a voice uttered by the user of the word based on the presented information, the control unit derives a second evaluation value, which is a value based on a comparison of the second voice with the reference value, and when the difference between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple second evaluation values ​​corresponding to each first evaluation value is within a predetermined range and the number or percentage of words for which the difference is determined to be within the predetermined range is greater than a default value, the control unit determines that a predetermined condition is met and executes the at least one content process. In addition, in order to solve the above-mentioned problem, an authentication method of a third aspect of an embodiment of the present invention is an authentication method in an electronic device having a control unit, wherein when the control unit acquires a first voice, which is a voice uttered by a user of at least one word, it derives a first evaluation value, which is a value based on a comparison of the first voice with a predetermined reference value, and registers the first evaluation value as authentication judgment information for determining whether at least one content process can be executed, and when a request to execute the at least one content process is accepted, it presents the word as information for the user to speak, and when the control unit acquires a second voice, which is a voice uttered by the user of the word based on the presented information, it derives a second evaluation value, which is a value based on a comparison of the second voice with the reference value, and when the difference between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple second evaluation values ​​corresponding to each first evaluation value is within a predetermined range and the number or percentage of words for which the difference is determined to be within the predetermined range is greater than a default value, it determines that a predetermined condition is met and executes the at least one content process. In addition, in order to solve the above-mentioned problem, an authentication program of a third aspect of an embodiment of the present invention causes a computer to function as follows: when a first voice, which is a voice in which a user speaks at least one word, is acquired, derive a first evaluation value, which is a value based on a comparison of the first voice with a predetermined reference value; register the first evaluation value as authentication judgment information for determining whether at least one content processing can be executed; when a request to execute the at least one content processing is received, present the word as information for the user to speak; when a second voice, which is a voice in which the user speaks the word based on the presented information, derive a second evaluation value, which is a value based on a comparison of the second voice with the reference value; and when a difference between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple second evaluation values ​​corresponding to each of them is within a predetermined range, and the number or percentage of words in which the difference is determined to be within the predetermined range is greater than a default value, determine that a predetermined condition is met and execute the at least one content processing. [Effects of the Invention]

[0013] According to the present invention, personal authentication can be easily performed while ensuring security. [Brief explanation of the drawings]

[0014] [Figure 1] FIG. 1 is a functional block diagram showing the configuration of an electronic circuit of an electronic device according to an embodiment of the present invention. [Figure 2] FIG. 1 is a front view showing the external configuration of an electronic dictionary according to an embodiment of the present invention. [Figure 3] FIG. 12 shows an example of word learning data 12f stored by the pronunciation test process in this embodiment. [Figure 4] 6 is a flowchart showing a lock setting process in the present embodiment. [Figure 5] 5 is a flowchart showing a voice password registration process in the lock setting process shown in FIG. 4. [Figure 6]4 is a diagram showing an example of an operation screen displayed on a touch panel display unit of the electronic dictionary according to the embodiment. FIG. [Figure 7] FIG. 10 is a diagram showing an example of a password setting method selection screen according to the embodiment. [Figure 8] FIG. 10 is a diagram showing an example of a voice password registration guidance screen according to the embodiment. [Figure 9] FIG. 4 is a diagram showing an example of a voice password registration screen according to the embodiment. [Figure 10] FIG. 4 is a diagram showing an example of a password setting content selection screen according to the embodiment. [Figure 11] FIG. 10 is a diagram showing an example of a setting completion screen in the embodiment. [Figure 12] FIG. 4 is a diagram showing an example of password setting content data according to the embodiment. [Figure 13] 6 is a flowchart showing an unlocking process in the present embodiment. [Figure 14] 4 is a diagram showing an example of an operation screen displayed on a touch panel display unit of the electronic dictionary according to the embodiment. FIG. [Figure 15] FIG. 4 is a diagram showing an example of an authentication method selection screen according to the embodiment. [Figure 16] FIG. 4 is a diagram showing an example of a voice authentication screen according to the embodiment. [Figure 17] FIG. 10 is a diagram showing an initial screen of a content processing function for which an execution request has been instructed on the operation screen according to the embodiment. [Figure 18] FIG. 10 is a diagram showing an example of a retry confirmation screen according to the embodiment. DETAILED DESCRIPTION OF THE INVENTION

[0015] Hereinafter, an embodiment of the present invention will be described with reference to the drawings.

[0016] FIG. 1 is a functional block diagram showing the configuration of an electronic circuit of an electronic device according to an embodiment of the present invention.

[0017] In this embodiment, an example is shown in which the electronic device is configured as, for example, an electronic dictionary 10. Note that the electronic device can be realized by various devices such as the electronic dictionary 10, a personal computer, a smartphone, a tablet PC, a game device, and the like.

[0018] The electronic dictionary 10 stores multiple types of dictionary content as dictionary data. The dictionary content stores information about at least one word meaning associated with each of multiple headwords. The dictionary content is not limited to dictionaries related to languages ​​such as English and Japanese, but also includes content such as dictionaries in various fields.

[0019] The electronic dictionary 10 has the configuration of a computer that reads programs recorded on various recording media or transmitted programs and whose operation is controlled by the read programs, and its electronic circuitry is equipped with a CPU (central processing unit) 11.

[0020] The CPU 11 functions as a control unit that controls the entire electronic dictionary 10. The CPU 11 controls the operation of each circuit unit in accordance with a control program that is pre-stored in the memory 12, or a control program that is read into the memory 12 from a recording medium 13 such as a ROM card via the recording medium reading unit 14, or a control program that is downloaded from an external device (such as a server) and read into the memory 12 via a network (not shown) such as the Internet.

[0021] The control program stored in the memory 12 is activated in response to an input signal from the key input unit 16 in response to a user operation, an input signal from the touch panel display unit 17 in response to a user operation, or a connection communication signal with an external recording medium 13 such as an EEPROM (registered trademark), RAM, or ROM connected via the recording medium reading unit 14.

[0022] The CPU 11 is connected to a memory 12, a recording medium reading unit 14, a communication unit 15, a key input unit 16, a touch panel display unit 17, a voice input unit (microphone) 18, and the like.

[0023] The control programs stored in the memory 12 include a dictionary control program 12a, a learning processing program 12b, an authentication program 12c, and a content processing program 12d for each of a plurality of different content processing functions. The dictionary control program 12a is a program for controlling the overall operation of the electronic dictionary 10. The dictionary control program 12a includes a handwritten character recognition program for recognizing characters handwritten on the touch panel display unit 17.

[0024] The learning processing program 12b is a program for realizing, for example, a language learning function. For example, the learning processing program 12b realizes, as a language learning function, a speaking learning function for learning pronunciation and a listening learning function for listening to speech by a native speaker (speaker of the native language). The learning processing program 12b includes a pronunciation testing program 12b1 for executing a process for determining whether the user's utterance of a word or sentence (text) in speaking learning is correct pronunciation (for example, pronunciation in the same way as a native speaker).

[0025] The pronunciation test program 12b1 recognizes the speech of at least one word (word by word, sentence by sentence, paragraph by paragraph, etc.) uttered by the user, which is input from the speech input unit 18, and derives an evaluation value based on a comparison with a preset reference value (for example, speech by a native speaker (third speech)). In the electronic dictionary 10 of this embodiment, the result of the pronunciation test by the pronunciation test program 12b1, including the evaluation value (word learning data 12f, described later), is registered as a speech password (authentication determination information) in the process for personal authentication, and is used to determine whether content processing can be executed.

[0026] The pronunciation test program 12b1 derives an evaluation value not only during speaking practice but also in conjunction with other processes. For example, when a lock is set using a security function (described later), speaking practice is not performed, and there are insufficient pronunciation test results (word learning data 12f) available for personal authentication processing. The pronunciation test program 12b1 inputs the voice of at least one word uttered by the user via the voice input unit 18, and derives an evaluation value in the same manner as during speaking practice. The evaluation value (first evaluation value) is registered as authentication determination information for determining whether or not at least one content process can be performed.

[0027] The authentication program 12c is a program for realizing a security function using personal authentication. In the electronic dictionary 10 of this embodiment, personal authentication is performed based on the results of a pronunciation test (word learning data 12f) performed by the pronunciation test program 12b1, i.e., an evaluation value derived for the voice of at least one word spoken by the user. In addition to the evaluation value derived for the voice, authentication can also be performed using a multi-digit character (e.g., a four-digit number) as a password.

[0028] Authentication program 12c includes lock setting program 12c1 that executes lock setting processing and lock release program 12c2 that executes lock release processing. In the lock setting processing, the target of management by the security function, that is, the content processing (content) that is the target of authentication based on the evaluation value derived by the pronunciation test processing, is set. In the lock release processing, the lock is released for the target that has been locked (password) in the lock setting processing.

[0029] The content processing program 12d is a program for implementing various content processing functions. The content processing program 12d includes programs for each of a plurality of different content processing functions. Examples of content handled by the content processing program 12d include calendars, flashcards, notebooks, sticky notes, flashcards, timetables, and download histories. Because some content data processed by the content processing program 12d includes personal information, the electronic dictionary 10 of this embodiment uses a security function to enable locking and unlocking for each piece of content.

[0030] The memory 12 stores dictionary data 12e, word learning data 12f, password setting content data 12g, lock number data 12h, password word data 12k, and the like.

[0031] The dictionary data 12e includes a database that collects dictionary contents such as a plurality of dictionaries, such as an English-Japanese dictionary, a Japanese-English dictionary, an English-English dictionary, an English-Chinese dictionary, and a Japanese dictionary, as well as a plurality of types of encyclopedias, etc. The dictionary data 12e associates semantic information that explains the meaning (word meaning) of each headword with each dictionary.

[0032] The dictionary data 12e may not be built into the main body of the electronic dictionary 10, but may be stored in an external device (such as a server) that can be accessed via a network.

[0033] The word learning data 12f is data showing the processing results of the learning processing program 12b (pronunciation test program 12b1). That is, the word learning data 12f includes an evaluation value for each word derived based on the voice input by vocalizing the word. The word learning data 12f is used for processing (voice password) for personal authentication of the security function by the authentication program 12c.

[0034] Password setting content data 12g is data indicating content (content processing) that has been set as a target for management by the security function through the lock setting process of authentication program 12c (lock setting program 12c1).

[0035] The lock number data 12h is data (password) used for processing for personal authentication of the security function by the authentication program 12c. The lock number data 12h is designated by a user using a multi-digit character (for example, a four-digit number) during lock setting processing, for example.

[0036] The password word data 12k is data indicating words used in processing for personal authentication of the security function by the authentication program 12c. At least one word is set in the password word data 12k for an object managed by the security function through the unlocking processing of the authentication program 12c (unlocking program 12c2). For example, one word or multiple words are selected based on predetermined conditions for one or multiple contents indicated by the password setting content data 12g.

[0037] The communication unit 15 performs communication control to communicate with other electronic devices via a network such as the Internet or a LAN (Local Area Network), or to perform communication control to perform short-range wireless communication such as Bluetooth (registered trademark) or Wi-Fi (registered trademark) between other electronic devices in the vicinity.

[0038] In the electronic dictionary 10 configured in this manner, the CPU 11 controls the operation of each part of the circuit in accordance with the instructions written in various programs such as the dictionary control program 12a and the learning processing program 12b, and the software and hardware work together to realize the functions described in the following operational explanation.

[0039] FIG. 2 is a front view showing the external configuration of electronic dictionary 10 according to this embodiment.

[0040] In the case of electronic dictionary 10 in Figure 2, the lower section of the device body, which can be opened and closed, houses CPU 11, memory 12, recording medium reading section 14, communication section 15, and voice input section 18, as well as key input section 16, and the upper section houses touch panel display section 17.

[0041] The key input section 16 is provided with character input keys 16a, a dictionary selection key 16b for selecting various dictionaries and various functions, a [Translate / Confirm] key 16c, a [Back] key 16d, cursor keys (up, down, left, right keys) 16e, a delete key 16f, a power button, and various other function keys.

[0042] Various menus, buttons 17a, etc. are displayed on touch panel display unit 17 in accordance with the execution of various functions. Touch panel display unit 17 allows touch operations to select various menus or buttons 17a, and handwritten character input to input characters, using, for example, a pen.

[0043] In addition, in the case of handwritten character input, when a pattern representing a character is handwritten with a pen in a handwritten character input area displayed on touch panel display unit 17, character recognition processing is performed on the pattern. Characters corresponding to the pattern obtained by the character recognition processing are displayed in the character input area of ​​touch panel display unit 17 in the same way as characters input by operating character input keys 16a of key input unit 16. Therefore, character strings for dictionary search can be input by handwriting characters on touch panel display unit 17.

[0044] The electronic dictionary 10 also allows characters to be input by voice. The voice input unit 18 inputs the voice spoken by the learner. A voice recognition process is performed on the input voice, and a character string corresponding to the spoken voice is input. The characters corresponding to the utterance obtained by the voice recognition process are displayed in the character input area of ​​the touch panel display unit 17, in the same way as characters input by operating the character input keys 16a of the key input unit 16. Therefore, a character string for dictionary search can be input by handwriting on the touch panel display unit 17.

[0045] Next, the operation of the electronic dictionary 10 in this embodiment will be described.

[0046] First, the language learning function provided in the electronic dictionary 10 will be described.

[0047] The learning function of the electronic dictionary 10 can be started by operating the keys on the key input unit 16 or by selecting a button 17 a displayed on the touch panel display unit 17 , for example.

[0048] When the CPU 11 detects an operation to start the learning function, it executes the learning processing program 12b to operate the language learning function. The language learning function executes, for example, speaking learning for learning pronunciation or listening learning for listening to speech by a native speaker (speaker of the native language) in response to a user's instruction.

[0049] When instructed to perform speaking learning, the CPU 11 executes the pronunciation test program 12b1 to perform a pronunciation test process for determining whether the user's utterance of words and sentences (text) is correct pronunciation (e.g., pronunciation in the same way as a native speaker).

[0050] For example, as a pronunciation test process, the CPU 11 displays a pre-prepared word or sentence (text) on the touch panel display unit 17, inputs the voice spoken by the user within a predetermined time from the voice input unit 18, and derives an evaluation value (first evaluation value) indicating whether the word or sentence (text) displayed by the input voice is pronounced correctly.

[0051] For example, the CPU 11 converts the voice input from the voice input unit 18 into voice data, calculates physical quantities such as power and frequency for each frame divided by a fixed time, and determines the section in which a word or sentence (text) was spoken, consonant / vowel, etc., using the temporal transition of these physical quantities. Based on the determination results, the CPU 11 scores multiple items (determination parameters), comprehensively evaluates the scores for the multiple items (item scores), and calculates a total score for the speech. The multiple items (determination parameters) include, for example, voice time, vowel determination, consonant determination, sharpness, smoothness, etc., and calculates the total score based on the item scores for each item.

[0052] The CPU 11 stores the multiple item scores and total score calculated by the pronunciation test as evaluation values ​​for the words or sentences (sentences) that were the subject of the pronunciation test, for example, as word learning data 12f for each word or sentence (sentence). In the pronunciation test program 12b1, multiple words or sentences (sentences) are prepared for the pronunciation test, and one word or sentence is arbitrarily selected from them and provided as the subject of the pronunciation test.

[0053] The CPU 11 stores the evaluation scores (item scores, total scores) of the pronunciation test for each word or sentence provided as the subject of the pronunciation test as the word learning data 12f. Furthermore, when the pronunciation test is conducted multiple times for the same word or sentence, the CPU 11 stores each evaluation value as the word learning data 12f.

[0054] FIG. 3 is a diagram showing an example of word learning data 12f stored by the pronunciation test process in this embodiment.

[0055] In the example shown in Figure 3, for example, pronunciation testing has been performed multiple times (four times in Figure 3) for the word "interest," and the total scores calculated based on the pronunciation testing results, "50," "49," "38," and "25," are stored in chronological order (in the example of "insterest" in Figure 3, the total score "50" is the most recent calculated data, and the total score "25" is the oldest calculated data). Note that, although not shown, item scores are also stored in the same way.

[0056] As shown in Figure 3, by storing the evaluation values ​​calculated from multiple pronunciation tests as a history, it is possible to estimate pronunciation trends from changes in the evaluation values, i.e., whether the score is increasing and pronunciation is improving, or whether the score is trending flat.

[0057] The above-described pronunciation test process is merely an example, and other pronunciation test process methods can be used as long as they can derive evaluation values ​​for the utterance of words or sentences (text) by the user.

[0058] Furthermore, in the above explanation, when the subject of the pronunciation test is a sentence (one sentence) or a sentence containing multiple sentences, an evaluation value is stored for each sentence / sentence, but it is also possible to derive an evaluation value for each word contained in the sentence / sentence and store the evaluation value for each word.

[0059] When an evaluation value is stored for each sentence / text, the utterance for unlocking can be a sentence or a sentence in the unlocking process described later. Also, when an evaluation value is stored for each word of the spoken sentence / text, the results of speaking training (pronunciation test) for the sentence / text can be used when the utterance for unlocking can be a word in the unlocking process described later.

[0060] Next, the locking process performed by the electronic dictionary 10 in this embodiment will be described.

[0061] Fig. 4 is a flowchart showing the lock setting process in this embodiment. Fig. 5 is a flowchart showing the voice password registration process in the lock setting process shown in Fig. 4. In the lock setting process, the content processing (content) to be managed by the security function, i.e., the content to be authenticated based on the evaluation value derived by the pronunciation test process, is set. In the electronic dictionary 10 of this embodiment, in order to protect the content data processed by the content processing program 12d, it is set as a management target (lock setting / unlocking) by the security function.

[0062] The object of management by the security function is not limited to content data processed by the content processing program 12d, but can also be other data handled in the electronic dictionary 10, such as user-specific data (personal information, etc.).

[0063] Fig. 6 is a diagram showing an example of an operation screen displayed on the touch panel display unit 17 of the electronic dictionary 10 in this embodiment. The operation screen shown in Fig. 6 is, for example, a home screen (initial screen), and includes an input area 31 for inputting an entry word (keyword) for searching dictionary content, a plurality of dictionary selection buttons 32 for selecting dictionary content to be used, and a plurality of content selection buttons 33 for selecting a content processing function to be used.

[0064] The content selection buttons 33 include buttons corresponding to content data (content processing functions) such as a calendar, a flashcard, a notebook, sticky notes, flashcards, a timetable, and a download history (DL history). The content selection buttons 33 also include a password setting button 33A for setting content data (content processing functions) to be managed by the security function.

[0065] When an operation to select any content data (content processing function) is performed using the content selection button 33, the CPU 11 executes content processing based on the content processing program 12d of the corresponding content processing function. The content data processed by the content processing is stored in the memory 12, for example.

[0066] Also, as shown in FIG. 6, when an operation to select password setting is performed, for example, by touching the password setting button 33A with a pen (or by specifying the password setting button 33A with the pointer P or cursor key 16e and operating the [Translate / Enter] key 16c), the CPU 11 starts the lock setting process using the lock setting program 12c1.

[0067] First, CPU 11 displays a password setting method selection screen on touch panel display unit 17 (step A1).

[0068] 7 is a diagram showing an example of a password setting method selection screen in this embodiment. As shown in Fig. 7, the password setting method selection screen includes a lock number password button 35 for setting "lock number" as a password (lock number password), and a voice password button 36 for setting "voice" as a password (voice password).

[0069] Here, when the voice password button 36 is selected (step A2, "Voice"), the CPU 11 determines whether word learning data 12f that can be used as a voice password exists. That is, the CPU 11 determines whether the pronunciation test process has been performed on the number of words used for voice authentication in the unlocking process described below. For example, if there is one word used for voice authentication, the CPU 11 determines that there is a voice password if there is word learning data 12f for at least that one word. Also, if there are multiple words used for voice authentication, the CPU 11 determines that there is a voice password if there is word learning data 12f for that number of words.

[0070] If it is not determined that the voice password exists (No in step A4), the CPU 11 executes a process for allowing the user to add word learning data 12f that can be used in the unlocking process.

[0071] First, the CPU 11 displays an audio password registration guidance screen on the touch panel display unit 17 (step A7).

[0072] 8 is a diagram showing an example of a voice password registration guidance screen in this embodiment. The voice password registration guidance screen shown in FIG. 8 displays, for example, a message informing the user that voice input of English words is required.

[0073] As shown in FIG. 8, when an operation to instruct execution of voice registration is performed, for example, by touching the registration button 37 displayed on the voice password registration guidance screen with a pen (or by specifying the registration button 37 with the pointer P and operating the [Translate / Confirm] key 16c), the CPU 11 starts the voice password registration process (flowchart shown in FIG. 5).

[0074] The CPU 11 displays a voice password registration screen on the touch panel display unit 17, and prompts the user to input a voice password.

[0075] FIG. 9 is a diagram showing an example of the voice password registration screen in this embodiment.

[0076] The CPU 11 displays one word ("condition" in FIG. 9) prepared in advance for voice password registration on the voice password registration screen, and also displays a message such as "Please pronounce" or "Recording" (step B1). After displaying the word, the CPU 11 waits for voice input from the voice input unit 18 for a preset time during which the user can speak the word (step B2).

[0077] When speech is input from speech input unit 18, CPU 11 executes pronunciation test processing for this speech using pronunciation test program 12b1. That is, CPU 11 executes pronunciation test processing to derive an evaluation value for the utterance of a word in the speaking learning described above, and adds the evaluation value for the word displayed on touch panel display unit 17 to word learning data 12f and stores it as authentication determination information for determining whether content processing can be executed (step B3).

[0078] In the voice password registration process, an evaluation value is derived and registered using a process similar to the pronunciation test process executed during speaking practice. Therefore, the evaluation value stored during speaking practice and the evaluation value stored by the voice password registration process can be used equally in the unlocking process described below.

[0079] If word learning data 12f for the number of words required for the unlocking process has not been registered (step B4, No), CPU 11 displays a word different from the previous word ("condition") on touch panel display 17, inputs a voice uttered by the user in response to this word, derives an evaluation value through pronunciation testing, and adds it to word learning data 12f (steps B2 and B3), as described above. In this way, CPU 11 repeats the process of deriving evaluation values ​​for different words through pronunciation testing and adding them to word learning data 12f until registration of word learning data 12f for the number of words required for the unlocking process is completed.

[0080] When the CPU 11 has completed registration of the word learning data 12f for the number of words required for the unlocking process (step B4, Yes), it terminates the voice password registration process and proceeds to processing for selecting content to be managed (subject to password setting) by the security function.

[0081] If the setting of the voice password is confirmed (step A2, "Voice") and it is determined that word learning data 12f usable as a voice password exists (voice password exists) (step A4, Yes), CPU 11 displays a re-registration confirmation screen on touch panel display 17 to ask the user whether to add word learning data 12f usable as a voice password (step A5). For example, the re-registration confirmation screen may display buttons for "Add / No" and prompt the user to select either option using pointer P.

[0082] If an instruction to "register additionally" is given (step A6, Yes), CPU 11 executes the process for registering the voice password described above (steps A7 and A8). On the other hand, if an instruction to "not register additionally" is given (step A6, No), CPU 11 proceeds to a process for selecting content (content processing that requires a determination of whether it can be executed) to be managed (subject to password setting) by the next security function (step A6, No).

[0083] When the process shifts to the process for selecting content to be managed by the security function, the CPU 11 displays a password setting content selection screen on the touch panel display unit 17 (step A9).

[0084] FIG. 10 is a diagram showing an example of a password setting content selection screen in this embodiment.

[0085] The password setting content selection screen is provided with a content selection area 41 for selecting the content to be managed by the security function from the content handled by the content processing program 12d, such as a calendar, a flashcard, a notebook, sticky notes, flashcards, a timetable, and a download (DL) history. The password setting content selection screen also has a decision button 42 and a cancel button 43.

[0086] A check box is provided for each piece of content displayed in the content selection area 41. The CPU 11 displays a check mark in the check box in response to a selection input operation by the user using, for example, the pointer P (or a pen), to indicate that the content has been selected as a target for management by the security function (step A10).

[0087] Here, if no operation is performed on the decision button 42 or the cancel button 43 (steps A11, A12, No), the CPU 11 executes the process of selecting or canceling content (selection operation on a selected check box) in accordance with the selection input operation on the check box in the content selection area 41 (steps A10 to A12).

[0088] The example shown in FIG. 10 indicates that the content notes, sticky notes, and download history are selected.

[0089] Here, if the cancel button 43 is operated (step A11, Yes), the CPU 11 determines that the user has instructed to cancel the password setting (lock setting process), and ends the process.

[0090] On the other hand, if the decision button 42 is operated (step A12, Yes), the CPU 11 stores the password setting content data 12g indicating the content selected (displaying a check mark) in the content selection area 41 in the memory 12 as content to be managed by the security function.

[0091] 12 is a diagram showing an example of password setting content data 12g in this embodiment. The password setting content data 12g shown in Fig. 12 indicates the contents of a note, a sticky note, and a download history selected by the user on the password setting content selection screen.

[0092] After storing the password setting content data 12g, the CPU 11 displays a setting completion screen on the touch panel display unit 17 (step A13).

[0093] FIG. 11 is a diagram showing an example of the setting completion screen in this embodiment.

[0094] The setting completion screen displays a message informing the user that password setting has been completed, along with a back button 45. CPU 11 ends the lock setting process in response to an operation on back button 45 using a pen or pointer P, and causes touch panel display unit 17 to display the home screen (initial screen).

[0095] Although the case of setting a voice password has been described, if an operation to select the lock number button 35 is performed in step A2 (step A2, "Lock No."), the CPU 11 can execute the lock number setting process (step A3).

[0096] In the lock number setting process, CPU 11, for example, displays a lock number setting screen on touch panel display unit 17, and has the user input a multi-digit character string (for example, a four-digit number) to be used as a "lock number" through this screen. For example, the user inputs a multi-digit character string to be used as a lock number password by operating a software keyboard (numeric keypad) on touch panel display unit 17 with a pen, or by operating keys on key input unit 16.

[0097] The CPU 11 stores the input multi-digit characters as lock No. data 12h in the memory 12. Thereafter, similar to the case of setting the voice password described above, a process for selecting content to be managed by the security function is executed (steps A9 to A13).

[0098] In the electronic dictionary 10 of this embodiment, a "lock number" or "voice" can be selected and used as a password, but for multiple contents, some contents use a "lock number" as a password and other contents use a "voice" as a password. Also, it is possible to set both a "lock number" and a "voice" as a password for one content.

[0099] This allows the user to use either the "lock number" password or the "voice" password according to their convenience, making content management by the security function more convenient.

[0100] In this way, in the lock setting process of this embodiment, the user can arbitrarily select and set the content to be managed by the security function from multiple contents. Also, if the voice password (word learning data 12f) required for the unlocking process is not available during the lock setting process, it is possible to register a voice password at that time. Therefore, prior speaking practice is not required to use the voice-based security function, which makes it easier to use the voice-based security function.

[0101] Next, the unlocking process in this embodiment will be described.

[0102] FIG. 13 is a flowchart showing the unlocking process in this embodiment.

[0103] The unlocking process is started when the content processing requested to be executed has been set as a target for management by the security function in the lock setting process.

[0104] FIG. 14 is a diagram showing an example of an operation screen displayed on the touch panel display unit 17 of the electronic dictionary 10 in this embodiment.

[0105] 14, for example, when the sticky note button 33B for selecting the content "sticky note" of the content selection buttons 33 is selected by operation of the pen or pointer P, the CPU 11 determines that a request to execute the "sticky note" content (content processing function) has been accepted. The CPU 11 refers to the password setting content data 12g and determines whether the "sticky note" content is set as a target for management by the security function.

[0106] Here, if the password-set content data 12g indicates the content of "sticky note", the CPU 11 starts the unlocking process by the unlocking program 12c2.

[0107] First, CPU 11 causes touch panel display unit 17 to display an authentication method selection screen (step C1).

[0108] Fig. 15 is a diagram showing an example of an authentication method selection screen in this embodiment. As shown in Fig. 15, the authentication method selection screen includes a lock number input area 51 (for example, an input area for four characters) for inputting a "lock number" to be used as a password, and a voice recognition button 52 for instructing execution of authentication using a voice password.

[0109] Here, when a multi-digit character string (for example, a four-digit number) to be used as the "lock number" is entered into the lock number input area 51 by operating the software keyboard (numeric keypad) with a pen or by operating the keys on the key input unit 16, the CPU 11 determines that authentication processing has been requested using the "lock number" as a password (lock number password) (step C2, "lock number"), and executes the lock number authentication processing (step C3).

[0110] In the lock number authentication process, the CPU 11 collates the characters (for example, a four-digit number) entered in the lock number input area 51 on the authentication method selection screen with the lock number password indicated by the lock number data 12h.

[0111] If the entered characters match the lock number password, the CPU 11 determines that authentication using the lock number password has been successful and that unlocking of content processing is possible (Yes in step C14).

[0112] The CPU 11 releases the lock on the content for which an execution request has been instructed on the operation screen, and proceeds to start up a content processing function that executes processing on this content (step C13).

[0113] If the entered characters do not match the lock number password, the CPU 11 determines that authentication using the lock number password has failed, invalidates the execution request instructed on the operation screen, and terminates the processing (does not start the content processing function processing instructed by the execution request) (step C14, No).

[0114] On the other hand, if an operation to select the voice recognition button 52 is performed on the authentication method selection screen shown in FIG. 15 (step C2, "voice"), the CPU 11 proceeds to authentication processing using a voice password.

[0115] First, the CPU 11 selects at least one word for voice authentication from words whose evaluation values ​​have already been derived by the pronunciation test process, that is, from words whose evaluation values ​​have been stored as the word learning data 12f (step C4).

[0116] CPU 11 displays a voice authentication screen including at least one word selected for voice authentication on touch panel display unit 17, and prompts the user to input voice (step C5).

[0117] Fig. 16 is a diagram showing an example of a voice authentication screen in this embodiment. In the example shown in Fig. 16, three words, "condition," "taught," and "occupy," are selected as words for voice authentication.

[0118] The CPU 11 displays the word selected for voice authentication on the voice authentication screen, and also displays messages such as "Release by voice authentication. Please pronounce in order" and "Recording." From the time when the word is displayed on the voice authentication screen, the CPU 11 waits for voice input from the voice input unit 18 for a preset time during which the user can speak the word (step C6).

[0119] When the CPU 11 receives input of at least one word uttered by the user from the speech input unit 18, the CPU 11 executes a pronunciation test process for the speech corresponding to the one word using the pronunciation test program 12b1. That is, the CPU 11 executes a pronunciation test process for deriving an evaluation value (second evaluation value) for the utterance of the word in the same manner as in the speaking learning described above (step C7). The CPU 11 temporarily stores the evaluation value (second evaluation value) derived for the utterance of the word in the memory 12.

[0120] If the CPU 11 has not input the voice for the number of words displayed on the voice authentication screen (step C8, No), it performs the pronunciation verification process in the same manner as described above on the voice for the next word that is subsequently input from the voice input unit 18 (steps C6, C7).

[0121] When the input of voice for the number of words displayed on the voice authentication screen is completed and an evaluation value (second evaluation value) for each utterance is derived (step C8, Yes), the CPU 11 proceeds to the next authentication process.

[0122] Now, the selection of words for voice authentication will be described.

[0123] The above explanation uses an example of selecting three words, but the number of words for voice authentication or the target words can be selected for each voice recognition process (unlock process) as follows.

[0124] (A1) The number of words for voice authentication is preset within a range of about 1 to 10. The number of words may be predetermined in the unlock program 12c2, or may be set in response to an instruction from the user in a setting process that is executed separately.

[0125] Furthermore, the number of words may be varied. For example, in the lock setting process, the number of words for voice authentication may be changed depending on the number of contents selected as targets for security function management. For example, if the number of contents is 1 or 2, the number of words may be 1, and if the number of contents is 3 or 4, the number of words may be 2.

[0126] The number of words may also be changed depending on the type of content selected as the object of security management. For example, if content containing important personal information (such as a calendar or download history) is selected, the number of words may be increased.

[0127] In this way, by changing the number of words for voice authentication depending on the content, it is possible to adjust (vary) the security level using a voice password depending on the content.

[0128] (A2) When selecting words for speech authentication from the words (word learning data 12f) for which evaluation values ​​have already been derived, the words are basically selected randomly. This allows the learning results of speaking practice to be effectively utilized.

[0129] Additionally, words are selected from those for which evaluation values ​​have been stored within a predetermined recent period. For example, words for which speaking practice has been conducted within the past month are selected. This allows for appropriate authentication processing to be performed on words for which speaking practice has been conducted using utterances close to the user's current speaking level.

[0130] Alternatively, words that have been practiced for a predetermined number of times may be selected as targets. For example, words that have been practiced for a large number of times may be selected as targets. This allows authentication to be performed on words with highly reliable evaluation values.

[0131] Furthermore, it is possible to select words whose evaluation value falls within a predetermined range, for example, a range that excludes words with high or low scores. This makes it possible to exclude words with evaluation values ​​that are easy for people other than the user to obtain in the pronunciation test, thereby improving the reliability of authentication.

[0132] Alternatively, the evaluation value may be divided into multiple categories and words may be selected from each category. This allows for a mixture of words with high, medium, and low evaluation values ​​to be selected, making it more difficult for someone other than the user to impersonate the user, thereby increasing the security level.

[0133] (A3) The evaluation value of a word selected for voice authentication is the evaluation value from the last time it was spoken and learned. Alternatively, the highest (or lowest) evaluation value of a word selected for voice authentication may be used. This allows the user, when speaking a word during the unlock process, to speak it in the same way as when it was last spoken and learned, or in the same way as when it had the highest (or lowest) evaluation value. In other words, it is difficult for anyone other than the user to imitate the user and speak in a way that will result in successful authentication, thereby improving the security level.

[0134] Furthermore, if the selected word has been learned through speaking practice multiple times, the evaluation value for authentication (second evaluation value) may be derived based on the multiple evaluation values. For example, the evaluation value for authentication (second evaluation value) may be determined by calculating the average of the evaluation values ​​from all speaking practice sessions, calculating the average of the evaluation values ​​from the most recent multiple speaking practice sessions (e.g., three sessions), or calculating the average of the evaluation values ​​from all speaking practice sessions that are higher (or lower) than a preset reference value. This allows the authentication process to be performed with a stable evaluation value for the word used for voice authentication, even if the same user speaks the same word and the evaluation value derived for each pronunciation test process varies.

[0135] Furthermore, if the selected word has been studied for speaking multiple times, multiple evaluation values ​​derived from the multiple speaking lessons are used. This allows the evaluation values ​​derived from multiple pronunciation tests to be used as a history, and authentication processing can be performed using pronunciation trends based on changes in the evaluation values, i.e., whether the score is increasing and pronunciation is improving, or whether the score is remaining stable.

[0136] Next, the CPU 11 compares the evaluation value (first evaluation value) of the word displayed on the voice authentication screen with the evaluation value (second evaluation value) derived for the voice input corresponding to the word displayed on the voice authentication screen (step C9), and determines whether the comparison result satisfies a predetermined condition (step C10). That is, the CPU 11 executes an authentication process to perform personal authentication to determine whether the utterance in response to the word displayed on the voice authentication screen is made by a legitimate user.

[0137] Here, we will explain the method and judgment conditions for comparing the evaluation value (first evaluation value) of the word displayed on the voice authentication screen with the evaluation value (second evaluation value) derived for the voice input according to the word displayed on the voice authentication screen.

[0138] (B1) When there is one word for voice authentication, if the difference between the first evaluation value and the second evaluation value is within a predetermined range, it is determined that the authentication is successful (the user is legitimate).

[0139] (B2) When there are multiple words for voice authentication, if the sum (or average) of the differences between the multiple first evaluation values ​​for each word and the multiple corresponding second evaluation values ​​is within a predetermined range, the authentication is determined to be successful (the user is legitimate).

[0140] (B3) If there are multiple words for voice authentication, the authentication is determined to be successful (the user is legitimate) if the difference between the multiple first evaluation values ​​and the corresponding second evaluation values ​​is within a predetermined range, and the number or percentage of words whose difference is determined to be within the predetermined range is greater than a preset value. In this way, by making a judgment using the evaluation values ​​of a plurality of words, it is possible to maintain a security level and to realize stable personal authentication that can appropriately identify legitimate users.

[0141] (B4) When multiple first evaluation values ​​derived from multiple pronunciation tests for a single word are used, the pronunciation trend is determined based on changes in the first evaluation values, i.e., whether the score is increasing, indicating improvement in pronunciation, or whether the score is trending flat. For example, if the multiple first evaluation values ​​are trending upward, and the second evaluation value is higher than the most recent first evaluation value among the multiple first evaluation values, and the difference between the most recent first evaluation value and the second evaluation value is within a predetermined reference range, authentication is determined to be successful (the user is legitimate). Note that the reference range may be changed depending on the upward trend. For example, if the upward trend is significant, the reference range is widened to make it easier for the second evaluation value to fall within the reference range. Furthermore, if the multiple first evaluation values ​​are trending flat, the reference range is narrowed to make it easier to determine that authentication is successful when the first evaluation value and the second evaluation value are closer. In this way, by utilizing the pronunciation trend indicated by the multiple first evaluation values, personal authentication can be performed according to the user's current speaking level.

[0142] (B5) The evaluation value includes a total score and an item score, but both or either one may be used. When using item scores, all scores of multiple items (judgment parameters) including speech time, vowel judgment, consonant judgment, sharpness, and smoothness may be used, or only some of the item scores may be used.

[0143] (B6) In the above explanation, a word is displayed on the voice authentication screen, and the voice input for this word is subjected to authentication processing. However, a single sentence or a paragraph (multiple sentences) may be displayed on the voice authentication screen, and authentication processing may be performed for the voice uttered for the sentence or paragraph. In this case, authentication processing can be performed as described above based on evaluation values ​​for multiple words contained in the sentence / paragraph. Furthermore, when an evaluation value for each sentence / paragraph is derived through a pronunciation test, authentication processing can be performed using the evaluation value for each sentence / paragraph in the same manner as when the evaluation value for each word is used.

[0144] The CPU 11 determines whether to unlock the content based on whether the comparison result between the first evaluation value and the second evaluation value satisfies a predetermined condition, and if it determines that the content processing can be unlocked (step C11, Yes), it unlocks the content for which an execution request has been instructed on the operation screen and proceeds to start up the content processing function that executes processing on this content (step C13).

[0145] The CPU 11 starts content processing by the content processing program 12d of the content processing function whose execution request has been instructed on the operation screen.

[0146] 17 is a diagram showing an initial screen of a content processing function for which an execution request is instructed on the operation screen in this embodiment. Fig. 17 shows the initial screen on which content processing of "sticky note" has started.

[0147] On the other hand, if unlocking of content processing is not possible (No in step C11), CPU 11 displays a retry confirmation screen on touch panel display unit 17 to ask the user whether to perform voice authentication again.

[0148] FIG. 18 is a diagram showing an example of a retry confirmation screen in this embodiment.

[0149] The retry confirmation screen displays a message such as "Voice authentication failed. Please try again or select back." along with a retry button 55 and a back button 56.

[0150] Here, if the retry button 55 is selected (step C12, Yes), the CPU 11 selects a new voice authentication word (step C4) and displays a voice authentication screen on the touch panel display unit 17 (step C5). That is, the CPU 11 uses a voice authentication word different from the previous time and executes the same authentication process as described above (steps C4 to C11).

[0151] On the other hand, if the back button 56 is selected (No in step C12), the CPU 11 determines that an instruction to stop the authentication process using the voice password has been given, and ends the authentication process.

[0152] In this way, electronic dictionary 10 in this embodiment performs personal authentication based on the evaluation value obtained as a result of a pronunciation test for speech input during speaking practice, eliminating the need to register a password for personal authentication in advance and reducing the burden on the user because the user does not need to remember the registered password. Furthermore, since personal authentication is based on the evaluation value of the pronunciation test, all that is required is an environment in which speech can be input from speech input unit 18, and the usage conditions are not strict, making it easy to use.

[0153] Furthermore, personal authentication is based on the evaluation score of the user's utterance of a word (or sentence / phrase), so it is difficult for anyone other than the user to utter a word that will derive an evaluation score equivalent to that of the user through the pronunciation test process, thereby ensuring security. In particular, by combining multiple words into a voice password, it is easy to change the security level.

[0154] Therefore, electronic dictionary 10 according to the present embodiment can easily perform personal authentication while ensuring security.

[0155] Furthermore, the methods described in the embodiments, i.e., the methods such as the processes shown in the flowcharts of Figures 4, 5, and 13, can be stored as a program that can be executed by a computer and distributed on a recording medium such as a memory card (a ROM card, a RAM card, etc.), a magnetic disk (a flexible disk, a hard disk, etc.), an optical disk (a CD-ROM, a DVD, etc.), a semiconductor memory, etc. Then, the computer reads the program recorded on the recording medium, and the operation is controlled by this program, thereby realizing the same processes as the functions described in the embodiments.

[0156] In addition, the program data for realizing each technique can be transmitted over a network (such as the Internet) in the form of program code, and the program data can be imported from a computer (such as a server device) connected to this network to realize functions similar to those of the above-mentioned embodiments.

[0157] The present invention is not limited to the embodiments, and various modifications can be made in the implementation stage without departing from the gist of the invention. Furthermore, the embodiments include inventions at various stages, and various inventions can be extracted by appropriately combining the disclosed multiple constituent elements. For example, even if some constituent elements are deleted from all the constituent elements shown in the embodiments or some constituent elements are combined, if the problem stated in the "Problem to be Solved by the Invention" section can be solved and the effect stated in the "Effects of the Invention" section can be obtained, the configuration in which these constituent elements are deleted or combined can be extracted as an invention.

[0158] The inventions described in the original claims of this application are set forth below.

[0159] [1] An electronic device including a control unit, The control unit When acquiring at least one word uttered by a user, deriving a first evaluation value based on a comparison with a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the at least one word is presented as information to be uttered by a user; When a voice of the at least one word uttered by the user is acquired based on the presented information, a second evaluation value is derived based on a comparison with the reference value; an electronic device that executes the at least one content process when it is determined that a comparison result between the first evaluation value registered as the authentication judgment information and the second evaluation value satisfies a predetermined condition.

[0160] [2] The control unit The electronic device according to [1], wherein the first evaluation value is derived for at least one word uttered by a user based on a third speech uttered by a speaker whose native language is the language of the at least one word.

[0161] [3] The control unit An electronic device according to [1] or [2], wherein, when setting up content processing that requires a determination of whether it can be executed, at least one word is displayed as information for the user to speak, and the first evaluation value is derived for the voice obtained in response to the at least one word.

[0162] [4] The control unit An electronic device described in any of [1] to [4], which compares a plurality of first evaluation values ​​corresponding to each of a plurality of sounds uttered by a user with a plurality of second evaluation values ​​derived for each of a plurality of sounds corresponding to a plurality of words uttered by the user based on the information, and determines whether content processing can be executed based on the comparison results between the plurality of first evaluation values ​​and the second evaluation values.

[0163] [5] The control unit The electronic device according to any one of [1] to [3], wherein if the difference between the first evaluation value and the second evaluation value is within a predetermined range, it is determined that the comparison result satisfies the predetermined condition, and the at least one content processing is executed.

[0164] [6] The control unit An electronic device as described in [5], which determines that the comparison result satisfies the predetermined condition and executes at least one content processing when the sum of the differences between the multiple first evaluation values ​​and the corresponding multiple second evaluation values ​​is within a predetermined range.

[0165] [7] The control unit An electronic device as described in [5], which determines that the comparison result satisfies the predetermined condition and executes at least one content processing when the difference between the multiple first evaluation values ​​and the corresponding multiple second evaluation values ​​is within a predetermined range and the number or percentage of words for which the difference is determined to be within the predetermined range is greater than a default value.

[0166] [8] The control unit The electronic device according to any one of [1] to [5], wherein the number of words to be presented is changed in accordance with the content process for which an execution request has been received.

[0167] [9] The control unit The electronic device according to any one of [1] to [5], wherein the word to be presented is changed each time a request to execute the content processing is received.

[0168]

[10] An authentication method in an electronic device including a control unit, The control unit When acquiring at least one word uttered by a user, deriving a first evaluation value based on a comparison with a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the at least one word is presented as information to be uttered by a user; When a voice of the at least one word uttered by the user is acquired based on the presented information, a second evaluation value is derived based on a comparison with the reference value; An authentication method that executes the at least one content process when it is determined that the comparison result of comparing the first evaluation value registered as the authentication judgment information with the second evaluation value satisfies a predetermined condition.

[0169]

[11] Computer, When acquiring at least one word uttered by a user, deriving a first evaluation value based on a comparison with a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the at least one word is presented as information to be uttered by a user; When a voice of the at least one word uttered by the user is acquired based on the presented information, a second evaluation value is derived based on a comparison with the reference value; An authentication program that functions to execute the at least one content processing when it is determined that the comparison result between the first evaluation value registered as the authentication judgment information and the second evaluation value satisfies a predetermined condition. [Explanation of symbols]

[0170] 10...electronic dictionary, 11...CPU, 12...memory, 12a...dictionary control processing program, 12b...learning processing program, 12b1...pronunciation test program, 12c...authentication program, 12c1...lock setting program, 12c2...lock release program, 12d...content program, 12e...dictionary data, 12f...word learning data, 12g...password setting content data, 12h...lock number data, 12k...password word data, 13...external recording medium, 14...recording medium reading unit, 15...communication unit, 16...key input unit, 17...touch panel display unit, 18...voice input unit.

Claims

1. An electronic device including a control unit, The control unit When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a numerical value based on a comparison of pronunciation between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a numerical value based on a comparison of pronunciation between the second speech and the reference value; an electronic device that executes the at least one content process when it is determined that a comparison result between the first evaluation value registered as the authentication judgment information and the second evaluation value satisfies a predetermined condition.

2. An electronic device equipped with a control unit, The control unit When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a value based on a comparison between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a value based on a comparison between the second speech and the reference value; an electronic device that determines that a predetermined condition is satisfied when a sum of differences between the plurality of first evaluation values ​​registered as the authentication judgment information and the plurality of second evaluation values ​​corresponding thereto is within a predetermined range, and executes the at least one content processing.

3. An electronic device equipped with a control unit, The control unit When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a value based on a comparison between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a value based on a comparison between the second speech and the reference value; an electronic device that determines that a predetermined condition is met and executes at least one content process when the difference between the plurality of first evaluation values ​​registered as the authentication judgment information and the corresponding plurality of second evaluation values ​​is within a predetermined range and the number or percentage of words for which the difference is determined to be within the predetermined range is greater than a default value.

4. The control unit calculating the first evaluation value by comparing the first voice with the reference value in terms of voice time, vowel determination, consonant determination, sharpness, and smoothness; calculating the second evaluation value by comparing the second voice with the reference value in terms of voice time, vowel determination, consonant determination, sharpness, and smoothness; 4. The electronic device according to claim 1.

5. The control unit 5. The electronic device according to claim 1, wherein the first evaluation value is derived for the first speech based on a third speech uttered by a speaker whose native language is a language of the at least one word.

6. The control unit An electronic device as described in any one of claims 1 to 5, wherein when setting up content processing that requires a determination of whether it can be executed, the at least one word is displayed as information for the user to utter, and the first evaluation value is derived for the first voice.

7. The control unit An electronic device as described in any one of claims 1 to 6, which compares a plurality of first evaluation values ​​corresponding to each of a plurality of sounds uttered by a user with a plurality of second evaluation values ​​derived for each of a plurality of sounds corresponding to a plurality of words uttered by the user based on the information, and determines whether content processing can be executed based on the comparison results between the plurality of first evaluation values ​​and the second evaluation values.

8. The control unit 7. The electronic device according to claim 1, wherein, when a difference between the first evaluation value and the second evaluation value is within a predetermined range, it determines that the predetermined condition is satisfied and executes the at least one content processing.

9. The control unit 9. The electronic device according to claim 1, wherein the number of words to be presented is changed in accordance with the content process for which an execution request has been accepted.

10. The control unit 9. The electronic device according to claim 1, wherein the words to be presented are changed each time a request for executing the content processing is received.

11. An authentication method in an electronic device including a control unit, The control unit When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a numerical value based on a comparison of pronunciation between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a numerical value based on a comparison of pronunciation between the second speech and the reference value; An authentication method comprising: comparing the first evaluation value registered as the authentication judgment information with the second evaluation value; and executing the at least one content processing if it is determined that the comparison result satisfies a predetermined condition.

12. Computer, When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a numerical value based on a comparison of pronunciation between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a numerical value based on a comparison of pronunciation between the second speech and the reference value; An authentication program for causing the program to function to execute at least one content processing when it is determined that the comparison result between the first evaluation value registered as the authentication judgment information and the second evaluation value satisfies a predetermined condition.

13. An authentication method in an electronic device equipped with a control unit, comprising: The control unit When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a value based on a comparison between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a value based on a comparison between the second speech and the reference value; An authentication method in which, if the sum of the differences between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple corresponding second evaluation values ​​is within a predetermined range, it is determined that a predetermined condition is met and the at least one content processing is executed.

14. A computer, When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a value based on a comparison between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a value based on a comparison between the second speech and the reference value; An authentication program for determining that a predetermined condition is met when the sum of the differences between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple corresponding second evaluation values ​​is within a predetermined range, and for executing at least one content processing.

15. An authentication method in an electronic device equipped with a control unit, comprising: The control unit When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a value based on a comparison between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a value based on a comparison between the second speech and the reference value; An authentication method in which, if the difference between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple corresponding second evaluation values ​​is within a predetermined range, and the number or percentage of words for which the difference is judged to be within the predetermined range is greater than a default value, it is judged that the predetermined condition is met and at least one content processing is executed.

16. A computer, When a first speech is acquired, the first speech being a speech in which at least one word is spoken by a user, a first evaluation value is derived, the first evaluation value being a value based on a comparison between the first speech and a preset reference value; registering the first evaluation value as authentication determination information for determining whether or not at least one content process can be executed; When a request to execute the at least one content process is received, the word is presented as information to be uttered by a user; When a second speech is acquired, the second speech being a speech in which the user utters the word based on the presented information, a second evaluation value is derived, the second evaluation value being a value based on a comparison between the second speech and the reference value; An authentication program that determines that a predetermined condition is met when the difference between the multiple first evaluation values ​​registered as the authentication judgment information and the multiple corresponding second evaluation values ​​is within a predetermined range, and the number or percentage of words for which the difference is determined to be within the predetermined range is greater than a default value, and functions to execute at least one content processing.

Citation Information

Patent Citations

  • Login verification method and device based on voiceprint recognition, equipment and storage medium

    CN109711129A

  • Method and device for deciding threshold value in speaker collation

    JP1999249684A

  • Voice evaluating device, method, and program

    JP2017126004A

  • Electronic device, acoustic device, control method of electronic device, and control program

    JP2019062377A

  • Voice authentication device, voice authentication system and voice authentication method

    JP2021064110A