Display device, display method, program, recording medium, and subtitle generation method

The display device addresses the issue of inappropriate subtitle display by using user attribute recognition and conversion to adapt subtitle content, ensuring appropriate word selection based on user attributes.

JP2025105127APending Publication Date: 2025-07-10SHARP KK
View PDF 1 Cites 0 Cited by

Patent Information

Application Number
JP2023223443
Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
Filing Date
2023-12-28
Publication Date
2025-07-10

AI Technical Summary

Technical Problem

Existing display devices fail to display subtitles appropriately based on the attributes of the user, often displaying words that may be difficult to understand or inappropriate for the viewer's age or educational level.

Method used

The display device includes an acquisition unit for video and subtitle data, a recognition unit to identify the user, an attribute-specific database to determine the user's attributes, a conversion unit to adjust subtitle words based on the user's attributes, and a display unit to superimpose the converted subtitles on the video.

Benefits of technology

Enables the display of subtitles that are appropriate for the user's attributes, such as age or educational level, ensuring that the content is accessible and suitable for the viewer.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2025105127000001_ABST
    Figure 2025105127000001_ABST
Patent Text Reader

Abstract

To provide a display device, and the like, capable of displaying a subtitle so as to include appropriate words in accordance with attributes of a viewing user, for example.SOLUTION: A display device includes: an acquisition unit configured to acquire video data and subtitle data from content; a recognition unit which recognizes a user who views the content; an attribute-specific database which stores words for each attribute of users; a determination unit which determines the attribute of the recognized user; a conversion unit which converts, by referring to the attribute-specific database, words included in the subtitle data in accordance with the attribute of the user recognized by the recognition unit; and a display unit which superimposes a subtitle based on the subtitle data converted by the conversion unit on the video data.SELECTED DRAWING: Figure 1
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present disclosure relates to a display device and the like.

Background Art

[0002] For example, as disclosed in Patent Document 1, there is disclosed a technique for providing a subtitle display device having a method for dealing with a case where word data superimposed on a video signal includes words that cannot be understood by viewers, words inappropriate for education, and copyright information.

Prior Art Documents

Patent Documents

[0003]

Patent Document 1

Summary of the Invention

Problems to be Solved by the Invention

[0004] One object of the present disclosure is to provide, for example, a display device or the like capable of displaying subtitles as appropriate words according to the attributes of a user who views the content.

Means for Solving the Problems

[0005] The display device of the present disclosure includes an acquisition unit that acquires video data and subtitle data from content, a recognition unit that recognizes a user who views the content, an attribute-specific database that stores words for each attribute of the user, a determination unit that determines the attribute of the recognized user, a conversion unit that converts words included in the subtitle data with reference to the attribute-specific database according to the attribute of the user recognized by the recognition unit, and a display unit that superimposes and displays subtitles based on the subtitle data converted by the conversion unit on the video data.

[0006] The display method of the present disclosure is a display method in a display device that stores an attribute-specific database for storing words for each user attribute, and includes an acquisition step of acquiring video data and subtitle data from content, a recognition step of recognizing a user who views the content, a determination step of determining the attribute of the recognized user, a conversion step of referring to the attribute-specific database and converting the words included in the subtitle data according to the attribute of the user recognized in the recognition step, and a display step of superimposing and displaying subtitles based on the subtitle data converted in the conversion step on the video data.

[0007] The program of the present disclosure realizes, in a computer that stores an attribute-specific database for storing words for each user attribute, an acquisition function of acquiring video data and subtitle data from content, a recognition function of recognizing a user who views the content, a determination function of determining the attribute of the recognized user, a conversion function of referring to the attribute-specific database and converting the words included in the subtitle data according to the attribute of the user recognized by the recognition function, and a display function of superimposing and displaying subtitles based on the subtitle data converted by the conversion function on the video data.

[0008] The subtitle creation method of the present disclosure is a subtitle generation method for generating subtitles to be displayed on a video of content in a device having an attribute-specific database for storing words for each user attribute, and includes an acquisition step of acquiring video data and subtitle data from content, a recognition step of recognizing a user who views the content, a determination step of determining the attribute of the recognized user, and a generation step of referring to the attribute-specific database and generating subtitles obtained by converting the words included in the subtitle data according to the attribute of the user recognized by the recognition unit.

Effect of the Invention

[0009] According to the present disclosure, for example, it is possible to display subtitles so as to be appropriate words according to the attribute of the viewing user.

Brief Description of the Drawings

[0010]

Figure 1

Figure 2

Figure 3

Figure 4

Figure 5

Figure 6

Figure 7

Figure 8

Figure 9

Figure 10

Figure 11

Figure 12

Figure 13

Figure 14

Figure 15

Figure 16

Figure 17

Figure 18

Mode for Carrying Out the Invention

[0011] Generally, a display device capable of displaying contents such as a program received as a broadcast wave, a video distributed on a video distribution site, etc., or an image output by a computer or the like is known. When displaying the contents, the display device is also configured to display subtitles.

[0012] Since subtitles include various words and phrases, there may be cases where they contain something unfavorable to the user. For example, there is a known technique in which a device stores "difficult-to-understand character strings" and "strings inappropriate for education" and displays a string that replaces the inappropriate string.

[0013] However, this character string simply replaces "difficult-to-understand character strings" and "strings inappropriate for education", and it was not always the case that an appropriate character string was displayed according to the attributes of the user.

[0014] Therefore, the display device and the like of the present disclosure make it possible to display subtitles so as to be appropriate words according to the attributes of the viewing user.

[0015] Hereinafter, the display device of the present disclosure will be described in the following embodiments with reference to the drawings. Note that the following embodiments are described as an example of the invention described in the claims, and the technical scope of the present invention is not limited to the description of the following embodiments.

[0016] [1. First Embodiment] Hereinafter, the first embodiment will be described. In the first embodiment, the following will be described as an example.

[0017] The characters and character strings displayed by the display device are obtained from the content and, as an example, display the subtitles (subtitle data) superimposed on the content. Note that the displayed characters may be, for example, the characters of teletext, the characters displayed by the display device such as a program guide, etc. Also, the characters include words consisting of one character or a plurality of characters. Further, the characters include characters and symbols that can be displayed by the display device, in addition to Japanese and English. Also, the characters include characters based on foreign character data included in the subtitle data described later.

[0018] The user attributes are displayed, taking the school year based on the user's age as an example. Note that the user attributes may be, for example, simply the age, generation (child, adult, elderly), etc.

[0019] Also, the characters displayed by the display device are displayed as a sentence including one or more words and phrases. Words are described taking words such as nouns and verbs as examples, but for example, any one unit (word, word, phrase) when decomposing a sentence may be sufficient. Also, in the following embodiments, words are described taking Japanese as an example, but it is not limited to Japanese.

[0020] [1.1 Entire Display Device] FIG. 1 is a diagram showing the entire display device 10. The display device 10 is a device capable of displaying content. For example, in this embodiment, it is a television capable of receiving broadcast waves and displaying programs. Also, the display device 10 may be a display capable of displaying video input from the outside. Also, the display device 10 may be, for example, a projector that projects a display screen onto a screen.

[0021] The display device 10 may have a device for acquiring information for identifying the user who is viewing. For example, the display device 10 may have, for example, a device for acquiring the voice of the viewing user, and for example, may have a microphone 12 as a voice input device. The microphone 12 may be one microphone or a microphone array composed of a plurality of microphones.

[0022] Further, the display device 10 may have a camera 14 for acquiring an image of the user who is viewing. The camera 14 can photograph, for example, an environment including the user who is viewing the display device 10. One or more cameras 14 may be arranged.

[0023] Here, the content may be anything that can be displayed on the display device 10. For example, the content may be a program that receives, demodulates, and displays terrestrial / BS / CS broadcasts, a video selected by the user on a video distribution site, or a video input from the outside via HDMI (registered trademark), D-SUB, etc.

[0024] [1.2 Hardware Configuration] The hardware configuration will be described with reference to FIG. 2.

[0025] The control unit 100 controls the entire display device 10. The control unit 100 realizes various functions by reading and executing various programs stored in a storage device (for example, the storage 110 and the ROM 120). The control unit 100 may be realized by one or more control devices / arithmetic devices (CPU (Central Processing Unit), SoC (System on a Chip)). Further, the control unit 100 may be composed of a control circuit.

[0026] The operation control unit 102 receives an operation from the user, gives an operation instruction to each functional unit, or notifies the control unit 100 of an operation signal corresponding to the received operation. For example, the operation control unit 102 may be an operation signal receiving unit from a remote control or an operation switch provided on the main body of the display device 10.

[0027] Storage 110 is a non-volatile storage device capable of storing programs and data. For example, it may be composed of storage devices such as HDD (Hard Disk Drive) or SSD (Solid State Drive). Also, Storage 110 may be configured as an externally connectable USB memory. Further, Storage 110 may be, for example, a storage area on the cloud.

[0028] ROM 120 is a non-volatile memory capable of retaining programs and data even when the power is turned off.

[0029] RAM 130 is the main memory mainly used by the control unit 100 when executing processes. RAM 130 is a rewritable memory that temporarily holds programs read from Storage 110 or ROM 120 and data including execution results.

[0030] The broadcast control unit 140 receives the broadcast wave transmitted by the broadcast station selected by the user, decodes video and characters from the broadcast wave and outputs them to the display unit 150, and decodes audio from the broadcast wave and outputs it to the audio output unit 165. The configuration of the broadcast control unit 140 will be described in more detail later.

[0031] The display unit 150 is a display device capable of displaying the video of the received program and various information. The display unit 150 may be, for example, a device capable of displaying video such as a liquid crystal display (LCD; Liquid Crystal Display) or an organic EL (Organic Electro Luminescence) display. Also, the display unit 150 includes an interface to which a display device can be connected. For example, it may be composed of an external display device connected via HDMI (registered trademark) (High-Definition Multimedia Interface), DVI (Digital Visual Interface), or Display Port. Further, the display unit 150 may be, for example, a projection device such as a projector.

[0032] The imaging unit 155 is an imaging device that images the surroundings where the display device 10 is installed, and is, for example, a camera or the like. The imaging unit 155 may be composed of one or a plurality of imaging devices. The imaging unit 155 outputs the captured image as an image signal. Further, the imaging unit 155 may output one or a plurality of images as a continuous video.

[0033] The voice input unit 160 is an input device capable of inputting the sounds around where the display device 10 is installed, and is, for example, a microphone or the like. Further, the voice input unit 160 may be composed of a plurality of input devices (for example, microphones). Also, the voice input unit 160 is mainly used when inputting the voice of the user who is mainly watching, but can also input general sounds such as environmental sounds, for example.

[0034] The voice output unit 165 outputs the voice included in the content. The voice output unit 165 may be a device such as a speaker or headphones, for example. Also, the voice output unit 165 only needs to output sound, and can output general sounds such as music and environmental sounds, for example.

[0035] The communication unit 170 is a communication interface for communicating with other devices. For example, the communication unit 170 may be a network interface capable of connecting to a wireless LAN, or may be a network interface capable of wired connection to Ethernet (registered trademark). Also, the communication unit 170 may be a communication device capable of connecting to a mobile communication network such as LTE / 4G / 5G / 6G, for example.

[0036] Here, one or a plurality of components shown in FIG. 2 may be composed of external devices connected to the display device 10. For example, the voice input unit 160 may be a microphone connected by USB. Also, one or a plurality of components shown in FIG. 2 may be realized by a terminal device connected by wireless communication such as Bluetooth (registered trademark) or a terminal device connected via a network. For example, a smartphone, tablet, smart speaker, or wearable terminal device connected to the display device 10 may use the display device, voice input / output device, and imaging device that it has.

[0037] Also, the configuration in FIG. 2 only needs to include the configurations necessary in the embodiments. For example, in the first embodiment, the voice input unit 160 may not be provided.

[0038] (Broadcast control unit) A simple configuration of the broadcast control unit 140 will be described with reference to FIG. 3. For example, broadcast data of digital broadcasts such as terrestrial / BS / CS is acquired by the tuner unit 200. Then, the OFDM demodulation unit 210 demodulates the broadcast data by OFDM, and after error correction and the like are performed, a TS packet is output.

[0039] The separation unit 220 separates the TS packet and outputs the video packet to the video decoding unit 230 and the audio packet to the audio decoding unit 260. Also, the separation unit 220 outputs a data packet (for example, a subtitle TS packet) including information related to subtitles to the subtitle decoding unit 240.

[0040] The video decoding unit 230 decodes video data from the input video packet and outputs it to the image processing unit 250. Also, the subtitle decoding unit 240 decodes subtitle data from the input data packet and outputs it to the image processing unit 250. Here, the subtitle data shall include information related to subtitles. For example, the subtitle data in this embodiment may include management data of subtitles included in subtitle PES data and data of subtitle texts to be displayed.

[0041] The image processing unit 250 outputs a video with subtitle data superimposed on the video data to the display unit 150 as necessary.

[0042] Also, the audio decoding unit 260 decodes audio data from the input audio packet and outputs the audio to the audio output unit 165.

[0043] Note that FIG. 3 schematically illustrates a general broadcast control unit 140, and other configurations may also be possible. For example, the video decoding unit 230 and the audio decoding unit 260 may be the same decoding unit.

[0044] [1.3 Software Configuration] The software configuration will be described with reference to FIG. 3. Each function is realized by the control unit 100 executing a program stored in, for example, a storage unit (storage 110, ROM 120, RAM 130).

[0045] The user processing unit 1010 executes processing on user information corresponding to the user. For example, it stores user information including identification information for identifying the user. The user processing unit 1010 stores the user information in the user information storage area 1102.

[0046] Here, an example of the user information stored in the user information storage area 1102 will be described with reference to FIG. 5(a). The user information stores a uniquely assigned user ID (for example, "Usr01"), an arbitrarily input or selected user name (for example, "A"), user identification information for identifying the user (for example, the image file "001.jpg" of the user's face image), an age which is an example of the user's attribute (for example, "6"), and a school year which is an example of the user's attribute (for example, "Grade 1 in primary school").

[0047] Here, in the present embodiment, the user's identification information stores the user's face image. For example, when identifying the user by other means, the information necessary for identifying the user may be stored. For example, when recognizing the user by voice, information related to the voice may be stored. Also, the user identification information may store not the face image but a feature amount based on the face image.

[0048] Also, the school year is an attribute set by the administrator based on the user's age. The school year is set by the administrator or the like, and may be linked to the age.

[0049] The age acquisition unit 1020 acquires the age of the user using the content. For example, the age acquisition unit 1020 uses the face age DB stored in the face age DB storage area 1104 to acquire or estimate the age of the user.

[0050] At this time, when the user who is viewing the content can be identified by the user processing unit 1010, the age acquisition unit 1020 acquires the age of the viewing user from the user information storage area 1102. Also, when the user information cannot be acquired, the age acquisition unit 1020 uses the face age DB or a learning model for age determination on the cloud, etc., to acquire the age from the face image of the user included in the image data captured by the imaging unit 155.

[0051] The subtitle determination unit 1030 determines the subtitle to be displayed. Then, the control unit 100 synthesizes and outputs the subtitle to the video data based on the display mode of the subtitle determined by the subtitle determination unit 1030. The subtitle determination unit 1030 determines the subtitle to be superimposed and displayed on the video based on the display timing (reference time, start time), position (coordinates), color, character size, font, decoration, language selection, display direction (vertical writing, horizontal writing), etc. of the subtitle included in the subtitle data. Also, when the display mode of the subtitle is set for each user, the subtitle determination unit 1030 determines the subtitle to be displayed in the set display mode.

[0052] The face age DB storage area 1106 is a database related to face age. For example, the control unit 100 can estimate the age from the face of the user included in the captured image by referring to the face age DB.

[0053] The synonym DB 1106 is a DB that stores synonyms of languages. Fig. 5(b) is a diagram showing an example of the synonym DB 1106. The synonym DB contains a plurality of words that are synonyms for each group of synonyms. Also, the synonym DB 1106 may store the association between a predetermined word and the ambiguity of the word. Note that the synonym DB 1106 includes words that can be replaced by a certain word and may include those that are not strictly synonyms.

[0054] The attribute-based DB 1108 stores one or more appropriate DBs corresponding to the attributes of the user. For example, FIG. 4 shows that the grade-level word DB 1110 is stored. The grade-level word DB 1110 stores appropriate words (for example, words that should be known for each grade of the user, headwords in a dictionary) for each grade of the user as an attribute.

[0055] FIG. 5(c) is a diagram showing an example of the grade-level word DB 1110. Each grade includes the language used in that grade. In the present embodiment, the grade-level word DB 1110 is stored as an example of the attribute-based DB 1108, but other DBs based on words for kanji proficiency tests by level, English proficiency test words by level, and common-use kanji words may also be stored. Also, these DBs may be stored on the cloud.

[0056] Also, the grade-level word DB 1110 may store grade-level words (characters) corresponding to specific symbols. For example, as shown in FIG. 5(d), words to be converted may be stored by grade corresponding to a certain word. When it comes to the data of the grade-level word DB 1108 in FIG. 5(d), the synonym DB 1106 may not be stored.

[0057] [1.4 Process flow] Hereinafter, the process flow in the present embodiment will be described. Note that the following processes will be described as being executed by the control unit 100, but each configuration described in FIGS. 2 to 4 may execute the processes of each step.

[0058] [1.4.1 User registration process] First, the case of registering a user will be described. In the present embodiment, the control unit 100 executes a user registration process based on the registered user. However, when the user is not specifically registered, the control unit 100 may not execute the user registration process.

[0059] Based on the operation from the user, the control unit 100 activates the user management screen (S102). By selecting the user management screen, the control unit 100 can newly store, update, or delete user information.

[0060] The control unit 100 starts face detection from the image captured by the imaging unit 155 (S104). Here, the control unit 100 may perform face detection according to the operation from the user, or may perform face detection once when the user management screen is activated. Also, during the execution of this process, the control unit 100 may constantly detect the user's face.

[0061] Here, when a face is detected in the image, that is, when face information which is an image of the user's face is detected, the control unit 100 temporarily stores the face information (S106; Yes → S108). Here, when two or more faces are detected, the control unit 100 may temporarily store each face. Also, when two or more faces are detected, the control unit 100 may store the face selected by the user from among the multiple detected faces.

[0062] Subsequently, the control unit 100 executes user identification processing based on the stored face image. Here, the control unit 100 determines whether the face image included in the imaging unit 155 matches the face image of the user included in the user information. The method by which the control unit 100 determines whether the face images match is, for example, to detect feature points in each image. Then, when the feature points match by a predetermined ratio or more in the two face images, it is determined that the two images match. Also, the control unit 100 may use an external service (for example, an image similarity determination service by AI) to determine the similarity of the two images.

[0063] Here, when the user of the detected face image is not stored in the user information storage area 1102, the control unit 100 stores it as a new user (S112; No → S114). Here, the control unit 100 may prompt the user to input a user name. Also, the control unit 100 may assign a user name according to a predetermined rule.

[0064] Subsequently, the control unit 100 executes age determination processing (S116). The control unit 100 may determine the age from the face image using a face age DB. Also, the control unit 100 may determine the age from the face image using, for example, an external service. Also, the control unit 100 may prompt the user to input an age.

[0065] Then, the control unit 100 updates the user information (S118). That is, when an existing user is registered, the control unit 100 updates the attributes of the stored user. Also, when this is a new user, the control unit 100 adds attributes to the new user.

[0066] Here, in this embodiment, as attributes of the user, the control unit 100 stores the user's age and the user's school year. The control unit 100 stores the age of the user obtained in S116 in the attributes of the user information.

[0067] Also, the control unit 100 may obtain the user's school year by the user's selection. Also, the control unit 100 may obtain the user's school year based on the age of the user obtained in S116.

[0068] For example, when the attribute of the user is the school year, the control unit 100 may obtain, for example, any school year from "elementary school grade 1 to grade 6, junior high school grade 1 to grade 3, high school grade 1 to grade 3".

[0069] [1.4.2 Subtitle display processing] FIG. 7 is a diagram for explaining the process of displaying subtitles. The control unit 100 executes a viewing user acquisition process (S132). By executing the viewing user acquisition process, the control unit 100 acquires the user who views the content. For example, the control unit 100 may acquire the age of the viewing user based on the image captured by the imaging unit 155. Alternatively, the control unit 100 may acquire the age of the viewing user set by the user using a remote controller or the like.

[0070] Subsequently, the control unit 100 acquires the school year of the viewing user (S134). Here, the control unit 100 acquires the school year (actual school year) corresponding to the actual age of the viewing user. The control unit 100 may refer to the user information, or may have the administrator or the like set the actual school year again.

[0071] Subsequently, the control unit 100 sets the school year-specific word DB corresponding to the school year acquired in S134. That is, the control unit 100 sets a list (conversion list) for converting subtitle data in the school year-specific DB of the set school year.

[0072] Subsequently, the control unit 100 temporarily outputs subtitle data (S138). Then, the control unit 100 converts the words included in the temporarily output subtitle data into appropriate words by executing a word conversion process (S140). Then, the control unit 100 outputs the subtitle data converted into appropriate words (S142).

[0073] For example, in the image processing unit 250 in the broadcast control unit 140, subtitles based on the subtitle data in which the words are converted are superimposed on the video based on the video data (S144). Then, the control unit 100 or the broadcast control unit 140 displays the video with the converted subtitles superimposed on the display unit 150 (S146).

[0074] [1.4.3 Word Conversion Process] FIG. 8 is a diagram showing an example of the process executed by the control unit 100 in S138 of FIG. 7. First, the control unit 100 decomposes one sentence of the temporarily output subtitle data into words (S152). Here, the control unit 100 decomposes the sentence of the subtitle data into phrases such as nouns, verbs, and conjunctions, for example. Here, the control unit 100 may decompose and extract at least words such as those that become the headwords of the dictionary (for example, nouns, etc.) as words. Further, the control unit 100 may decompose and extract in units such as conjunctions and particles.

[0075] The control unit 100 selects one word (S154), and when the word exists in the grade-specific word DB 1110, the control unit 100 converts the word (S156; Yes→S158). Here, various methods can be considered for converting words. For example, the following methods can be considered.

[0076] (1) Use the synonym DB 1106 For example, the control unit 100 acquires the synonyms of the word selected in S154 from the synonym DB 1106. Then, the control unit 100 determines whether the acquired synonym word is stored in the grade-specific word DB 1110. Then, the control unit 100 selects and converts the word included in the grade-specific word DB 1110 of the set grade among the synonyms. (2) Use the grade-specific word DB 1110 For example, the control unit 100 uses the grade-specific word DB 1110 shown in FIG. 5(d). For example, the control unit 100 determines whether the word selected in S154 is included in the grade-specific word DB 1110. Then, when the word selected in S154 is included in the grade-specific word DB 1110, the control unit 100 extracts the corresponding word for each grade and changes the word in the subtitle data.

[0077] If the control unit 100 is not the last word, the control unit 100 continues to return to S154 to select the next word (S154). If it is the last word of the subtitle data, the control unit 100 reassembles the sentence and outputs it as the subtitle data of the converted word (S162).

[0078] [1.5 Operation Example] FIG. 9(a) is a diagram showing an example of a display screen W10 for storing user information. On the display screen W10, a face image of the user is displayed at G10. Also, a name is displayed at P10, and the user's age is displayed at P12. Here, when the user selects button B10, the user name can be changed to an arbitrary name. Also, when the user selects button B12, the user's age can be arbitrarily changed. Also, when the user selects button B14, the control unit 100 can recognize the age from the user's face image and set the age.

[0079] Also, the user's attributes are displayed at P14. In the present embodiment, the grade is displayed as an attribute. When the user selects button B16, the user's attributes can be arbitrarily changed.

[0080] Also, the display mode of the subtitle can be set at P16. The display screen W10 can set the size, color, etc. as the display mode of the subtitle.

[0081] FIG. 9(b) is an example of a display screen W20 showing a scene in which a user can make a selection when viewing content. The user's selection may be that the user is photographed by the photographing unit 155 and recognized by the control unit 100. Also, as shown by the area R20, the user's selection may be to select the user who views the content.

[0082] FIG. 10 is a diagram schematically showing an example of the display of the user and the subtitle. For example, FIG. 10(a) is a diagram schematically showing a situation in which a content is being viewed by Usr03 (the user to be registered), who is not the registered user. Since the user is the user to be registered, there is no set grade. Therefore, at this time, the control unit 100 displays the subtitle M22, "I am a robot," based on the subtitle data obtained from the content.

[0083] FIG. 10(b) is a diagram schematically showing a situation where, for example, a child (first grader), Usr01, is viewing content. The control unit 100 displays the character M24, "I am a robot," based on the character data converted using the grade-specific word DB 1110 for first graders.

[0084] [1.6 Effects, etc.] As described above, according to the present embodiment, the control unit 100 can estimate the age of the user and convert the subtitles of the content corresponding to that age group. The display device 10 will display subtitles according to the age and grade of the user viewing the content, and appropriate subtitles will be displayed for the user.

[0085] In the above-described embodiment, the age of the user is obtained based on an image (e.g., the face image of the user), but other methods may also be used. For example, the age of the user may be obtained based on voice.

[0086] Also, in the above-described embodiment, the attribute-specific DB 1108 has been described as the grade-specific word DB 1110, but words may be converted based on other attributes. For example, a DB according to the proficiency level of language ability such as English words may be used.

[0087] [2. Second Embodiment] The second embodiment will be described. The second embodiment is an embodiment where the actual grade of the user viewing the content is different from the actually set grade (learning grade).

[0088] In the second embodiment, the description of the parts having the same hardware and software configuration as those in the first embodiment will be omitted, and the description will focus on the points different from the first embodiment.

[0089] For example, in the first embodiment, in S134 of FIG. 7, the control unit 100 obtained the grade as the attribute of the user based on the user obtained in S132. In this embodiment, in S132, it is an embodiment where any grade can be changed.

[0090] The display screen W30 in Fig. 11(a) is selectable by the user in the area R30. Here, when one user is selected from the users displayed in the area R30, the screen transitions to the display screen W32 in Fig. 11(b), and the area R32 is displayed. The area R32 enables the selected user to choose whether to use the actual school year specified by age or an arbitrarily set learning school year different from the actual school year.

[0091] Here, when a learning school year is selected in the area R32, the screen transitions to the display screen W34 in Fig. 11(c), and the area R34 is displayed. The area R34 displays items for which any school year can be selected. When the user selects one arbitrary school year from the displayed school years, that school year is set. The control unit 100 converts the words of the content based on the school year set as the learning school year.

[0092] Thus, according to this embodiment, it is possible to set the learning school year regardless of the user's age.

[0093] Also, the control unit 100 may automatically set the learning school year. For example, when the user selects "easy" after watching the content, the control unit 100 may increase the learning school year by one. Also, when the user selects "difficult" after watching the content, the control unit 100 may decrease the learning school year by one.

[0094] [3. Third Embodiment] The third embodiment will be described. The third embodiment is an embodiment when there are multiple users watching.

[0095] Regarding the third embodiment, the description of the parts with the same hardware and software configuration as the first embodiment will be omitted, and the description will focus on the differences from the first embodiment.

[0096] FIG. 12 is a diagram in which the user information in FIG. 5(a) is replaced. The user information in FIG. 12 further stores a priority for each user. The priority may be stored as, for example, "high", "medium", or "low", or may be stored as a numerical value.

[0097] FIG. 13 is a diagram in which the process in FIG. 7 is replaced. When the control unit 100 acquires the viewing user, it determines whether there is a registered user among them (S302). For example, the control unit 100 determines whether there is a user in the image captured by the imaging unit 155 who is the same as the user in the image stored in the user information, that is, a registered user.

[0098] Here, when there is no registered user (S302; No), the control unit 100 may output the subtitle as it is without converting the subtitle data of the content. Therefore, the control unit 100 transitions the process to S142 and outputs the subtitle data.

[0099] When there is a registered user (S302; Yes), the control unit 100 determines whether there are multiple registered users (S304). Here, when there is one registered user who is viewing, the control unit 100 acquires the grade of the user (S304; No → S134).

[0100] Also, when there are multiple registered users who are viewing, the control unit 100 acquires the priority of the viewing users from the user information (S306). Then, the control unit 100 acquires the grade of the user with the highest priority among the viewing users (S308).

[0101] Note that the grade acquired by the control unit 100 in S134 and S308 may be the actual grade or the learning grade.

[0102] An operation example will be described with reference to FIG. 14. In FIG. 14(a), registered user Usr01 and registered user Usr02 view the content on the display device 10. Here, referring to FIG. 12, the priority of Usr01 is high. Therefore, the control unit 100 converts the subtitle data based on the grade "first grade of elementary school" of Usr01.

[0103] In FIG. 14(b), registered user Usr02 and unregistered user Usr03 view content on display device 10. Here, there is one registered user viewing the content, not a plurality of users. Therefore, control unit 100 converts the subtitle data based on the school year "fifth grade" of Usr02.

[0104] Note that when recognizing a plurality of users as described above, for example, control unit 100 may notify to that effect on the display screen. For example, in the case of FIG. 14(a), control unit 100 may perform a confirmation display on the display screen saying "Subtitles will be converted according to the settings of Usr01. Is that okay?"

[0105] Also, when the user performs a cancel process, control unit 100 may output subtitles without converting the subtitle data as it is. Also, when the user performs a cancel process, control unit 100 may convert the subtitles according to the settings of Usr02 this time.

[0106] Thus, according to the present embodiment, appropriate subtitle data can be output even when a plurality of users are viewing.

[0107] [4. Fourth Embodiment] The fourth embodiment will be described. The fourth embodiment is an embodiment in which words are converted based on one DB. For the fourth embodiment, descriptions of parts having the same hardware and software configurations as those of the first embodiment will be omitted, and the description will focus on the points different from the first embodiment.

[0108] This embodiment is an embodiment in which FIG. 4 of the first embodiment is replaced with FIG. 15. FIG. 15 is a diagram showing an example of the software configuration of this embodiment. In this embodiment, conversion word DB 1120 is stored instead of the synonym DB 1106 (attribute-specific DB 1108) of the first embodiment.

[0109] Also, in the first embodiment, the attribute was described as the school year, but in this embodiment, it will be described as the age group. The age group may be set as, for example, "child", "adult", "elderly".

[0110] Here, the conversion word is a DB in which the word to be converted is stored. For example, as shown in FIG. 16, for a word, the converted word (conversion word) is stored. Also, for each word, an attribute that is the target of converting the word is stored.

[0111] Here, in S156 of FIG. 8, it was determined whether the word selected in S154 is in the school-year word DB1110, but in this embodiment, it is determined whether it is in the conversion word DB1120. And when the word selected in S154 is in the conversion word DB1120 and the attribute of the viewing user matches, the word is converted to the word of the conversion word (S158).

[0112] For example, when the user is a "child", the control unit 100 converts the subtitle data "There is no evidence" to the subtitle data "There is no proof" and outputs it. Also, when the user is an "elderly", the control unit 100 converts the subtitle data "Outsourcing the business" to "Entrusting the business externally" and outputs it.

[0113] Also, the control unit 100 may delete and output words "XXX" (where XXX is, as a specific example, a word that harms public order or good customs, etc.) that are not suitable for children.

[0114] [5. Fifth Embodiment] The fifth embodiment will be described. The fifth embodiment is an embodiment in which Chinese characters not learned by school year are converted to hiragana.

[0115] This embodiment is an embodiment in which FIG. 4 of the first embodiment is replaced with FIG. 17 and FIG. 8 is replaced with FIG. 18. Hereinafter, description will be made with reference to FIGS. 17 and 18. And, for the fifth embodiment, description of the parts having the same hardware and software configurations as those of the first embodiment will be omitted, and the description will be centered on the points different from the first embodiment.

[0116] FIG. 17 is a diagram showing an example of the software configuration of this embodiment. In this embodiment, a grade-based Chinese character DB 1122 is stored instead of the synonym DB 1106 (attribute-based DB 1108) of the first embodiment.

[0117] The grade-based Chinese character DB 1122 stores Chinese characters for each grade. For example, Chinese characters may be stored for each grade (attribute) such as "first grade of primary school" and "second grade of primary school".

[0118] As shown in FIG. 18, the control unit 100 decomposes one sentence into characters (S502). Then, the control unit 100 selects one character (S504). The control unit 100 determines whether the selected character is in the grade-based DB (S506).

[0119] Note that the grade-based Chinese character DB 1122 for "fifth grade of primary school" may be a DB that includes all Chinese characters up to the fifth grade (from the first grade to the fifth grade). Also, when the grade-based Chinese character DB 1122 is one DB, the grades including the target Chinese characters may be from the first grade to the fifth grade (that is, up to the fifth grade).

[0120] Then, when the character selected in S504 is in the grade-based Chinese character DB 1122 (stored as a Chinese character up to the fifth grade as an attribute), the control unit 100 outputs the character as it is (S506; Yes → S510). Also, when the character selected in S504 is not in the grade-based Chinese character DB 1122, the control unit 100 converts the character into hiragana and outputs it (S506; No → S508).

[0121] The control unit 100 repeatedly executes the process until all the characters included in one sentence are processed (S510; No → S504).

[0122] [6. Modification Example] The present disclosure is not limited to the above-described embodiments, and various modifications are possible. That is, embodiments obtained by appropriately combining technical means modified within the scope not departing from the gist of the present disclosure are also included in the technical scope.

[0123] In addition, although the above-described embodiments are described separately for convenience of explanation, they can be executed in combination within a possible range. Also, any technology described in the specification has the intention of obtaining rights in amendments, divisional applications, etc.

[0124] Also, although the above-described database (DB) has been described as being stored in a storage device in a display device, it may be an external DB. For example, the database may be a database stored in the cloud. Also, the database may be provided in an external service.

[0125] In addition, the program operating in each device in each embodiment is a program (a program that functions a computer) that controls a CPU or the like so as to realize the functions of the above-described embodiments. And the information handled by these devices is temporarily stored in a temporary storage device (for example, RAM) during its processing, and then stored in storage devices such as various ROMs and HDDs, and read out by the CPU as necessary for correction and writing.

[0126] Here, the recording medium for storing the program may be any of a semiconductor medium (for example, ROM, non-volatile memory card, etc.), an optical recording medium, a magneto-optical recording medium (for example, DVD (Digital Versatile Disc), CD (Compact Disc), BD (Blu-ray (registered trademark) Disc), etc.), a magnetic recording medium (for example, magnetic tape, flexible disk, etc.).

[0127] When distributing it in the market, the program can be stored on a portable recording medium for distribution, or transferred to a server computer connected via a network such as the Internet. In this case, the storage device of the server apparatus is of course also included in the present disclosure.

[0128] Also, the above-described data does not have to be stored in the device, but may be stored in an external device and appropriately called. For example, the data may be stored in a NAS (Network Attached Storage) or stored on the cloud.

[0129] Note that the scope of the present disclosure is not limited to the configurations explicitly described in the specification, and combinations of the technologies disclosed in this specification are also included in the scope. Among the present disclosure, the configuration for which a patent is sought is described in the appended claims, but it is not intended to exclude it from the technical scope for the reason that it is not described in the claims.

[0130] Also, in the above specification, the descriptions of "in the case of ~" and "when ~" are explanations as one example, and are not configurations limited to the described content. Configurations that are not these cases or times are also disclosed to the extent that they are obvious to those skilled in the art and have the intention of obtaining rights.

[0131] Also, regarding the descriptions with an order for the processes and data flows described in the specification, the order is not limited to the described order. For example, configurations in which a part of the process is deleted or the order is changed are also disclosed and have the intention of obtaining rights.

[0132] Also, although the functions described in the embodiments are described as being executed by each device, they may be realized by one device or further by using an external server.

[0133] Moreover, each functional block or various features of the apparatus used in the above-described embodiments can be implemented or executed by an electric circuit, for example, an integrated circuit or a plurality of integrated circuits. An electric circuit designed to execute the functions described in this specification may include a general-purpose use processor, a digital signal processor (DSP), an application specific integrated circuit (ASIC), a field programmable gate array (FPGA), or other programmable logic devices, discrete gates or transistor logic, discrete hardware components, or a combination thereof. The general-purpose use processor may be a microprocessor, or may be a conventional type processor, controller, microcontroller, or state machine. The above-described electric circuit may be composed of a digital circuit or an analog circuit. Further, when an integrated circuit technology that replaces the current integrated circuit appears due to the progress of semiconductor technology, one or more aspects of the present disclosure can also use a new integrated circuit by such technology.

Explanation of Signs

[0134] 10 Display device 100 Control unit 102 Operation control unit 110 Storage 120 ROM 130 RAM 140 Broadcast control unit 150 Display unit 155 Photographing unit 160 Voice input unit 165 Voice output unit 170 Communication unit

Claims

1. An acquisition unit that acquires video data and subtitle data from content; A recognition unit that recognizes a user who views the content; An attribute-specific database that stores words for each user attribute; A determination unit that determines the attribute of the recognized user; A conversion unit that refers to the attribute-specific database and converts the words included in the subtitle data according to the attribute of the user recognized by the recognition unit; A display unit that superimposes and displays subtitles based on the subtitle data converted by the conversion unit on the video data; A display device comprising the above.

2. Further comprising a thesaurus database that stores synonyms, The conversion unit reads out synonyms of the words included in the subtitle data and, when the synonyms are included in the attribute-specific database, replaces the words with the synonyms The display device according to claim 1.

3. Further comprising a photographing unit that photographs a user based on a camera device, The determination unit determines the attribute of the user based on an image including the user photographed by the photographing unit The display device according to claim 1.

4. The attribute is age, The determination unit determines the age of the user, The conversion unit converts the subtitle data into words corresponding to the age of the user The display device according to claim 1.

5. Further comprising a storage unit that stores user information that stores the school year of the user as an attribute in association with the age of the user, The determination unit determines the school year of the user by referring to the user information from the determined age of the user, The conversion unit converts the subtitle data into words corresponding to the school year of the user The display device according to claim 4.

6. The attribute is school year, The determination unit determines the school year of the user, The conversion unit converts the subtitle data into words corresponding to the school year of the user The display device according to claim 1.

7. A display method in a display device that stores an attribute-specific database that stores words for each user attribute, An acquisition step of acquiring video data and subtitle data from content; A recognition step of recognizing a user who views the content; A determination step of determining the attribute of the recognized user; A conversion step of referring to the attribute-specific database and converting the words included in the subtitle data according to the attribute of the user recognized in the recognition step; A display step of superimposing and displaying a caption based on the caption data converted by the conversion step on the video data; A display method including the above.

8. In a computer that stores an attribute-specific database that stores words for each user attribute, An acquisition function for acquiring video data and caption data from content; A recognition function for recognizing a user who views the content; A determination function for determining the attribute of the recognized user; A conversion function for converting the words included in the caption data with reference to the attribute-specific database according to the attribute of the user recognized by the recognition function; A display function for superimposing and displaying a caption based on the caption data converted by the conversion function on the video data; A program for realizing the above.

9. A recording medium on which the program of Claim 8 is recorded.

10. In an apparatus having an attribute-specific database that stores words for each user attribute, a caption generation method for generating a caption to be displayed on a video of content, comprising: An acquisition step of acquiring video data and caption data from content; A recognition step of recognizing a user who views the content; A determination step of determining the attribute of the recognized user; A generation step of generating a caption obtained by converting the words included in the caption data with reference to the attribute-specific database according to the attribute of the user recognized by the recognition unit; A caption generation method including the above.

Citation Information

Patent Citations

  • Title display

    JP2002142168A