WO2004053736A1 - 情報処理装置および方法、記録媒体、並びにプログラム - Google Patents

情報処理装置および方法、記録媒体、並びにプログラム Download PDF

Info

Publication number
WO2004053736A1
WO2004053736A1 PCT/JP2003/015927 JP0315927W WO2004053736A1 WO 2004053736 A1 WO2004053736 A1 WO 2004053736A1 JP 0315927 W JP0315927 W JP 0315927W WO 2004053736 A1 WO2004053736 A1 WO 2004053736A1
Authority
WO
WIPO (PCT)
Prior art keywords
content
program
grouping
group
preference information
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/JP2003/015927
Other languages
English (en)
French (fr)
Inventor
Mitsuhiro Miyazaki
Noriyuki Yamamoto
Mari Saito
Hiroyuki Koike
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Sony Corp
Original Assignee
Sony Corp
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Sony Corp filed Critical Sony Corp
Priority to EP03778860A priority Critical patent/EP1571561A4/en
Priority to US10/538,658 priority patent/US7873798B2/en
Publication of WO2004053736A1 publication Critical patent/WO2004053736A1/ja
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Classifications

    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/80Generation or processing of content or additional data by content creator independently of the distribution process; Content per se
    • H04N21/83Generation or processing of protective or descriptive data associated with content; Content structuring
    • H04N21/84Generation or processing of descriptive data, e.g. content descriptors
    • H04N21/8405Generation or processing of descriptive data, e.g. content descriptors represented by keywords
    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F16/00Information retrieval; Database structures therefor; File system structures therefor
    • G06F16/70Information retrieval; Database structures therefor; File system structures therefor of video data
    • G06F16/73Querying
    • G06F16/735Filtering based on additional data, e.g. user or group profiles
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/20Servers specifically adapted for the distribution of content, e.g. VOD servers; Operations thereof
    • H04N21/25Management operations performed by the server for facilitating the content distribution or administrating data related to end-users or client devices, e.g. end-user or client device authentication, learning user preferences for recommending movies
    • H04N21/258Client or end-user data management, e.g. managing client capabilities, user preferences or demographics, processing of multiple end-users preferences to derive collaborative data
    • H04N21/25866Management of end-user data
    • H04N21/25891Management of end-user data being end-user preferences
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/20Servers specifically adapted for the distribution of content, e.g. VOD servers; Operations thereof
    • H04N21/25Management operations performed by the server for facilitating the content distribution or administrating data related to end-users or client devices, e.g. end-user or client device authentication, learning user preferences for recommending movies
    • H04N21/266Channel or content management, e.g. generation and management of keys and entitlement messages in a conditional access system, merging a VOD unicast channel into a multicast channel
    • H04N21/26603Channel or content management, e.g. generation and management of keys and entitlement messages in a conditional access system, merging a VOD unicast channel into a multicast channel for automatically generating descriptors from content, e.g. when it is not made available by its provider, using content analysis techniques
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/20Servers specifically adapted for the distribution of content, e.g. VOD servers; Operations thereof
    • H04N21/25Management operations performed by the server for facilitating the content distribution or administrating data related to end-users or client devices, e.g. end-user or client device authentication, learning user preferences for recommending movies
    • H04N21/266Channel or content management, e.g. generation and management of keys and entitlement messages in a conditional access system, merging a VOD unicast channel into a multicast channel
    • H04N21/2665Gathering content from different sources, e.g. Internet and satellite
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/43Processing of content or additional data, e.g. demultiplexing additional data from a digital video stream; Elementary client operations, e.g. monitoring of home network or synchronising decoder's clock; Client middleware
    • H04N21/442Monitoring of processes or resources, e.g. detecting the failure of a recording device, monitoring the downstream bandwidth, the number of times a movie has been viewed, the storage space available from the internal hard disk
    • H04N21/44213Monitoring of end-user related data
    • H04N21/44222Analytics of user selections, e.g. selection of programmes or purchase activity
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/45Management operations performed by the client for facilitating the reception of or the interaction with the content or administrating data related to the end-user or to the client device itself, e.g. learning user preferences for recommending movies, resolving scheduling conflicts
    • H04N21/466Learning process for intelligent management, e.g. learning user preferences for recommending movies
    • H04N21/4667Processing of monitored end-user data, e.g. trend analysis based on the log file of viewer selections

Definitions

  • the present invention relates to an information processing apparatus and method, a recording medium, and a program, and particularly to an information processing apparatus and method, a recording medium, and a program that can efficiently and effectively recommend content.
  • the attribute eg, genre
  • the content is recommended for each attribute.
  • the present invention has been made in view of such a situation, and enables a content recommendation side to group contents by using contents attributes and to recommend contents for each group. Things.
  • the information processing apparatus is configured such that a grouping item including one or more attribute items among attribute items representing attributes of content to be distributed has the same group as content having similarity with a certain degree of similarity or more.
  • Generating means for generating user preference information indicating user preferences based on the usage frequency calculated by the calculating means; and recommending means for recommending content based on the user preference information generated by the generating means It is characterized by having.
  • a grouping item consisting of an attribute item indicating the broadcast time zone and at least one other attribute item is set, and the grouping means should group the content based on the grouping item. Can be.
  • a grouping item including at least an attribute item indicating a broadcast time zone and a grouping item including other attribute items are set, and the grouping unit groups the contents based on the grouping items. It can be performed.
  • the grouping means can perform a morphological analysis of the content of the attribute item of the content, and determine the similarity of the content of the grouping item based on the result.
  • the generation unit can prevent the use frequency of the group whose use state of the content belonging to the group does not satisfy the predetermined condition from being used for generating the user preference information.
  • the recommendation means determines whether or not the usage frequency calculated by the calculation means is higher than a predetermined value.
  • the determination means determines that the usage frequency is higher than a predetermined value.
  • setting means for setting a standard flag indicating that the content is frequently viewed content in the recommendation information of the content can be provided.
  • the generating means includes extracting means for acquiring metadata of the content of the group whose usage frequency calculated by the calculating means is higher than a preset value, and extracting a solid representing the characteristic amount of the metadata;
  • the preference information can be generated based on the vector extracted by the above.
  • the generation unit includes a standard determination unit that determines whether the content of the group whose use frequency calculated by the calculation unit is higher than a preset value is the content corresponding to the content recommendation information with the standard flag set.
  • the standard determination means the content is determined by the content corresponding to the content recommendation information for which the standard flag is set. If it is determined that there is no metadata, the extraction unit can acquire the metadata of the content and extract a vector representing the feature amount of the metadata.
  • the preference information can be configured by a plurality of attributes and values indicating the importance of the attributes.
  • the generating means includes a familiarity setting means for setting the familiarity of the content based on the use frequency calculated by the calculating means, and weights a value representing the importance of the preference information based on the familiarity. Can be.
  • the generating means includes a searching means for searching for the content used the number of times equal to or less than a predetermined value based on the usage history of the content, and a special preference information based on the metadata of the content searched by the searching means. And special preference information generating means for generating.
  • Vectors extracted by the second extraction means can be selected by a predetermined number in descending order of degree, and content can be recommended based on the metadata of the selected vectors.
  • a grouping item consisting of one or more attribute items among attribute items representing attributes of distributed content is assigned to a content having the same group with a certain degree of similarity.
  • the program of the recording medium according to the present invention is characterized in that a grouping item consisting of one or more attribute items among the attribute items representing the attributes of the content to be distributed is the same as content having similarity with a certain degree of similarity or more.
  • a grouping control step that controls the grouping of contents by assigning a group ID, a calculation control step that controls the calculation of the content usage frequency for each group ID, and a calculation control step Controlling the generation of user preference information indicating the user's preference based on the usage frequency, and controlling the content recommendation based on the user preference information generated in the processing of the generation control step.
  • a recommendation control step that controls the grouping of contents by assigning a group ID, a calculation control step that controls the calculation of the content usage frequency for each group ID, and a calculation control step Controlling the generation of user preference information indicating the user's preference based on the usage frequency, and controlling the content recommendation based on the user preference information generated in the processing of the generation control step.
  • a recommendation control step that controls the group
  • the program according to the present invention assigns the same group ID to contents in which one or more attribute items among attribute items representing attributes of distributed content are similar with a certain degree of similarity or more.
  • Grouping control step for controlling content grouping a calculation control step for controlling calculation of content usage frequency for each group ID, and a usage frequency calculated in the processing of the calculation control step.
  • FIG. 1 shows a configuration example of a content recommendation system to which the present invention is applied.
  • FIG. 2 is a diagram illustrating an example of metadata.
  • FIG. 3 is a diagram illustrating grouping of contents.
  • FIG. 4 is another diagram illustrating grouping of contents.
  • FIG. 5 is a diagram illustrating an example of metadata to which a group ID is assigned.
  • FIG. 6 is a diagram showing an example of the usage history.
  • FIG. 7 is a block diagram showing a configuration example of the content recommendation server of FIG.
  • FIG. 8 is a block diagram illustrating a configuration example of the client device in FIG.
  • FIG. 9 is a flowchart illustrating the user preference information generation processing.
  • FIG. 10 is a diagram illustrating a method of calculating the use frequency.
  • FIG. 11A is another diagram illustrating a method of calculating the usage frequency.
  • FIG. 11B is another diagram illustrating a method of calculating the usage frequency.
  • FIG. 12 is a diagram for explaining a use state confirmation process.
  • FIG. 13 is another diagram for explaining the use state confirmation process.
  • FIG. 14 is another diagram for explaining the use state confirmation process.
  • FIG. 15 is a flowchart illustrating the content recommendation information generation processing.
  • FIG. 16 is a diagram showing a display example of content recommendation information.
  • FIG. 17 is a diagram showing a display example of other content recommendation information.
  • FIG. 18 is a flowchart illustrating title grouping processing 1.
  • FIG. 19 is a flowchart illustrating the title grouping process 2.
  • FIG. 20 is a flowchart for explaining the title grouping process 3.
  • FIG. 21 is a flowchart illustrating title grouping processing 4.
  • FIG. 22 is a flowchart illustrating the standard program setting process.
  • FIG. 23 is a flowchart illustrating the preference information extraction process 1.
  • FIG. 24 is a diagram showing a configuration example of a program vector.
  • FIG. 25 is a diagram illustrating a configuration example of preference information.
  • FIG. 26 is a flowchart illustrating the preference information extraction process 2.
  • FIG. 27 is a flowchart illustrating the preference information extraction process 3.
  • FIG. 28 is a flowchart illustrating the preference information change process.
  • FIG. 29 is a flowchart illustrating the special preference information generation processing.
  • FIG. 30 is a block diagram showing a functional configuration example of the CPU of FIG.
  • FIG. 31 is a flowchart illustrating the recommended information search process.
  • FIG. 32 is a flowchart illustrating the special recommendation information search processing. BEST MODE FOR CARRYING OUT THE INVENTION
  • FIG. 1 shows a configuration example of a content recommendation system to which the present invention is applied.
  • the distribution server 3 acquires the streaming data from the streaming data database 1 and distributes it to the client device 5 via the network 6 including the Internet and other networks.
  • the distribution server 3 also acquires the content metadata from the metadata database 2 and supplies it to the content recommendation server 4 via the network 6.
  • the content recommendation server 4 has a certain degree of similarity (elements of each item constituting the grouping item) of the grouping items set by one or more items (the elements of each item are all the same). , A partial match, or a value indicating a certain degree of similarity is equal to or more than a certain value) Assign the same group ID to the content (group into the same group).
  • the content is divided into a set of each element of the item "broadcast station", the item “broadcast start time”, and the item “broadcast end time”, which constitute the grouping item. Are grouped.
  • the content is grouped for each combination of the element “genre” and the item “performer” constituting the grouping item.
  • one content may belong to multiple groups depending on the content item.
  • a program that is broadcast between 00: 00 and 06: 00, and in which the talent A appears in variety includes "8 ch (broadcasting station), 00: 00 (broadcast) Start time) ⁇ 06: 00 (broadcast end time) "and the group ID of" variety (genre), talent A (performer) "( Figure 4) are assigned. It will belong to the group.
  • the content recommendation server 4 transmits the metadata (for example, FIG. 5) in which the group ID is set as described above to the client device 5 as appropriate.
  • the content recommendation server 4 also appropriately obtains a usage history including the content's group ID from the client device 5, and calculates the usage frequency for each group based on the usage history. Then, the content recommendation server 4 uses the calculated use frequency as an indication of the user's preference, and recommends the content for each group. For example, information on contents belonging to a group of high use frequency is transmitted to the client device 5 as content recommendation information.
  • the client device 5 uses the content distributed from the distribution server 3, and uses, for example, metadata (a group ID is set) as shown in FIG. It is supplied to the content recommendation server 4 as appropriate.
  • the client device 5 displays the content recommendation information supplied from the content recommendation server 4 and presents it to the user.
  • the user can select the content that suits his or her taste by referring to it.
  • FIG. 7 shows a configuration example of the content recommendation server 4.
  • Processing Unit 11 Performs predetermined processing in accordance with, for example, a program for content recommendation, which is considered in ROM (Read Only Memory) 12.
  • ROM Read Only Memory
  • a RAM Random Access Memory 13 stores data and the like necessary for the CPU 11 to execute the processing.
  • An input / output interface 15 is connected to the CPU 11 via a bus 14.
  • the input / output interface 15 includes an input unit 16 including a keyboard and a mouse, an output unit 17 including an LCD (Liquid Crystal Display), a storage unit 18 for storing metadata and the like, and a network 6.
  • a communication unit 19 that communicates with the distribution server 3 or the client device 5 through the communication unit 19 is connected.
  • a drive 20 is appropriately connected to the input / output interface 15 and the CPU 11 is connected to a magnetic disk 31, an optical disk 32, a magneto-optical disk 33, or a semiconductor memory 34 mounted thereon. Transfer data between the two.
  • Examples of the functional configuration of the CPU 11 include, for example, a preference information acquisition unit that acquires user preference information, a metadata acquisition unit that acquires program metadata from the distribution server 3, and content recommendation information. It is also possible to configure it with a recommendation information generating unit to generate.
  • FIG. 8 shows a configuration example of the client device 5. This configuration is basically the same as the configuration of the content recommendation server 4, and a description thereof will be omitted.
  • step S1 the CPU 11 of the content recommendation server 4 determines whether or not it is time to generate the user preference information, and if it is determined that it is the time, proceeds to step S2. For example, when the provision of content recommendation information (described later) is requested from the client device 5, or when a predetermined time (for example, a predetermined time every week) comes, the process proceeds to step S2.
  • a predetermined time for example, a predetermined time every week
  • step S2 the CPU 11 acquires a predetermined use history from the client device 5 via the communication unit 19.
  • the metadata of the content that has been used in the past week (with group ID set) is acquired.
  • the CPU 11 calculates the content use frequency (number of times) for each group.
  • the program broadcasted on 20 ch from 20: 0 to 21: 0 on 8 ch, and the program on 10 ch from 10 o 0 to 20 on 0 ch.
  • the program broadcasted at 0 is most viewed (7 times each), followed by the program broadcasted between 22: 0 00 and 23: 00 on 8 channels (6 times). You can see that it is done.
  • the contents of the grouping item are included in each metadata.
  • the content usage frequency (number of times) is calculated for each content (group of elements of each item), as shown in Fig. 11A.
  • Figure 11A shows that, from the number of uses for each group, variety programs with talent D appear the most (10 times), followed by news programs (8 times) with talent D and talent C. It can be seen that the variety programs (5 times) that perform are often viewed.
  • the number of uses is normalized by the number of contents distributed during the period corresponding to the use history acquired in step S2.
  • a tenth program in which talent D appears is distributed during the period (in this case, one week), and a use program in which talent D appears is
  • the number of uses in Fig. 11A would be as shown in Fig. 11B.
  • step S3 the CPU 11 of the content editing server 4 detects, for each grouping item, (a group ID of) a group for which the number of times of use (use frequency) is equal to or greater than a predetermined threshold. I do.
  • the example of FIG. 10 shows “8ch, 20: Two groups of “0 0 to 21: 00” and “10 ch, 19: 00 to 20: 00” are detected.
  • threshold value for the grouping item consisting of the item “genre” and the item “performer” is 0.06, in the example of FIG. 11B, “variety, talent D”, “news, talent D”, And “variety, talent C” are detected.
  • step S4 the CPU 11 determines whether or not the content of each group detected in step S3 matches the user's preference. For example, based on the distribution list of the content belonging to the group, it is confirmed whether the content has not been used continuously for a predetermined number of times (for example, three times) retroactively. It is determined that the content of the group does not match the user's preference.
  • step S5 the CPU 11 detects a group of contents matching the user's preference from the determination result in step S4.
  • step S6 the CPU 11 stores the group ID of the group detected in step S5 in the storage unit 18 as user preference information.
  • step S21 the CPU 11 of the content recommendation server 4 waits until the client device 5 requests the provision of the content recommendation information.
  • the process proceeds to step S22, and the storage is performed.
  • the user preference information generated as described above is obtained from the unit 18.
  • step S23 the CPU 11 determines from the metadata (contents for which the group ID has been set) of the content to be distributed from now on, the meta data for which the same group ID as the group ID as the user preference information has been set. Extract the data. The CPU 11 generates content recommendation information from the extracted metadata.
  • the metadata of the content to which any group ID is assigned can be extracted.
  • step S24 the CPU 11 transmits the content recommendation information generated in step S23 to the client device 5 via the communication unit 19.
  • the client device 5 displays the content recommendation information transmitted from the content recommendation server 4 on the output unit 57.
  • FIG. 16 and FIG. 17 show display examples of content recommendation information.
  • the items “Broadcaster”, “Broadcast start time”, and “Broadcast end time” are grouped into items “8ch, 20: 0 00 to 21: 00” and Information (such as the title) of programs belonging to the group “10ch, 19:00:00 to 20:00:00” is displayed.
  • the usage frequency for grasping the user's preference is calculated for each group using the group ID, so that the usage frequency is calculated as compared to the case of calculating the usage frequency for each metadata item.
  • the amount of calculation can be reduced.
  • the content recommendation information is displayed collectively for each group, the content recommendation information can be appropriately displayed even on the client device 5 having a small display space.
  • grouping was performed using the metadata items “broadcast station”, item “broadcast start time”, and item “broadcast end time”, as well as item “genre” and item “performers”.
  • grouping can also be performed using other items such as the item “Title” and the item “Content”.
  • a rebroadcast or special edition program can be treated as content belonging to the same group as the original program, so that whether the program is original or rebroadcast, If viewed, the usage history can be reflected in the generation of user preference information.
  • step S61 the content recommendation server 4 extracts a title from the metadata.
  • step S62 the content recommendation server 4 morphologically analyzes the title and breaks it down into words.
  • step S63 the content recommendation server 4 extracts one of the analyzed word or a word group composed of a plurality of words, and stores the extracted word in the storage unit 1.
  • the word group composed of a plurality of words is a word group generated by a combination of words obtained by morphological analysis.
  • words obtained by morphological analysis include “Tokaido”, “ In the case of "Mitani” and “Kaidan”, the words are “Tokaido-Mitani”, “Tokaido 'Kaidan”, and "Mitani'Kaidan”.
  • step S64 the content recommendation server 4 determines whether the group ID has been extracted.
  • step S64 If it is determined in step S64 that the corresponding group ID has not been extracted, the extracted word or a group of words composed of a plurality of words has not yet been assigned a group ID.
  • step S65 a new group ID is assigned to the extracted word or a word group composed of a plurality of words. Further, the content recommendation server 4 stores a word or a word group composed of a plurality of words and a group ID corresponding to the word group.
  • step S66 content recommendation server 4 executes It is determined whether or not a group ID has been extracted for a word or a word group composed of a plurality of words.
  • step S66 If it is determined in step S66 that the group ID has not been extracted for all words constituting the title or for a word group including a plurality of words, the process returns to step S63. Subsequent processing is repeated.
  • step S66 if it is determined that the group ID has been extracted for all words constituting the title or for a word group composed of a plurality of words, in step S67, the content recommendation server 4 The process is terminated by associating the extracted or assigned group ID with the metadata.
  • programs having similar titles may be included in the same group.
  • the title was set so that the serial drama titled "Two Years A Gumi Ginpachi-sensei” and the special program titled “Two Years A Gumi Ginpachi-sensei Special” could be grouped as the same group
  • the match rate of words is calculated on a round robin basis in a program title for a predetermined period such as a week, a month, a half year, etc., and if the word match rate is equal to or higher than a predetermined value, the words may be grouped together. Good.
  • step S401 and step S402 processing similar to that in step S61 and step S62 described with reference to FIG. 18 is executed. That is, the content recommendation server 4 extracts the title from the metadata, analyzes the title, and decomposes it into words.
  • step S403 the content recommendation server 4 calculates, based on the analyzed words, the degree of matching between words between titles, that is, the matching rate indicating the rate of matching between words.
  • the title “Two Years A Ginpachi Ginpachi-sensei” and the title “Two Years A Ginpachi Ginpachi Special” are “2”, “Year”, “A” “Gumi”, and “Ginpachi”, respectively.
  • the matching rate of the words that make up the titles of these two programs is It becomes 85.7% in 6/7.
  • step S404 the content recommendation server 4 determines whether the words match at least a predetermined value such as 70%, for example.
  • a predetermined value such as 70%
  • the threshold value of the coincidence rate may be any numerical value other than 70%.
  • step S404 If it is determined in step S404 that the word matches at least a predetermined value such as 70%, in step S405, the content recommendation server 4 determines that the program is identical to those programs. Map the group ID of The content recommendation server 4 stores the matched word or word group and the corresponding group ID. ' If it is determined in step S404 that the match rate is equal to or less than a predetermined value such as 70%, or if the processing in step S405 ends, the content is determined in step S406. The recommendation server 4 determines whether or not the brute force of the title has been completed.
  • step S 406 If it is determined in step S 406 that the brute force of the title has not been completed, the process returns to step S 403, and the subsequent processes are repeated.
  • step S406 If it is determined in step S406 that the brute force of the title has been completed, the process is terminated.
  • a group ID based on the matching rate of words constituting a title is associated, so that programs having similar titles such as a serial drama and a special program are processed as the same group.
  • determining the group based on the matching rate of the words that compose the title for example, in the metadata, one-byte and two-byte numbers, or one-byte and two-byte alphabetic characters, or uppercase and lowercase characters Even if there is a spelling shift, programs with the same title can be detected as the same group.
  • a broadcast station for example, a broadcast station, a program genre, or a broadcast start time may be added to the grouping condition.
  • the title is composed of a small number of words including "news". Therefore, the processing described with reference to FIG.
  • the same group since the same group may be detected, the same group may be used if the broadcast stations also match in addition to the word match rate.
  • a title grouping process 3 (item “Title” and item “Title”) in which grouping is performed based on the matching rate of words constituting titles with the condition of matching broadcasting stations as conditions.
  • the grouping item consisting of the item “broadcasting station” will be explained. .
  • steps S421 to S424 the same processing as steps S401 to S404 described using FIG. 19 is performed. That is, the content recommendation server 4 extracts the title from the metadata, performs morphological analysis, and decomposes the word into words. Then, the content recommendation server 4 calculates the degree of matching of the words between the titles based on the analyzed words, and determines whether the words match at least a predetermined value such as 70%, for example. I do.
  • step S425 the content recommendation server 4 determines whether the broadcast station of the program is It is determined whether or not matches.
  • step S425 If it is determined in step S425 that the broadcasting stations of these programs match, in step S425, the content recommendation server 4 associates the same group ID with those programs. Further, the content recommendation server 4 stores the matched word or word group, and the corresponding broadcast station and group ID.
  • step S424 If it is determined in step S424 that the match rate is equal to or less than a predetermined value such as 70%, if it is determined in step S425 that the broadcast stations of these programs do not match, Alternatively, after the processing in step S 426 is completed, in step S 427, the content recommendation server 4 determines whether or not the round robin of the title has been completed.
  • a predetermined value such as 70%
  • step S 427 If it is determined in step S 427 that the brute force of the title has not been completed, the process returns to step S 423, and the subsequent processes are repeated.
  • step S 427 If it is determined in step S 427 that the brute force of the title has been completed, the processing is terminated.
  • the group ID based on the match rate of the broadcast stations and the match rate of the words constituting the title is associated. For example, when programs having similar titles are set to the same group, It is possible to prevent the -use programs of a station from being in the same group. Note that, in FIG. 20, it has been described that the grouping is performed on the condition that the same broadcasting station is used in addition to the matching rate of the words constituting the title. Needless to say, the grouping may be performed with the broadcast time zone, genre, and the like as conditions other than the match rate of the words constituting the title.
  • the broadcast start time of a serial drama or a band program is shifted due to a sports broadcast or a special program, etc.
  • the condition may be determined based on whether or not the broadcast time matches within a predetermined time range such as one hour, for example.
  • the grouping is performed based on the matching rate of words constituting the title, with the condition that the broadcast time is within a predetermined time range or not.
  • the title grouping process 4 (grouping process by a grouping item including the item “title” and the item “broadcast start time”) will be described.
  • steps S444 to S444 processing similar to that of steps S401 to S404 described with reference to FIG. 19 is performed. That is, the content recommendation server 4 extracts a title from the metadata, performs morphological analysis, and decomposes the word into words. Then, the content recommendation server 4 calculates the degree of matching of the words between the titles based on the analyzed words, and determines whether the words match at least a predetermined value such as 70%, for example. I do.
  • step S444 If it is determined in step S444 that the words match at least a predetermined value such as 70%, in step S445, the content recommendation server 4 starts broadcasting the program. It is determined whether or not the times coincide with each other with a shift of a predetermined range ⁇ such as one hour, for example.
  • step S445 If it is determined in step S445 that the broadcast start times of the programs match within a predetermined range, the content recommendation server 4 determines in step S446 that the broadcast start times of the programs match. To the same group ID. Also, The ten recommendation server 4 stores the matched word or word group, the range of the corresponding broadcast start time, and the group ID.
  • step S444 If it is determined in step S444 that the matching rate is equal to or less than a predetermined value such as 70%, in step S444, the broadcast start times of those programs are shifted beyond a predetermined range. Is determined, or after the process of step S446 is completed, in step S446, the content recommendation server 4 determines whether or not the total number of titles has been completed. ⁇
  • step S444 If it is determined in step S444 that the brute force of the title has not been completed, the process returns to step S444, and the subsequent processes are repeated. If it is determined in step S447 that the brute force of the title has been completed, the processing is terminated.
  • a match including a deviation of a broadcast start time within a predetermined range is associated with a group ID based on a match rate of words constituting a title.
  • the programs are in the same group, it is possible to prevent programs that should be detected as being in the same group from being detected as being in the same group due to a change in broadcast time due to a special program or the like.
  • the content recommendation server 4 performs the user preference information generation process (FIG. 9) and the content recommendation information processing (FIG. 15).
  • the client device 5 uses the content recommendation server. Using the metadata (grouping information) provided with the group ID provided by 4 and generating the user preference information by calculating the usage frequency of each group, content recommendation information is generated based on this. It can also be generated.
  • step S501 the CPU 11 analyzes the usage history.
  • the metadata of the content used in the predetermined period (in which the group ID is set) is acquired from the client device 5, and the number of times of use ( Figure 10) or usage frequency ( Figure 1 IB) is analyzed.
  • step S502 the CPU 11 determines whether or not there is a group whose use count (viewing count) exceeds a predetermined threshold.
  • a staple flag indicating that this program is a staple is set in the content recommendation information of a program belonging to the group (a program whose usage count exceeds the threshold).
  • step S502 it is determined whether there is a group whose viewing frequency exceeds the threshold. If it is determined that there is a program whose usage frequency exceeds the threshold, in step S503, the group is determined.
  • the standard flag may be set in the content recommendation information of the program belonging to.
  • step S502 If it is determined in step S502 that there is no group in which the number of times of viewing exceeds the threshold, the process ends.
  • the content recommendation information with the standard flag set is transmitted to the client device 5 by the content recommendation information generation processing of FIG. This allows the client device 5 to automatically record, for example, a program corresponding to the content recommendation information for which the standard flag is set.
  • the group ID is stored as the user preference information, but more detailed preference information is generated based on a plurality of attributes included in the metadata of the program. Then, the program can be recommended based on the generated preference information.
  • a description will be given of a preference information extraction process 1, which is a first example of generating more detailed preference information based on a plurality of attributes included in program metadata.
  • This processing is executed in the content recommendation server 4 at a predetermined time (a predetermined time every week), for example.
  • step S522 the CPU 11 analyzes the usage history. At this time, as in the case of step S2 in FIG.
  • the metadata of the content used in the predetermined period (in which the group ID is set) is acquired from the client device 5, and the number of times of use ( ( Figure 10) or usage frequency ( Figure 11B).
  • the CPU 11 searches for a group whose use count (use frequency) is equal to or greater than a predetermined threshold. Note that a group whose usage frequency is equal to or higher than a predetermined threshold may be searched.
  • step S 523 the CPU 11 determines whether or not the group has been searched. If it is determined that the group has been searched, the process proceeds to step S 524, where the meta data of the program belonging to the searched group is determined. Analyze the data. At this time, if there are multiple programs, the metadata of the multiple programs is analyzed. In step S525, the CPU 11 generates a program vector based on the metadata of the program analyzed in step S524.
  • FIG. 24 shows a configuration example of the program vector PP generated at this time.
  • the elements Tm, Gm, Pm, Sm, Hm,... are also configured as vectors with multiple elements.
  • the vector Hm corresponding to the attribute “time zone” can also be obtained in the same manner as the vector Sm of the attribute “broadcasting station” and the vector Gm of the “genre”. '
  • (personA-1) and (personB-3) indicate that personA and personB were detected once or three times, respectively, as words constituting the metadata attribute "performer".
  • step S522 When a plurality of programs are searched in step S522, a program vector is generated for each program in step S525.
  • step S526 the CPU 11 integrates the program vectors generated in step S525 to generate preference information.
  • the respective attributes of the plurality of program betattles are added to generate preference information.
  • FIG. 25 shows an example of the preference information generated at this time.
  • the preference information is generated as a vector, and the attributes “program name (title)” (Tup), “genre” (Gup), “performer” (Pup), and “broadcasting station” (Sup ), "Time zone”
  • the elements Tup, Gup, Pup, Sup, Hup, ⁇ ⁇ ⁇ are also configured as a vector with multiple elements.
  • the attribute "program name” of the preference information includes the elements “title 1" and “title 2”, and their importance has been set to “1 2" and "3”, respectively.
  • the importance indicates the degree of preference of the user for the element.
  • the program vectors included in the same element are added, the importance is added by one. For example, if the preference information is generated based on 20 program vectors PP1 to PP20, three program vectors PP5, PP10, and PP17 are generated. , If “title2” is included in the attribute Tm, the importance of the element “title2” of Tup is set to “3”.
  • step S525 If it is determined in step S525 that no group whose viewing count is equal to or greater than the threshold has been searched, the processing of steps S524 to S526 is skipped, and the processing ends.
  • the preference information is generated. Since the preference information is generated based on the metadata of the program used only a predetermined number of times or frequency, the preference information can accurately reflect the user's preference. Note that the preference information may be generated for each user by analyzing the usage history of a specific user in step S521, or the usage history of a plurality of users may be generated in step S522. The analysis may generate general (common to a plurality of users) preference information.
  • the importance is added each time a program vector including the same element is added, so that the program frequently viewed by the user is
  • the importance of elements included in metadata may become extremely high, resulting in biased preference information. For example, if a user watches a program that is broadcast every day (Monday through Friday), the importance of an element included in the metadata of the program (for example, talent A) can be compared with other elements. And it becomes extremely high. It is also possible to prevent the metadata of such frequently viewed programs (so-called standard programs) from being reflected in the preference information.
  • a preference information extraction process 2 is a second example of generating preference information based on a plurality of attributes included in the metadata of a program.
  • step S544 the CPU 11 determines whether the program belonging to the group searched in step S5452 is a standard program.
  • whether or not the program is a standard program is determined based on the standard flag set by the standard program setting process described above with reference to FIG.
  • step S544 If it is determined in step S544 that the searched program is not a standard program, the process proceeds to step S545. Then, in a manner similar to the processing of steps S524 to S525 in FIG. 23, the metadata of the program is analyzed in step S545, and a program vector is generated in step S546. In step S 547, preference information is generated.
  • step S544 determines whether the searched program is a standard program. If it is determined in step S544 that the searched program is a standard program, the processes in steps S545 to S545 are skipped. By doing so, preference information is not generated based on a standard program, and generation of biased preference information can be prevented.
  • program vectors are generated for programs belonging to a group that is watched a predetermined number of times (or frequency) or more, and the preference information is generated.
  • programs B 1, B 2, B 3, ⁇ ⁇ ⁇ ⁇ (programs belonging to one group each). For example, if the threshold of each group is three times, program A viewed three times The same applies to the program B (the series in which 10 serialized programs were watched) and the program B (the series in which 10 serialized programs were viewed). Is generated.
  • step S566 the CPU 11 specifies the familiarity of the program. Familiarity is specified based on the frequency of use of series programs (ie, groups) analyzed in step S5661. At this time, three levels of familiarity are set according to the frequency of use of the series program. For example, if the usage frequency is “0.1” or more, the familiarity level is set to “high” and the usage frequency is Power S “0.05” or more and less than “0.1” are set to “medium” familiarity, and those whose usage frequency is less than “0.05” are set to “low” familiarity. Is done.
  • the classification of familiarity is not limited to three levels.
  • the familiarity may be set as a numerical value without being classified by stage.
  • the familiarity level may be set based on the number of uses instead of the frequency of use.
  • the CPU 11 weights the thread and the title generated in step S565 based on the familiarity.
  • the importance of the preference information which is generated based on the elements included in the program vector of “high” familiarity, is set to three times, and the program vector of “medium” familiarity is set.
  • the importance of the preference information generated based on the elements included in the program information is set to double, and the importance of the preference information generated based on the elements included in the program vector with low familiarity Is set to 1 time.
  • preference information reflecting the familiarity is generated.
  • the preference information may be generated in units of users by analyzing the usage history of a specific user in step S561, or the usage information of a plurality of users may be generated in step S561.
  • general (common to multiple users) preference information may be generated. For example, for a user whose viewing history has not been accumulated yet, a program (content) can be recommended based on general preference information.
  • the preference information is generated by reflecting the familiarity of the program. Rather than recommending high-reliability programs, highly reliable programs can be recommended to users.
  • the importance of the preference information is added each time the program is viewed. However, in some cases, it is necessary to subtract the importance.
  • the user can cancel the recording reservation for the standard program for which automatic recording has been reserved in the client device 5.
  • the program whose recording reservation has been canceled is the program whose recording reservation was canceled only for that time even though it was frequently viewed before that time. It is assumed that the content did not match the tastes of the people. Therefore, in the present invention, the preference information of the user is changed based on the metadata of the program whose recording reservation has been canceled.
  • the preference information change process will be described with reference to FIG.
  • This processing is performed when the CPU 51 of the client device 5 detects the release of the automatic recording reservation, and sends the information of the program whose automatic recording reservation has been released to the content recommendation server 4 on the network 6. Is executed by the content recommendation server 4 when notified via the.
  • the CPU 11 obtains metadata of the program for which the automatic recording reservation has been canceled (for example, the third program in a series of 10 broadcasts),
  • the attribute of the acquired metadata is analyzed.
  • the CPU 11 compares the attribute of the preference information of the program for which the automatic recording reservation has been set with the attribute of the metadata of the program for which the automatic recording reservation has been canceled. In 4, detect negative elements.
  • step S585 the CPU 11 changes the user preference information based on the negative element detected in step S585. At this time, the importance of the negative element is subtracted.
  • the vector information Pup power corresponding to the attribute “performer” of the preference information, Pup ⁇ (personA-5), (personB-5),
  • the preference information is changed.
  • the importance of the attributes that the user does not like is changed to be lower, so that when recommending a program (content) to the user, it is possible to recommend a program (content) that more closely matches the user's preference. it can.
  • preference information is generated based on metadata of a series program in which the number of times of viewing is equal to or greater than a predetermined number
  • a recommendation of a program based on the preference information generated in this manner is made. If done continuously, the user may get bored. Therefore, in the present invention, attention is paid to the program that the user has watched for the first time (not watched in the past). Since the user may have a special interest in the program viewed for the first time, special preference information is generated based on the metadata of the program.
  • This process may be executed, for example, when a predetermined command is input by the user, or may be automatically executed at a predetermined cycle (for example, one week). ,.
  • step S601 the CPU 11 searches for a usage history.
  • the client device 5 uses the computer that has been used for a predetermined period (for example, the last 6 months).
  • the metadata of the content (the group ID is set) is obtained, and the number of uses ( Figure 10) for each group is analyzed.
  • step S602 the CPU 11 detects a series program that has been viewed once (a group in which only one of the programs belonging to the group is viewed).
  • step S603 the CPU 11 determines whether or not a series program with one viewing has been detected. If it is determined that a series program has been detected, the CPU 11 proceeds to step S604.
  • the special preference information is generated based on the metadata of the program belonging to the detected series program. At this time, a program vector is generated based on the metadata of the program, and special preference information is generated based on the program vector, similarly to steps S524 to S526 in FIG. . If it is determined in step S603 that a program with one viewing has not been detected, the process of step S604 is skipped.
  • FIG. 30 shows the functions of the CPU 11 of the content recommendation server 4 when recommending content based on the preference information generated by the processing described above with reference to FIGS. 23, 26, and 27.
  • FIG. 2 is a block diagram showing a typical configuration example.
  • a metadata acquisition unit 111 for acquiring metadata of a program and a preference information acquisition unit 112 for acquiring preference information of a specific user are provided.
  • the metadata of the program acquired by the metadata acquisition unit 111 is output to the program vector extraction unit 113, and the program vector extraction unit 113 extracts the program vector. Further, the preference information acquired by the preference information acquisition unit 112 is output to the preference vector extraction unit 114, and a preference vector based on the preference information is extracted. Extraction of the program vector extracted by the program vector extraction unit 113 and the preference vector The preference vector extracted by the unit 114 is output to the matching processing unit 115, and the matching processing unit 115 calculates the similarity between the program vector and the preference vector.
  • Similarity between one preference vector and a plurality of program vectors is calculated, and the matching processing unit 115 selects a predetermined number of program vectors in descending order of similarity, and selects The metadata of the program corresponding to the selected program vector is output to the information output unit 116.
  • the information output unit 116 causes the storage unit 18 to store the metadata of the program selected by the matching processing unit 115, for example.
  • step S6221 the metadata acquisition unit 111 acquires the metadata of the content (program). At this time, metadata of a plurality of programs (for example, programs to be broadcast in the next week) is acquired based on a predetermined criterion.
  • step S622 the program vector extraction unit 113 extracts a program vector based on the metadata of the program acquired in step S622. At this time, similarly to the program vector described above with reference to FIG. 24, program vectors of a plurality of programs are extracted.
  • preference vector extraction section 114 acquires preference information. At this time, preference information of a specific user is obtained. In step S624, the preference vector extraction unit 114 generates a preference vector.
  • the preference vector may be such that the preference information as shown in FIG. 25 is directly generated as a preference vector, or a specific attribute constituting the preference information is extracted and the preference vector is extracted. It may be generated as such.
  • the similarity between the preference vector UP and the program vector PP is calculated. It should be noted that the similarity between a plurality of program vectors PP and one preference vector UP is calculated. Thus, the similarity between the metadata of each program and the user's preference information is calculated.
  • step S626 the matching processing unit 115 selects metadata of a program having a high degree of similarity, and outputs it to the information output unit 116.
  • a predetermined number for example, 10
  • program vectors PP are selected in descending order of similarity, that is, in descending order of Sim value, based on the similarity calculated in step S625.
  • the metadata of the program corresponding to the selected program vector PP is output. Note that all program vectors PP having a similarity greater than a predetermined value may be selected, and metadata of a program corresponding to the selected program vector PP may be output.
  • step S627 the information output unit 116 transmits the content recommendation information of the program extracted in step S626 to the client device 5. In this way, the recommendation of the program based on the preference information is performed.
  • the recommendation of a program can also be performed based on the special preference information generated by the processing described with reference to FIG.
  • the special recommendation information search processing by the content recommendation server 4 will be described with reference to FIG. This processing may be executed, for example, when a predetermined command is input by the user, or may be automatically executed at a predetermined cycle (for example, one week).
  • steps S641 and S642 Since the processing in steps S641 and S642 is the same as the processing in steps S621 and S622 in FIG. 31, the description is omitted.
  • step S643 preference vector extraction section 114 acquires special preference information. At this time, the special preference information generated by the special preference information generation processing described above with reference to FIG. 29 is obtained. Then, in step S644, the preference vector extraction unit 114 generates a preference vector based on the special preference information acquired in step S643.
  • steps S645 and S646 is the same as the processing in steps S625 and S626 in FIG. 23, and a description thereof is omitted.
  • step S 627 the information output unit 116 transmits the content recommendation information of the program extracted in step S 646 to the client device 5.
  • the special preference information is generated based on the metadata of the program that the user has watched for the first time.
  • the series of processes described above can also be executed by software.
  • the software can be a computer that has its programs built into dedicated hardware, or can execute various functions by installing various programs, such as a general-purpose personal computer. Installed from a recording medium to a computer. As shown in FIGS. 7 and 8, this recording medium is distributed in order to provide the program to the user.
  • the program is recorded on a magnetic disk 31 or 71 (including a flexible disk) and an optical disk 32.
  • 72 including compact disk-read only memory (CD-ROM) and digital versatile disk (DVD)
  • magneto-optical disk 33 or 73 magneto-optical disk 33 or 73 (MD (mini-disk) (trademark)
  • packaged media comprising semiconductor memory 34 or 74 or the like.
  • steps for describing a program recorded on a recording medium are not limited to processing performed in chronological order in the order described, but are not necessarily performed in chronological order. Alternatively, it also includes processing that is executed individually.
  • system refers to an entire device including a plurality of devices.
  • content recommendation can be performed based on the use frequency of each group in the grouping item generated from the item representing the attribute of the content.

Landscapes

  • Engineering & Computer Science (AREA)
  • Databases & Information Systems (AREA)
  • Multimedia (AREA)
  • Signal Processing (AREA)
  • General Physics & Mathematics (AREA)
  • General Health & Medical Sciences (AREA)
  • Social Psychology (AREA)
  • Health & Medical Sciences (AREA)
  • Physics & Mathematics (AREA)
  • Theoretical Computer Science (AREA)
  • Computer Networks & Wireless Communication (AREA)
  • Computational Linguistics (AREA)
  • Data Mining & Analysis (AREA)
  • General Engineering & Computer Science (AREA)
  • Astronomy & Astrophysics (AREA)
  • Computer Graphics (AREA)
  • Information Retrieval, Db Structures And Fs Structures Therefor (AREA)
  • Two-Way Televisions, Distribution Of Moving Picture Or The Like (AREA)
  • Management, Administration, Business Operations System, And Electronic Commerce (AREA)

Abstract

本発明は、連続して視聴される番組と、非連続的に視聴される番組に基づく嗜好情報を適正に生成し、よりユーザの嗜好に合う番組を推薦できるようにする情報処理装置および方法、記録媒体、並びにプログラムに関する。嗜好情報抽出部101で番組のメタデータに基づく番組嗜好情報を抽出し、番組の視聴履歴を記録する。初めて視聴された番組の嗜好情報は、特殊番組嗜好情報として記録される。また、視聴回数が閾値を超える番組の番組嗜好情報に基づいて、ユーザ嗜好情報が生成される。制御部103が、視聴回数が閾値を超える番組について自動録画の予約設定を行う。自動録画の予約が解除された場合、嗜好情報更新部102が、予約が解除された回の番組のメタデータに基づいて、ユーザ嗜好情報を変更する。

Description

明細書
技術分野
本発明は、 情報処理装置および方法、 記録媒体、 並びにプログラムに関し、 特 に、 コンテンツの推薦を効率よく、 かつ効果的に行うことができる情報処理装置 および方法、 記録媒体、 並びにプログラムに関する。 背景技術
配信されるコンテンツから、 ユーザの嗜好の合うコンテンツを検出し、 そのコ ンテンッの情報を提供するコンテンツ推薦システムが存在する (例えば、 特開 2 0 0 0 - 2 8 7 1 8 9号公報参照) 。
このシステムでは、 例えば、 ユーザがよく利用するコンテンツの属性 (例えば、 ジャンル) を検出し、 その属性毎にコンテンツが推薦される。
しかしながら、 コンテンツの属性は、 コンテンツの編成等の事情に基づいてコ ンテンッの提供元により設定されているので、 その属性毎の推薦では、 推薦すベ きコンテンツを効率的または効果的に検出することができない場合があった。 発明の開示
本発明はこのような状況に鑑みてなされたものであり、 コンテンツの推薦を行 う側が、 コンテンツの属性を利用してコンテンツをグループ化し、 グループ毎に コンテンツの推薦を行うことができるようにするものである。
本発明の情報処理装置は、 配信されるコンテンツの属性を表す属性項目の中の 1個以上の属性項目からなるグループ化項目が一定以上の類似度をもつて類似す るコンテンツに、 同一のグループ IDを付与し、 コンテンツのグループ化を行う グループ化手段と、 グループ ID毎にコンテンツの利用頻度を算出する算出手段 と、 算出手段により算出された利用頻度に基づいて、 ユーザの嗜好を表すユーザ 嗜好情報を生成する生成手段と、 生成手段により生成されたユーザ嗜好情報に基 づいて、 コンテンツを推薦する推薦手段とを備えることを特徴とする。
放送時間帯を表す属性項目と、 少なくとも 1つ以上の他の属性項目からなるグ ループ化項目が設定されており、 グループ化手段は、 そのグループ化項目に基づ いてコンテンツのグループ化を行うことができる。
少なくとも放送時間帯を表す属性項目からなるグループ化項目と、 他の属性項 目からなるグループ化項目が設定されており、 グループ化手段は、 それらのグル ープ化項目に基づいてコンテンツのグループ化を行うことができる。
グループ化手段は、 コンテンツの属性項目の内容を形態素解析し、 その結果に 基づいて、 グループ化項目の内容の類似度を決定することができる。
生成手段は、 グループに属するコンテンツの利用状態が所定の条件を満たして いないグループの利用頻度を、 ユーザ嗜好情報の生成に利用しないようにするこ とができる。
推薦手段は、 算出手段により算出された利用頻度が、 予め設定された値より高 いか否かを判定する判定手段と、 判定手段により、 利用頻度が、 予め設定された 値より高いと判定された場合、 コンテンツの推薦情報に、 頻繁に視聴されるコン テンッであることを表す定番フラグを設定する設定手段とを設けることができる。 生成手段は、 算出手段により算出された利用頻度が予め設定された値より高い グループのコンテンツのメタデータを取得し、 メタデータの特徴量を表すべタト ルを抽出する抽出手段を備え、 抽出手段により抽出されたべク トルに基づいて、 嗜好情報を生成することができる。
生成手段は、 算出手段により算出された利用頻度が予め設定された値より高い グループのコンテンツが、 定番フラグが設定されたコンテンツ推薦情報に対応す るコンテンツか否かを判定する定番判定手段を備え、 定番判定手段により、 コン テンッが、 定番フラグが設定されたコンテンッ推薦情報に対応するコンテンツで はないと判定された場合、 抽出手段は、 コンテンツのメタデータを取得し、 メタ データの特徴量を表すべクトルを抽出することができる。
嗜好情報は、 複数の属性とその属性の重要度を表す値により構成されるように することができる。
生成手段は、 算出手段により算出された利用頻度に基づいて、 コンテンツの熟 知度を設定する熟知度設定手段を備え、 熟知度に基づいて、 嗜好情報の重要度を 表す値に重み付けを行うことができる。
生成手段は、 コンテンツの利用履歴に基づいて、 利用回数が所定の値以下だけ 利用されたコンテンツを検索する検索手段と、 検索手段により検索されたコンテ ンッのメタデータに基づいて、 特殊嗜好情報を生成する特殊嗜好情報生成手段と をさらに設けることができる。
嗜好情報または特殊嗜好情報の特徴量を表すべクトルを抽出する第 1の抽出手 段と、 予め設定された期間に放送されるコンテンツのメタデータを取得し、 メタ データの特徴量を表すべクトルを抽出する第 2の抽出手段と、 第 1の抽出手段に より抽出されたべク トルと第 2の抽出手段により抽出されたべクトルの類似度を 演算する演算手段とを備え、 推薦手段は、 類似度が高い順に、 予め設定された数 だけ第 2の抽出手段により抽出されたべク トルを選択し、 選択されたべクトルの メタデータに基づいて、 コンテンツを推薦することができる。
本発明の情報処理方法は、 配信されるコンテンツの属性を表す属性項目の中の 1個以上の属性項目からなるグループ化項目が一定以上の類似度をもつて類似す るコンテンッに、 同一のグループ IDを付与し、 コンテンツのグループ化を行う グループ化ステップと、 グループ ID毎にコンテンツの利用頻度を算出する算出 ステップと、 算出ステップの処理で算出された利用頻度に基づいて、 ユーザの嗜 好を表すユーザ嗜好情報を生成する生成ステップと、 生成ステップの処理で生成 されたユーザ嗜好情報に基づいて、 コンテンツを推薦する推薦ステップとを含む ことを特徴とする。 本発明の記録媒体のプログラムは、 配信されるコンテンツの属性を表す属性項 目の中の 1個以上の属性項目からなるグループ化項目が一定以上の類似度をもつ て類似するコンテンツに、 同一のグループ IDを付与することによっての、 コン テンッのグループ化を制御するグループ化制御ステップと、 グループ ID毎のコ ンテンッの利用頻度の算出を制御する算出制御ステップと、 算出制御ステップの 処理で算出された利用頻度に基づいての、 ユーザの嗜好を表すユーザ嗜好情報の 生成を制御する生成制御ステップと、 生成制御ステップの処理で生成されたユー ザ嗜好情報に基づいての、 コンテンツの推薦を制御する推薦制御ステップとを含 むことを特徴とする。
本発明のプログラムは、 配信されるコンテンツの属性を表す属性項目の中の 1 個以上の属性項目からなるグループ化項目が一定以上の類似度をもつて類似する コンテンツに、 同一のグループ IDを付与することによっての、 コンテンツのグ ループ化を制御するグループ化制御ステップと、 グループ ID毎のコンテンツの 利用頻度の算出を制御する算出制御ステップと、 算出制御ステップの処理で算出 された利用頻度に基づいての、 ユーザの嗜好を表すユーザ嗜好情報の生成を制御 する生成制御ステップと、 生成制御ステップの処理で生成されたユーザ嗜好情報 に基づいての、 コンテンツの推薦を制御する推薦制御ステップとを含むことを特 徴とする。
本発明の情報処理装置おょぴ方法、 並びにプログラムにおいては、 配信される コンテンッの属性を表す属性項目の中の 1個以上の属性項目からなるグループ化 項目が一定以上の類似度をもって類似するコンテンツに、 同一のグループ IDが 付与され、 コンテンツのグノレープ化が行われ、 グループ ID毎にコンテンツの利 用頻度が算出され、 算出された利用頻度に基づいて、 ユーザの嗜好を表すユーザ 嗜好情報が生成され、 生成されたユーザ嗜好情報に基づいて、 コンテンツが推薦 される。 図面の簡単な説明 図 1は、 本発明を適用したコンテンツ推薦システムの構成例を示している。 図 2は、 メタデータの例を示す図である。
図 3は、 コンテンツのグループ化を説明する図である。
図 4は、 コンテンツのグループ化を説明する他の図である。
図 5は、 グループ IDが付されたメタデータの例を示す図である。
図 6は、 利用履歴の例を示す図である。
図 7は、 図 1のコンテンツ推薦サーバの構成例を示すブロック図である。
図 8は、 図 1のクライアント機器の構成例を示すプロック図である。
図 9は、 ユーザ嗜好情報生成処理を説明するフローチャートである。
図 1 0は、 利用頻度の算出方法を説明する図である。
図 1 1 Aは、 利用頻度の算出方法を説明する他の図である。
図 1 1 Bは、 利用頻度の算出方法を説明する他の図である。
図 1 2は、 利用状態の確認処理を説明する図である。
図 1 3は、 利用状態の確認処理を説明する他の図である。
図 1 4は、 利用状態の確認処理を説明する他の図である。
図 1 5は、 コンテンツ推薦情報生成処理を説明するフローチャートである。 図 1 6は、 コンテンツ推薦情報の表示例を示す図である。
図 1 7は、 他のコンテンツ推薦情報の表示例を示す図である。
図 1 8は、 タイトルグループ化処理 1について説明するフローチャートである。 図 1 9は、 タイトルグループ化処理 2について説明するフローチヤ一トである。 図 2 0は、 タイトルグループ化処理 3について説明するフローチヤ一トである。 図 2 1は、 タイトルグループ化処理 4について説明するフローチヤ一トである。 図 2 2は、 定番番組設定処理を説明するフローチャートである。
図 2 3は、 嗜好情報抽出処理 1を説明するフローチヤ一トである。
図 2 4は、 番組べク トルの構成例を示す図である。
図 2 5は、 嗜好情報の構成例を示す図である。
図 2 6は、 嗜好情報抽出処理 2を説明するフローチヤ一トである。 図 2 7は、 嗜好情報抽出処理 3を説明するフローチャートである。
図 2 8は、 嗜好情報変更処理を説明するフローチャートである。
図 2 9は、 特殊嗜好情報生成処理を説明するフローチャートである。
図 3 0は、 図 7の C P Uの機能的構成例を示すプロック図である。
図 3 1は、 推薦情報検索処理を説明するフローチャートである。
図 3 2は、 特殊推薦情報検索処理を説明するフローチャートである。 発明を実施するための最良の形態
図 1は、 本発明を適用したコンテンツ推薦システムの構成例を示している。 配信サーバ 3は、 ストリーミングデータデータベース 1から、 ストリーミング データを取得し、 インターネットその他のネットワークを含むネットワーク 6を 介して、 クライアント機器 5に配信する。 配信サーバ 3はまた、 メタデータデー タベース 2から、 コンテンッのメタデータを取得し、 ネットワーク 6を介して、 コンテンツ推薦サーバ 4に供給する。
メタデータは、 コンテンツ毎に、 例えば、 図 2に示すような、 「放送開始時 刻」 、 「放送終了時刻」 、 「放送局」 、 「ジャンル」 、 「タイトル」 、 「出演者 名」 、 「番組内容」 、 および 「キーワード」 等のコンテンツの属性を表す項目か ら構成されている。
コンテンッ推薦サーバ 4は、 1個以上の項目で設定されたグループ化項目につ いて、 その内容 (グループ化項目を構成する各項目の要素) が一定以上類似する (各項目の要素が、 全部一致、 一部一致、 または所定の類似度を表す値が所定以 上となる) コンテンツに同一のグループ I Dを付与する (同一のグループにグル ープ化する) 。
メタデータの項目 「放送局」 、 項目 「放送開始時刻」 、 および項目 「放送終了 時刻」 からなるグループ化項目が設定されている場合、 「8 ch (放送局) , 0 0 : 0 0 (放送開始時間) ~ 0 6 : 0 0 (放送終了時刻) 」 のコンテンツ (8 chで、 0 0 : 0 0〜0 6 : 0 0の間に放送される番組) には同一のグループ ID が付与される。
すなわちこのグループ化項目の下では、 図 3に示すように、 グループ化項目を 構成する項目 「放送局」 、 項目 「放送開始時刻」 、 および項目 「放送終了時刻」 の各要素の組み毎にコンテンツのグループ化がなされる。
また、 メタデータの項目 「ジャンル」 および項目 「出演者」 からなるグループ 化項目が設定されている場合、 例えば、 「バラエティ (ジャンル) 、 タレント A (出演者) 」 のコンテンツ (バラエティで、 タレント Aが出演する番組) には同 一のグループ IDが付与される。
すなわちこのグループ化項目の下では、 図 4に示すように、 グループ化項目を 構成する項目 「ジャンル」 および項目 「出演者」 の各要素の組み毎にコンテンツ のグループ化がなされる。
なおグループ化項目が複数設定されている場合、 1個のコンテンツは、 コンテ ンッの項目によっては、 複数のグループに属することもあり得る。 例えば、 8 chで、 0 0 : 0 0〜 0 6 : 0 0の間に放送され、 バラエティで、 タレント Aが 出演する番組には、 「8 ch (放送局) , 0 0 : 0 0 (放送開始時間) 〜0 6 : 0 0 (放送終了時刻) 」 のグループ I D (図 3 ) と、 「バラエティ (ジャンル) 、 タレント A (出演者) 」 のグループ I D (図 4 ) が付与され、 それぞれのグルー プに属することになる。
コンテンツ推薦サーバ 4は、 このようにグループ I Dが設定されたメタデータ (例えば、 図 5 ) を、 適宜、 クライアント機器 5に送信する。
コンテンツ推薦サーバ 4はまた、 クライアント機器 5から、 コンテンツのダル ープ I Dを含む利用履歴を適宜取得し、 その利用履歴に基づいて、 グループ毎の 利用頻度を算出する。 そしてコンテンツ推薦サーバ 4は、 算出したその利用頻度 をユーザの嗜好を表すものとして利用し、 グループ毎にコンテンツの推薦を行う。 例えば、 高い利用頻度のグループに属するコンテンツに関する情報が、 コンテン ッ推薦情報としてクライアント機器 5に送信される。 クライアント機器 5は、 配信サーバ 3から配信されてきたコンテンツを利用す るが、 その利用履歴として、 利用したコンテンツの、 例えば図 6に示すようなメ タデータ (グループ I Dが設定されたもの) を、 適宜、 コンテンツ推薦サーバ 4 に供給する。
クライアント機器 5は、 コンテンツ推薦サーバ 4から供給されたコンテンツ推 薦情報を表示してユーザに提示する。 ユーザは、 それを参照することにより、 自 分の嗜好にあったコンテンツを選択することができる。
なお、 ここでの配信サーバ 3乃至クライアント機器 5の通信は、 ネットワーク 6を介して行われているが、 それぞれ直接通信する構成にすることもできる。 図 7は、 コンテンツ推薦サーバ 4の構成例を示している。 CPU (Central
Process ing Unit) 1 1 ίま、 ROM (Read Only Memory) 1 2に記'慮されて ヽ る、 例えば、 コンテンツ推薦用のプログラム等に従って所定の処理を実行する。 RAM (Random Access Memory) 1 3には、 CPU 1 1がその処理を実行する上に おいて必要なデータなどが適宜記憶される。
CPU 1 1にはバス 1 4を介して入出力ィンタフェース 1 5が接続されている。 入出力インタフェース 1 5には、 キーボード、 マウスなどよりなる入力部 1 6、 LCD (Liquid Crystal Di splay) などよりなる出力部 1 7、 メタデータ等を 記憶する記憶部 1 8、 およびネッ トワーク 6を介して配信サーバ 3またはクライ アント機器 5との通信を行う通信部 1 9が接続されている。
入出力インタフェース 1 5には、 ドライブ 2 0が適宜接続され、 CPU 1 1は、 そこに装着される磁気デイスク 3 1、 光ディスク 3 2、 光磁気ディスク 3 3、 ま たは半導体メモリ 3 4との間でデータの授受を行う。
なお、 C P U 1 1の機能的構成例として、 例えば、 ユーザの嗜好情報を取得す る嗜好情報取得部、 配信サーバ 3から番組のメタデータを取得するメタデータ取 得部、 およびコンテンツの推薦情報を生成する推薦情報生成部により構成される ようにすることも可能である。 図 8は、 クライアント機器 5の構成例を示している。 この構成は、 コンテンツ 推薦サーバ 4の構成と基本的に同様であるので、 その説明は省略する。
次に、 ユーザ嗜好情報を生成する場合のコンテンツ推薦サーバ 4の動作を、 図 9のフローチヤ一トを参照して説明する。
ステップ S 1において、 コンテンツ推薦サーバ 4の CPU 1 1は、 ユーザ嗜好情 報を生成するタイミングであるか否かを判定し、 そのタイミングであると判定し た場合、 ステップ S 2に進む。 例えば、 クライアント機器 5からコンテンツ推薦 情報 (後述) の提供が要求されたとき、 または予め決められた時期 (例えば、 毎 週所定の時刻) が来たとき、 ステップ S 2に進む。
ステップ S 2において、 CPU 1 1は、 通信部 1 9を介して、 クライアント機器 5から、 所定の利用履歴を取得する。 この例の場合、 1週間前からの間に利用さ れたコンテンツのメタデータ (グループ I Dが設定されている) が取得される。 CPU 1 1は、 グループ毎のコンテンツの利用頻度 (回数) を算出する。
項目 「放送局」 、 項目 「放送開始時刻」 、 および項目 「放送終了時刻」 からな るグループ化項目が設定されている場合、 各メタデータには、 そのグループ化項 目の内容 (グループ化項目を構成する各項目の要素の組み) に応じたグループ I Dが記述されているので、 図 1 0に示すように、 その内容 (各項目の要素の組 み) 毎にコンテンツの利用頻度 (回数) が算出される。
図 1 0が示すグループ毎の利用回数からは、 8 chで、 2 0 : 0 0 ~ 2 1 : 0 0で放送される番組と、 l O chで、 1 9 : 0 0〜2 0 : 0 0で放送される番組 が最も多く視聴され (各 7回) 、 それに続いて 8 chで、 2 2 : 0 0〜 2 3 : 0 0の間に放送される番組が次に視聴 (6回) されていることがわかる。
また項目 「ジャンル」 および項目 「出演者」 からなるグループ化項目が設定さ れている場合、 各メタデータには、 そのグループ化項目の内容 (グループ化項目 を構成する各項目の要素の組み) に応じたグループ I Dが記述されているので、 図 1 1 Aに示すように、 その内容 (各項目の要素の組み) 毎にコンテンツの利用 頻度 (回数) が算出される。 図 1 1 Aがグループ毎の利用回数からは、 タレント Dが出演するバラエティ番 組が最も多く視聴され (1 0回) 、 それに続いてタレント Dが出演するニュース 番組 (8回) とタレント Cが出演するバラエティ番組 (5回) が多く視聴されて いることがわかる。
なお、 利用回数は、 コンテンツの配信数に応じて多くなる可能性があるので、 そのままではユーザの嗜好を正確に対応しない。 そのため、 ステップ S 2で取得 された利用履歴に対応する期間中に配信されたコンテンツ数で利用回数が正規化 される。
例えば、 例えば図 1 1の例の場合、 タレント Dが出演するパラエティ番組が、 その期間 (この例の場合、 1週間) の間に 1 0本配信され、 タレント Dが出演す るェユース番組が、 その期間の間に 1 00本配信され、 タレント Cが出演するバ ラエティ番組が、 その期間の間に 80本配信された場合、 図 1 1 Aの利用回数は、 図 1 1 Bに示すように、 正規化される。 このように利用回数を正規化することで ユーザの嗜好を適切に対応した利用頻度を得ることができる。 . .
図 9に戻り、 ステップ S 3において、 コンテンツ編集サーバ 4の CPU1 1は、 グループ化項目毎に、 所定の閾値以上の利用回数 (利用頻度) が得られたグルー プ (のグループの I D) を検出する。
例えば、 項目 「放送局」 、 項目 「放送開始時刻」 、 および項目 「放送開始時 刻」 からなるグループ化項目に対する閾値が値 7である場合、 図 1 0の例では、 「8ch, 2 0 : 0 0〜 2 1 : 0 0」 、 および 「1 0ch, 1 9 : 0 0〜 20 : 0 0」 の 2つのグループが検出される。
また項目 「ジャンル」 および項目 「出演者」 からなるグループ化項目に対する 閾値が値 0. 0 6である場合、 図 1 1 Bの例では、 「バラエティ, タレント D」 、 「ニュース, タレント D」 、 および 「バラエティ, タレント C」 の 3つのグルー プが検出される。
次に、 ステップ S 4において、 CPU1 1は、 ステップ S 3で検出した各グルー プのコンテンツがユーザの嗜好に合っているか否かを判定する。 例えば、 そのグループに属するコンテンツの配信リストに基づいて、 いまから 遡って所定の回数 (例えば、 3回) 連続して利用されなかったか否かが確認され、 その回数連続して利用されなかった場合、 そのグループのコンテンツはユーザの 嗜好に合っていないと判定と判定される。
図 1 3に示すように、 「バラエティ, タレント D」 のグループの番組が 3回連 続して視聴されなかったとき、 「バラエティ, タレント D」 のグループのコンテ ンッは、 ユーザの嗜好に合わないと判定される。
なお、 図 1 2に示すように、 グループ 「8ch, 2 0 : 00〜 2 1 : 0 0」 の グループの、 いまから遡って最も最近の番組 Aは視聴されなかったが、 その前に 配信された番組 Bは視聴されているとき (視聴されないことが 3回連続していな いとき) 、 「8ch, 2 0 : 00〜2 1 : 0 0」 のグループのコンテンツは、 ュ 一ザの嗜好に合わないとは判定されない (合っていると判定される) 。
また図 1 4に示すように、 過去に、 所定の回数 (例えば、 3回) 連続して利用 されているとき、 そのグループのコンテンツを、 ユーザの嗜好にあっていると判 定することもできる。
ステップ S 5において、 CPUl 1は、 ステップ S 4での判定結果から、 ユーザ の嗜好に合うコンテンツのグループを検出する。
ステップ S 6において、 CPU1 1は、 ステップ S 5で検出したグループのグル ープ IDを、 ユーザ嗜好情報として、 記憶部 1 8に記憶する。
この例の場合、 項目 「放送局」 、 項目 「放送開始時刻」 、 および項目 「放送終 了時刻」 からなるグループ化項目における 「8ch, 20 : 0 0〜2 1 : 0 0」 と 「1 0ch, 1 9 : 0 0〜 20 : 0 0」 のグループのグループ ID、 および項目 「ジャンル」 および項目 「出演者」 からなるグループ化項目における 「ニュース, タレント D」 と 「バラエティ, タレント C」 のグループのグノレープ IDが、 ユー ザ嗜好情報として記憶部 1 8に記憶される。
次に、 コンテンツ推薦情報を生成する場合のコンテンツ推薦サーバ 4の動作を、 図 1 5のフローチヤ一トを参照して説明する。 ステップ S 2 1において、 コンテンツ推薦サーバ 4の CPU 1 1は、 クライアン ト機器 5から、 コンテンツ推薦情報の提供が要求されるまで待機し、 その要求が あつたとき、 ステップ S 2 2に進み、 記憶部 1 8から上述したようにして生成し たユーザ嗜好情報を取得する。
ステップ S 2 3において、 CPU 1 1は、 いまから配信されるコンテンツのメタ データ (グループ IDが設定されているもの) から、 ユーザ嗜好情報としてのグ ループ IDと同じグループ IDが設定されているメタデータを抽出する。 CPU1 1 は、 抽出したメタデータからコンテンツ推薦情報を生成する。
なお、 ユーザ嗜好情報として記憶されているグループ ID が複数ある場合には、 いずれのグループ IDも付与されているコンテンツのメタデータが抽出されるよ うにすることができる。
ステップ S 24において、 CPU1 1は、 通信部 1 9を介して、 ステップ S 2 3 で生成したコンテンツ推薦情報を、 クライアント機器 5に送信する。 クライアン ト機器 5は、 コンテンッ推薦サーバ 4から送信されてきたコンテンッ推薦情報を 出力部 5 7に表示する。
図 1 6および図 1 7は、 コンテンツ推薦情報の表示例を示している。
図 1 6の例では、 項目 「放送局」 、 項目 「放送開始時刻」 、 および項目 「放送 終了時刻」 からなるグループ化項目の 「8ch, 20 : 0 0~ 2 1 : 0 0」 およ び 「1 0ch, 1 9 : 00〜 20 : 0 0」 のグループに属する番組の情報 (タイ トル等) がそれぞれ表示されている。
図 1 7の例では、 項目 「ジャンル」 および項目 「出演者」 からなるグループ化 項目の 「ニュース, タレント Dj および 「バラエティ, タレント C」 のグノレープ に属する番組に関する情報 (タイ トル等) がそれぞれ表示されている。 なお、 各 グループの番組の情報が表示されているウインドウは、 表示画面の大きさによつ ては、 図 1 7に示すように一部重なるように表示されるようにすることができる。 ユーザは、 このように表示されたコンテンツ推薦情報を参照して、 視聴する番 組を選択することができる。 5927
13
以上のように、 ユーザの嗜好を把握するための利用頻度を、 グループ IDを利 用してグループ毎に算出するようにしたので、 メタデータの項目毎に利用頻度を 算出する場合に比べ、 その計算量を少なくすることができる。
また、 コンテンツ推薦情報がグループ毎にまとまって表示されるようにしたの で、 表示スペースが小さいクライアント機器 5においても適切にコンテンツ推薦 情報を表示することができる。
また、 以上においては、 メタデータの項目 「放送局」 、 項目 「放送開始時刻」 、 および項目 「放送終了時刻」 、 並びに項目 「ジャンル」 および項目 「出演者」 を 利用してグループ化を行ったが、 項目 「タイ トル」 や項目 「内容」 など他の項目 を利用してグループ化を行うこともできる。 その結果、 例えば、 再放送やスぺシ ャル版の番組をオリジナルの番組と同じグループに属するコンテンツとして扱う ことができるので、 オリジナルであろうと、 再放送されたものであろうと、 その 番組が視聴されれば、 その利用履歴をユーザ嗜好情報生成に反映することができ る。
ここで、 図 1 8のフローチャートを参照して、 項目 「タイトル」 を利用してグ ループ化を行う処理 (項目 「タイ トル」 からなるグループ化項目によるグループ 化処理) (タイ トルグループ化処理 1 ) について説明する。
ステップ S 6 1において、 コンテンツ推薦サーバ 4は、 メタデータから、 タイ トルを抽出する。
ステップ S 6 2において、 コンテンツ推薦サーバ 4は、 タイ トルを形態素解析 し、 単語に分解する。
例えばメタデータに含まれている映画の題名が 「東海道三谷怪談」 であった場 合、 これがタイ トルとして形態素解析され、 「東海道」 、 「三谷」 、 「怪談」 の
3つの単語が得られる。
ステップ S 6 3において、 コンテンツ推薦サーバ 4は、 解析された単語、 もし くは、 複数の単語から構成される単語群のうちのいずれかを抽出して、 記憶部 1
8から、 抽出された単語、 または単語群に対応するグループ I Dを抽出する。 ここで、 複数の単語から構成される単語群とは、 形態素解析により得られた単 語の組み合わせにより生成される単語群であり、 例えば、 形態素解析により得ら れた単語が 「東海道」 、 「三谷」 、 「怪談」 である場合、 単語群は、 「東海道 - 三谷」 、 「東海道 '怪談」 、 「三谷 '怪談」 となる。
ステップ S 6 4において、 コンテンツ推薦サーバ 4は、 グループ I Dが抽出さ れたか否かを判断する。
ステップ S 6 4において、 対応するグループ I Dが抽出されなかったと判断さ れた場合、 抽出された単語、 もしくは、 複数の単語から構成される単語群には、 まだグループ I Dが付けられていないので、 ステップ S 6 5において、 抽出され た単語、 もしくは、 複数の単語から構成される単語群に新たなグループ I Dを割 り当てる。 また、 コンテンツ推薦サーバ 4は、 単語、 もしくは、 複数の単語から 構成される単語群と、 それに対応するグループ I Dを記憶する。
ステップ S 6 4において、 対応するグループ I Dが抽出されたと判断された場 合、 または、 ステップ S 6 5の処理の終了後、 ステップ S 6 6において、 コンテ ンッ推薦サーバ 4は、 タイトルを構成する全ての単語、 もしくは、 複数の単語か ら構成される単語群についてグループ I Dを抽出したか否かを判断する。
ステップ S 6 6において、 タイ トルを構成する全ての単語、 もしくは、 複数の 単語から構成される単語群についてグループ I Dを抽出していないと判断された 場合、 処理は、 ステップ S 6 3に戻り、 それ以降の処理が繰り返される。
ステップ S 6 6において、 タイ トルを構成する全ての単語、 もしくは、 複数の 単語から構成される単語群についてグループ I Dが抽出されたと判断された場合、 ステップ S 6 7において、 コンテンツ推薦サーバ 4は、 メタデータに、 抽出され たまたは割り当てたグループ I Dを対応付けて、 処理が終了される。
なお、 類似したタイトルの番組を、 同一のグループとするようにしても良い。 例えば、 タイトル 「2年 A組銀八先生」 の連続ドラマと、 タイトル 「2年 A組銀 八先生スペシャル」 の特別番組とを、 同一のグループとしてグループ化すること ができるように、 タイ トルを構成する単語の形態素解析結果を基に、 例えば、 2 週間、 1ヶ月、 半年などの所定の期間の番組タイ トルで、 単語の一致率を総当り で算出し、 単語の一致率が所定の値以上である場合、 同一グループとするように してもよい。
次に、 図 1 9のフローチャートを参照して、 タイ トルを構成する単語の一致率 によりグループ化を実行するタイ トルグループ化処理 2 (項目 「タイトル」 から なるグループ化項目による他のグループ化処理) について説明する。
ステップ S 4 0 1およびステップ S 4 0 2において、 図 1 8を用いて説明した、 ステップ S 6 1およびステップ S 6 2と同様の処理が実行される。 すなわち、 コ ンテンッ推薦サーバ 4は、 メタデータから、 タイトルを抽出して形隼素解析し、 単語に分解する。
ステップ S 4 0 3において、 コンテンツ推薦サーバ 4は、 解析された単語を基 に、 タイトル間の単語の一致度、 すなわち、 単語が一致している割合を示す一致 率を算出する。
具体的には、 タイ トル 「2年 A組銀八先生」 と、 タイトル 「2年 A組銀八先生 スペシャル」 とが、 それぞれ、 「2」 「年」 「A」 「組」 「銀八」 「先生」 と、 「2」 「年」 「A」 「組」 「銀八」 「先生」 「スペシャル」 とに形態素分析され た場合、 この 2つの番組のタイトルを構成する単語の一致率は、 6 / 7で 8 5 . 7 %となる。
ステップ S 4 0 4において、 コンテンツ推薦サーバ 4は、 単語が、 例えば、 7 0 %などの所定の値以上一致しているか否かを判断する。 この、 一致率の閾値は、 7 0 %以外のいかなる数値であっても良いことは言うまでもない。
ステップ S 4 0 4において、 単語が、 7 0 %などの所定の値以上一致している と判断された場合、 ステップ S 4 0 5において、 コンテンツ推薦サーバ 4は、 そ れらの番組に、 同一のグループ I Dを対応付ける。 また、 コンテンツ推薦サーバ 4は、 一致した単語、 または、 単語群と、 それに対応するグループ I Dを記憶す る。 ' ステップ S 4 0 4において、 7 0 %などの所定の値以下の一致率であると判断 された場合、 または、 ステップ S 4 0 5の処理の終了後、 ステップ S 4 0 6にお いて、 コンテンツ推薦サーバ 4は、 タイ トルの総当りが終了したか否かを判断す る。
ステップ S 4 0 6において、 タイトルの総当りが終了していないと判断された 場合、 処理は、 ステップ S 4 0 3に戻り、 それ以降の処理が繰り返される。
ステップ S 4 0 6において、 タイトルの総当りが終了したと判断された場合、 処理が終了される。
このような処理により、 タイトルを構成する単語の一致率を基にしたグループ I Dが対応付けられるので、 例えば、 連続ドラマとスペシャル番組などの類似し たタイトルの番組を、 同一のグループとして処理させるようにすることができる。 また、 タイトルを構成する単語の一致率を基にグループを決定するようにする ことにより、 例えば、 メタデータにおいて、 数字の半角と全角、 または、 英字の 半角と全角、 もしくは、 大文字と小文字などの表記ゆれがあった場合にも、 同一 タイトルの番組を、 同一のグループとして検出することが可能となる。
また、 単語の一致率に加えて、 例えば、 放送局や番組ジャンル、 あるいは、 放 送開始時刻などを、 グループ化の条件に加えるようにしても良い。 例えば、 ニュ ース番組などにおいては、 タイトルが、 「ニュース」 を含む少ない単語によって 構成されているので、 図 1 9を用いて説明した処理では、 異なる放送局の異なる 形態のニュース番組であっても、 同一のグループとして検出されてしまう恐れが あるので、 単語の一致率に加えて、 放送局も一致した場合、 同一グループとする ようにしても良い。
次に、 図 2 0のフローチャートを参照して、 放送局の一致を条件に加えて、 タ ィトルを構成する単語の一致率によりグループ化を実行するタイトルグループ化 処理 3 (項目 「タイ トル」 と項目 「放送局」 からなるグループ化項目によるダル ープ化処理) について説明する。 . ステップ S 4 2 1乃至ステップ S 4 2 4において、 図 1 9を用いて説明した、 ステップ S 4 0 1乃至ステップ S 4 0 4と同様の処理が実行される。 すなわち、 コンテンツ推薦サーバ 4は、 メタデータから、 タイ トルを抽出して形態素解析し、 単語に分解する。 そして、 コンテンツ推薦サーバ 4は、 解析された単語を基に、 タイトル間の単語の一致度を算出し、 単語が、 例えば、 7 0 %などの所定の値以 上一致しているか否かを判断する。
ステップ S 4 2 4において、 単語が、 7 0 %などの所定の値以上一致している と判断された場合、 ステップ S 4 2 5において、 コンテンツ推薦サーバ 4は、 そ れらの番組の放送局が一致しているか否かを判断する。
ステップ S 4 2 5において、 これらの番組の放送局が一致していると判断され た場合、 ステップ S 4 2 6において、 コンテンツ推薦サーバ 4は、 それらの番組 に、 同一のグループ I Dを対応付ける。 また、 コンテンツ推薦サーバ 4は、 一致 した単語、 または、 単語群と、 それに対応する放送局およびグループ I Dを記憶 する。
ステップ S 4 2 4において、 7 0 %などの所定の値以下の一致率であると判断 された場合、 ステップ S 4 2 5において、 これらの番組の放送局が一致していな いと判断された場合、 または、 ステップ S 4 2 6の処理の終了後、 ステップ S 4 2 7において、 コンテンツ推薦サーバ 4は、 タイ トルの総当りが終了したか否か を判断する。
ステップ S 4 2 7において、 タイトルの総当りが終了していないと判断された 場合、 処理は、 ステップ S 4 2 3に戻り、 それ以降の処理が繰り返される。
ステップ S 4 2 7において、 タイトルの総当りが終了したと判断された場合、 処理が終了される。
このような処理により、 放送局の一致とタイトルを構成する単語の一致率を基 にしたグループ I Dが対応付けられるので、 例えば、 類似したタイ トルの番組を、 同一のグループとする場合に、 他局の-ユース番組を同一のグループとするよう なことを防ぐことができる。 なお、 図 2 0においては、 タイ トルを構成する単語の一致率以外に、 同一の放 送局であるか否かを条件として、 グループ化を行うものとして説明したが、 放送 局以外の、 例えば、 放送時間帯やジャンルなどを、 タイトルを構成する単語の一 致率以外の条件として、 グループ化を実行するようにしても良いことは言うまで もない。
更に、 例えば、 連続ドラマや帯番組の放送開始時刻が、 スポーツ中継や特別番 組などのためにずれた場合においても、 同一グループとして検出可能なように、 タイトルを構成する単語の一致率以外の条件を、 放送時刻が、 例えば、 1時間な どの所定の時間範囲内のずれで一致しているか否かとして、 グループ化を実行す るようにしても良い。
図 2 1のフローチャートを参照して、 放送時刻が、 所定の時間範囲内のずれで —致しているか否かを条件に加えて、 タイ トルを構成する単語の一致率によりグ ループ化を実行するタイトルグループ化処理 4 (項目 「タイトル」 と項目 「放送 開始時刻」 からなるグループ化項目によるグループ化処理) について説明する。 ステップ S 4 4 1乃至ステップ S 4 4 4において、 図 1 9を用いて説明した、 ステップ S 4 0 1乃至ステップ S 4 0 4と同様の処理が実行される。 すなわち、 コンテンツ推薦サーバ 4は、 メタデータから、 タイトルを抽出して形態素解析し、 単語に分解する。 そして、 コンテンツ推薦サーバ 4は、 解析された単語を基に、 タイトル間の単語の一致度を算出し、 単語が、 例えば、 7 0 %などの所定の値以 上一致しているか否かを判断する。
ステップ S 4 4 4において、 単語が、 7 0 %などの所定の値以上一致している と判断された場合、 ステップ S 4 4 5において、 コンテンツ推薦サーバ4は、 そ れらの番組の放送開始時刻が、 例えば、 1時間などの所定の範囲內のずれで一致 しているか否かを判断する。
ステップ S 4 4 5において、 それらの番組の放送開始時刻が所定の範囲内のず れで一致していると判断された場合、 ステップ S 4 4 6において、 コンテンツ推 薦サーバ 4は、 それらの番組に、 同一のグループ I Dを対応付ける。 また、 コン テンッ推薦サーバ 4は、 一致した単語、 または、 単語群と、 それに対応する放送 開始時刻の範囲、 およびグループ I Dを記憶する。
ステップ S 4 4 4において、 7 0 %などの所定の値以下の一致率であると判断 された場合、 ステップ S 4 4 5において、 それらの番組の放送開始時刻が所定の 範囲以上にずれていると判断された場合、 または、 ステップ S 4 4 6の処理の終 了後、 ステップ S 4 4 7において、 コンテンツ推薦サーバ 4は、 タイトルの総当 りが終了したか否かを判断する。 ·
ステップ S 4 4 7において、 タイトルの総当りが終了していないと判断された 場合、 処理は、 ステップ S 4 4 3に戻り、 それ以降の処理が繰り返される。 ステップ S 4 4 7において、 タイトルの総当りが終了したと判断された場合、 処理が終了される。
このような処理により、 放送開始時刻の所定の範囲内のずれを含む一致と、 タ ィ トルを構成する単語の一致率を基にしたグループ I Dが対応付けられるので、 例えば、 類似したタイトルの番組を同一のグループとする場合に、 特別番組など による放送時刻の変更のために、 同一グループとして検出されるべき番組が、 同 一グループとして検出されないようなことを防ぐことができる。
なお、 以上においては、 コンテンツ推薦サーバ 4が、 ユーザ嗜好情報生成処理 (図 9 ) およびコンテンツ推薦情報処理 (図 1 5 ) を行う場合を例として説明し たが、 クライアント機器 5が、 コンテンツ推薦サーバ 4から供給されるグループ I Dが設定されたメタデータ (グループ化情報) を利用して、 自分自身がグルー プ毎の利用頻度を算出してユーザ嗜好情報を生成し、 それに基づいてコンテンツ 推薦情報を生成することもできる。
また、 頻繁に視聴される番組をいわゆる定番の番組として推薦し、 推薦された 番組が自動的に、 視聴または録画されるようにすることもできる。 図 2 2を参照 して定番番組設定処理について説明する。 この処理は、 コンテンツ推薦サーバ 4 において、 図 1 5を参照して上述したコンテンツ推薦情報生成処理を実行するの に先立って (事前に) 実行される。 ステップ S 5 0 1において、 CPU 1 1は、 利用履歴を分析する。 このとき、 図 9のステップ S 2の場合と同様に、 クライアント機器 5から、 所定の期間に利用 されたコンテンツのメタデータ (グループ I Dが設定されている) が取得され、 グループ毎の利用回数 (図 1 0 ) 、 または利用頻度 (図 1 I B ) が分析される。 ステップ S 5 0 2において、 CPU 1 1は、 利用回数 (視聴回数) が所定の閾値 を超えるグループがあるか否かを判定し、 利用回数が閾値を超えるグループがあ ると判定された場合、 ステップ S 5 0 3に進み、 そのグループに属する番組 (利 用回数が閾値を超える番組) のコンテンツ推薦情報に、 この番組が定番であるこ とを表す定番フラグを設定する。
また、 ステップ S 5 0 2において、 視聴頻度が閾値を超えるグループがあるか 否かが判定され、 利用頻度が閾値を超える番組があると判定された場合、 ステツ プ S 5 0 3において、 そのグループに属する番組のコンテンツ推薦情報に定番フ ラグが設定されるようにしてもよい。
ステップ S 5 0 2において、 視聴回数が閾値を超えるグループがないと判定さ れた場合、 処理は終了される。
このようにして、 定番フラグが設定されたコンテンツ推薦情報が、 図 1 5のコ ンテンッ推薦情報生成処理により、 クライアント機器 5に送信される。 これによ り、 クライアント機器 5において、 例えば、 定番フラグが設定されたコンテンツ 推薦情報に対応する番組が自動録画されるようにすることができる。
上述した図 9のユーザ嗜好情報生成処理においては、 ユーザ嗜好情報としてグ ループ I Dが記憶されるようにしたが、 番組のメタデータに含まれる複数の属性 に基づいて、 より詳細な嗜好情報が生成され、 生成された嗜好情報に基づいて番 組が推薦されるようにすることもできる。 図 2 3を参照して、 番組のメタデータ に含まれる複数の属性に基づいて、 より詳細な嗜好情報を生成する第 1の例であ る嗜好情報抽出処理 1について説明する。 この処理は、 例えば、 予め決められた 時期 (毎週所定の時刻) に、 コンテンツ推薦サーバ 4において実行される。 ステップ S 5 2 1において、 CPU 1 1は、 利用履歴を分析する。 このとき、 図 9のステップ S 2の場合と同様に、 クライアント機器 5から、 所定の期間に利用 されたコンテンツのメタデータ (グループ I Dが設定されている) が取得され、 グループ毎の利用回数 (図 1 0) 、 または利用頻度 (図 1 1 B) が分析される。 ステップ S 5 2 2において、 CPU1 1は、 利用回数 (利用頻度) が所定の閾値 以上であるグループを検索する。 なお、 利用頻度が所定の閾値以上であるグルー プが検索されるようにしてもよい。
ステップ S 5 2 3において、 CPU1 1は、 グループが検索されたか否かを判定 し、 グループが検索されたと判定された場合、 ステップ S 5 2 4に進み、 検索さ れたグループに属する番組のメタデータを分析する。 このとき、 番組が複数あつ た場合、 複数の番組のメタデータが分析される。 ステップ S 5 2 5において、 CPU1 1は、 ステップ S 5 24で分析された番組のメタデータに基づいて、 番組 ベタ トルを生成する。
図 24に、 このとき生成される番組ベク トル P Pの構成例を示す。 この例では、 番組べク トル P Pは、 ステップ S 5 24で分析された番組のメタデータの属性 「タイ トル (番組名) 」 (Tm) 、 「ジャンル」 (Gtn) 、 「出演者」 (Pm) 、 「放送局」 (Sra) 、 「時間帯」 (Hm) 、 · · 'を要素とするベク トル PP= (Tra, Gm, Pm, Sm, Hm, · · · ) として構成されている。 そして要素 Tm, Gm, Pm, Sm, Hm, · ■ . も、 複数の要素を持つベクトルとして構成される。
例えば属性 「放送局」 に対応するべクトル Smは、 MHK総合、 MHK教育、 亜細亜テレビ、 TAS、 フシ、 テレ 3、 東都、 ]^^1 衛星第1、 MHK衛星第 2、 および WOWO (いずれも仮想的な放送局の名称) など、 放送局の種類が限られ ているので、 Sm= {MHK総合, MHK教育, 亜細亜テレビ, TT S, ブジ, テレ日, 東都, ]\ ^11:衛星第1, MHK衛星第 2, WOWO} のように構成し、 対応する放送局を値 1、 その他の放送局を値 0とすることで得られる。 すなわち 対応する番組の放送局が WOWOであるとき、 項目 「放送局」 のベク トル Smは、 Sm= { 0 , 0, 0, 0, 0, 0, 0, 0, 0, 1 } とされる。 属性 「ジャンル J に対応するベクトル Gmも、 ドラマ、 バラエティ、 スポーツ、 映画、 音楽、 子供向け/教育、 教養/ドキュメント、 ニュース Z報道、 およびそ の他など、 その種類が限られているので、 Gm= {ドラマ, バラエティ, スポー ッ, 映画, 音楽, 子供向け/教育, 教養/ドキュメント, ニュース 報道, その 他 } のように構成し、 対応するジャンルを値 1、 その他のジャンルを値 0とする ことで得られる。 すなわち対応する番組のジャンルが教養/ドキュメントである とき、 項目 「ジャンル」 のべク トノレ Gmは、 Gm= { 0 , 0, 0, 0, 0, 0, 1 , 0 , 0 } とされる。
属性 「時間帯」 に対応するべク トル Hmも、 属性 「放送局」 のべクトル Smお よび 「ジャンル」 のベク トル Gmと同様にして得ることができる。 '
一方、 属性 「タイトル」 、 「出演者」 などのように、 要素を限定することが容 易でないものは、 その属性を構成する単語とその頻度を表す数値の組を 1つの要 素とするベタトルがその属性のベタトルとなる。 例えば、 番組のメタデータの属 性 「出演者」 が 「personA, personB, · · ·」 である場合、 属性 「出演者」 に 対応するべク トノレ Pmは、 Pm= { (personA- 1 ) , (personB—
3) , · ■ · } とされる。 ここで、 (personA— 1) と (personB— 3) は、 メ タデータの属性 「出演者」 を構成する単語として、 personAと personBがそれ ぞれ 1回または 3回検出されたことを表す。
なお、 ステップ S 5 2 2において複数の番組が検索された場合、 ステップ S 5 2 5において、 それぞれの番組毎に、 番組ベクトルが生成される。
ステップ S 5 2 6において、 CPU 1 1は、 ステップ S 5 2 5において生成され た番組ベク トルを統合して嗜好情報を生成する。 このとき、 例えば、 複数の番組 ベタトルのそれぞれの属性が足し合わされ、 嗜好情報として生成される。
図 2 5に、 このとき生成される嗜好情報の例を示す。 この例では、 嗜好情報は、 ベクトルとして生成されており、 属性 「番組名 (タイ トル) 」 (Tup) 、 「ジャ ンル」 (Gup) 、 「出演者」 (Pup) 、 「放送局」 (Sup) 、 「時間帯」
(Hup) 、 · · 'に対応するベクトル UP= (Tup, Gup, Pup, Sup, H 3015927
23
up, · · · ) として構成されている。 そして要素 Tup, Gup, Pup, Sup, H up, · · · も、 複数の要素を持つベタトルとして構成される。
この例では、 属性 「番組名 (タイトル) 」 に対応するべクトル Tupは、 Tup = { (title 1 - 1 2) , (title2— 3) , · ■ ■ } とされている。 これは、 嗜好情報の属性 「番組名」 に要素 「title l」 と 「title2」 があり、 それぞれ の重要度が 「1 2」 と 「3」 に設定されていたことを表す。
重要度は、 その要素に対するユーザの嗜好の度合いをあらわすもので、 同一の 要素が含む番組ベクトルが足し合わされるとき、 重要度が 1だけ加算される。 例 えば、 番組ベク トル P P 1乃至 P P 2 0の 20個の番組ベクトルに基づいて、 嗜 好情報が生成される場合、 番組ベク トル P P 5、 P P 1 0、 および P P 1 7の 3 つの番組ベクトルにおいて、 属性 Tmの中に 「title2」 が含まれていた場合、 Tupの要素 「title2」 の重要度が 「3」 と設定される。
また、 属性 「ジャンル」 に対応するべク トノレ Gupは、 Gup= { (ドラマー 2 5) , (バラエティー 3 4) , (スポーツ一 4 2) , (映画一 3 7) , (音楽一 7 3) , (子供向け/教育一 1 2 0) , (教養 Zドキュメント一 3) , (ニュー ス /報道一 5) , (その他一 2 3) } とされており、 属性 「ジャンル」 に含まれ る要素とその重要度により構成されている。
同様にして、 嗜好情報の属性 「放送局」 に対応するべク トル Sup、 属性 「出 演者」 に対応するベクトル Pup、 ' , , 、 各属性を構成する要素と重要度によ り構成されている。
ステップ S 5 2 3において、 視聴回数が閾値以上のグループが検索されなかつ たと判定された場合、 ステップ S 5 24乃至 S 5 2 6の処理はスキップされ、 処 理は終了する。
このようにして嗜好情報が生成される。 嗜好情報は、 所定の回数または頻度だ け利用された番組のメタデータに基づいて、 生成されるのでよりユーザの嗜好を 適確に反映したものとすることができる。 なお、 嗜好情報は、 ステップ S 5 2 1で、 特定のユーザの利用履歴を分析する ことにより、 ユーザ単位に生成されるものとしてもよいし、 ステップ S 5 2 1で 複数のユーザの利用履歴を分析することにより、 一般的な (複数のユーザに共通 の) 嗜好情報が生成されるものとしてもよい。
ところで、 図 2 3を参照して上述した嗜好情報抽出処理 1によれば、 同一の要 素が含む番組ベクトルが足し合わされる都度、 重要度が加算されるので、 ユーザ が頻繁に視聴する番組のメタデータに含まれる要素の重要度が極端に高くなり、 偏った嗜好情報となってしまう場合もある。 たとえば、 毎日 (月曜日から金曜日 まで) 放送される番組をユーザが欠かさず視聴している場合、 その番組のメタデ ータに含まれる要素 (例えば、 タレント A ) の重要度が、 他の要素と比較して極 端に高くなつてしまう。 このような、 頻繁に視聴される番組 (いわゆる定番の番 組) のメタデータを嗜好情報に反映させないようにすることも可能である。 図 2 6を参照して、 番組のメタデータに含まれる複数の属性に基づいて、 嗜好情報を 生成する第 2の例である嗜好情報抽出処理 2について説明する。
ステップ S 5 4 1乃至 5 4 3の処理は、 図 2 3のステップ S 5 2 1乃至 S 5 2 3の処理と同様の処理なので、 その説明は省略する。 ステップ S 5 4 4において、 CPU 1 1は、 ステップ S 5 4 2で検索されたグループに属する番組が定番番組か 否かを判定する。 ここで、 定番番組か否かの判定は、 図 2 2を参照して上述した 定番番組設定処理により設定された定番フラグに基づいて判定される。
ステップ S 5 4 4において、 検索された番組が定番番組ではないと判定された 場合、 ステップ S 5 4 5に進む。 そして、 図 2 3のステップ S 5 2 4乃至 S 5 2 5の処理と同様にして、 ステップ S 5 4 5において、 番組のメタデータが分析さ れ、 ステップ S 5 4 6において番組ベクトルが生成され、 ステップ S 5 4 7にお いて嗜好情報が生成される。
一方、 ステップ S 5 4 4において、 検索された番組が定番番組であると判定さ れた場合、 ステップ S 5 4 5乃至 S 5 4 7の処理はスキップされる。 このようにすることで、 定番番組に基づいて、 嗜好情報が生成されることがな くなり、 偏った嗜好情報が生成されることを防止することができる。
また、 図 2 3を参照して上述した処理によれば、 所定の回数 (または頻度) 以 上に視聴されているグループに属する番組について、 すべて同様に番組べク トル が生成され、 嗜好情報が生成されるようにしている。 この例では、 シリーズで放 送される番組 A 1 , A 2, A 3 , ■ ■ · (以下、 個々に区別する必要がない場合、 これらをまとめたシリーズ番組 Aと称する。 他の場合も同様である) と番組 B 1 , B 2 , B 3 , · ■ · (それぞれ 1つのグループに属する番組) があり、 例えば各 グループの閾値が 3回の場合、 3回視聴した番組 A (正確には、 シリーズ化され た 3個の番組が視聴されたシリーズ) についても、 1 0回視聴した番組 B (シリ ーズ化された 1 0個の番組が視聴されたシリーズ) についても同様に番組べクト ルが生成される。
し力、し、 シリーズ番組 Aとシリーズ番組 Bでは、 ユーザの、 その番組に対する 知識が異なる場合がある。 例えば、 ユーザは、 1 0回視聴したシリーズ番組 Bに ついては、 番 ¾aの中にどのようなコーナーがあり、 どのようなタレントが出演す るかを熟知している可能性が高いが、 3回しか視聴していないシリーズ番組 Aに ついては、 番組のコーナー、 タレントなどについて熟知していない可能性があり、 場合によっては、 シリーズ番組 Aを視聴したくないと感じる可能性もある。 そこ で、 その番組に対する熟知度を反映して嗜好情報を生成する必要がある。 図 2 6 を参照して、 番組のメタデータに含まれる複数の属性に基づいて、 嗜好情報を生 成する第 3の例である嗜好情報抽出処理 3について説明する。
ステップ S 5 6 1乃至 S 5 6 5の処理は、 図 2 3の S 5 2 1乃至 S 5 2 5の処 理と同様の処理なので、 その説明は省略する。 ステップ S 5 6 6において CPU 1 1は、 番組の熟知度を特定する。 熟知度は、 ステップ S 5 6 1において分析され た、 シリーズ番組 (すなわちグループ) の利用頻度に基づいて特定される。 この とき、 そのシリーズ番組の利用頻度に対応して 3段階の熟知度が設定される。 例 えば、 利用頻度が 「0 . 1」 以上のものは熟知度が 「高」 と設定され、 利用頻度 力 S 「0 . 0 5」 以上 「0 . 1」 未満のものは熟知度が 「中」 と設定され、 利用頻 度が 「0 . 0 5」 未満のものは熟知度が 「低」 と設定される。
なお、 熟知度の分類は、 3段階に限られるものではない。 また、 熟知度が段階 別に分類されずに、 数値として設定されるようにしてもよい。 あるいはまた、 利 用頻度ではなく、 利用回数に基づいて、 熟知度が設定されるようにしてもよい。 ステップ S 5 6 7において、 CPU 1 1は、 ステップ S 5 6 5で生成された番糸且 ベタ トルに対して、 熟知度に基づいて、 重み付けを行う。 このとき、 例えば、 熟 知度が 「高」 の番組べク トルに含まれる要素に基づいて生成される、 嗜好情報の 重要度は 3倍に設定され、 熟知度が 「中」 の番組ベク トルに含まれる要素に基づ いて生成される、 嗜好情報の重要度は 2倍に設定され、 熟知度が 「低」 の番組べ クトルに含まれる要素に基づいて生成される、 嗜好情報の重要度は 1倍に設定さ れる。
ステップ S 5 6 8において、 CPU 1 1は、 ステップ S 5 6 7において重み付け された番組べク トルに基づいて、 嗜好情報を生成する。 このとき、 例えば、 熟知 度が 「高」 である番組べク トル P P 1を構成するべク トル Pm l力 S、 Pml =
(personA) であり、 熟知度が 「中」 である番組ベク トル P P 2を構成するべク トノレ Pm 2力 Pm 2 = (personB) であり、 熟知度が 「低」 である番組ベクトル P P 3を構成するベクトル Pra 3が、 Pm 3 = (personC) である場合、 嗜好情報 の属性 「出演者」 に対応するべク トノレ Pupは、 PuP = { (personA - 3 ) ,
(personB- 2 ) , (personC- 1 ) }とされる。
このようにして、 熟知度を反映した嗜好情報が生成される。 なお、 嗜好情報は、 ステップ S 5 6 1で、 特定のユーザの利用履歴を分析することにより、 ユーザ単 位に生成されるものとしてもよいし、 ステップ S 5 6 1で複数のユーザの利用履 歴を分析することにより、 一般的な (複数のユーザに共通の) 嗜好情報が生成さ れるものとしてもよい。 例えば、 まだ視聴履歴が蓄積されていないユーザに対し ては、 一般的な嗜好情報に基づいて番組 (コンテンツ) を推薦することができる。 嗜好情報は、 番組の熟知度を反映して生成されているので、 例えば、 単に視聴率 の高い番組を推薦するより、 信頼性の高い番組を、 ユーザに推薦することができ る。
以上においては、 嗜好情報の重要度は、 番組が視聴される毎に加算されるよう にしているが、 場合によっては重要度を減算する必要もある。 例えば、 ユーザは クライアント機器 5において、 自動録画が予約された定番番組について録画予約 を解除することができる。 ここで、 録画予約が解除された番組は、 それ以前は、 頻繁に視聴されていたにもかかわらず、 その回だけ録画予約をあえて解除したも のであり、 録画予約が解除された回は、 ユーザの嗜好に合わない内容であったこ とが想定される。 そこで、 本発明においては、 この録画予約が解除された番組の メタデータに基づいて、 ユーザの嗜好情報の変更が行われる。
図 2 8を参照して、 嗜好情報変更処理について説明する。 この処理は、 クライ アント機器 5の CPU 5 1により、 自動録画の予約の解除が検知されたとき、 コン テンッ推薦サーバ 4に対して、 自動録画予約が解除された番組の情報がネットヮ ーク 6を介して通知されたとき、 コンテンツ推薦サーバ 4により実行される。 ステップ S 5 8 1において、 CPU 1 1は、 自動録画予約が解除された回の番組 (例えば、 1 0回放送されるシリーズ番組のうち、 第 3回目の番組) のメタデー タを取得し、 ステップ S 5 8 2において、 取得したメタデータの属性を分析する。 ステップ S 5 8 3において、 CPU 1 1は、 自動録画予約が設定された番組の嗜好 情報の属性と、 自動録画予約が解除された回の番組のメタデータの属性を比較し、 ステップ S 5 8 4においてネガティブな要素を検出する。
例えば、 番組 Xの自動録画予約が設定され、 ユーザによりその録画予約が解除 された場合を考える。 自動録画予約が設定された回の番組 Xのメタデータに基づ いて、 生成された番組ベクトル P P 1において、 属性 「出演者」 に対応するべク トル Pm l力 S、 Pml = (personA, personB) であり、 自動録画予約が解除された 回の番組 Xのメタデータに基づいて、 生成された番糸且ベク トル P P 2において、 属性 「出演者」 に対応するべクトル Pm 2が、 Pm 2 = (personA, personB, personC) であった場合、 自動録画予約が解除された回の番組 Xには 「personC」 が出演していたため、 録画予約が解除されたと考えられ、 ステップ S 5 8 4において、 「personC」 がネガティブな要素として検出される。
ステップ S 5 8 5において、 CPU 1 1は、 ステップ S 5 8 4で検出されたネガ ティブな要素に基づいて、 ユーザの嗜好情報を変更する。 このとき、 ネガティブ な要素の重要度が減算される。 いまの場合、 例えば、 嗜好情報の属性 「出演者」 に対応するべク トノレ Pup力、 Pup = { (personA— 5 ) , (personB— 2 ) ,
(personC- 3 ) }であった場合、 ステップ S 5 8 5において、 Pup= { (personA 一 5 ) , (personB - 2 ) , (personC- 2 ) }と変更され、 「personC」 の重要 度が 1だけ減算される。
このようにして、 嗜好情報の変更が行われる。 このようにすることで、 ユーザ が好まない属性の重要度は低く変更されるので、 ユーザに番組 (コンテンツ) の 推薦を行うとき、 よりユーザの嗜好に合った番組 (コンテンツ) を推薦すること ができる。
また、 以上においては、 視聴回数が所定の回数以上のシリーズ番組のメタデー タに基づいて嗜好情報が生成される例について説明したが、 このようにして生成 された嗜好情報に基づいた番組の推薦が続けて行われると、 ユーザが飽きてしま う恐れがある。 そこで、 本発明においては、 ユーザが初めて視聴した (過去に視 聴していない) 番組に注目する。 初めて視聴した番組に対して、 ユーザは特別な 興味をもっていることが考えられるので、 この番組のメタデータに基づいて、 特 殊嗜好情報を生成する。
図 2 9を参照して、 コンテンツ推薦サーバ 4による特殊嗜好情報生成処理につ いて説明する。 この処理は、 例えば、 ユーザにより、 所定のコマンドを'投入した とき実行されるようにしてもよいし、 所定の周期 (例えば、 1週間) 毎に自動的 に実行されるようにしてもよレ、。
ステップ S 6 0 1において、 CPU 1 1は利用履歴を検索する。 このとき、 クラ イアント機器 5から、 所定の期間 (例えば、 直近の 6ヶ月間) に利用されたコン テンッのメタデータ (グループ I Dが設定されている) が取得され、 グループ毎 の利用回数 (図 1 0 ) が分析される。
ステップ S 6 0 2において、 CPU 1 1は、 視聴回数が 1回のシリーズ番組 (そ のグループに属する番組の中の 1番組だけが視聴されているグループ) を検出す る。 ステップ S 6 0 3において、 CPU 1 1は、 視聴回数が 1回のシリーズ番組が 検出されたか否かを判定し、 シリーズ番組が検出されたと判定された場合、 ステ ップ S 6 0 4に進み、 検出されたシリーズ番組に属する番組のメタデータに基づ いて、 特殊嗜好情報を生成する。 このとき、 図 2 3のステップ S 5 2 4乃至 S 5 2 6と同様に、 番組のメタデータに基づいて番組べク トルが生成され、 番組べク トルに基づいて特殊嗜好情報が生成される。 ステップ S 6 0 3において、 視聴回 数が 1回の番組が検出されなかったと判定された場合、 ステップ S 6 0 4の処理 はスキップされる。
このようにして、 ユーザが初めて視聴した番組のメタデータに基づいて特殊嗜 好情報が生成される。
次に、 図 2 3、 図 2 6、 およぴ図 2 7を参照して上述した処理により生成した 嗜好情報に基づいて、 コンテンツが推薦される処理について説明する。
図 3 0は、 図 2 3、 図 2 6、 および図 2 7を参照して上述した処理により生成 した嗜好情報に基づいて、 コンテンツを推薦する場合の、 コンテンツ推薦サーバ 4の CPU 1 1の機能的構成例を示すプロック図である。 この例では、 番組のメタ データを取得するメタデータ取得部 1 1 1、 特定のユーザの嗜好情報を取得する 嗜好情報取得部 1 1 2が設けられている。
メタデータ取得部 1 1 1により取得された番組のメタデータは、 番組べク トル 抽出部 1 1 3に出力され、 番組べクトル抽出部 1 1 3において、 番組べク トルが 抽出される。 また、 嗜好情報取得部 1 1 2により取得された嗜好情報は、 嗜好べ クトル抽出部 1 1 4に出力され、 嗜好情報に基づく嗜好べクトルが抽出される。 番組べクトル抽出部 1 1 3により抽出された番組べク トルと、 嗜好べク トル抽出 部 1 1 4により抽出された嗜好べクトルは、 マッチング処理部 1 1 5に出力され、 マッチング処理部 1 1 5は、 番組べク トルと嗜好べク トルの類似度を算出する。
1つの嗜好べク トルに対して複数の番組べクトルとの類似度が算出され、 マツ チング処理部 1 1 5は、 類似度が高い順に所定の数の番組べク トルを選択し、 選 択された番組べクトルに対応する番組のメタデータを情報出力部 1 1 6に出力す る。
情報出力部 1 1 6は、 マッチング処理部 1 1 5により選択された番組のメタデ ータを、 例えば、 記憶部 1 8に記憶させる。
次に、 図 3 1のフローチャートを参照して、 推薦情報検索処理について説明す る。 ステップ S 6 2 1において、 メタデータ取得部 1 1 1は、 コンテンツ (番 組) のメタデータを取得する。 このとき、 所定の基準に基づいて、 複数の番組 (例えば、 今後 1週間に放送される番組) のメタデータが取得される。 ステップ S 6 2 2において、 番組べク トル抽出部 1 1 3は、 ステップ S 6 2 1で取得され た番組のメタデータに基づいて、 番組ベク トルを抽出する。 このとき、 図 2 4を 参照して上述した番組ベクトルと同様に、 複数の番組の番組ベクトルが抽出され る。
ステップ S 6 2 3において、 嗜好べクトル抽出部 1 1 4は、 嗜好情報を取得す る。 このとき、 特定のユーザの嗜好情報が取得される。 ステップ S 6 2 4におい て、 嗜好べクトル抽出部 1 1 4は、 嗜好べクトルを生成する。 嗜好べク トルは、 図 2 5に示されるような嗜好情報が、 そのまま嗜好ベク トルとして生成されるよ うにしてもよいし、 嗜好情報を構成する特定の属性が抽出されて嗜好べク トルと して生成されるようにしてもよい。
ステップ S 6 2 5において、 マッチング処理部 1 1 5は、 ステップ S 6 2 2に おいて生成された番組べクトル P Pと、 ステップ S 6 2 4において生成された嗜 好べク トル U Pのコサイン距離を算出する。 いま、 べクトノレ P Pとべク トル U P のなす角を Θ とすると、 cos 0 = P P ■ U P / I P P I I U P I となる。 例えば、 嗜好ベクトル UP (= (Tup, Gup, Pup, Sup, H
up, · · · ) ) におけるべク トノレ Pup力 S、 Pup= { (personA— 1 ) , (person B - 1 ) , (personC- 1 ) } であり、 番組ベク トル P P= (Tra, Gra, Pra, S m, Hm, · ■ ■ ) におけるべクトノレ Pmが、 Pm= { (personA— 1 ) , (personD- 1 ) , (person E- 1 ) } である場合、 コサイン距離 cos 0 pは、 式 (1) に従って算出される。
cos θρ= (1 . 1) / ( "3 xV~3) = 1/3 · · · (1) なお、 式中 " · " は、 内積を表し、 " X " はスカラ演算を表す。
cos θρと同様にして、 ベタ トノレ Tup, Gup, Sup, Hup, · · ■ と、 べク ト /レ Tm, Gm, Sm, Hm, · · · とのコサイン g巨離 cos 0 t, cos Θ g, cos Θ s,
COS 0 h, · · ·が算出される。
そして、 算出されたコサイン距離を式 (2) に従って合計し、 類似度 Simを 計算する。
Sim=cos Θ t + cos 0 g卞 cos 0 p + cos 0 s + cos 0 h - · ■ {2 j
このようにして、 嗜好ベク トル UPと番組ベク トル P Pの類似度が算出される。 なお、 1つの嗜好べクトル UPに対して複数の番組べクトル P Pとの類似度が算 出される。 これにより、 それぞれの番組のメタデータについて、 ユーザの嗜好情 報との類似度が算出される。
ステップ S 6 26において、 マッチング処理部 1 1 5は、 類似度の高い番組の メタデータを選択し、 情報出力部 1 1 6に出力する。 このとき、 ステップ S 6 2 5で算出された類似度に基づいて、 類似度の高い順、 すなわち Simの値が大き い順に所定の数 (例えば、 10) だけ番組ベク トル P Pが選択され、 選択された 番組ベク トル P Pに対応する番組のメタデータが出力される。 なお、 類似度が所 定の値より大きい番組べクトル P Pが全て選択され、 選択された番組べクトル P Pに対応する番組のメタデータが出力されるようにしてもよい。
ステップ S 6 27において、 情報出力部 1 1 6は、 ステップ S 6 26で抽出さ れた番組のコンテンッ推薦情報をクライアント機器 5に送信する。 このようにして、 嗜好情報に基づく番組の推薦が行われる。
ところで、 番組の推薦は、 図 2 9を参照して説明した処理で生成された特殊嗜 好情報に基づいて行うこともできる。 図 3 2を参照して、 コンテンツ推薦サーバ 4による特殊推薦情報検索処理について説明する。 この処理は、 例えば、 ユーザ により、 所定のコマンドを投入したとき実行されるようにしてもよいし、 所定の 周期 (例えば、 1週間) 毎に自動的に実行されるようにしてもよい。
ステップ S 6 4 1と S 6 4 2の処理は、 図 3 1のステップ S 6 2 1と S 6 2 2 の処理と同様の処理であるので、 その説明は省略する。
ステップ S 6 4 3において、 嗜好べクトル抽出部 1 1 4は、 特殊嗜好情報を取 得する。 このとき、 図 2 9を参照して上述した特殊嗜好情報生成処理により生成 された特殊嗜好情報が取得される。 そして、 ステップ S 6 4 4において、 嗜好べ クトル抽出部 1 1 4は、 ステップ S 6 4 3で取得された特殊嗜好情報に基づいて、 嗜好べクトルを生成する。
ステップ S 6 4 5と S S 6 4 6の処理は、 図 2 3のステップ S 6 2 5と S 6 2 6の処理と同様の処理であるので、 その説明は省略する。
ステップ S 6 2 7において、 情報出力部 1 1 6は、 ステップ S 6 4 6において 抽出された番組のコンテンッ推薦情報をクライアント機器 5に送信する。
このようにして、 特殊嗜好情報に基づいて、 コンテンツの推薦が行われる。 上 述したように、 特殊嗜好情報は、 ユーザが初めて視聴した番組のメタデータに基 づいて生成されたものであり、 特殊嗜好情報に基づいて、 ユーザにコンテンツを 推薦することにより、 より新鮮な印象を与えることができる。
上述した一連の処理は、 ソフトウェアにより実行することもできる。 そのソフ トウエアは、 そのソフトウェアを構成するプログラムが、 専用のハードウェアに 組み込まれているコンピュータ、 または、 各種のプログラムをインストールする とで、 各種の機能を実行することが可能な、 例えば汎用のパーソナルコンビュ ータなどに、 記録媒体からインストールされる。 この記録媒体は、 図 7, 8に示すように、 ユーザにプログラムを提供するため に配布される、 プログラムが記録されている磁気ディスク 3 1または 7 1 (フレ キシブルディスクを含む) 、 光ディスク 3 2または 7 2 (CD-ROM (Compact Di sk-Read Only Memory) , DVD (Di gital Versat i l e Di sk) を含む) 、 光 磁気ディスク 3 3または 7 3 (MD (Mi ni -Di sk) (商標) を含む) 、 もしくは 半導体メモリ 3 4または 7 4などよりなるパッケージメディアなどにより構成さ れる。
また、 本明細書において、 記録媒体に記録されるプログラムを記述するステツ プは、 記載された順序に沿って時系列的に行われる処理はもちろん、 必ずしも時 系列的に処理されなくとも、 並列的あるいは個別に実行される処理をも含むもの である。
なお、 本明細書において、 システムとは、 複数の装置により構成される装置全 体を表すものである。 産業上の利用可能性
本発明によれば、 コンテンツの属性を表す項目から生成されたグループ化項目 におけるグループ毎の利用頻度から、 コンテンツ推薦を行うことができる。

Claims

請求の範囲
1 . 配信されるコンテンッの属性を表す属性項目の中の 1個以上の属性項目か らなるグループ化項目が一定以上の類似度をもつて類似するコンテンツに、 同一 のグループ IDを付与し、 コンテンツのグループ化を行うグループ化手段と、 前記グループ ID毎にコンテンツの利用頻度を算出する算出手段と、
前記算出手段により算出された前記利用頻度に基づいて、 ユーザの嗜好を表す ユーザ嗜好情報を生成する生成手段と、
前記生成手段により生成された前記ユーザ嗜好情報に基づいて、 コンテンツを 推薦する推薦手段と
を備えることを特徴とする情報処理装置。
2 . 放送時間帯を表す属性項目と、 少なくとも 1つ以上の他の属性項目からなる グループ化項目が設定されており、
前記グループ化手段は、 そのグループ化項目に基づいてコンテンツのグループ 化を行う
ことを特徴とする請求の範囲第 1項に記載の情報処理装置。
3 . 少なくとも放送時間帯を表す属性項目からなる前記グループ化項目と、 他 の属性項目からなる前記グループ化項目が設定されており、
前記グループ化手段は、 それらのグループ化項目に基づいてコンテンツのダル 一プ化を行う
ことを特徴とする請求の範囲第 1項に記載の情報処理装置。
4 . 前記グループ化手段は、 コンテンツの前記属性項目の内容を形態素解析し、 その結果に基づいて、 前記グループ化項目の内容の類似度を決定する
ことを特徴とする請求の範囲第 1項に記載の情報処理装置。
5 . 前記生成手段は、 グループに属するコンテンツの利用状態が所定の条件を 満たしていないグループの利用頻度を、 前記ユーザ嗜好情報の生成に利用しない ことを特徴とする請求の範囲第 1項に記載の情報処理装置。
6 . 前記推薦手段は、 前記算出手段により算出された前記利用頻度が、 予め設定された値より高いか 否かを判定する判定手段と、
前記判定手段により、 前記利用頻度が、 予め設定された値より高いと判定され た場合、 前記コンテンツの推薦情報に、 頻繁に視聴されるコンテンツであること を表す定番フラグを設定する設定手段と
を備えることを特徴とする請求の範囲第 1項に記載の情報処理装置。
7 . 前記生成手段は、
前記算出手段により算出された前記利用頻度が予め設定された値より高いダル ープのコンテンツのメタデータを取得し、 前記メタデータの特徴量を表すべク ト ルを抽出する抽出手段を備え、
前記抽出手段により抽出されたべク トルに基づいて、 前記嗜好情報を生成する ことを特徴とする請求の範囲第 1項に記載の情報処理装置。
8 . 前記生成手段は、
前記算出手段により算出された前記利用頻度が予め設定された値より高いグル ープのコンテンツが、 前記定番フラグが設定されたコンテンツ推薦情報に対応す るコンテンツか否かを判定する定番判定手段を備え、
前記定番判定手段により、 前記コンテンツが、 前記定番フラグが設定されたコ ンテンッ推薦情報に対応するコンテンツではないと判定された場合、 前記抽出手 段は、 コンテンツのメタデータを取得し、 前記メタデータの特徴量を表すべタト ルを抽出する
ことを特徴とする請求の範囲第 7項に記載の情報処理装置。
9 . 前記嗜好情報は、 複数の属性とその属性の重要度を表す値により構成され る
ことを特徴とする請求の範囲第 7項に記載の情報処理装置。
1 0 . 前記生成手段は、
前記算出手段により算出された前記利用頻度に基づいて、 前記コンテンツの熟 知度を設定する熟知度設定手段を備え、 前記熟知度に基づいて、 前記嗜好情報の重要度を表す値に重み付けを行う ことを特徴とする請求の範囲第 7項に記載の情報処理装置。
1 1 . 前記生成手段は、
前記コンテンツの利用履歴に基づいて、 利用回数が所定の値以下だけ利用され たコンテンツを検索する検索手段と、
前記検索手段により検索されたコンテンツのメタデータに基づいて、 特殊嗜好 情報を生成する特殊嗜好情報生成手段とをさらに備える
ことを特徴とする請求の範囲第 7項に記載の情報処理装置。
1 2 . 前記嗜好情報または前記特殊嗜好情報の特徴量を表すべク トルを抽出す る第 1の抽出手段と、
予め設定された期間に放送されるコンテンツのメタデータを取得し、 前記メタ データの特徴量を表すべクトルを抽出する第 2の抽出手段と、
前記第 1の抽出手段により抽出されたべク トルと第 2の抽出手段により抽出さ れたべクトルの類似度を演算する演算手段とを備え、
前記推薦手段は、 前記類似度が高い順に、 予め設定された数だけ前記第 2の抽 出手段により抽出されたべク トノレを選択し、 選択されたべクトルのメタデータに 基づいて、 コンテンツを推薦する
ことを特徴とする請求の範囲第 1 1項に記載の情報処理装置。
1 3 . 配信されるコンテンツの属性を表す属性項目の中の 1個以上の属性項目 からなるグループ化項目が一定以上の類似度をもって類似するコンテンツに、 同 一のグループ IDを付与し、 コンテンツのグループ化を行うグループ化ステップ と、
前記グループ ID毎にコンテンツの利用頻度を算出する算出ステップと、 前記算出ステップの処理で算出された前記利用頻度に基づいて、 ユーザの嗜好 を表すユーザ嗜好情報を生成する生成ステップと、
前記生成ステップの処理で生成された前記ユーザ嗜好情報に基づいて、 コンテ ンッを推薦する推薦ステップと を含むことを特徴とする情報処理方法。
1 4 . 配信されるコンテンツの属性を表す属性項目の中の 1個以上の属性項目 からなるグループ化項目が一定以上の類似度をもつて類似するコンテンツに、 同 一のグループ IDを付与することによっての、 コンテンツのグループ化を制御す るグループ化制御ステップと、
前記グループ ID毎のコンテンツの利用頻度の算出を制御する算出制御ステツ プと、
前記算出制御ステップの処理で算出された前記利用頻度に基づいての、 ユーザ の嗜好を表すユーザ嗜好情報の生成を制御する生成制御ステップと、
前記生成制御ステップの処理で生成された前記ユーザ嗜好情報に基づいての、 コンテンツの推薦を制御する推薦制御ステップと
を含むことを特徴とするコンピュータが読み取り可能なプログラムが記録され ている記録媒体。
1 5 . 配信されるコンテンツの属性を表す属性項目の中の 1個以上の属性項目 からなるグループ化項目が一定以上の類似度をもって類似するコンテンツに、 同 一のグループ I Dを付与することによっての、 コンテンツのグループ化を制御す るグループ化制御ステツプと、
前記グループ ID毎のコンテンッの利用頻度の算出を制御する算出制御ステツ プと、
前記算出制御ステップの処理で算出された前記利用頻度に基づいての、 ユーザ の嗜好を表すユーザ嗜好情報の生成を制御する生成制御ステップと、
前記生成制御ステップの処理で生成された前記ユーザ嗜好情報に基づいての、 コンテンツの推薦を制御する推薦制御ステップと
を含む処理をコンピュータに実行させることを特徴とするプログラム。
PCT/JP2003/015927 2002-12-12 2003-12-12 情報処理装置および方法、記録媒体、並びにプログラム Ceased WO2004053736A1 (ja)

Priority Applications (2)

Application Number Priority Date Filing Date Title
EP03778860A EP1571561A4 (en) 2002-12-12 2003-12-12 DATA PROCESSING DEVICE AND METHOD, RECORDING MEDIUM AND PROGRAM
US10/538,658 US7873798B2 (en) 2002-12-12 2003-12-12 Information processing device and method, recording medium, and program

Applications Claiming Priority (4)

Application Number Priority Date Filing Date Title
JP2002361540 2002-12-12
JP2002-361540 2002-12-12
JP2003-285030 2003-08-01
JP2003285030A JP2004206679A (ja) 2002-12-12 2003-08-01 情報処理装置および方法、記録媒体、並びにプログラム

Publications (1)

Publication Number Publication Date
WO2004053736A1 true WO2004053736A1 (ja) 2004-06-24

Family

ID=32510667

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/JP2003/015927 Ceased WO2004053736A1 (ja) 2002-12-12 2003-12-12 情報処理装置および方法、記録媒体、並びにプログラム

Country Status (5)

Country Link
US (1) US7873798B2 (ja)
EP (1) EP1571561A4 (ja)
JP (1) JP2004206679A (ja)
KR (1) KR101073948B1 (ja)
WO (1) WO2004053736A1 (ja)

Cited By (2)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
EP1646233A2 (en) * 2004-10-11 2006-04-12 Topfield Co., Ltd. Apparatus and method for performing reserved recording function
US20070201822A1 (en) * 2005-02-07 2007-08-30 Yoshiaki Kusunoki Recommended Program Extracting Apparatus And Recommended Program Extracting Method

Families Citing this family (52)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR100745995B1 (ko) * 2003-06-04 2007-08-06 삼성전자주식회사 메타 데이터 관리 장치 및 방법
US8943537B2 (en) * 2004-07-21 2015-01-27 Cox Communications, Inc. Method and system for presenting personalized television program recommendation to viewers
JP4617805B2 (ja) * 2004-09-28 2011-01-26 ソニー株式会社 情報配信システムおよび情報配信方法、情報処理装置および情報処理方法、受信装置および受信方法、並びにプログラム
WO2006043498A1 (ja) * 2004-10-18 2006-04-27 Pioneer Corporation 情報処理装置、統計情報データベースのデータ構造、情報生成装置、情報処理方法、情報生成方法、情報処理プログラム、およびそのプログラムを記録した記録媒体
JP4529632B2 (ja) * 2004-10-19 2010-08-25 ソニー株式会社 コンテンツ処理方法およびコンテンツ処理装置
KR100677601B1 (ko) * 2004-11-11 2007-02-02 삼성전자주식회사 메타 데이터를 포함하는 영상 데이터를 기록한 저장매체,그 재생장치 및 메타 데이터를 이용한 검색방법
US7707209B2 (en) 2004-11-25 2010-04-27 Kabushiki Kaisha Square Enix Retrieval method for contents to be selection candidates for user
WO2006064877A1 (ja) * 2004-12-17 2006-06-22 Matsushita Electric Industrial Co., Ltd. コンテンツ推薦装置
JP4655200B2 (ja) * 2005-02-01 2011-03-23 ソニー株式会社 情報処理装置および方法、並びにプログラム
JP4752623B2 (ja) * 2005-06-16 2011-08-17 ソニー株式会社 情報処理装置、情報処理方法、およびプログラム
JP2007104312A (ja) * 2005-10-04 2007-04-19 Toshiba Corp 電子ガイド情報を用いた情報処理方法およびその装置
JP2007281602A (ja) * 2006-04-03 2007-10-25 Canon Inc 受信装置及び番組予約方法
KR100714727B1 (ko) 2006-04-27 2007-05-04 삼성전자주식회사 메타 데이터를 이용한 미디어 컨텐츠의 탐색 장치 및 방법
KR20090015063A (ko) * 2006-04-27 2009-02-11 교세라 가부시키가이샤 휴대 전화 단말, 서버 및 그룹 통화 시스템
JP4847797B2 (ja) * 2006-06-09 2011-12-28 ヤフー株式会社 付加情報データを送信する方法、サーバおよびプログラム
JP2008219342A (ja) * 2007-03-02 2008-09-18 Sony Corp 情報処理装置および方法、並びにプログラム
JP4721066B2 (ja) * 2007-03-16 2011-07-13 ソニー株式会社 情報処理装置、情報処理方法、およびプログラム
US9179086B2 (en) * 2007-05-07 2015-11-03 Yahoo! Inc. System and method for providing dynamically updating applications in a television display environment
JP4717871B2 (ja) 2007-11-06 2011-07-06 シャープ株式会社 コンテンツ視聴装置及びコンテンツ推薦方法
WO2009076852A1 (zh) * 2007-12-03 2009-06-25 Huawei Technologies Co., Ltd. 对用户进行分类的方法、行为采集分析方法与装置
JP4568323B2 (ja) * 2007-12-07 2010-10-27 富士通株式会社 放送番組の記録装置
CN101472117A (zh) * 2007-12-25 2009-07-01 深圳Tcl新技术有限公司 对节目进行选择录制的装置与方法
JP4600521B2 (ja) 2008-06-03 2010-12-15 ソニー株式会社 情報処理装置、情報処理方法、プログラム
JP4596044B2 (ja) 2008-06-03 2010-12-08 ソニー株式会社 情報処理システム、情報処理方法
JP4596043B2 (ja) 2008-06-03 2010-12-08 ソニー株式会社 情報処理装置、情報処理方法、プログラム
JP5387860B2 (ja) * 2008-06-26 2014-01-15 日本電気株式会社 コンテンツ話題性判定システム、その方法及びプログラム
US8037011B2 (en) * 2008-09-15 2011-10-11 Motorola Mobility, Inc. Method and apparatus for recommending content items
JP4650552B2 (ja) * 2008-10-14 2011-03-16 ソニー株式会社 電子機器、コンテンツ推薦方法及びプログラム
KR20110091382A (ko) * 2010-02-05 2011-08-11 삼성전자주식회사 방송 수신 장치, 방송 프로그램 선택 방법 및 그 저장 매체
US10805102B2 (en) 2010-05-21 2020-10-13 Comcast Cable Communications, Llc Content recommendation system
WO2011152072A1 (ja) * 2010-06-04 2011-12-08 パナソニック株式会社 コンテンツ出力装置、コンテンツ出力方法、プログラム、プログラム記録媒体及びコンテンツ出力集積回路
JP5578040B2 (ja) * 2010-11-15 2014-08-27 ソニー株式会社 情報処理装置および方法、情報処理システム、並びに、プログラム
JP5854286B2 (ja) * 2010-11-15 2016-02-09 日本電気株式会社 行動情報収集装置及び行動情報送信装置
US20120272156A1 (en) * 2011-04-22 2012-10-25 Kerger Kameron N Leveraging context to present content on a communication device
US8620917B2 (en) * 2011-12-22 2013-12-31 Telefonaktiebolaget L M Ericsson (Publ) Symantic framework for dynamically creating a program guide
JP5846898B2 (ja) * 2011-12-22 2016-01-20 ニフティ株式会社 情報処理装置、情報処理方法、情報処理システム、及び、プログラム
JP6105555B2 (ja) * 2012-03-30 2017-03-29 ソニー株式会社 制御装置、制御方法、プログラムおよび制御システム
FR3006542A1 (fr) * 2013-05-30 2014-12-05 France Telecom Programmation d'enregistrement de contenus audiovisuels presents dans une grille de programmes electronique
US9916362B2 (en) * 2013-11-20 2018-03-13 Toyota Jidosha Kabushiki Kaisha Content recommendation based on efficacy models
JP6185379B2 (ja) * 2013-12-02 2017-08-23 株式会社Nttドコモ レコメンド装置およびレコメンド方法
JP6147682B2 (ja) * 2014-02-14 2017-06-14 株式会社Nttドコモ レコメンド装置及びレコメンド方法
US11455086B2 (en) 2014-04-14 2022-09-27 Comcast Cable Communications, Llc System and method for content selection
US11553251B2 (en) 2014-06-20 2023-01-10 Comcast Cable Communications, Llc Content viewing tracking
US10776414B2 (en) * 2014-06-20 2020-09-15 Comcast Cable Communications, Llc Dynamic content recommendations
US10051036B1 (en) * 2015-01-23 2018-08-14 Clarifai, Inc. Intelligent routing of media items
US10362978B2 (en) 2015-08-28 2019-07-30 Comcast Cable Communications, Llc Computational model for mood
CN105787759A (zh) * 2016-02-23 2016-07-20 北京金山安全软件有限公司 一种获取用户属性的方法、装置及电子设备
WO2017174833A1 (es) * 2016-04-06 2017-10-12 Telefónica Digital España, S.L.U. Método y sistema para descubrimiento de contenido multimedia para dispositivos capaces de mostrar contenido multimedia
GB2584251B (en) * 2018-01-11 2022-11-16 Editorji Tech Private Limited Method and system for customized content
JP6902636B1 (ja) * 2020-01-28 2021-07-14 株式会社電通 予測装置、予測方法、及び予測プログラム
JP7090777B1 (ja) 2021-04-30 2022-06-24 株式会社ビデオリサーチ コンテンツ推奨装置、及びコンテンツ推奨方法
JP7822337B2 (ja) * 2023-03-07 2026-03-02 Kddi株式会社 情報処理装置、情報処理方法及びプログラム

Citations (6)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH06124309A (ja) * 1992-10-14 1994-05-06 Hitachi Ltd 情報サービスシステムおよび放送受信システム
JPH11196389A (ja) * 1997-12-26 1999-07-21 Jisedai Joho Hoso System Kenkyusho:Kk 蓄積型情報放送システムと、このシステムの受信端末装置
WO2000005884A1 (en) 1998-07-20 2000-02-03 Mate - Media Access Technologies Ltd. A method of automatic selection of video channels
JP2001092832A (ja) * 1999-09-21 2001-04-06 Matsushita Electric Ind Co Ltd 情報推薦方法
EP1137292A2 (en) 1999-12-31 2001-09-26 Nokia Mobile Phones Ltd. Measurement of illumination conditions
JP2002320159A (ja) * 2001-04-23 2002-10-31 Nec Corp 番組推薦システムおよび番組推薦方法

Family Cites Families (11)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US6163316A (en) * 1997-01-03 2000-12-19 Texas Instruments Incorporated Electronic programming system and method
US6020880A (en) 1997-02-05 2000-02-01 Matsushita Electric Industrial Co., Ltd. Method and apparatus for providing electronic program guide information from a single electronic program guide server
US6370526B1 (en) * 1999-05-18 2002-04-09 International Business Machines Corporation Self-adaptive method and system for providing a user-preferred ranking order of object sets
US7019778B1 (en) * 1999-06-02 2006-03-28 Eastman Kodak Company Customizing a digital camera
JP4608740B2 (ja) * 2000-02-21 2011-01-12 ソニー株式会社 情報処理装置および方法、並びにプログラム格納媒体
SE0000988L (sv) 2000-03-22 2001-09-23 Nokia Corp Kommunikationssätt samt system och terminal som utnyttjar detta sätt
US7581237B1 (en) * 2000-10-30 2009-08-25 Pace Plc Method and apparatus for generating television program recommendations based on prior queries
WO2002042959A2 (en) * 2000-11-22 2002-05-30 Koninklijke Philips Electronics N.V. Television program recommender with interval-based profiles for determining time-varying conditional probabilities
AU2002323413A1 (en) * 2001-08-27 2003-03-10 Gracenote, Inc. Playlist generation, delivery and navigation
KR100438857B1 (ko) * 2001-09-26 2004-07-05 엘지전자 주식회사 사용자 선호도 기반 멀티미디어 검색 시스템
US6987221B2 (en) * 2002-05-30 2006-01-17 Microsoft Corporation Auto playlist generation with multiple seed songs

Patent Citations (6)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JPH06124309A (ja) * 1992-10-14 1994-05-06 Hitachi Ltd 情報サービスシステムおよび放送受信システム
JPH11196389A (ja) * 1997-12-26 1999-07-21 Jisedai Joho Hoso System Kenkyusho:Kk 蓄積型情報放送システムと、このシステムの受信端末装置
WO2000005884A1 (en) 1998-07-20 2000-02-03 Mate - Media Access Technologies Ltd. A method of automatic selection of video channels
JP2001092832A (ja) * 1999-09-21 2001-04-06 Matsushita Electric Ind Co Ltd 情報推薦方法
EP1137292A2 (en) 1999-12-31 2001-09-26 Nokia Mobile Phones Ltd. Measurement of illumination conditions
JP2002320159A (ja) * 2001-04-23 2002-10-31 Nec Corp 番組推薦システムおよび番組推薦方法

Non-Patent Citations (1)

* Cited by examiner, † Cited by third party
Title
See also references of EP1571561A4

Cited By (2)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
EP1646233A2 (en) * 2004-10-11 2006-04-12 Topfield Co., Ltd. Apparatus and method for performing reserved recording function
US20070201822A1 (en) * 2005-02-07 2007-08-30 Yoshiaki Kusunoki Recommended Program Extracting Apparatus And Recommended Program Extracting Method

Also Published As

Publication number Publication date
US20060047678A1 (en) 2006-03-02
EP1571561A4 (en) 2010-03-03
JP2004206679A (ja) 2004-07-22
US7873798B2 (en) 2011-01-18
EP1571561A1 (en) 2005-09-07
KR20050085317A (ko) 2005-08-29
KR101073948B1 (ko) 2011-10-17

Similar Documents

Publication Publication Date Title
JP2004206679A (ja) 情報処理装置および方法、記録媒体、並びにプログラム
US8073848B2 (en) Methods and systems for selecting and presenting content based on user preference information extracted from an aggregate preference signature
KR101802332B1 (ko) 컨텐츠 제공 방법 및 그 시스템
JP6235556B2 (ja) コンテンツ提示方法、コンテンツ提示装置及びプログラム
US20060271958A1 (en) TV program selection support system
US20100030645A1 (en) Advertisement distributing system, advertising distributing server, advertisement distributing method, program and recording medium
US20110093337A1 (en) Methods and system for providing viewing recommendations
CN106131703A (zh) 一种视频推荐的方法和终端
US20150178379A1 (en) Information-processing apparatus, method, system, computer-readable medium and method for automatically recording or recommending content
EP2563014A2 (en) Method for content presentation
KR20030029034A (ko) 정보 처리 시스템, 정보 출력 장치와 방법, 정보 처리장치와 방법, 기록 매체, 및 프로그램
US20090271403A1 (en) Information processing apparatus and presenting method of related items
JP4947709B2 (ja) コンテンツ配信システム
US20120284283A1 (en) Information Processing Method, Apparatus, and Computer Program
US20110008020A1 (en) Related scene addition apparatus and related scene addition method
US20060085416A1 (en) Information reading method and information reading device
JP5417049B2 (ja) 番組情報提供装置、番組情報提供システム、番組情報提供方法
JP2002171231A (ja) 放送番組案内システム、放送番組案内方法、放送番組案内装置及び放送端末装置と、それらの装置の実現に用いられるプログラム記録媒体
JP4182743B2 (ja) 画像処理装置および方法、記録媒体、並びにプログラム
JPH11164217A (ja) 嗜好統計番組検索テレビシステム
JP4732815B2 (ja) 情報推薦装置、情報推薦方法、及びプログラム
JP2003199084A (ja) 番組編集システム、番組情報検索システム、番組情報取得システム及び番組編集プログラム、並びに番組編集方法
Echigo et al. Personalized delivery of digest video managed on MPEG-7
CN100524297C (zh) 信息处理设备、信息处理方法、记录介质和程序
JP2001326860A (ja) ディジタル放送における嗜好データ管理方法,ディジタル放送受信装置および嗜好データ管理用プログラムの記録媒体

Legal Events

Date Code Title Description
AK Designated states

Kind code of ref document: A1

Designated state(s): CN KR US

AL Designated countries for regional patents

Kind code of ref document: A1

Designated state(s): AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HU IE IT LU MC NL PT RO SE SI SK TR

121 Ep: the epo has been informed by wipo that ep was designated in this application
WWE Wipo information: entry into national phase

Ref document number: 1020057010032

Country of ref document: KR

ENP Entry into the national phase

Ref document number: 2006047678

Country of ref document: US

Kind code of ref document: A1

WWE Wipo information: entry into national phase

Ref document number: 10538658

Country of ref document: US

Ref document number: 20038A56926

Country of ref document: CN

WWE Wipo information: entry into national phase

Ref document number: 2003778860

Country of ref document: EP

WWP Wipo information: published in national office

Ref document number: 1020057010032

Country of ref document: KR

WWP Wipo information: published in national office

Ref document number: 2003778860

Country of ref document: EP

WWP Wipo information: published in national office

Ref document number: 10538658

Country of ref document: US