摘要:
A computer implemented method, apparatus, and computer usable program code for automatically selecting an optimal control cohort. Attributes are selected based on patient data. Treatment cohort records are clustered to form clustered treatment cohorts. Control cohort records are scored to form potential control cohort members. The optimal control cohort is selected by minimizing differences between the potential control cohort members and the clustered treatment cohorts.
摘要:
A computer implemented method, apparatus, and computer usable program code for automatically selecting an optimal control cohort. Attributes are selected based on patient data. Treatment cohort records are clustered to form clustered treatment cohorts. Control cohort records are scored to form potential control cohort members. The optimal control cohort is selected by minimizing differences between the potential control cohort members and the clustered treatment cohorts.
摘要:
Method, system, and article of manufacture for selecting prospects for a product promotion though data mining. An initial set of prospects in a customer database is identified, by data mining, as initially identified prospects based on predetermined selection criteria. The number of initially identified prospects is compared to a target number of prospects. When the number of initially identified prospects matches the target number of prospects, the initially identified prospects are utilized as the final selection of prospects. When the number of initially identified prospects mismatches the target number of prospects, the final selection of prospects is determined by performing a culling process or an augmenting process to reduce or increase, respectively, the initial set of prospects using a heuristic measure H, until the number of prospects in the initial set of prospects matches the target number of prospects.
摘要:
A method, apparatus, and computer instructions for direct linkage of relational database table to a data preparation tool for data preparation. In a preferred embodiment, the mechanism of the present invention allows data to be read directly from one or more relational database tables to a data preparation tool into datasets without generating output flat files. Multiple datasets from different relational database table are merged into one dataset if more than one relational database table is read. Upon completion of necessary data preparation on the dataset by the data preparation tool, the present invention creates a new relational database table and loads resulting data from the prepared dataset into the new relational database table.
摘要:
An initial set of prospects in a customer database is identified, by data mining, as initially identified prospects based on predetermined selection criteria. The number of initially identified prospects is compared to a target number of prospects. When the number of initially identified prospects matches the target number of prospects, the initially identified prospects are utilized as the final selection of prospects. When the number of initially identified prospects mismatches the target number of prospects, the final selection of prospects is determined by performing a culling process or an augmenting process to reduce or increase, respectively, the initial set of prospects using a heuristic measure H, until the number of prospects in the initial set of prospects matches the target number of prospects.
摘要:
Method, system, and article of manufacture for selecting prospects for a product promotion though data mining. An initial set of prospects in a customer database is identified, by data mining, as initially identified prospects based on predetermined selection criteria. The number of initially identified prospects is compared to a target number of prospects. When the number of initially identified prospects matches the target number of prospects, the initially identified prospects are utilized as the final selection of prospects. When the number of initially identified prospects mismatches the target number of prospects, the final selection of prospects is determined by performing a culling process or an augmenting process to reduce or increase, respectively, the initial set of prospects using a heuristic measure H, until the number of prospects in the initial set of prospects matches the target number of prospects.
摘要:
Methods and apparatus, including computer program products, implementing and using techniques for populating a data cache on a server. Data requests received by the server are collected in a repository. A data mining algorithm is applied to the collected data requests to predict a set of data that is likely to be requested during an upcoming time period. It is determined whether the complete set of predicted data exists in the data cache. If the complete set of predicted data does not exist in the data cache, the missing data is retrieved from a database and added to the data cache.
摘要:
Methods and apparatus, including computer program products, implementing and using techniques for populating a data cache on a server. Data requests received by the server are collected in a repository. A data mining algorithm is applied to the collected data requests to predict a set of data that is likely to be requested during an upcoming time period. It is determined whether the complete set of predicted data exists in the data cache. If the complete set of predicted data does not exist in the data cache, the missing data is retrieved from a database and added to the data cache.