-
公开(公告)号:US20240419981A1
公开(公告)日:2024-12-19
申请号:US18750363
申请日:2024-06-21
Applicant: ByteDance Inc.
Inventor: Mark Thomas Daly , Shawn Ryan Jeffery , Matthew DeLand , Nick Pendar , Andrew James , David Johnston
IPC: G06N5/02 , G06F16/215 , G06F16/23 , G06N20/00
Abstract: In general, embodiments of the present invention provide systems, methods and computer readable media for automated dynamic data quality assessment. One aspect of the subject matter described in this specification includes the actions of receiving a data quality job including a new data sample; and, if the new data sample is determined to be added to a reservoir of data samples, sending a quality verification request to an oracle; receiving a new data sample quality estimate from the oracle; and adding the new data sample and estimate to the reservoir. A second aspect of the subject matter includes the actions of receiving, from a predictive model, a judgment associated with a new data sample; analyzing the new data sample based in part on the judgment to determine whether to send a new data sample quality verification request to an oracle; and, if a new data sample quality estimate is received from the oracle, determining whether to add the new data sample and the judgment to the reservoir.
-
公开(公告)号:US12045732B2
公开(公告)日:2024-07-23
申请号:US17684935
申请日:2022-03-02
Applicant: ByteDance Inc.
Inventor: Mark Thomas Daly , Shawn Ryan Jeffery , Matthew DeLand , Nick Pendar , Andrew James , David Johnston
IPC: G06F16/00 , G06F16/215 , G06F16/23 , G06N5/02 , G06N20/00
CPC classification number: G06N5/02 , G06F16/215 , G06F16/2358 , G06F16/2365 , G06N20/00
Abstract: In general, embodiments of the present invention provide systems, methods and computer readable media for automated dynamic data quality assessment. One aspect of the subject matter described in this specification includes the actions of receiving a data quality job including a new data sample; and, if the new data sample is determined to be added to a reservoir of data samples, sending a quality verification request to an oracle; receiving a new data sample quality estimate from the oracle; and adding the new data sample and estimate to the reservoir. A second aspect of the subject matter includes the actions of receiving, from a predictive model, a judgment associated with a new data sample; analyzing the new data sample based in part on the judgment to determine whether to send a new data sample quality verification request to an oracle; and, if a new data sample quality estimate is received from the oracle, determining whether to add the new data sample and the judgment to the reservoir.
-