-
公开(公告)号:US20210383225A1
公开(公告)日:2021-12-09
申请号:US17338777
申请日:2021-06-04
Applicant: DeepMind Technologies Limited
Inventor: Jean-Bastien François Laurent Grill , Florian Strub , Florent Altché , Corentin Tallec , Pierre Richemond , Bernardo Avila Pires , Zhaohan Guo , Mohammad Gheshlaghi Azar , Bilal Piot , Remi Munos , Michal Valko
Abstract: A computer-implemented method of training a neural network. The method comprises processing a first transformed view of a training data item, e.g. an image, with a target neural network to generate a target output, processing a second transformed view of the training data item, e.g. image, with an online neural network to generate a prediction of the target output, updating parameters of the online neural network to minimize an error between the prediction of the target output and the target output, and updating parameters of the target neural network based on the parameters of the online neural network. The method can effectively train an encoder neural network without using labelled training data items, and without using a contrastive loss, i.e. without needing “negative examples” which comprise transformed views of different data items.