JOINT REPRESENTATION LEARNING FROM IMAGES AND TEXT

    公开(公告)号:US20210056353A1

    公开(公告)日:2021-02-25

    申请号:US17000048

    申请日:2020-08-21

    Abstract: The disclosure provides a framework or system for learning visual representation using a large set of image/text pairs. The disclosure provides, for example, a method of visual representation learning, a joint representation learning system, and an artificial intelligence (AI) system that employs one or more of the trained models from the method or system. The AI system can be used, for example, in autonomous or semi-autonomous vehicles. In one example, the method of visual representation learning includes: (1) receiving a set of image embeddings from an image representation model and a set of text embeddings from a text representation model, and (2) training, employing mutual information, a critic function by learning relationships between the set of image embeddings and the set of text embeddings.

Patent Agency Ranking