Machine-Learned Attention Models Featuring Omnidirectional Processing
Abstract:
Provided are machine-learned attention models that feature omnidirectional processing, example implementations of which can be referred to as Omnidirectional Representations from Transformers (OMNINET). In example models described in the present disclosure, instead of maintaining a strictly horizontal receptive field, each token is allowed to attend to all tokens in some or all of the other tokens across the entire network.
Information query
Patent Agency Ranking
0/0