Google Replaces BERT Self-Attention with Fourier Transform: 92% Accuracy, 7 Times Faster on GPUs

transformersfourier-transformbertnlpefficiency

Abstraction: FNet replaces transformer self-attention with Fourier Transform for faster training

Key points:

Connections: Google · Bert · Transformers · Attention Mechanism · Large Language Models

Source: https://syncedreview.com/2021/05/14/deepmind-podracer-tpu-based-rl-frameworks-deliver-exceptional-performance-at-low-cost-19/