Linear attention architecture model combining RNN efficiency with Transformer quality, enabling infinite context length at constant memory.
No takes yet.
Try another filter or be the first to share your honest opinion.