SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning
Main Authors: | , , |
---|---|
Format: | Article |
Language: | English |
Published: |
Institute of Electrical and Electronics Engineers (IEEE),
2022-07-12T14:08:01Z.
|
Subjects: | |
Online Access: | Get fulltext |