Attention-Head-Pruning

(โ˜… 22)

Layer-wise Pruning of Transformer Heads for Efficient Language Modeling

Attention-Head-Pruning Latest Version Download

Download Latest Version (.zip)
// repository documentation