4 ms·Single-layer transformer model "HarEmb" showcasing PII SOTA performance1 points by fblgit 5mo agofblgit 5mo agoone of a kind single-transformer block layer, high throughput. The new generation of transformer-based lightweight models for common NLP tasks?