Single-layer transformer model "HarEmb" showcasing PII SOTA performance
huggingface.co
1 thread
one of a kind single-transformer block layer, high throughput. The new generation of transformer-based lightweight models for common NLP tasks?