Pinned
On a personal note:
- I was quite surprised by how much information can be compressed and decoded accurately, but ONLY if sufficient compute is used
- Using more parameters can result in more quality AND efficiency, and compression is a way to unlock this interesting interplay
New paper alert π¨
What if I told you there is an architecture that provides a _knob_ to control quality-efficiency trade-offs directly at test-time?
Introducing Compress & Attend Transformers (CATs) that provide you exactly this!
π§΅(1/n) π




