I'm hiring engineers at Boardy!
@boardyai is an AI Superconnector that exchanges millions of messages with people around the world and has made 160k+ introductions.
Those introductions have led to investments, hires, and companies being started.
We've built all of this with a
Boardy hit @twilio rate limits and their support is legitimately terrible.
Looking for alternatives ASAP of a provider who actually wants high-growth startups as customers.
Sorry if you’ve been trying to call/text @boardyai this week and he’s been a bit unreliable. Will fix it
undoing positional encoding before compressing the K cache....simple idea that nobody was doing, unlocks 10x KV cache compression with zero accuracy loss
i just beat @GoogleDeepMind's turboquant
introducing Shard. 10x KV cache compression on Llama-3.1-8B. zero quality loss
- 10x @ 8K context, 11.2x @ 32K
- NIAH recall 1.000 across 4K-32K
- LongBench Δ ≈ 0 vs FP16
turboquant tops out at 4-6x at the same quality. we doubled it.