My second short story release in English is ready: Tales of Illustrious Computer Scientists: Iola Varga, nun and computer scientist. invece.org/iola.html
Follow my simple reasoning here. Large sparse LLMs are more powerful then dense one *while* being faster and thus more energy efficient. Dense models are a need that is artificially created by VRAM scarcity. Today they are practically useful, but not for long.
Meta released Muse Glimmer 30B: multimodal model for your Claw/Pi setups 🔥
we tested and fine-tuned the model for you, and shipped day-0 support in transformers and llama.cpp, including DFlash for 2-4x speed-ups 🥵
read our blog huggingface.co/blog/muse-glim…
Fast H3 implementation for Metal. Enjoy, modify, and so forth: github.com/antirez/h3.c Contains code from @liuliu which is welcomed in taking back whatever parts he likes for @drawthingsapp in case there are H3 plans there.