The reality is what we are seeing unfold is Nvidia speedrunning the creation of a synthetic hyperscaler.
Apologies in advance to all the investors who are stuck in their priors that this will trigger.
But what is a hyperscaler? Strip it down and it’s a scaled infrastructure
DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe.
On Ollama, this model runs with high performance (100tps+) and zero data retention. Your data stays yours.
ollama run deepseek-v4-flash:0731-cloud