Pinned48@💙💛@bleu48Oct 2, 2021連盟モバイルの勝率表示の件 - 48's diary 連盟モバイルの勝率表示の件 - 48's diaryFrom bleu48.hatenablog.com12457
48@💙💛@bleu488hオフロード最適化が流行り出したらすぐこんなの出てくるのね。高速化ってオリンピック競技みたいな競争になるのよね。やかもち@ちもろぐ@Yacamochi_db19h今さらだけど「Strata」凄すぎない? 容量87GBもある特大LLMモデル「Qwen3.8 Flash Next」が普通のパソコンで動くんだが? ✅RTX 5090単騎:120~150 t/s(Prefill 4Kt/s) しかも特大版Qwen3.8が普通に賢いし、LibreChatに入れてWeb検索を与えたらちゃんと検索するし、何よりガードレールが緩い😌11196
やかもち@ちもろぐ@Yacamochi_db19h今さらだけど「Strata」凄すぎない? 容量87GBもある特大LLMモデル「Qwen3.8 Flash Next」が普通のパソコンで動くんだが? ✅RTX 5090単騎:120~150 t/s(Prefill 4Kt/s) しかも特大版Qwen3.8が普通に賢いし、LibreChatに入れてWeb検索を与えたらちゃんと検索するし、何よりガードレールが緩い😌
48@💙💛@bleu4815hGo言語のこういうのプラグマとかで指定させてくれたらなあと先日手動でループ展開してて思った。 bleu48.hatenablog.com/entry/2026/09/…柴田 淳 - 「みんなのPython」の著者・機械学習講師@ats18hUberがGoのgoroutineの初期スタックを2KBから32KBに増やしたところ、スタック拡張に費やしていたCPUが約10%から1%未満に。メモリ節約のための仕組みをあえて止め、CPUを節約するという、巨大サービスならではの最適化です。 Show more2241
柴田 淳 - 「みんなのPython」の著者・機械学習講師@ats18hUberがGoのgoroutineの初期スタックを2KBから32KBに増やしたところ、スタック拡張に費やしていたCPUが約10%から1%未満に。メモリ節約のための仕組みをあえて止め、CPUを節約するという、巨大サービスならではの最適化です。 Show more
48@💙💛@bleu4823h先週3.8Flash出なかったっけ?Google DeepMind@GoogleDeepMindSep 30Introducing Gemini 4 Argon – our new frontier model. It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.140
Google DeepMind@GoogleDeepMindSep 30Introducing Gemini 4 Argon – our new frontier model. It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.