1. X
  2. Modal
Log inSign up
Modal
1,579 posts
Image
user avatar
Modal
@modal
AI infrastructure that developers love 💚 Run inference, sandboxes, batch processing, training, and many other things on Modal
New York City
modal.com
Joined July 2022
161
Following
34.4K
Followers
AffiliatesAffiliatesRepliesRepliesArticlesArticlesMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    user avatar
    Modal
    @modal
    Jun 23
    It is not too late to _actually_ own your inference. Introducing: Modal Auto Endpoints.
    Image
    00:00
    234K
  • user avatar
    Modal
    @modal
    16h
    Day 0 support for Inkling-Small on Modal. - 276B parameter MoE with 12B active - 1M context - Variable thinking effort - Native image & audio understanding - NVFP4 checkpoint fits on a single @NVIDIAAI B300 Great to work with @thinkymachines and @sgl_project on this one.
    Image
    00:00
    user avatar
    Thinking Machines
    @thinkymachines
    17h
    Today, we are releasing Inkling-Small. Inkling-Small achieves comparable performance to Inkling at a quarter of its size. It features 276B total parameters, 12B active. We are making the full weights available. thinkingmachines.ai/news/inkling-s… Fine-tune it on Tinker today, or chat with
    7.5K
  • user avatar
    Modal
    @modal
    Jul 29
    Our Head of Inference, @_gongy , sat down with @cognition's Head of Research, @silasalberti, for a wide-ranging conversation on RL, inference, and the infrastructure work our two teams share, from training frontier coding models to running inference at scale. Watch here:
    user avatar
    gongy
    Modal
    @_gongy
    Jul 29
    I sat down with @silasalberti, Head of Research @cognition, to chat about the intersection of RL and inference -- from training frontier coding models to running inference at scale. Watch till the end for bonus doggo! :o 0:00 — Intros 1:04 — What's hardest to get right in an RL
    Image
    00:00
    26K
  • user avatar
    Modal
    @modal
    Jul 28
    Registration is live for Runtime. Apply to attend our inaugural conference for engineers running AI in production. October 1st live at The Midway in San Francisco.
    Image
    00:00
    30K
  • user avatar
    Modal
    @modal
    Jul 27
    Serving a 2.8T model well takes a village. Grateful to @simon_mo_, @rogerw0108 and everyone at @inferact, who spent five days trading configs and optimizations with us, right up into the early hours of day zero. And to @Kimi_Moonshot for bringing us together. We're excited to
    user avatar
    vLLM
    @vllm_project
    Jul 27
    With Kimi K3 Day-0 on vLLM: Open Frontier Intelligence for Everyone 🚀 At 2.8 trillion parameters, Moonshot AI's Kimi K3 is one of the most powerful open-weight models ever released. Starting today, you can serve it on vLLM the moment the weights are public. What K3 brings: 🧠
    Image
    00:00
    41K
  • See @modal's full profile

    Sign up
    Log in
Advertisement
Advertisement