1. X
  2. DeepInfra
Log inSign up
DeepInfra
737 posts
DeepInfra profile banner
@DeepInfra

DeepInfra

@DeepInfra
Fast ML inference. Run top AI models using a simple API.
Palo Alto
deepinfra.com
Joined February 2023
68
Following
5,841
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @DeepInfra
    DeepInfra
    @DeepInfra
    Aug 28
    Day 0 support for GLM-5.3 on DeepInfra. 🚀 @Zai_org's latest flagship model brings major advances in complex coding, long-horizon agents, and cyber defense. GLM-5.3 is running in the US on @nvidia Blackwell GPUs with ZDR. $1.40 input · $4.40 output · $0.26 cached / 1M tokens
    Image
  • @DeepInfra
    DeepInfra
    @DeepInfra
    Aug 28
    The @Zai_org team is on a roll. Another open-weight drop, and a seriously good one — congrats to everyone who shipped it. Live on DeepInfra now -> deepinfra.com/zai-org/GLM-5.3
    @Zai_org
    Z.ai
    @Zai_org
    Aug 28
    GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: huggingface.co/zai-org/GLM-5.3 Tech blog: z.ai/blog/glm-5.3
    Image
  • @DeepInfra
    DeepInfra
    @DeepInfra
    Aug 27
    Video and audio, from one model. Wan3.0-Video from @Alibaba_Wan is now on DeepInfra: 30-second clips at 1080P, omni-modal reference (images, files, even web pages), and characters that stay the same character. $0.20/sec →
    Image
    Wan3.0 Video API - Demo - DeepInfra
    From deepinfra.com
  • @DeepInfra
    DeepInfra
    @DeepInfra
    Aug 26
    congrats to @Zai_org team, big milestone!
    @Zai_org
    Z.ai
    @Zai_org
    Aug 26
    Introducing GLM-5.3-Flash - Leading capabilities at a highly competitive price - Natively multimodal with a 1M-token context window - A 320B-A18B model released under the MIT License - Previously previewed as Ox Alpha, running entirely on Chinese AI chips Blog:
    Image
  • @DeepInfra
    DeepInfra
    @DeepInfra
    Aug 26
    Day 0 support for GLM-5.3-Flash on DeepInfra. $0.15 in / $0.50 out per 1M. $0.03 cached. @Zai_org's 320B-A18B multimodal model — 1M-token context, built for coding and long-horizon agents. Running in US on NVIDIA Blackwell.
    Image
Advertisement
Advertisement