Log inSign up
Baseten
2,883 posts
Baseten profile banner
@baseten

Baseten

@baseten
Inference is everything.
San Francisco and New York
baseten.co
Joined March 2021
85
Following
18.8K
Followers
RepliesRepliesRepostsRepostsMediaMediaArticlesArticles

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @baseten
    Baseten
    @baseten
    Aug 28
    GLM-5.3 is live on Baseten Model APIs, day 0. - The smartest open-weight model at 743B params - 1M token context - US only - ZDR Try it here: baseten.co/library/glm-53/
    Image
    4
  • @baseten
    Baseten
    @baseten
    14h
    "If you're not taking care of yourself, you're not writing good code." We were honored to host @bryan_johnson and @saranormous at our office for a live discussion with some of our customers and friends. Thank you to everyone who joined us. 💚
    @bryan_johnson
    Bryan Johnson
    @bryan_johnson
    Aug 29
    Amazing longevity event with @baseten > over 3,500 applied for 150 spots > ppl really into the bioage tests, lots of fun > talked about doing epic things > using health to power greatness > @saranormous was perfect vibe fit > thanks baseten @saltyph @amiruci @DannieHerz
    Image
    Image
    Image
    1
  • @baseten
    Baseten
    @baseten
    Aug 29
    Our kernel engineers built an agentic framework to automatically find, build, validate, and ship optimized kernels into production. The new framework cut latency on Qwen-Image by 42.3%, and FLUX.2 by 15.2%.
    @BrianLi23
    Brian Li
    Baseten
    @BrianLi23
    Aug 28
    Article cover image
    Article
    Agentic Kernels in Production
    TL;DR: We’ve built an agentic kernel development framework that identifies model-level optimization opportunities, generates improved kernels, and validates them in our serving stack. On our current...
    3
  • @baseten
    Baseten
    @baseten
    Aug 28
    Post-train GLM-5.3 and GLM-5.3-Flash on Baseten Loops. Inference + training support on day 0. docs.baseten.co/loops/overview
    Image
    6
  • @baseten
    Baseten
    @baseten
    Aug 27
    We're proud to be the fastest inference provider on Artificial Analysis, OpenRouter, and Hugging Face for GLM-5.3-Flash, at 122+ TPS. All served from the US only, starting on day 0, with ZDR by default. Stay tuned for updates as our engineers continue to optimize GLM-5.3-Flash
    Image
    00:00
    14
Advertisement
Advertisement