1. X
  2. Peter Gostev (SF 24-28 August)
Log inSign up
Peter Gostev (SF 24-28 August)
Arena.ai
2,214 posts
Peter Gostev (SF 24-28 August) profile banner
user avatar

Peter Gostev (SF 24-28 August)

Arena.ai
@petergostev
London 🇬🇧 AI Capability @arena linkedin.com/in/peter-goste…
Joined June 2025
1,036
Following
24.1K
Followers
RepliesRepliesMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    user avatar
    Peter Gostev (SF 24-28 August)
    Arena.ai
    @petergostev
    Feb 24
    I've got a fun new benchmark for you where most LLMs are doing pretty badly - "Bullshit Benchmark". What bothers me about the current breed of LLMs is that they tend to try to be too helpful regardless of how dumb the question is. So I've built 55 'bullshit' questions that don't
    Image
    00:00
  • user avatar
    Peter Gostev (SF 24-28 August)
    Arena.ai
    @petergostev
    Aug 27
    ThursdayAI - Qwen 3.8, GLM 5.3, Navigator n2 & H3 Max | ThursdAI Aug 27
  • user avatar
    Peter Gostev (SF 24-28 August)
    Arena.ai
    @petergostev
    Aug 27
    Interesting, if you take the last couple of quarters, Hyperscaler revenue has been ok, but the Neocloud side has doubled
    Image
  • user avatar
    Peter Gostev (SF 24-28 August)
    Arena.ai
    @petergostev
    Aug 26
    This is interesting and explains why Chinese labs with much less compute are still competing well, perhaps for now, before training runs get actually big Dylan: “When Anthropic trains Mythos, it’s sub-200 megawatts.”
    Image
    00:00
    Image
    01:16:52
    user avatar
    Dwarkesh Patel
    @dwarkesh_sp
    Aug 25
    Had a lot of fun chatting again with my twin brother @dylan522p We went through lab economics over the next few years - the shift from inference to training as RSI draws near; and how Anthropic and OpenAI are on track to control most of the world’s usable FLOPs within the next
  • user avatar
    Peter Gostev (SF 24-28 August)
    Arena.ai
    @petergostev
    Aug 26
    I don't particularly believe we are in an automated research era, but I totally believe that engineering powered by AI will be eating the world
    user avatar
    Patrick C Toulme
    @PatrickToulme
    Aug 25
    OpenAI Jalapeño is truly the first AI silicon developed by GPT-Astra and other OpenAI internal models. I believe they heavily used reinforcement learning on internal models like GPT-Astra to achieve these SOTA results. Here is what I think they did: 1. RL an internal GPT to
Image
REPLAY
user avatar
Peter Gostev (SF 24-28 August)
Arena.ai
@petergostev
ThursdayAI - Qwen 3.8, GLM 5.3, Navigator n2 & H3 Max | ThursdAI Aug 27
Advertisement
Advertisement