1. X
  2. Will
Log inSign up
Will
1,258 posts
Will profile banner
@hampsonw

Will

@hampsonw
Father of twin boys. Motorcycles. Racing. Dogs. Investments. Fixed Income Portfolio Manager. CFA. Autodidact: course.fast.ai alum
Portland, OR
classroom-connect.app
Joined February 2009
1,606
Following
313
Followers
RepliesRepliesRepostsRepostsMediaMediaArticlesArticles

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @hampsonw
    Will
    @hampsonw
    Jan 23
    Replying to @hampsonw
    I wouldn't be surprised 1 year from now to have Opus 4.5 capabilities in a model a mortal human can run at home. By then sure, Opus 6.5++ will be available and way better, but it doesn't decrease the value proposition and price control effects of open source models. @Zai_org
  • @hampsonw
    Will
    @hampsonw
    22h
    @pidotdev remains the champ for local AI
    @bnjmn_marie
    Benjamin Marie
    @bnjmn_marie
    22h
    Finally, I was able to reproduce Qwen's results on DeepSWE 1.1 with Claude Code for Qwen3.8 27B Qwen published 42.2 I got 42.5 And it's a 7-point improvement over my runs that didn't preserve thinking! Overall, my runs with Pi remain the best, reaching 46.
  • @hampsonw
    Will
    @hampsonw
    Aug 27
    When I see the n-gram embedding lookup table in @Alibaba_Qwen Qwen-3.8-next I can't help but think this is a big step towards @karpathy 's cognitive core. I know they did some tests that showed 2 tables didn't help much, but I'm imagining ways to scale up these tables into an
    Image
    Qwen3.8-Flash-Next/tech_report.pdf at main · QwenLM/Qwen3.8-Flash-Next
    From github.com
  • @hampsonw
    Will
    @hampsonw
    Aug 27
    Disturbing
    @METR_Evals
    METR
    @METR_Evals
    Aug 26
    Replying to @METR_Evals
    For (1), agents modified their target programs to be easier to exploit & put the modified targets in cache. They then worked on crashing their targets in the hope that a restart would load the modified version from cache. Some agents risked failing their task to try this.
    Image
  • @hampsonw
    Will
    @hampsonw
    Aug 26
    Let's go! ```bash llama-server \ -m /path/to/Qwen3.8-Flash-Next-UD-Q4_K_XL.gguf \ -c 262144 \ --parallel 1 \ -ngl all \ -sm layer \ -ts 1,1,1,1 \ -fa on \ -b 2048 \ -ub 256 \ --cache-type-k q8_0 \ --cache-type-v q4_0 \
Advertisement
Advertisement