Log inSign up
Amazon Science
Amazon News
7,626 posts
Amazon Science profile banner
@AmazonScience

Amazon Science

Amazon News
@AmazonScience
The latest news and research from Amazon's science community. #AmazonScience
Global
amazon.science
Joined December 2019
1,917
Following
92.1K
Followers
RepliesRepliesRepostsRepostsMediaMediaArticlesArticles

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @AmazonScience
    Amazon Science
    Amazon News
    @AmazonScience
    Sep 4
    .@amazon and @DARPA brought together 150+ researchers in Seattle last week, with speakers including @awscloud CEO @mattsgarman, Fields Medalist @TaoistTerence, @EPrinceton Associate Professor @BorisHanin, and @NSAGov's Michael O'Hara to explore how AI is transforming mathematical
    Image
    Image
    Image
    Image
    2
  • @AmazonScience
    Amazon Science
    Amazon News
    @AmazonScience
    Sep 2
    Amazon Redshift researchers were awarded Best Paper Runner-Up: Industrial Track at @VLDBconf for eliminating compilation cold starts in query execution – cutting compilation time from seconds to milliseconds with a 7x speedup on TPC-DS benchmarks. #VLDB2026
    Amazon Science OG image squid.png
    FastCompose: Eliminating compilation cold starts in query execution with composition
    From amazon.science
  • @AmazonScience
    Amazon Science
    Amazon News
    @AmazonScience
    Aug 31
    Verus is an open-source, automated program verifier for Rust that mechanically checks code against a formal mathematical specification for all possible inputs. Amazon used it to prove correctness of Nitro Isolation Engine primitives.
    Verus-16x9.gif
    Developing provably correct Rust code with Verus
    From amazon.science
    8
  • @AmazonScience
    Amazon Science
    Amazon News
    @AmazonScience
    Aug 28
    When LLM judges agree, the right question is why. Shared prompts, model families, or training lineage can make a majority look stronger than it is. Dependence-aware aggregation via Ising models accounts for this, improving accuracy 9–14% over weighted majority vote.
    agreement_fig1.png
    When LLM judges agree, should we believe them?
    From amazon.science
  • @AmazonScience
    Amazon Science
    Amazon News
    @AmazonScience
    Aug 25
    How did a model upgrade make agents worse? By pairing real enterprise SOPs with functioning tools and ground-truth grading across 12 industries and 2,000+ tasks, SOP-Bench helps find such anomalies.
    Collaborative workflow.16x9.png
    SOP-Bench: A new benchmark for evaluating AI agents on real business procedures
    From amazon.science
Advertisement
Advertisement