"generated 6 billion tokens of output while proving 29,500 theorems"
Imagine now these learnings to be shared across projects... Or even better across companies!!
AI4Science in collaborative mode via shared memory is the future!
Anthropic says its first Fermat runs failed not on math, but because dozens of Claude agents lost track of shared project state. They only succeeded after moving to Prove2Me's DAG of statements and proofs. At that scale, "memory" must be external, structured infrastructure, not



