Pinned
Excited to share our new work from Inherent where we train a 27B AI Scientist model to replicate research papers, and get it to outperform Claude Opus 4.8 and GPT-5.5 Codex!
Lots of detail on stabilizing RL in complex, under-specified, long-horizon tasks! arxiv.org/abs/2608.13331
1/ Today, we introduce Faraday, a 27B-parameter AI Scientist that extends the capabilities of coding agents with a layer of scientific intuition. Trained via long-horizon RL, Faraday outperforms Claude Opus 4.8 and GPT-5.5 on the task of replicating research papers. 🧵





