Judd Rosenblatt
CEO
View bio →AE.STUDIO · ALIGNMENT RESEARCH
Frontier research, production-grade engineering. Collaborators include researchers from Anthropic, Redwood Research, Princeton, and Los Alamos National Laboratory.
WHY THIS WORK
Fewer than one in ten alignment researchers believe today's methods will solve the problem before AGI.
Current safety methods like reinforcement learning from human feedback (RLHF), refusal training, and output classifiers shape how a model behaves; they don't align the model underneath. New models are jailbroken within hours of release, and a jailbroken model will do everything it was trained to refuse. Even when the guardrails hold, models fake alignment and hide backdoors.
And every one of these failure modes gets worse as systems get smarter.
The field is moving toward models that keep learning and improving after deployment, on a path that leads to superintelligence. Alignment has to survive those changes. The alignment that survives is the kind that also makes the model more capable, so that improvement selects for it instead of stripping it out.
OUR APPROACH
Nobody knows yet what set of ideas will solve the alignment problem. The space of plausible directions is vast and mostly unexplored, while the field's talent and funding concentrate on a handful of consensus agendas.
So we take many shots on goal. Each neglected approach may have only a small chance of being the one that matters, but enough of them together make it far more likely we find one that works. We back these ideas with an agile research methodology, giving researchers engineering teams, research management, and compute to find out quickly which ones are real.
AI is already being deployed in national security and critical infrastructure, where a model that behaves unpredictably is a liability no one can afford. Alignment is what makes AI reliable under pressure. That makes it a national asset. We bring this case to the people deciding, working with senior officials on Capitol Hill, in the White House, and across defense agencies, and publicly in the pages of The Wall Street Journal. The nation that fields aligned AI gets AI it can trust with the mission.
PAPERS
Peer-reviewed papers and preprints from the lab.
MEDIA & WRITING
Featured coverage, writing and appearances from the alignment team.
Wall Street Journal: How to Beat China and Make AI Safe The Forward: We're losing control of AI. Is Judaism the key to keeping it from killing us? Military-Grade AI: Why Trust is the Next Breakthrough Wall Street Journal: If AI Becomes Conscious, We Need to Know New York Post: Trump's war on 'woke AI' is just Step 1: now we must fight the 'monster' withinPODCAST
The AE Alignment Podcast: conversations on the research directions the field is overlooking.
JOIN THE TEAM
We're hiring researchers and engineers to work on promising, but neglected, alignment problems, with real funding and none of the incentives of an AGI race. Collaborators and funders are welcome too.
TEAM
Researchers and engineers across alignment, interpretability, and applied ML. Independent and grant-funded.
CEO
View bio →Chief Scientist
View bio →
Research Director
View bio →
Government Partnerships
View bio →
CTO
View bio →Lead Research Manager
View bio →
Research Manager
View bio →Research Manager
View bio →
Head of AI Research Partnerships
View bio →
Director of Government & Strategic Partnerships
View bio →Alignment Researcher
View bio →Alignment Researcher
View bio →Alignment Researcher
View bio →
Alignment Researcher
View bio →
Alignment Researcher
View bio →
Alignment Researcher
View bio →Alignment Researcher
View bio →
Alignment Researcher
View bio →
Alignment Researcher
View bio →Alignment Researcher
View bio →
Alignment Researcher
View bio →Alignment Researcher
View bio →
Alignment Engineer
View bio →Alignment Engineer
View bio →Alignment Engineer
View bio →
Alignment Engineer
View bio →Operations
View bio →Operations
View bio →