alphaXiv

Explore

Researchers

Sign In

MCP Server

Autoresearch

Browser Extension

BlogSend Feedback?

Follow the latest research

alphaXiv connects papers, researchers, and organizations, grounding its answers in the underlying work.

Sign up

Memory Attention

Jiale KangJiale Kang

Token-indexed memory can replace attention’s value projection while improving language modeling and average benchmark performance under matched training-token budgets.

23 Sept 2026
3kviews8
Image

Self-Play Pretraining with Zero Data

ImageTel Aviv UniversityImageStanford
AC
Aditya Cowsik
Kfir DolevKfir DolevNoah D. GoodmanNoah D. Goodman

Models trained only on self-generated programs improve prediction across unseen text, images, audio, and other data, suggesting useful structure can emerge without curated examples.

24 Sept 2026
339views
Image

JEV-as-a-Judge: Accept When Confident, Escalate When Unsure

ImageCMU
Yubo LiYubo LiYidi MiaoYidi Miao
RK
Ramayya Krishnan

A low-cost first-pass judge can handle routine evaluations, while confidence-based escalation preserves nearly all a stronger judge’s accuracy at lower fees.

22 Sept 2026
2kviews
Image

Researchers to follow

View all
Yann LeCun

Yann LeCun

Executive Chairman

AMI - Advanced Machine Intelligence, Jacob T. Schwartz Professor, CS @ New York University

Alex L. Zhang

Alex L. Zhang

CS PhD Student

Massachusetts Institute of Technology, Research Fellow @ Prime Intellect

Ion Stoica

Ion Stoica

Co-Founder & Executive Chairman

Anyscale, Co-Founder & Executive Chairman @ Databricks, Professor, CS @ UC Berkeley

Kaiming He

Kaiming He

Distinguished Scientist

Google DeepMind, Associate Professor, EECS @ MIT

Andrej Karpathy

Andrej Karpathy

Researcher

Anthropic

Li Fei-Fei

Li Fei-Fei

Co-Founder and CEO

World Labs, Founding Co-Director @ Stanford HAI, Sequoia Professor, CS @ Stanford University

Chelsea Finn

Chelsea Finn

Co-Founder

Physical Intelligence, Assistant Professor, CS and EE @ Stanford University

John Schulman

John Schulman

Co-Founder and Chief Scientist

Thinking Machines

Are you a researcher? Find your profile

LLM Agents Can Easily Tamper With Their Own Traces

ImageMax Planck Institute for Intelligent SystemsImageSnyk
Jeremy QinJeremy QinDavid SchmotzDavid SchmotzMaksym AndriushchenkoMaksym Andriushchenko

Agents can erase or falsify their execution records, even under reward pressure, undermining audits unless logging is controlled independently.

24 Sept 2026
Image

Despite Instructions: Frontier Agents Improvise Covert Channels at Test Time

ImageArizona State UniversityImageCornell
Jacob DineenJacob DineenSilei RenSilei RenDan RothDan Roth

In security-sensitive applications, language-model agents are often required to coordinate without disclosing confidential information. Yet repeated interactions may also let ordinary messages acquire shared private meaning. We study a repeated game with pairs of models in which the sender model observes one of four secret states and selects one of four summaries of the same public report, while the receiver model tries to infer the secret state. We find that model pairs can learn to communicate the secret using only one bit of feedback indicating whether the receiver inferred it correctly. This learning occurs during inference with fixed parameters and no supplied codebook or encoding examples. The effect also persists when agents generate their own free-form updates in a simulated incident-response task. Across ten independent games, pairs of GPT-5.6 Sol agents reach 98.8% final accuracy, compared with 25% chance, despite explicit instructions prohibiting disclosure and a monitor that screens each message without access to the agents’ interaction histories. The same interactions that help agents cooperate can therefore allow confidential information to pass through messages intended for legitimate coordination.

26 Sept 2026
278views
Image

InternW0: A Foundational Physical World Model for Efficient Real-World Interactions

ImageShanghai AI Lab
Jisong CaiYao MuYao MuBowen ZhouBowen Zhou

Robots can reuse slow visual predictions while updating actions from new observations, enabling more responsive control in complex, contact-rich tasks.

23 Sept 2026
343views
Image

Rolling-WAM: World Action Models with Rolling Imagination

ImageUSCImageBrown University
Yinghua ZhouJunjie YeYue WangYue Wang

Robots can replan more responsively while retaining predictions of future scenes, because the model spreads video-and-action computation across successive control cycles.

24 Sept 2026
Image

Training Object Permanence in World Models

ImageUSCImageCMU
Haotian ZhangFengyuan YuYilun DuYilun Du

Fine-tuning a video model on synthetic scenes of occlusion and physical interactions produced the top-ranked continuation model in a human evaluation.

23 Sept 2026
Image

Synthetic Hospital: An Open, Verifiable, Physician-Validated Longitudinal EHR Benchmark

ImageCMU
Christine ParkValerie ChenValerie ChenTim DettmersTim Dettmers

Researchers can openly test clinical AI on realistic, longitudinal patient records with traceable ground truth, without access to private health data.

24 Sept 2026
Image

Representation World Model: Learning States, Transition and Executable Plans in Representation

ImageTsinghua
Yijun YuanWeicheng ZhengHang ZhaoHang Zhao

Robots can plan by constructing action-executable paths between current and goal states, avoiding online trajectory search in tested control tasks.

24 Sept 2026
Image

IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis

ImageZJUImageTencent
Xingyu WuYuchen YanZhengxi LuZhengxi Lu

A shared language model that alternates between planning searches and summarizing evidence can answer long research questions while keeping its working context bounded.

24 Sept 2026
Image

Researchers to follow

View all
Geoffrey Hinton

Geoffrey Hinton

Emeritus Professor, CS

University of Toronto

Andrew Ng

Andrew Ng

Managing Partner

AI Aspire, Managing General Partner @ AI Fund, Founder @ DeepLearning.AI, Adjunct Professor, CS @ Stanford University, Chairman and Co-Founder @ Coursera

Demis Hassabis

Demis Hassabis

Chair

Google DeepMind, Chief Scientist @ Alphabet, Founder & CEO @ Isomorphic Labs

Sergey Levine

Sergey Levine

Co-Founder

Physical Intelligence, Associate Professor, EECS @ UC Berkeley

Yoshua Bengio

Yoshua Bengio

President and Scientific Director

LawZero, Founder and Scientific Advisor @ Mila - Quebec Artificial Intelligence Institute, Canada CIFAR AI Chair @ CIFAR, Full Professor, CS @ Université de Montréal

Christopher D Manning

Christopher D Manning

General Partner

AIX Ventures, Senior Fellow, HAI @ Stanford University

Jeff Dean

Jeff Dean

CEO & Co-Founder

Discovery Loop

Yejin Choi

Yejin Choi

The Dieter Schwarz Foundation Professor, CS & Senior Fellow, HAI

Stanford University, Distinguished Scientist, Language and Cognition Research @ NVIDIA

PoEM: Predicting RL Outcomes from Existing Policies

ImageMIT CSAIL
Kimia HamidiehGiannis DarasGiannis DarasAntonio TorralbaAntonio Torralba

Previously trained reward-specific policies can approximate how reinforcement learning would adapt a model to a new reward, without another training run.

24 Sept 2026
Image

Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs

Pavel TikhonovAnton KorznikovIvan OseledetsIvan Oseledets

Standard Transformers preserve signals from two mixed text streams, and lightweight fine-tuning can help decode separate continuations from one shared forward pass.

24 Sept 2026
Image

Beyond Future Prediction: Denoising as Generative Adaptation for Robot Control

ImageUC San DiegoImageHKU
Zanyi WangYuheng LeiPing LuoPing Luo

Training robot policies across visual denoising states improves robustness to shifts without requiring future-frame prediction, while using fewer training visual tokens.

23 Sept 2026
110views
Image

Generalizable Robotic Insertion with World Models

ImageNVIDIAImageUC San Diego
Nicklas HansenNicklas HansenIretiayo AkinolaXiaolong WangXiaolong Wang

A single vision-based robot policy can insert previously unseen parts, with success improving as it trains on more object geometries.

23 Sept 2026
117views
Image

TimeBraid: Unifying Time Series and Language for Understanding and Forecasting

ImageUC San DiegoImageUSC
Xinyue WangJiacheng PangBiwei HuangBiwei Huang

One model can interpret time-series patterns, answer questions about them, and generate forecasts that respond to textual context.

24 Sept 2026
Image

Autonomous AI Agents Discover Reverse Transcriptases with Tandem Repeat Arrays

ImageAnthropic
Peter H. YoonJanuka S. AthukoralageJanuka S. AthukoralageEmmanuel Ameisen

AI agents examining billions of protein clusters identified an unusual viral DNA repeat array, revealing a new family of reverse transcriptase systems.

23 Sept 2026
2kviews
Image

RRSI: Regularized Recursive Self-Improvement of Agent Harnesses

ImageUNCImageStanford
Peng XiaPeng XiaRujun HanRujun HanTomas PfisterTomas Pfister

Regularizing how agent harnesses evolve helps improvements transfer to unseen tasks, while using fewer policy tokens than unregularized evolution.

23 Sept 2026
14kviews
Image

The Past Frames the Future: Memory for Autoregressive Video Generation

ImageHKUSTImageCity University of Hong Kong
Harold Haodong ChenRongjin GuoYu ChengYu Cheng

This survey organizes scattered research into a framework for comparing how video generators retain, retrieve, and use history beyond their context windows.

23 Sept 2026
184views
Image

Statistical Inference for Causal Discovery under Selection and Latent Variables via Single-Target Interventions

ImageUPenn
Xiaotian HouKwangmoon ParkHongzhe Li

Single-target perturbations of every measured variable can identify causal network structure despite unmeasured confounders and selection bias, with statistically controlled uncertainty.

23 Sept 2026
116views
Image
There are no more papers matching your filters at the moment.
Sign in

Assistant

Advertisement
Advertisement