alphaXiv

Explore

Researchers

Sign In

MCP Server

Autoresearch

BlogSend Feedback?

Follow the latest research

alphaXiv connects papers, researchers, and organizations, grounding its answers in the underlying work.

Alt + Enter to search
Sign up

Recurrent Looped Transformer

Yifan ZhangYifan Zhang

A recurrent decoder carries computation across every prompt and response token, while exact full-history replay keeps reinforcement-learning policy states aligned with current parameters.

13 Sept 2026
8kviews
Paper thumbnail
View PDF

The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

ImageSJTUImageTsinghua
Yi DuanYing LiuBowen ZhouBowen Zhou

The survey’s five-level framework distinguishes persistent learning from genuine recursive improvement, clarifying which decisions AI systems control and which remain human-governed.

10 Sept 2026
541views
Paper thumbnail
View PDF

Thinking with Looped Flows

ImageEPFLImageKAIST
Ayhan SuleymanzadeChanhyuk LeeJinwoo KimJinwoo Kim

Training recurrent denoisers on progressively cleaner states enables reasoning models to improve with additional inference computation and generate diverse valid solutions.

10 Sept 2026
873views
Paper thumbnail
View PDF

Researchers to follow

View all
Kaiming He

Kaiming He

Distinguished Scientist

Google DeepMind, Associate Professor, EECS @ MIT

Percy Liang

Percy Liang

Cofounder

Simile AI, Professor, CS @ Stanford University, Cofounder @ Together AI

Andrej Karpathy

Andrej Karpathy

Researcher

Anthropic

Pieter Abbeel

Pieter Abbeel

Head, Frontier Model Research

Amazon, Professor, EECS @ UC Berkeley

Ian Goodfellow

Ian Goodfellow

Co-Founder

Stealth Startup

Andrew Ng

Andrew Ng

Managing Partner

AI Aspire, Managing General Partner @ AI Fund, Founder @ DeepLearning.AI, Adjunct Professor, CS @ Stanford University, Chairman and Co-Founder @ Coursera

Demis Hassabis

Demis Hassabis

Chair

Google DeepMind, Chief Scientist @ Alphabet, Founder & CEO @ Isomorphic Labs

Jeff Dean

Jeff Dean

CEO & Co-Founder

Discovery Loop

Are you a researcher? Find your profile

DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression

ImageDeepseek

Cross-layer cache reuse, low-precision storage, and bounded replay make million-token multimodal agents substantially easier to serve within limited memory.

10 Sept 2026
28kviews
Paper thumbnail
View PDF

Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data

ImageStanfordImageUW
AJ
Atindra Jha
Margaret LiMargaret LiPercy LiangPercy Liang

Repeated training data erodes Mixture-of-Experts language models faster than dense models, while dropout can partly restore generalization.

10 Sept 2026
1kviews
Paper thumbnail
View PDF

Why Does Post-Training Quantization Work?

ImageTsinghua
Yuxiang ChenMichael BeyerJun ZhuJun Zhu

Pretraining teaches language models to counteract quantization errors across layers, while output geometry preserves their highest-confidence token predictions.

10 Sept 2026
187views
Paper thumbnail
View PDF

SenseNova-U1.5: Towards Native Unified Visual Intelligence

Haiwen DiaoHaiwen DiaoJiahao WangZiwei LiuZiwei Liu

A shared pixel-space representation lets one model understand, reason about, generate, and edit images while preserving strong multimodal comprehension.

10 Sept 2026
375views6k
Paper thumbnail
View PDF

World in World: Explore the World with World Models

ImageWestlake University
Chenxi SongYanming YangChi ZhangChi Zhang

A frozen video world model can explore new camera paths while preserving an event’s appearance, timing, and previously generated scene states.

10 Sept 2026
373views51
Paper thumbnail
View PDF

Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity Linking

ImageCMU
Parinthapat PengpunParinthapat PengpunSimran KhanujaSimran KhanujaGraham NeubigGraham Neubig

Multidimensional rarity measures expose culturally specific entities that popularity metrics miss, while reasoning-guided Wikipedia retrieval improves multilingual linking on this neglected long tail.

09 Sept 2026
542views2
Paper thumbnail
View PDF

Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation

ImageTsinghua
Jintao ZhangJintao ZhangKai JiangJun ZhuJun Zhu

Incoming video streams can be edited in real time while preserving source motion and timing, enabling style transfer, virtual try-on, and scene replacement.

10 Sept 2026
141views
Paper thumbnail
View PDF

UniMPA: A Unified Memory-Prediction-Action Model via Action-Grounded Transition Modeling

ImageNTU
Wei LiWei LiRui ShaoRui ShaoZiwei LiuZiwei Liu

Grounding predicted state changes in temporally aligned robot experience helps manipulation policies adapt executable actions to new scenes and recover from disturbances.

10 Sept 2026
131views
Paper thumbnail
View PDF

Researchers to follow

View all
Stefano Ermon

Stefano Ermon

CEO & Co-Founder

Inception, Associate Professor, CS @ Stanford University

Dario Amodei

Dario Amodei

CEO and Co-Founder

Anthropic

Oriol Vinyals

Oriol Vinyals

Co-Founder

Discovery Loop

Quoc V. Le

Quoc V. Le

Co-Founder

Discovery Loop

Shuran Song

Shuran Song

Assistant Professor, EE, by courtesy of CS

Stanford University

Linxi "Jim" Fan

Linxi "Jim" Fan

Director & Distinguished Research Scientist

NVIDIA

Yuke Zhu

Yuke Zhu

Associate Professor, CS

The University of Texas at Austin, Director and Distinguished Research Scientist @ NVIDIA Research

Junyang Lin

Junyang Lin

Independent Researcher

Unaffiliated

Exact Purification Rates for Alternating Qubit Measurements and Their Blind Spot at Right Angles

Agus Mulia Bakti

Repeated measurements in two qubit bases yield closed-form purification rates, while right-angle settings cannot reveal noise through this diagnostic.

13 Sept 2026
Paper thumbnail
View PDF

BenchShield: Formal Model-Backed Instrumentation for Reward Integrity in LLM-Agent Evaluation Infrastructure

ImageDartmouth CollegeImageOSU
Shenghan ZhengZonglin DiDawn SongDawn Song

Infrastructure-side evidence lets benchmarks distinguish merely exposed reward-hacking paths from exploits agents actually use during evaluation.

10 Sept 2026
144views
Paper thumbnail
View PDF

Learning Realistic Athletic Sprinting Without Demonstrations

ImageStanford
William WangNicholas BiancoC. Karen LiuC. Karen Liu

Muscle-driven reinforcement learning generates biomechanically realistic sprints and athletic drills without motion-capture demonstrations, enabling predictive studies of altered muscle properties.

10 Sept 2026
125views
Paper thumbnail
View PDF

MindTopo: Can Foundation Models Reason in Topological Space?

ImageNorthwestern UniversityImageMicrosoft Research
Yunfei GeAnbang LiuJiajun WuJiajun Wu

The benchmark reveals that multimodal models can recognize topological relations in static scenes but struggle to preserve them while planning actions.

10 Sept 2026
1
Paper thumbnail
View PDF

Mathematical inverse problem for the world data inference of the parton distribution functions of the proton

Henri Hänninen

The framework recasts proton PDF extraction as a coupled linear inverse problem, enabling model-agnostic reconstruction from diverse deep-inelastic-scattering measurements.

10 Sept 2026
Paper thumbnail
View PDF

Finite Time Blowup for Navier–Stokes

ImageOpenAI

A smooth, compactly forced three-dimensional flow can develop unbounded velocity in finite time while retaining uniformly bounded kinetic energy.

08 Sept 2026
9kviews
Paper thumbnail
View PDF

Semigroup-JEPA: Latent Dynamics Consistency for Zero-Shot Physics Generalization

ImageYaleImageBrown University
Andy Zeyi LiuAndy Zeyi LiuHaoran SunHaoran SunRandall BalestrieroRandall Balestriero

Recursive latent-rollout training helps visual world models preserve dynamics-relevant information, enabling long-horizon prediction and robot control across unseen gravitational conditions.

09 Sept 2026
1kviews1
Paper thumbnail
View PDF

Memory as Plans: World-Action Modeling with Memory-Grounded Planning

ImageHITImageNTU
Sizhe ZhaoHaozhe XieWeiyu Zhao

Robots can use sparse visual memories to plan long-horizon manipulations while executing with fixed context, preserving historical evidence without growing inference inputs.

10 Sept 2026
Paper thumbnail
View PDF

NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction

ImageShanghai AI LabImageSJTU
Intern-NCP TeamJiaqi CaoDahua LinDahua Lin

Jointly predicting tokens and multi-token concepts enables language models to reach comparable training loss with substantially fewer tokens than standard architectures.

09 Sept 2026
170views8k
Paper thumbnail
View PDF
There are no more papers matching your filters at the moment.
Sign in

Assistant

Advertisement
Advertisement