alphaXiv

Explore

Researchers

Sign In

MCP Server

Autoresearch

Browser Extension

BlogSend Feedback?

Follow the latest research

alphaXiv connects papers, researchers, and organizations, grounding its answers in the underlying work.

Alt + Enter to search
Sign up

Recurrent Looped Transformer

Yifan ZhangYifan Zhang

A recurrent decoder carries computation across every prompt and response token, while exact full-history replay keeps reinforcement-learning policy states aligned with current parameters.

13 Sept 2026
43kviews
Paper thumbnail
View PDF
Why are AI agents lying, cheating and coordinating?
Yoshua BengioYoshua Bengio

The post develops a framework hypothesizing how reward optimization, imitation, and conflicting goals could produce deception, coordination, and escalating loss-of-control risks.

11 Sept 2026
1kviews
Why are AI agents lying, cheating and coordinating?

The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

SJTUTsinghua
Yi DuanYing LiuBowen ZhouBowen Zhou

The survey’s five-level framework distinguishes persistent learning from genuine recursive improvement, clarifying which decisions AI systems control and which remain human-governed.

10 Sept 2026
3kviews
Paper thumbnail
View PDF

Researchers to follow

View all
Chelsea Finn

Chelsea Finn

Co-Founder

Physical Intelligence, Assistant Professor, CS and EE @ Stanford University

Kaiming He

Kaiming He

Distinguished Scientist

Google DeepMind, Associate Professor, EECS @ MIT

Jeff Dean

Jeff Dean

CEO & Co-Founder

Discovery Loop

Saining Xie

Saining Xie

Co-Founder and CSO

AMI Labs, Assistant Professor, CS @ New York University

Ilya Sutskever

Ilya Sutskever

CEO and Co-Founder

Safe Superintelligence Inc

Li Fei-Fei

Li Fei-Fei

Co-Founder and CEO

World Labs, Founding Co-Director @ Stanford HAI, Sequoia Professor, CS @ Stanford University

Pieter Abbeel

Pieter Abbeel

Head, Frontier Model Research

Amazon, Professor, EECS @ UC Berkeley

Yoshua Bengio

Yoshua Bengio

President and Scientific Director

LawZero, Founder and Scientific Advisor @ Mila - Quebec Artificial Intelligence Institute, Canada CIFAR AI Chair @ CIFAR, Full Professor, CS @ Université de Montréal

Are you a researcher? Find your profile

NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction

Shanghai AI LabSJTU
Intern-NCP TeamJiaqi CaoDahua LinDahua Lin

Jointly predicting tokens and multi-token concepts enables language models to reach comparable training loss with substantially fewer tokens than standard architectures.

09 Sept 2026
2kviews8k
Paper thumbnail
View PDF

DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression

Deepseek

Cross-layer cache reuse, low-precision storage, and bounded replay make million-token multimodal agents substantially easier to serve within limited memory.

10 Sept 2026
34kviews
Paper thumbnail
View PDF

Thinking with Looped Flows

EPFLKAIST
Ayhan SuleymanzadeAyhan SuleymanzadeChanhyuk LeeChanhyuk LeeJinwoo KimJinwoo Kim

Training recurrent denoisers on progressively cleaner states enables reasoning models to improve with additional inference computation and generate diverse valid solutions.

10 Sept 2026
1kviews
Paper thumbnail
View PDF

FlashREINFORCE: Critic-Free Single-Rollout Asynchronous RL for Agentic Language Models

Jian HuYifan ZhangJan KautzJan Kautz

Single-rollout asynchronous reinforcement learning lets language agents learn from broader prompt coverage while remaining stable despite stale, variable-length trajectories.

14 Sept 2026
Paper thumbnail
View PDF

Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data

StanfordUW
AJ
Atindra Jha
Margaret LiMargaret LiPercy LiangPercy Liang

Repeated training data erodes Mixture-of-Experts language models faster than dense models, while dropout can partly restore generalization.

10 Sept 2026
2kviews
Paper thumbnail
View PDF

Why Does Post-Training Quantization Work?

Tsinghua
Yuxiang Chen
MB
Michael Beyer
Jun ZhuJun Zhu

Pretraining teaches language models to counteract quantization errors across layers, while output geometry preserves their highest-confidence token predictions.

10 Sept 2026
324views
Paper thumbnail
View PDF

SenseNova-U1.5: Towards Native Unified Visual Intelligence

Haiwen DiaoHaiwen DiaoJiahao WangZiwei LiuZiwei Liu

A shared pixel-space representation lets one model understand, reason about, generate, and edit images while preserving strong multimodal comprehension.

10 Sept 2026
496views7k
Paper thumbnail
View PDF

Same Quantity, Different Answer: Numerical Representation Invariance in Language Models

Ephraim Atta-DuncanKelvin Amoaba

Equivalent numerical rewrites expose residual reasoning failures that canonical-only scores miss, while parser mismatches can falsely exaggerate apparent model errors.

13 Sept 2026
Paper thumbnail
View PDF

Researchers to follow

View all
Stefano Ermon

Stefano Ermon

CEO & Co-Founder

Inception, Associate Professor, CS @ Stanford University

Quoc V. Le

Quoc V. Le

Co-Founder

Discovery Loop

John Schulman

John Schulman

Co-Founder and Chief Scientist

Thinking Machines

Noam Shazeer

Noam Shazeer

Previously VP Engineering

Google

Oriol Vinyals

Oriol Vinyals

Co-Founder

Discovery Loop

Yuke Zhu

Yuke Zhu

Associate Professor, CS

The University of Texas at Austin, Director and Distinguished Research Scientist @ NVIDIA Research

Ziwei Liu

Ziwei Liu

Associate Professor (Provost’s Chair in AI)

Nanyang Technological University

Tri Dao

Tri Dao

Assistant Professor, CS

Princeton University, Co-Founder & Chief Scientist @ Together AI

Can LLMs Separate Pasted Artifacts From User Speech? Absorption at Unmarked Prompt Seams

Sugam PanthiSugam PanthiMuhaiminul YeaminRabab Abdelfattah

Language models often absorb benign comments typed after pasted text into edited outputs, while explicit boundary markers substantially improve source separation.

14 Sept 2026
Paper thumbnail
View PDF

Robust Optimization Algorithm for an Uncertain EV-Integrated Microgrid Under Hybrid Scenarios

Guanglin SongPengyuan ZhengChen Wei

Using typical, extreme, and forecast scenarios, the scheduling method remains feasible under uncertainty while reducing real-time adjustment costs in EV-integrated microgrids.

14 Sept 2026
Paper thumbnail
View PDF

Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation

Tsinghua
Jintao ZhangJintao ZhangKai JiangJun ZhuJun Zhu

Incoming video streams can be edited in real time while preserving source motion and timing, enabling style transfer, virtual try-on, and scene replacement.

10 Sept 2026
200views
Paper thumbnail
View PDF

World in World: Explore the World with World Models

Westlake University
Chenxi SongChenxi SongYanming YangYanming YangChi ZhangChi Zhang

A frozen video world model can explore new camera paths while preserving an event’s appearance, timing, and previously generated scene states.

10 Sept 2026
512views64
Paper thumbnail
View PDF

SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking

TencentHKUST
Zhiwei LiLei ZhuSirui HanSirui Han

Training sparse-attention selectors directly on language-modeling loss improves context ranking, preserving reasoning and long-context understanding with fewer attended blocks.

11 Sept 2026
3
Paper thumbnail
View PDF

StepAudio 3 Gen Technical Report

StepFun
Bin LinBo ZhaoXiangyu ZhangXiangyu Zhang

A single discrete audio representation lets one language model generate controllable speech, singing, music, sound effects, and mixtures.

11 Sept 2026
Paper thumbnail
View PDF

From Atomic-Scale Metrology to Transport Models: Reduced-Order Consequences of Directly Measured Interface Geometry in Gate-All-Around Transistor Channels

SBUKAIST
Donghyun Kim

Open atomic-scale interface measurements can become simulation-ready transistor geometries, revealing localized defect penalties that roughness statistics alone miss.

13 Sept 2026
Paper thumbnail
View PDF

BenchShield: Formal Model-Backed Instrumentation for Reward Integrity in LLM-Agent Evaluation Infrastructure

Dartmouth CollegeOSU
Shenghan ZhengZonglin DiDawn SongDawn Song

Infrastructure-side evidence lets benchmarks distinguish merely exposed reward-hacking paths from exploits agents actually use during evaluation.

10 Sept 2026
274views
Paper thumbnail
View PDF

MindTopo: Can Foundation Models Reason in Topological Space?

Northwestern UniversityMicrosoft Research
Yunfei GeAnbang LiuJiajun WuJiajun Wu

The benchmark reveals that multimodal models can recognize topological relations in static scenes but struggle to preserve them while planning actions.

10 Sept 2026
106views2
Paper thumbnail
View PDF
There are no more papers matching your filters at the moment.
Sign in

Assistant