Related Experiment Video
Updated: Jul 10, 2026

Gradient Echo Quantum Memory in Warm Atomic Vapor
Published on: November 11, 2013
Reinforcement learning control of quantum error correction
Volodymyr Sivak1, Alexis Morvan2, Michael Broughton3
1Google Quantum AI, Santa Barbara, CA, USA. vladsivak@google.com.
None:
Quantum error correction (QEC) is the primary strategy for protecting a quantum computer from the environment1,2. The prerequisite of QEC is that errors must remain sufficiently rare, which requires perpetually adapting the control parameters of the computer to the drifting environmental conditions. The current solution to this problem is to terminate the entire quantum computation for recalibration, but it is incompatible with the long runtimes of future quantum algorithms3,4. Here we address this challenge by unifying calibration with computation. We grant the QEC process5-11 a dual role: its error-detection events are not only used to correct the logical quantum state but are also repurposed as a learning signal, teaching a reinforcement learning agent12-16 to continuously steer the control parameters and stabilize the quantum system during computation. We experimentally demonstrate this framework on a Willow superconducting processor, improving the logical stability of the surface code 3.5-fold against injected drift. By synthesizing our full suite of technological advances, we achieve record performance of the surface and colour codes, with average logical error per cycle of 7.72(9) × 10-4 and 8.19(14) × 10-3, respectively. Numerical simulations of large codes with tens of thousands of control parameters confirm the scalability of our RL framework, revealing an optimization speed that is independent of system size. This work thus enables a new paradigm: a quantum computer that learns from its errors and never stops computing.
Related Concept Videos
Propagation of Uncertainty from Random Error
Propagation of Uncertainty from Systematic Error
Control Systems
At the heart...
Detection of Gross Error: The Q Test
Conservation of Energy in Control Volume
For steady flow systems, the time derivative of the stored energy becomes zero since there is no energy accumulation within the control volume. This simplifies the energy equation to:
Feedback control systems
Linear feedback systems are theoretical models that simplify analysis and design. These systems operate under the principle that their output is directly proportional to their input within certain ranges. For instance, an amplifier in a control system behaves linearly as long as the input signal remains within a specific range. However, most physical systems exhibit inherent nonlinearity...