Adaptive and Robust Feedback-Based Quantum Optimization using Reinforcement Learning
摘要
The paper introduces a novel hybrid Feedback-Based Quantum Optimization (FBQO) framework that integrates reinforcement learning (RL) for adaptive parameter control and Kalman filters for noise mitigation to enhance quantum optimization on noisy intermediate-scale quantum (NISQ) devices. Unlike traditional quantum Approximate optimization algorithm (QAOA), the proposed method dynamically tunes parameters through learned policies and statistically filtered feedback, enabling faster convergence and greater noise resilience. Key contributions include the formulation of quantum optimization as a Markov Decision Process, integration of deep quantum networks (DQNs) for adaptive control, and the use of Kalman filtering for robust state estimation. Experimental results on the Max-Cut problem demonstrate superior performance in convergence rate, stability, and optimization accuracy over existing techniques. The main finding is that RL-FBQO achieves convergence in 10–20 iterations, compared to 20–30 for FBQO and 40–50 for QAOA, with superior stability.