A Variational Quantum Soft Actor-Critic Algorithm for Continuous Control Tasks
摘要
Quantum Computing promises the availability of computational resources and generalization capabilities well beyond the possibilities of classical computers. An interesting approach for leveraging the near-term, Noisy Intermediate-Scale Quantum Computers, is the hybrid training of Parameterized Quantum Circuits (PQCs), i.e. the optimization of a parameterized quantum algorithms as a function approximation with classical optimization techniques. When PQCs are used in Machine Learning models, they may offer some advantages over classical models in terms of memory consumption and sample complexity for classical data analysis. In this work we explore and assess the advantages of the application of Parametric Quantum Circuits to one of the state-of-art Reinforcement Learning algorithm for continuous control - namely Soft Actor-Critic. We investigate its performance on the control of a virtual robotic arm by means of digital simulations of quantum circuits. A quantum advantage over the classical algorithm has been found in terms of a significant decrease in the amount of required parameters for satisfactory model training, paving the way for further developments and studies. A quantum advantage over the classical algorithm has been found in terms of a significant decrease in the amount of required parameters for satisfactory model training, paving the way for further developments and studies.