Quantum agents in the Gym: a variational quantum algorithm for deep Q-learning

Andrea Skolik1,2, Sofiene Jerbi3, and Vedran Dunjko1

1Leiden University, Niels Bohrweg 1, 2333 CA Leiden, The Netherlands
2Volkswagen Data:Lab, Ungererstraße 69, 80805 Munich, Germany
3Institute for Theoretical Physics, University of Innsbruck, Technikerstr. 21a, A-6020 Innsbruck, Austria

Find this paper interesting or want to discuss? Scite or leave a comment on SciRate.

Abstract

Quantum machine learning (QML) has been identified as one of the key fields that could reap advantages from near-term quantum devices, next to optimization and quantum chemistry. Research in this area has focused primarily on variational quantum algorithms (VQAs), and several proposals to enhance supervised, unsupervised and reinforcement learning (RL) algorithms with VQAs have been put forward. Out of the three, RL is the least studied and it is still an open question whether VQAs can be competitive with state-of-the-art classical algorithms based on neural networks (NNs) even on simple benchmark tasks. In this work, we introduce a training method for parametrized quantum circuits (PQCs) that can be used to solve RL tasks for discrete and continuous state spaces based on the deep Q-learning algorithm. We investigate which architectural choices for quantum Q-learning agents are most important for successfully solving certain types of environments by performing ablation studies for a number of different data encoding and readout strategies. We provide insight into why the performance of a VQA-based Q-learning algorithm crucially depends on the observables of the quantum model and show how to choose suitable observables based on the learning task at hand. To compare our model against the classical DQN algorithm, we perform an extensive hyperparameter search of PQCs and NNs with varying numbers of parameters. We confirm that similar to results in classical literature, the architectural choices and hyperparameters contribute more to the agents' success in a RL setting than the number of parameters used in the model. Finally, we show when recent separation results between classical and quantum agents for policy gradient RL can be extended to inferring optimal Q-values in restricted families of environments.

Deep reinforcement learning has yielded remarkable successes over the past decade, achieving superhuman levels in many seminal "AI" benchmarks such as the game of Go, Chess, poker etc. In this work we explore how techniques from deep reinforcement learning can be transferred to the realm of variational quantum algorithms for a special type of reinforcement learning algorithm called Q-learning. In essence, we propose a quantum variant of deep reinforcement learning which substitutes the neural network with a quantum analogue – a parametrized quantum circuit. We show that with careful design choices such an architecture performs well on two classical benchmark tasks from the OpenAI Gym, perform comparisons of our model to a neural network-driven system on the same learning task, and analyse the theoretical perspectives and limitations of the model. We find that the quantum learner is competitive to its classical counterpart, and prove that in some task environments one can achieve a provable exponential separation between classical and quantum Q-learners.

► BibTeX data

► References

[1] Kishor Bharti, Alba Cervera-Lierta, Thi Ha Kyaw, Tobias Haug, Sumner Alperin-Lea, Abhinav Anand, Matthias Degroote, Hermanni Heimonen, Jakob S Kottmann, Tim Menke, et al. Noisy intermediate-scale quantum (nisq) algorithms. arXiv preprint arXiv:2101.08448, 2021 doi:10.1103/​RevModPhys.94.015004.
https:/​/​doi.org/​10.1103/​RevModPhys.94.015004
arXiv:2101.08448

[2] John Preskill. Quantum computing in the nisq era and beyond. Quantum, 2:79, 2018. doi:10.22331/​q-2018-08-06-79.
https:/​/​doi.org/​10.22331/​q-2018-08-06-79

[3] Kosuke Mitarai, Makoto Negoro, Masahiro Kitagawa, and Keisuke Fujii. Quantum circuit learning. Physical Review A, 98(3):032309, 2018. doi:10.1103/​PhysRevA.98.032309.
https:/​/​doi.org/​10.1103/​PhysRevA.98.032309

[4] Maria Schuld, Alex Bocharov, Krysta M Svore, and Nathan Wiebe. Circuit-centric quantum classifiers. Physical Review A, 101(3):032308, 2020. doi:10.1103/​PhysRevA.101.032308.
https:/​/​doi.org/​10.1103/​PhysRevA.101.032308

[5] Maria Schuld and Nathan Killoran. Quantum machine learning in feature hilbert spaces. Physical review letters, 122(4):040504, 2019. doi:10.1103/​PhysRevLett.122.040504.
https:/​/​doi.org/​10.1103/​PhysRevLett.122.040504

[6] Vojtěch Havlíček, Antonio D Córcoles, Kristan Temme, Aram W Harrow, Abhinav Kandala, Jerry M Chow, and Jay M Gambetta. Supervised learning with quantum-enhanced feature spaces. Nature, 567(7747):209–212, 2019. doi:10.1038/​s41586-019-0980-2.
https:/​/​doi.org/​10.1038/​s41586-019-0980-2

[7] Edward Farhi and Hartmut Neven. Classification with quantum neural networks on near term processors. arXiv preprint arXiv:1802.06002, 2018.
arXiv:1802.06002

[8] Mohammad H Amin, Evgeny Andriyash, Jason Rolfe, Bohdan Kulchytskyy, and Roger Melko. Quantum boltzmann machine. Physical Review X, 8(2):021050, 2018. doi:10.1103/​PhysRevX.8.021050.
https:/​/​doi.org/​10.1103/​PhysRevX.8.021050

[9] Brian Coyle, Daniel Mills, Vincent Danos, and Elham Kashefi. The born supremacy: Quantum advantage and training of an ising born machine. npj Quantum Information, 6(1):1–11, 2020. doi:10.1038/​s41534-020-00288-9.
https:/​/​doi.org/​10.1038/​s41534-020-00288-9

[10] Christa Zoufal, Aurélien Lucchi, and Stefan Woerner. Variational quantum boltzmann machines. Quantum Machine Intelligence, 3(1):1–15, 2021. doi:10.1007/​s42484-020-00033-7.
https:/​/​doi.org/​10.1007/​s42484-020-00033-7

[11] Seth Lloyd and Christian Weedbrook. Quantum generative adversarial learning. Physical review letters, 121(4):040502, 2018. doi:10.1103/​PhysRevLett.121.040502.
https:/​/​doi.org/​10.1103/​PhysRevLett.121.040502

[12] Christa Zoufal, Aurélien Lucchi, and Stefan Woerner. Quantum generative adversarial networks for learning and loading random distributions. npj Quantum Information, 5(1):1–9, 2019. doi:10.1038/​s41534-019-0223-2.
https:/​/​doi.org/​10.1038/​s41534-019-0223-2

[13] Shouvanik Chakrabarti, Huang Yiming, Tongyang Li, Soheil Feizi, and Xiaodi Wu. Quantum wasserstein generative adversarial networks. In Advances in Neural Information Processing Systems, pages 6781–6792, 2019.

[14] A Hamann, V Dunjko, and S Wölk. Quantum-accessible reinforcement learning beyond strictly epochal environments. arXiv preprint arXiv:2008.01481, 2020. doi:10.1007/​s42484-021-00049-7.
https:/​/​doi.org/​10.1007/​s42484-021-00049-7
arXiv:2008.01481

[15] Sofiene Jerbi, Lea M Trenkwalder, Hendrik Poulsen Nautrup, Hans J Briegel, and Vedran Dunjko. Quantum enhancements for deep reinforcement learning in large spaces. PRX Quantum, 2(1):010328, 2021. doi:10.1103/​PRXQuantum.2.010328.
https:/​/​doi.org/​10.1103/​PRXQuantum.2.010328

[16] Samuel Yen-Chi Chen, Chao-Han Huck Yang, Jun Qi, Pin-Yu Chen, Xiaoli Ma, and Hsi-Sheng Goan. Variational quantum circuits for deep reinforcement learning. IEEE Access, 8:141007–141024, 2020. doi:10.1109/​ACCESS.2020.3010470.
https:/​/​doi.org/​10.1109/​ACCESS.2020.3010470

[17] Owen Lockwood and Mei Si. Reinforcement learning with quantum variational circuit. In Proceedings of the AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment, pages 245–251, 2020.

[18] Shaojun Wu, Shan Jin, Dingding Wen, and Xiaoting Wang. Quantum reinforcement learning in continuous action space. arXiv preprint arXiv:2012.10711, 2020.
arXiv:2012.10711

[19] Marcello Benedetti, Erika Lloyd, Stefan Sack, and Mattia Fiorentini. Parameterized quantum circuits as machine learning models. Quantum Science and Technology, 4(4):043001, 2019. doi:10.1088/​2058-9565/​ab4eb5.
https:/​/​doi.org/​10.1088/​2058-9565/​ab4eb5

[20] Sofiene Jerbi, Casper Gyurik, Simon Marshall, Hans Briegel, and Vedran Dunjko. Parametrized quantum policies for reinforcement learning. Advances in Neural Information Processing Systems, 34, arXiv preprint arXiv:2103.05577 2021.
arXiv:2103.05577

[21] Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al. Human-level control through deep reinforcement learning. nature, 518(7540):529–533, 2015. doi:10.1038/​nature14236.
https:/​/​doi.org/​10.1038/​nature14236

[22] David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al. Mastering the game of go with deep neural networks and tree search. nature, 529(7587):484–489, 2016. doi:10.1038/​nature16961.
https:/​/​doi.org/​10.1038/​nature16961

[23] Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung, Przemyslaw Debiak, Christy Dennison, David Farhi, Quirin Fischer, Shariq Hashme, Chris Hesse, et al. Dota 2 with large scale deep reinforcement learning. arXiv preprint arXiv:1912.06680, 2019.
arXiv:1912.06680

[24] Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al. Grandmaster level in starcraft ii using multi-agent reinforcement learning. Nature, 575(7782):350–354, 2019. doi:10.1038/​s41586-019-1724-z.
https:/​/​doi.org/​10.1038/​s41586-019-1724-z

[25] Vijay R Konda and John N Tsitsiklis. Actor-critic algorithms. In Advances in neural information processing systems, pages 1008–1014, 2000.

[26] Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. Asynchronous methods for deep reinforcement learning. In International conference on machine learning, pages 1928–1937. PMLR, 2016.

[27] Christopher John Cornish Hellaby Watkins. Learning from delayed rewards. 1989.

[28] Leslie N Smith. A disciplined approach to neural network hyper-parameters: Part 1–learning rate, batch size, momentum, and weight decay. arXiv preprint arXiv:1803.09820, 2018.
arXiv:1803.09820

[29] Ziyu Ye, Andrew Gilman, Qihang Peng, Kelly Levick, Pamela Cosman, and Larry Milstein. Comparison of neural network architectures for spectrum sensing. In 2019 IEEE Globecom Workshops (GC Wkshps), pages 1–6. IEEE, 2019. doi:10.1109/​GCWkshps45667.2019.9024482.
https:/​/​doi.org/​10.1109/​GCWkshps45667.2019.9024482

[30] Hao Yu, Tiantian Xie, Michael Hamilton, and Bogdan Wilamowski. Comparison of different neural network architectures for digit image recognition. In 2011 4th International Conference on Human System Interactions, HSI 2011, pages 98–103. IEEE, 2011. doi:10.1109/​HSI.2011.5937350.
https:/​/​doi.org/​10.1109/​HSI.2011.5937350

[31] F Cordoni. A comparison of modern deep neural network architectures for energy spot price forecasting. Digital Finance, 2:189–210, 2020. doi:10.1007/​s42521-020-00022-2.
https:/​/​doi.org/​10.1007/​s42521-020-00022-2

[32] Tomasz Szandała. Review and comparison of commonly used activation functions for deep neural networks. In Bio-inspired Neurocomputing, pages 203–224. Springer, 2021.

[33] Chigozie Nwankpa, Winifred Ijomah, Anthony Gachagan, and Stephen Marshall. Activation functions: Comparison of trends in practice and research for deep learning. arXiv preprint arXiv:1811.03378, 2018.
arXiv:1811.03378

[34] Sebastian Urban. Neural network architectures and activation functions: A gaussian process approach. PhD thesis, Technische Universität München, 2018.

[35] Leslie N Smith. Cyclical learning rates for training neural networks. In 2017 IEEE winter conference on applications of computer vision (WACV), pages 464–472. IEEE, 2017. doi:10.1109/​WACV.2017.58.
https:/​/​doi.org/​10.1109/​WACV.2017.58

[36] Thomas Elsken, Jan Hendrik Metzen, and Frank Hutter. Neural architecture search: A survey. The Journal of Machine Learning Research, 20(1):1997–2017, 2019.

[37] Frank Hutter, Lars Kotthoff, and Joaquin Vanschoren. Automated machine learning: methods, systems, challenges. Springer Nature, 2019. doi:10.1007/​978-3-030-05318-5.
https:/​/​doi.org/​10.1007/​978-3-030-05318-5

[38] Jarrod R McClean, Sergio Boixo, Vadim N Smelyanskiy, Ryan Babbush, and Hartmut Neven. Barren plateaus in quantum neural network training landscapes. Nature communications, 9(1):1–6, 2018. doi:10.1038/​s41467-018-07090-4.
https:/​/​doi.org/​10.1038/​s41467-018-07090-4

[39] Bobak Toussi Kiani, Seth Lloyd, and Reevu Maity. Learning unitaries by gradient descent. arXiv preprint arXiv:2001.11897, 2020.
arXiv:2001.11897

[40] Roeland Wiersema, Cunlu Zhou, Yvette de Sereville, Juan Felipe Carrasquilla, Yong Baek Kim, and Henry Yuen. Exploring entanglement and optimization within the hamiltonian variational ansatz. PRX Quantum, 1(2):020319, 2020. doi:10.1103/​PRXQuantum.1.020319.
https:/​/​doi.org/​10.1103/​PRXQuantum.1.020319

[41] M Cerezo, Akira Sone, Tyler Volkoff, Lukasz Cincio, and Patrick J Coles. Cost function dependent barren plateaus in shallow parametrized quantum circuits. Nature Communications, 12(1):1–12, 2021. doi:10.1038/​s41467-021-21728-w.
https:/​/​doi.org/​10.1038/​s41467-021-21728-w

[42] Samson Wang, Enrico Fontana, Marco Cerezo, Kunal Sharma, Akira Sone, Lukasz Cincio, and Patrick J Coles. Noise-induced barren plateaus in variational quantum algorithms. Nature communications, 12(1):1–11, 2021. doi:10.1038/​s41467-021-27045-6.
https:/​/​doi.org/​10.1038/​s41467-021-27045-6

[43] Andrea Skolik, Jarrod R McClean, Masoud Mohseni, Patrick van der Smagt, and Martin Leib. Layerwise learning for quantum neural networks. Quantum Machine Intelligence, 3 (1):1–11, 2021. doi:10.1007/​s42484-020-00036-4.
https:/​/​doi.org/​10.1007/​s42484-020-00036-4

[44] Carlos Ortiz Marrero, Mária Kieferová, and Nathan Wiebe. Entanglement-induced barren plateaus. PRX Quantum, 2(4):040316, 2021. doi:10.1103/​PRXQuantum.2.040316.
https:/​/​doi.org/​10.1103/​PRXQuantum.2.040316

[45] Sukin Sim, Peter D Johnson, and Alán Aspuru-Guzik. Expressibility and entangling capability of parameterized quantum circuits for hybrid quantum-classical algorithms. Advanced Quantum Technologies, 2(12):1900070, 2019. doi:10.1002/​qute.201900070.
https:/​/​doi.org/​10.1002/​qute.201900070

[46] Sukin Sim, Jhonathan Romero Fontalvo, Jérôme F Gonthier, and Alexander A Kunitsa. Adaptive pruning-based optimization of parameterized quantum circuits. Quantum Science and Technology, 2021. doi:10.1088/​2058-9565/​abe107.
https:/​/​doi.org/​10.1088/​2058-9565/​abe107

[47] Xiaoyuan Liu, Anthony Angone, Ruslan Shaydulin, Ilya Safro, Yuri Alexeev, and Lukasz Cincio. Layer vqe: A variational approach for combinatorial optimization on noisy quantum computers. arXiv preprint arXiv:2102.05566, 2021. doi:10.1109/​TQE.2021.3140190.
https:/​/​doi.org/​10.1109/​TQE.2021.3140190
arXiv:2102.05566

[48] Maria Schuld, Ryan Sweke, and Johannes Jakob Meyer. Effect of data encoding on the expressive power of variational quantum-machine-learning models. Physical Review A, 103(3):032430, 2021. doi:10.1103/​PhysRevA.103.032430.
https:/​/​doi.org/​10.1103/​PhysRevA.103.032430

[49] Openai gym wiki, cartpole v0. URL: https:/​/​github.com/​openai/​gym/​wiki/​CartPole-v0.
https:/​/​github.com/​openai/​gym/​wiki/​CartPole-v0

[50] Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. Openai gym. arXiv preprint arXiv:1606.01540, 2016.
arXiv:1606.01540

[51] Adrián Pérez-Salinas, Alba Cervera-Lierta, Elies Gil-Fuster, and José I Latorre. Data re-uploading for a universal quantum classifier. Quantum, 4:226, 2020. doi:10.22331/​q-2020-02-06-226.
https:/​/​doi.org/​10.22331/​q-2020-02-06-226

[52] Kei Ota, Devesh K Jha, and Asako Kanezaki. Training larger networks for deep reinforcement learning. arXiv preprint arXiv:2102.07920, 2021.
arXiv:2102.07920

[53] Code used in this work https:/​/​github.com/​askolik/​quantum_agents. URL: https:/​/​github.com/​askolik/​quantum_agents.
https:/​/​github.com/​askolik/​quantum_agents

[54] Richard S Sutton and Andrew G Barto. Reinforcement learning: An introduction. MIT press, 2018. doi:10.1109/​TNN.1998.712192.
https:/​/​doi.org/​10.1109/​TNN.1998.712192

[55] Richard S Sutton, David A McAllester, Satinder P Singh, Yishay Mansour, et al. Policy gradient methods for reinforcement learning with function approximation. In NIPs, volume 99, pages 1057–1063. Citeseer, 1999.

[56] Evan Greensmith, Peter L Bartlett, and Jonathan Baxter. Variance reduction techniques for gradient estimates in reinforcement learning. Journal of Machine Learning Research, 5(9), 2004.

[57] Francisco S Melo. Convergence of q-learning: A simple proof. Institute Of Systems and Robotics, Tech. Rep, pages 1–4, 2001.

[58] Long-Ji Lin. Self-supervised Learning by Reinforcement and Artificial Neural Networks. PhD thesis, Carnegie Mellon University, School of Computer Science, 1992.

[59] Francisco S Melo and M Isabel Ribeiro. Q-learning with linear function approximation. In International Conference on Computational Learning Theory, pages 308–322. Springer, 2007. doi:10.1007/​978-3-540-72927-3_23.
https:/​/​doi.org/​10.1007/​978-3-540-72927-3_23

[60] Abhinav Kandala, Antonio Mezzacapo, Kristan Temme, Maika Takita, Markus Brink, Jerry M Chow, and Jay M Gambetta. Hardware-efficient variational quantum eigensolver for small molecules and quantum magnets. Nature, 549(7671):242–246, 2017. doi:10.1038/​nature23879.
https:/​/​doi.org/​10.1038/​nature23879

[61] Yunchao Liu, Srinivasan Arunachalam, and Kristan Temme. A rigorous and robust quantum speed-up in supervised machine learning. Nature Physics, pages 1–5, 2021. doi:10.1038/​s41567-021-01287-z.
https:/​/​doi.org/​10.1038/​s41567-021-01287-z

[62] Vedran Dunjko, Yi-Kai Liu, Xingyao Wu, and Jacob M Taylor. Exponential improvements for quantum-accessible reinforcement learning. arXiv preprint arXiv:1710.11160, 2017.
arXiv:1710.11160

[63] Peter W Shor. Polynomial-time algorithms for prime factorization and discrete logarithms on a quantum computer. SIAM review, 41(2):303–332, 1999. doi:10.1137/​S0036144598347011.
https:/​/​doi.org/​10.1137/​S0036144598347011

[64] Openai gym wiki, frozen lake v0. URL: https:/​/​github.com/​openai/​gym/​wiki/​FrozenLake-v0.
https:/​/​github.com/​openai/​gym/​wiki/​FrozenLake-v0

[65] Michael Broughton, Guillaume Verdon, Trevor McCourt, Antonio J Martinez, Jae Hyeon Yoo, Sergei V Isakov, Philip Massey, Murphy Yuezhen Niu, Ramin Halavati, Evan Peters, et al. Tensorflow quantum: A software framework for quantum machine learning. arXiv preprint arXiv:2003.02989, 2020.
arXiv:2003.02989

[66] Cirq, https:/​/​quantumai.google/​cirq. URL: https:/​/​quantumai.google/​cirq.
https:/​/​quantumai.google/​cirq

[67] Openai gym leaderboard. URL: https:/​/​github.com/​openai/​gym/​wiki/​Leaderboard.
https:/​/​github.com/​openai/​gym/​wiki/​Leaderboard

[68] Jin-Guo Liu and Lei Wang. Differentiable learning of quantum circuit born machines. Physical Review A, 98(6):062324, 2018. doi:10.1103/​PhysRevA.98.062324.
https:/​/​doi.org/​10.1103/​PhysRevA.98.062324

Cited by

[1] Bhagya Rekha Sangisetti and Suresh Pabboju, "Deep fit_predic: a novel integrated pyramid dilation EfficientNet-B3 scheme for fitness prediction system", Computer Methods in Biomechanics and Biomedical Engineering 27 14, 2009 (2024).

[2] Samuel Yen-Chi Chen, 2025 IEEE Symposium for Multidisciplinary Computational Intelligence Incubators (MCII Companion) 1 (2025) ISBN:979-8-3315-1966-7.

[3] Ruijiang Zhang, Siwei Liu, Qing-Shan Jia, and Xu Wang, 2023 China Automation Congress (CAC) 2774 (2023) ISBN:979-8-3503-0375-9.

[4] Xianchao Zhu and Xiaokai Hou, "Quantum architecture search via truly proximal policy optimization", Scientific Reports 13 1, 5157 (2023).

[5] Sohan Jayram Reddy K and Bhargavi K, 2025 International Conference on Intelligent and Innovative Technologies in Computing, Electrical and Electronics (IITCEE) 1 (2025) ISBN:979-8-3315-1591-1.

[6] Austin Braniff, Fengqi You, and Yuhe Tian, Proceedings of the 36th European Symposium on Computer Aided Process Engineering (ESCAPE 36) 6, 1659 (2026) ISBN:9781777940355.

[7] Daniel Hein, Simon Wiedemann, Markus Baumann, Patrik Felbinger, Justin Klein, Maximilian Schieder, Jonas Stein, Daniëlle Schuman, Thomas Cope, and Steffen Udluft, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 121 (2025) ISBN:979-8-3315-5736-2.

[8] Manuel Schonberger, Maja Franz, Stefanie Scherzinger, and Wolfgang Mauerer, 2022 IEEE 19th International Conference on Software Architecture Companion (ICSA-C) 164 (2022) ISBN:978-1-6654-9493-9.

[9] Naihua Ji, Rongyi Bao, Zhao Chen, Yiming Yu, and Hongyang Ma, "Hybrid Quantum Neural Network Image Anti-Noise Classification Model Combined with Error Mitigation", Applied Sciences 14 4, 1392 (2024).

[10] Jiahua Xu, Ningyi Xie, Xinwei Lee, Yoshiyuki Saito, Nobuyoshi Asai, and Dongsheng Cai, 2024 IEEE 17th International Symposium on Embedded Multicore/Many-core Systems-on-Chip (MCSoC) 490 (2024) ISBN:979-8-3315-3047-1.

[11] Namasi G. Sankar, Ankit Khandelwal, and M Girish Chandra, 2024 16th International Conference on COMmunication Systems & NETworkS (COMSNETS) 1058 (2024) ISBN:979-8-3503-8311-9.

[12] Joaquim M. Gaspar, Alexandre Bergerault, Vassilis Apostolou, and Arno Ricou, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 329 (2024) ISBN:979-8-3315-4137-8.

[13] Yifeng Peng, Xinyi Li, Zhemin Zhang, Samuel Yen-Chi Chen, Zhiding Liang, and Ying Wang, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 1715 (2025) ISBN:979-8-3315-5736-2.

[14] Samuel Yen-Chi Chen, "Quantum Artificial Intelligence: From Quantum Neural Networks to Self-Programming Architectures [Feature]", IEEE Circuits and Systems Magazine 26 1, 41 (2026).

[15] Samuel Yen-Chi Chen, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1516 (2024) ISBN:979-8-3315-4137-8.

[16] Lukas Schulte, Daniel Hein, Steffen Udluft, and Thomas A. Runkler, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 1727 (2025) ISBN:979-8-3315-5736-2.

[17] Kai Cao, Chenghu Zhou, Richard Church, Xia Li, and Wenwen Li, "Revisiting spatial optimization in the era of geospatial big data and GeoAI", International Journal of Applied Earth Observation and Geoinformation 129, 103832 (2024).

[18] Bhaskara Narottama, Zina Mohamed, and Sonia Aïssa, "Quantum Machine Learning for Next-G Wireless Communications: Fundamentals and the Path Ahead", IEEE Open Journal of the Communications Society 4, 2204 (2023).

[19] Hans Hohenfeld, Dirk Heimann, Felix Wiebe, and Frank Kirchner, "Quantum Deep Reinforcement Learning for Robot Navigation Tasks", IEEE Access 12, 87217 (2024).

[20] Samuel Yen-Chi Chen, 2024 15th International Conference on Information and Communication Technology Convergence (ICTC) 1139 (2024) ISBN:979-8-3503-6463-7.

[21] Gilberto Cunha, Alexandra Ramôa, André Sequeira, Michael de Oliveira, and Luís Barbosa, "Quantum Bayesian networks can speed up reinforcement learning in partially observable environments", Quantum Machine Intelligence 8 2, 65 (2026).

[22] Samuel Yen-Chi Chen, 2023 IEEE International Conference on Quantum Computing and Engineering (QCE) 31 (2023) ISBN:979-8-3503-4323-6.

[23] Junghoon Park, Jiook Cha, Samuel Yen-Chi Chen, Shinjae Yoo, and Huan-Hsin Tseng, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1583 (2024) ISBN:979-8-3315-4137-8.

[24] Samuel Yen-Chi Chen, Proceedings of the 2023 International Workshop on Quantum Classical Cooperative 17 (2023) ISBN:9798400701627.

[25] Samuel Yen-Chi Chen and Shinjae Yoo, Federated Learning 311 (2024) ISBN:9780443190377.

[26] Akshay Ajagekar and Fengqi You, "Demand Response in Building Microgrids with Variational Quantum Circuit Enabled Hybrid Control Strategy", IFAC-PapersOnLine 58 13, 182 (2024).

[27] Shiqin Di, Jinchen Xu, Guoqiang Shu, Congcong Feng, Xiaodong Ding, and Zheng Shan, "Amplitude transformed quantum convolutional neural network", Applied Intelligence 53 18, 20863 (2023).

[28] Nguyen Manh Dung, Nguyen Xuan Tung, and Nguyen Tien Hoa, 2025 10th International Conference on Applying New Technology in Green Buildings (ATiGB) 966 (2025) ISBN:979-8-3315-9548-7.

[29] Eli J. Laird and Mitchell A. Thornton, 2025 IEEE 18th Dallas Circuits and Systems Conference (DCAS) 1 (2025) ISBN:979-8-3315-9934-8.

[30] Kuan-Cheng Chen, Samuel Yen-Chi Chen, Chen-Yu Liu, and Kin K. Leung, 2025 International Wireless Communications and Mobile Computing (IWCMC) 337 (2025) ISBN:979-8-3315-0887-6.

[31] M. Cerezo, Guillaume Verdon, Hsin-Yuan Huang, Lukasz Cincio, and Patrick J. Coles, "Challenges and opportunities in quantum machine learning", Nature Computational Science 2 9, 567 (2022).

[32] Simon Eisenmann, Daniel Hein, Steffen Udluft, and Thomas A. Runkler, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1490 (2024) ISBN:979-8-3315-4137-8.

[33] Maida Shahid and Muhammad Awais Hassan, "Introducing Quantum Variational Circuit for Efficient Management of Common Pool Resources", IEEE Access 11, 110862 (2023).

[34] Siddhant Dutta, Nouhaila Innan, Alberto Marchisio, Sadok Ben Yahia, and Muhammad Shafique, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 341 (2024) ISBN:979-8-3315-4137-8.

[35] Tobias Rohe, Simon Grätz, Michael Kölle, Sebastian Zielinski, Jonas Stein, and Claudia Linnhoff-Popien, Lecture Notes in Networks and Systems 1285, 21 (2025) ISBN:978-3-031-84459-1.

[36] Satyander Kaushik and Ravneet Kaur, 2025 International Conference on Innovations and Emerging Technologies In AI & Communication Systems (IETACS) 470 (2025) ISBN:979-8-3315-7073-6.

[37] Eva Andrés, Manuel Pegalajar Cuéllar, and Gabriel Navarro, "On the Use of Quantum Reinforcement Learning in Energy-Efficiency Scenarios", Energies 15 16, 6034 (2022).

[38] Su Fong Chien, Samuel Yen-Chi Chen, Mau Luen Tham, Heng Siong Lim, Charilaos C. Zarakovitis, Michail A. Kourtis, and Yi Jie Wong, 2025 IEEE Globecom Workshops (GC Wkshps) 927 (2025) ISBN:979-8-3315-6741-5.

[39] Peigen Zeng, Ying He, F. Richard Yu, and Victor C.M. Leung, GLOBECOM 2023 - 2023 IEEE Global Communications Conference 01 (2023) ISBN:979-8-3503-1090-0.

[40] Roozbeh Razavi-Far, Mohammad Meymani, Erfan Mahmoudinia, Dorsa Vazirzade, Peyman Paknezhad, Fateme Ghasemi, Saeed Saravani, Somayeh Nikkhoo, and Kimia Haghjooei, "Quantum adversarial machine learning: from classical adaptations to quantum-native methods", Artificial Intelligence Review 59 8, 177 (2026).

[41] Xianchao Zhu, Yashuang Mu, Xuetao Wang, and William Zhu, "Efficient relation extraction via quantum reinforcement learning", Complex & Intelligent Systems 10 3, 4009 (2024).

[42] Amira Abbas, Andris Ambainis, Brandon Augustino, Andreas Bärtschi, Harry Buhrman, Carleton Coffrin, Giorgio Cortiana, Vedran Dunjko, Daniel J. Egger, Bruce G. Elmegreen, Nicola Franco, Filippo Fratini, Bryce Fuller, Julien Gacon, Constantin Gonciulea, Sander Gribling, Swati Gupta, Stuart Hadfield, Raoul Heese, Gerhard Kircher, Thomas Kleinert, Thorsten Koch, Georgios Korpas, Steve Lenk, Jakub Marecek, Vanio Markov, Guglielmo Mazzola, Stefano Mensa, Naeimeh Mohseni, Giacomo Nannicini, Corey O’Meara, Elena Peña Tapia, Sebastian Pokutta, Manuel Proissl, Patrick Rebentrost, Emre Sahin, Benjamin C. B. Symons, Sabine Tornow, Víctor Valls, Stefan Woerner, Mira L. Wolf-Bauwens, Jon Yard, Sheir Yarkoni, Dirk Zechiel, Sergiy Zhuk, and Christa Zoufal, "Challenges and opportunities in quantum optimization", Nature Reviews Physics 6 12, 718 (2024).

[43] Yize Sun, Mohamad Hagog, Marc Weber, Daniel Hein, Steffen Udluft, Yunpu Ma, and Volker Tresp, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 127 (2025) ISBN:979-8-3315-5736-2.

[44] Rodrigo Coelho, André Sequeira, and Luís Paulo Santos, "VQC-based reinforcement learning with data re-uploading: performance and trainability", Quantum Machine Intelligence 6 2, 53 (2024).

[45] Maja Franz, Lucas Wolf, Maniraman Periyasamy, Christian Ufrecht, Daniel D. Scherer, Axel Plinge, Christopher Mutschler, and Wolfgang Mauerer, "Uncovering instabilities in variational-quantum deep Q-networks", Journal of the Franklin Institute 360 17, 13822 (2023).

[46] Charles Moussa, Yash J. Patel, Vedran Dunjko, Thomas Bäck, and Jan N. van Rijn, "Hyperparameter importance and optimization of quantum neural networks across small datasets", Machine Learning 113 4, 1941 (2024).

[47] Pablo Rodriguez-Grasa, Yue Ban, and Mikel Sanz, "Neural quantum kernels: Training quantum kernels with quantum neural networks", Physical Review Research 7 2, 023269 (2025).

[48] Gyu Seon Kim, Samuel Yen-Chi Chen, Soohyun Park, and Joongheon Kim, Communications in Computer and Information Science 2724, 51 (2026) ISBN:978-981-95-7828-3.

[49] Akash Sinha, Antonio Macaluso, and Matthias Klusch, "Nav-Q: quantum deep reinforcement learning for collision-free navigation of self-driving cars", Quantum Machine Intelligence 7 1, 19 (2025).

[50] M. P. Cuéllar, C. Cano, L. G. B. Ruiz, and L. Servadei, "Time series quantum classifiers with amplitude embedding", Quantum Machine Intelligence 5 2, 45 (2023).

[51] Xianchao Zhu, "Temporally extended successor feature neural episodic control", Scientific Reports 14 1, 15103 (2024).

[52] André Sequeira, Luis Paulo Santos, and Luis Soares Barbosa, "Trainability issues in quantum policy gradients", Machine Learning: Science and Technology 5 3, 035037 (2024).

[53] Austin Braniff, Fengqi You, and Yuhe Tian, Proceedings of the 35th European Symposium on Computer Aided Process Engineering (ESCAPE 35) 4, 1403 (2025) ISBN:9781777940331.

[54] Han Qi, Lei Wang, Hongsheng Zhu, Abdullah Gani, and Changqing Gong, "The barren plateaus of quantum neural networks: review, taxonomy and trends", Quantum Information Processing 22 12, 435 (2023).

[55] Thet Htar Su, Shaswot Shresthamali, and Masaaki Kondo, "Quantum framework for reinforcement learning: Integrating the Markov decision process, quantum arithmetic, and trajectory search", Physical Review A 111 6, 062421 (2025).

[56] K.T. Mpofu and P. Mthunzi-Kufa, 2024 4th International Multidisciplinary Information Technology and Engineering Conference (IMITEC) 240 (2024) ISBN:979-8-3503-8798-8.

[57] Tailong Xiao, Jingzheng Huang, Hongjing Li, Jianping Fan, and Guihua Zeng, "Quantum generative adversarial imitation learning", New Journal of Physics 25 3, 033034 (2023).

[58] Sofiene Jerbi, Lukas J. Fiderer, Hendrik Poulsen Nautrup, Jonas M. Kübler, Hans J. Briegel, and Vedran Dunjko, "Quantum machine learning beyond kernel methods", Nature Communications 14 1, 517 (2023).

[59] Maximilian Zorn, Jonas Stein, Philipp Altmann, Michael Kölle, Claudia Linnhoff-Popien, and Thomas Gabor, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1721 (2024) ISBN:979-8-3315-4137-8.

[60] Bhaskara Narottama, Abrar Ul Haq, James Adu Ansere, Nidhi Simmons, Berk Canberk, Simon L. Cotton, Hyundong Shin, and Trung Q. Duong, "Quantum Deep Reinforcement Learning for Digital Twin-Enabled 6G Networks and Semantic Communications: Considerations for Adoption and Security", IEEE Transactions on Network Science and Engineering 13, 2053 (2026).

[61] El Amine Cherrat, Iordanis Kerenidis, and Anupam Prakash, "Quantum reinforcement learning via policy iteration", Quantum Machine Intelligence 5 2, 30 (2023).

[62] Jorge M. Ramirez, Elaine Wong, Caio Alves, Sarah Chehade, and Ryan Bennink, "Expressiveness of Commutative Quantum Circuits: A Probabilistic Approach", IEEE Transactions on Quantum Engineering 5, 1 (2024).

[63] Min Xu, Yize Tang, Yiliang Wang, Zilong Zhong, Ziheng Xu, Ting Li, and Fei Li, 2024 4th International Conference on Electronic Information Engineering and Computer (EIECT) 664 (2024) ISBN:979-8-3315-2885-0.

[64] Shaojun Wu, Shan Jin, Dingding Wen, Donghong Han, and Xiaoting Wang, "Quantum reinforcement learning in continuous action space", Quantum 9, 1660 (2025).

[65] Zengjing Chen, Lu Wang, and Chengzhi Xing, "A Quantum Technology for Reinforcement Learning on Channel Assignment", Advanced Quantum Technologies 7 8, 2300141 (2024).

[66] Maniraman Periyasamy, Marc Hölle, Marco Wiedmann, Daniel D. Scherer, Axel Plinge, and Christopher Mutschler, 2024 International Joint Conference on Neural Networks (IJCNN) 1 (2024) ISBN:979-8-3503-5931-2.

[67] Diego García-Vega, Fernando Plou Llorente, Alejandro Leal Castaño, Elías F. Combarro, and José Ranilla, "pLazyQML: A parallel package for efficient execution of QML models on classical computers", (2025).

[68] Andrii Kurkin, Jonas Hegemann, Mo Kordzanganeh, and Alexey Melnikov, "Forecasting steam mass flow in power plants using the parallel hybrid network", Engineering Applications of Artificial Intelligence 160, 111912 (2025).

[69] Andrea Skolik, Michele Cattelan, Sheir Yarkoni, Thomas Bäck, and Vedran Dunjko, "Equivariant quantum circuits for learning on weighted graphs", npj Quantum Information 9 1, 47 (2023).

[70] Akshay Ajagekar and Fengqi You, "Variational quantum circuit based demand response in buildings leveraging a hybrid quantum-classical strategy", Applied Energy 364, 123244 (2024).

[71] Chen-Yu Liu, En-Jui Kuo, Chu-Hsuan Abraham Lin, Jason Gemsun Young, Yeong-Jar Chang, Min-Hsiu Hsieh, and Hsi-Sheng Goan, "Quantum-Train: rethinking hybrid quantum-classical machine learning in the model compression perspective", Quantum Machine Intelligence 7 2, 80 (2025).

[72] Michael Kölle, Timo Witter, Tobias Rohe, Gerhard Stenzel, Philipp Altmann, and Thomas Gabor, 2024 IEEE International Conference on Quantum Software (QSW) 157 (2024) ISBN:979-8-3503-6847-5.

[73] Aleksandra Birut and Paweł Ksieniewicz, Lecture Notes in Networks and Systems 2014, 31 (2026) ISBN:978-3-032-27926-2.

[74] Xinliang Wei, Xitong Gao, Kejiang Ye, Cheng-Zhong Xu, and Yu Wang, "A Quantum Reinforcement Learning Approach for Joint Resource Allocation and Task Offloading in Mobile Edge Computing", IEEE Transactions on Mobile Computing 24 4, 2580 (2025).

[75] Georg Kruse, Rodrigo Coelho, Andreas Rosskopf, Robert Wille, and Jeanette Miriam Lorenz, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 1640 (2025) ISBN:979-8-3315-5736-2.

[76] Hrvoje Kukina and Clemens Heitzinger, Communications in Computer and Information Science 2872, 50 (2026) ISBN:978-3-032-17624-0.

[77] Alberto Di Meglio, Karl Jansen, Ivano Tavernelli, Constantia Alexandrou, Srinivasan Arunachalam, Christian W. Bauer, Kerstin Borras, Stefano Carrazza, Arianna Crippa, Vincent Croft, Roland de Putter, Andrea Delgado, Vedran Dunjko, Daniel J. Egger, Elias Fernández-Combarro, Elina Fuchs, Lena Funcke, Daniel González-Cuadra, Michele Grossi, Jad C. Halimeh, Zoë Holmes, Stefan Kühn, Denis Lacroix, Randy Lewis, Donatella Lucchesi, Miriam Lucio Martinez, Federico Meloni, Antonio Mezzacapo, Simone Montangero, Lento Nagano, Vincent R. Pascuzzi, Voica Radescu, Enrique Rico Ortega, Alessandro Roggero, Julian Schuhmacher, Joao Seixas, Pietro Silvi, Panagiotis Spentzouris, Francesco Tacchino, Kristan Temme, Koji Terashi, Jordi Tura, Cenk Tüysüz, Sofia Vallecorsa, Uwe-Jens Wiese, Shinjae Yoo, and Jinglei Zhang, "Quantum Computing for High-Energy Physics: State of the Art and Challenges", PRX Quantum 5 3, 037001 (2024).

[78] Ahmad Alomari and Sathish A. P. Kumar, 2025 IEEE Conference on Artificial Intelligence (CAI) 1375 (2025) ISBN:979-8-3315-2400-5.

[79] Anca Ioana Muscalagiu, 2023 25th International Symposium on Symbolic and Numeric Algorithms for Scientific Computing (SYNASC) 316 (2023) ISBN:979-8-3503-9412-2.

[80] Ying He, Chao Huang, F. Richard Yu, Zhiquan Liu, and Guang Zhou, "Quantum Reinforcement Learning for Cognitive Radio-Assisted UAV Traffic Offloading", IEEE Transactions on Vehicular Technology 74 9, 14620 (2025).

[81] Hsin-Yi Lin, Samuel Yen-Chi Chen, Huan-Hsin Tseng, and Shinjae Yoo, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 241 (2025) ISBN:979-8-3315-5736-2.

[82] Michael Kölle, Alexander Feist, Jonas Stein, Sebastian Wölckert, and Claudia Linnhoff-Popien, Lecture Notes in Computer Science 15906, 275 (2025) ISBN:978-3-031-97634-6.

[83] Sathish Kumar, Temitope Adeniyi, Ahmad Alomari, and Santanu Ganguly, 2023 IEEE International Conference on Quantum Computing and Engineering (QCE) 68 (2023) ISBN:979-8-3503-4323-6.

[84] Louis Schatzki, Martín Larocca, Quynh T. Nguyen, Frédéric Sauvage, and M. Cerezo, "Theoretical guarantees for permutation-equivariant quantum neural networks", npj Quantum Information 10 1, 12 (2024).

[85] Chen-Yu Liu, Samuel Yen-Chi Chen, Kuan-Cheng Chen, Wei-Jia Huang, Wei-Hao Huang, and Yen-Jui Chang, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 1700 (2025) ISBN:979-8-3315-5736-2.

[86] Oliver Sefrin, Manuel Radons, Lars Simon, and Sabine Wölk, "Quantum reinforcement learning in dynamic environments", Quantum Machine Intelligence 8 1, 58 (2026).

[87] Sara Aminpour, Sarah Sharif, and Mike Banad, "Development of hybrid quantum classifiers for realistic classification tasks", Journal of Physics Communications 10 2, 025004 (2026).

[88] Ni Made Erma Pratiwi Astiti and Byung Moo Lee, "A Survey of Quantum Deep Reinforcement Learning for Resource Allocation in Future Wireless Networks", IEEE Communications Surveys & Tutorials 28, 109 (2026).

[89] Aramchehr Zare, Michela Longo, Andrea Di Martino, and Seyedmahdi Miraftabzadeh, 2025 IEEE Vehicle Power and Propulsion Conference (VPPC) 1 (2025) ISBN:979-8-3315-9846-4.

[90] Samuel Yen-Chi Chen, ICASSP 2024 - 2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 13186 (2024) ISBN:979-8-3503-4485-1.

[91] Maja Franz, Tobias Winker, Sven Groppe, and Wolfgang Mauerer, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 409 (2024) ISBN:979-8-3315-4137-8.

[92] Manuel P. Cuéllar, M. C. Pegalajar, and C. Cano, "Automatic evolutionary design of quantum rule-based systems and applications to quantum reinforcement learning", Quantum Information Processing 23 5, 179 (2024).

[93] Hibah Agha, Samuel Yen-Chi Chen, Huan-Hsin Tseng, and Shinjae Yoo, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 1748 (2025) ISBN:979-8-3315-5736-2.

[94] Louyang Yu, Wenbin Yu, Yadang Chen, and Chengjun Zhang, "QCNN-Inspired Variational Circuits for Enhanced Noise Robustness in Quantum Deep Q-Learning", Information 17 3, 250 (2026).

[95] M. Sohaib Alam, Noah F. Berthusen, and Peter P. Orth, "Quantum logic gate synthesis as a Markov decision process", npj Quantum Information 9 1, 108 (2023).

[96] Dirk Heimann, Hans Hohenfeld, Gunnar Schönhoff, Elie Mounzer, and Frank Kirchner, "Learning Fourier series with parametrized quantum circuits", Physical Review Research 7 2, 023151 (2025).

[97] Samuel Yen-Chi Chen, "Asynchronous training of quantum reinforcement learning", Procedia Computer Science 222, 321 (2023).

[98] Francesco Rundo, "Quantum Hyperbolic Deep Learning for Foreign-Exchange Trading: A Hybrid Reinforcement-Learning Pipeline over Attractor-Aware Magnet-Price Manifolds", Big Data and Cognitive Computing 10 6, 191 (2026).

[99] R. Palanivel and P. Muthulakshmi, Lecture Notes in Networks and Systems 997, 409 (2024) ISBN:978-981-97-3241-8.

[100] Maheswari K.P., Kalaiselvi Thiruvenkadam, and Sriramakrishnan Pathmanaban, "A quantum computing approach for brain tumor image classification using amplitude encoding technique", Discover Quantum Science 2 1, 8 (2026).

[101] Haixu Yu, Xudong Zhao, and Chunlin Chen, "Quantum-Inspired Reinforcement Learning for Quantum Control", IEEE Transactions on Control Systems Technology 33 1, 61 (2025).

[102] Sofiene Jerbi, Casper Gyurik, Simon C. Marshall, Riccardo Molteni, and Vedran Dunjko, "Shadows of quantum machine learning", Nature Communications 15 1, 5676 (2024).

[103] Xin Dai, Tzu-Chieh Wei, Shinjae Yoo, and Samuel Yen-Chi Chen, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1525 (2024) ISBN:979-8-3315-4137-8.

[104] Samuel Yen-Chi Chen, Huan-Hsin Tseng, Hsin-Yi Lin, and Shinjae Yoo, 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing Workshops (ICASSPW) 1 (2025) ISBN:979-8-3315-1931-5.

[105] Hung Truong Thanh Nguyen, Truong Thinh Nguyen, and Hung Cao, Proceedings of the 41st ACM/SIGAPP Symposium on Applied Computing 1383 (2026) ISBN:9798400722943.

[106] Yize Sun, Zixin Wu, Volker Tresp, and Yunpu Ma, "Quantum Architecture Search with Unsupervised Representation Learning", Quantum 10, 1994 (2026).

[107] Samuel Yen-Chi Chen, Quantum Computational AI 3 (2026) ISBN:9780443302596.

[108] Shuhong Dai, Nishant Saurabh, Qingle Wang, Jiawei Nian, Shuwen Kan, Ying Mao, and Long Cheng, "Quantum Reinforcement Learning for QoS-Aware Real-Time Job Scheduling in Cloud Systems", IEEE Systems Journal 19 2, 471 (2025).

[109] Samuel Yen-Chi Chen, 2025 IEEE 31st International Symposium on On-Line Testing and Robust System Design (IOLTS) 1 (2025) ISBN:979-8-3315-3334-2.

[110] Junghoon Justin Park, Huan-Hsin Tseng, Shinjae Yoo, Samuel Yen-Chi Chen, and Jiook Cha, 2025 IEEE International Conference on Quantum Artificial Intelligence (QAI) 366 (2025) ISBN:979-8-3315-6986-0.

[111] Samuel Yen-Chi Chen, Huan-Hsin Tseng, Hsin-Yi Lin, and Shinjae Yoo, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 1826 (2025) ISBN:979-8-3315-5736-2.

[112] Javier Lazaro, Juan-Ignacio Vazquez, and Pablo García Bringas, Lecture Notes in Computer Science 16203, 239 (2026) ISBN:978-3-032-08461-3.

[113] Yize Sun, Mohamad Hagog, Marc Weber, Daniel Hein, Steffen Udluft, Yunpu Ma, and Volker Tresp, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 588 (2025) ISBN:979-8-3315-5736-2.

[114] Aramchehr Zare and Mehrdad Boroushaki, "Performance comparison of the quantum and classical deep Q-learning approaches in dynamic environments control", EPJ Quantum Technology 12 1, 74 (2025).

[115] Tong Xie, Taoyong Li, Xinxin Yuan, and Jiacheng Ni, "Quantum Deep Q-Network for Intelligent Packet Routing in 6G Heterogeneous Wireless Networks", Applied Sciences 16 12, 6096 (2026).

[116] Asel Sagingalieva, Mo Kordzanganeh, Andrii Kurkin, Artem Melnikov, Daniil Kuhmistrov, Michael Perelshtein, Alexey Melnikov, Andrea Skolik, and David Von Dollen, "Hybrid quantum ResNet for car classification and its hyperparameter optimization", Quantum Machine Intelligence 5 2, 38 (2023).

[117] Georg Kruse, Rodrigo Coelho, Andreas Rosskopf, Robert Wille, and Jeanette Miriam Lorenz, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1617 (2024) ISBN:979-8-3315-4137-8.

[118] Soohyun Park, Hyunsoo Lee, Seok Bin Son, Soyi Jung, and Joongheon Kim, "Quantum federated learning with pole-angle quantum local training and trainable measurement", Neural Networks 187, 107301 (2025).

[119] Seyed Shakib Vedaie, Archismita Dalal, Eduardo J. Páez, and Barry C. Sanders, "Framework for learning and control in the classical and quantum domains", Annals of Physics 458, 169471 (2023).

[120] Huan-Hsin Tseng, Hsin-Yi Lin, Samuel Yen-Chi Chen, and Shinjae Yoo, 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing Workshops (ICASSPW) 1 (2025) ISBN:979-8-3315-1931-5.

[121] Charles Moussa, Jan N. van Rijn, Thomas Bäck, and Vedran Dunjko, Lecture Notes in Computer Science 13601, 32 (2022) ISBN:978-3-031-18839-8.

[122] Peigen Zeng, Ying He, F. Richard Yu, and Jianbo Du, GLOBECOM 2024 - 2024 IEEE Global Communications Conference 4010 (2024) ISBN:979-8-3503-5125-5.

[123] Li Ding and Lee Spector, "Multi-Objective Evolutionary Architecture Search for Parameterized Quantum Circuits", Entropy 25 1, 93 (2023).

[124] Samuel Yen-Chi Chen and Zhiding Liang, 2025 IEEE International Symposium on Circuits and Systems (ISCAS) 1 (2025) ISBN:979-8-3503-5683-0.

[125] Maniraman Periyasamy, Axel Plinge, Christopher Mutschler, Daniel D. Scherer, and Wolfgang Mauerer, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1504 (2024) ISBN:979-8-3315-4137-8.

[126] Hon Wai Lau, Aoi Hayashi, Akitada Sakurai, William John Munro, and Kae Nemoto, "Modular quantum extreme reservoir computing", Physical Review A 113 1, 012429 (2026).

[127] Nico Meyer, Jakob Murauer, Alexander Popov, Christian Ufrecht, Axel Plinge, Christopher Mutschler, and Daniel D. Scherer, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1458 (2024) ISBN:979-8-3315-4137-8.

[128] Kuan-Cheng Chen, Huan-Hsin Tseng, Samuel Yen-Chi Chen, Chen-Yu Liu, and Kin K. Leung, 2025 IEEE International Conference on Quantum Artificial Intelligence (QAI) 351 (2025) ISBN:979-8-3315-6986-0.

[129] Serge Rainjonneau, Igor Tokarev, Sergei Iudin, Saaketh Rayaprolu, Karan Pinto, Daria Lemtiuzhnikova, Miras Koblan, Egor Barashov, Mo Kordzanganeh, Markus Pflitsch, and Alexey Melnikov, "Quantum Algorithms Applied to Satellite Mission Planning for Earth Observation", IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing 16, 7062 (2023).

[130] Ben Jaderberg, Antonio A. Gentile, Youssef Achari Berrada, Elvira Shishenina, and Vincent E. Elfving, "Let quantum neural networks choose their own frequencies", Physical Review A 109 4, 042421 (2024).

[131] Oliver Sefrin and Sabine Wölk, "A hybrid learning agent for episodic learning tasks with unknown target distance", Quantum Machine Intelligence 7 1, 52 (2025).

[132] André Sequeira, Luis Paulo Santos, and Luis Soares Barbosa, "On Quantum Natural Policy Gradients", IEEE Transactions on Quantum Engineering 5, 1 (2024).

[133] Agustin Silva, Omar Gustavo Zabaleta, and Constancio Miguel Arizmendi, "Maximizing Local Rewards on Multi-Agent Quantum Games through Gradient-Based Learning Strategies", Entropy 25 11, 1484 (2023).

[134] Prateek Karkera and Srivaramangai R, Lecture Notes in Networks and Systems 1047, 421 (2024) ISBN:978-3-031-64835-9.

[135] Samuel Yen-Chi Chen, Chen-Yu Liu, Kuan-Cheng Chen, Wei-Jia Huang, Yen-Jui Chang, and Wei-Hao Huang, 2025 IEEE International Conference on Quantum Computing and Engineering (QCE) 1525 (2025) ISBN:979-8-3315-5736-2.

[136] Yize Sun, Yunpu Ma, and Volker Tresp, 2023 IEEE International Conference on Quantum Computing and Engineering (QCE) 15 (2023) ISBN:979-8-3503-4323-6.

[137] Lokes S, Suresh Reddy, and Samuel Yen Chi Chen, 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing Workshops (ICASSPW) 1 (2025) ISBN:979-8-3315-1931-5.

[138] Manuel P. Cuéllar, "What we can do with one qubit in quantum machine learning: ten classical machine learning problems that can be solved with a single qubit", Quantum Machine Intelligence 6 2, 76 (2024).

[139] Andrea Skolik, Stefano Mangini, Thomas Bäck, Chiara Macchiavello, and Vedran Dunjko, "Robustness of quantum reinforcement learning under hardware errors", EPJ Quantum Technology 10 1, 8 (2023).

[140] Ruilin Liu, Sebastián V. Romero, Izaskun Oregi, Eneko Osaba, Esther Villar-Rodriguez, and Yue Ban, "Digital Quantum Simulation and Circuit Learning for the Generation of Coherent States", Entropy 24 11, 1529 (2022).

[141] Samuel Yen-Chi Chen and Joongheon Kim, Proceedings of the 33rd ACM International Conference on Information and Knowledge Management 5507 (2024) ISBN:9798400704369.

[142] Ahmad Alomari and Sathish A. P. Kumar, "GPA: Grover Policy Agent for Generating Optimal Quantum Sensor Circuits", IEEE Transactions on Artificial Intelligence 6 10, 2722 (2025).

[143] Irwindeep Singh, Sukhpal Singh Gill, Jinzhao Sun, and Jan Mol, "QAISim: a toolkit for modeling and simulation of AI in quantum cloud computing environments", Cluster Computing 29 2, 99 (2026).

[144] Tiffany Khou and Don Roosan, Quantum Computing in Medicine 343 (2026) ISBN:9780443341939.

[145] Chi-Sheng Chen and En-Jui Kuo, ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 22422 (2026) ISBN:979-8-3315-6701-9.

[146] Gyu Seon Kim, Samuel Yen-Chi Chen, Soohyun Park, and Joongheon Kim, ICASSP 2025 - 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 1 (2025) ISBN:979-8-3503-6874-1.

[147] Laura María Donaire, Gloria Ortega, Francisco Orts, Ester Martín Garzón, and Ernestas Filatovas, "A hybrid quantum-classical approach for liver disease detection using quantum machine learning", Engineering Applications of Artificial Intelligence 164, 113240 (2026).

[148] Dang Van Huynh, Sasinda C. Prabhashana, Hyundong Shin, and Trung Q. Duong, 2025 IEEE Globecom Workshops (GC Wkshps) 915 (2025) ISBN:979-8-3315-6741-5.

[149] Corey Trahan, Mark Loveland, and Samuel Dent, "Quantum Physics-Informed Neural Networks", Entropy 26 8, 649 (2024).

[150] Alberto B. de Palhares, Joab M. Varela, Diogo H. G. Duarte, and Rafael Chaves, "Exploring Clifford and near-Clifford initialization strategies for quantum variational optimization", International Journal of Modern Physics C 2643005 (2026).

[151] Yichen Xie, 2025 International Joint Conference on Neural Networks (IJCNN) 1 (2025) ISBN:979-8-3315-1042-8.

[152] A. Sannia, A. Giordano, N. Lo Gullo, C. Mastroianni, and F. Plastina, "A hybrid classical-quantum approach to speed-up Q-learning", Scientific Reports 13 1, 3913 (2023).

[153] Samuel Yen-Chi Chen, 2024 International Joint Conference on Neural Networks (IJCNN) 1 (2024) ISBN:979-8-3503-5931-2.

[154] Wenhan Yu and Jun Zhao, 2023 International Conference on Computer and Applications (ICCA) 1 (2023) ISBN:979-8-3503-0325-4.

[155] Samuel Yen-Chi Chen and Kuan-Cheng Chen, 2025 International Joint Conference on Neural Networks (IJCNN) 1 (2025) ISBN:979-8-3315-1042-8.

[156] James Chao, Ramiro Rodriguez, and Sean Crowe, Proceedings of the Companion Conference on Genetic and Evolutionary Computation 2179 (2023) ISBN:9798400701207.

[157] Eva Andrés, M. P. Cuéllar, and G. Navarro, "Efficient Dimensionality Reduction Strategies for Quantum Reinforcement Learning", IEEE Access 11, 104534 (2023).

[158] Ahmad Alomari and Sathish A. P. Kumar, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 1364 (2024) ISBN:979-8-3315-4137-8.

[159] Tomoaki Kimura, Kodai Shiba, Chih-Chieh Chen, Masaru Sogabe, Katsuyoshi Sakamoto, and Tomah Sogabe, "Quantum circuit architectures via quantum observable Markov decision process planning", Journal of Physics Communications 6 7, 075006 (2022).

[160] Samuel Yen-Chi Chen, ICASSP 2023 - 2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 1 (2023) ISBN:978-1-7281-6327-7.

[161] Yaofu Liu, Chang Xu, and Siyuan Jin, 2023 IEEE International Conference on Quantum Software (QSW) 83 (2023) ISBN:979-8-3503-0479-4.

[162] André Sequeira, Luis Paulo Santos, and Luis Soares Barbosa, 2024 IEEE International Conference on Quantum Computing and Engineering (QCE) 335 (2024) ISBN:979-8-3315-4137-8.

[163] Manjunath T D and Biswajit Bhowmik, 2024 IEEE 4th International Conference on VLSI Systems, Architecture, Technology and Applications (VLSI SATA) 1 (2024) ISBN:979-8-3503-6226-8.

[164] Manqoba Q. Hlatshwayo, Manav Babel, Dalila Islas-Sanchez, and Konstantinos Georgopoulos, "A Technical Review of Quantum Computing Use Cases for Finance and Economics", Quantum Reports 8 1, 26 (2026).

[165] Ahmad Alomari and Sathish A. P. Kumar, "ReLAQA: Reinforcement Learning-Based Autonomous Quantum Agent for Quantum Applications", IEEE Transactions on Artificial Intelligence 6 3, 549 (2025).

[166] Michael Broughton, Guillaume Verdon, Trevor McCourt, Antonio J. Martinez, Jae Hyeon Yoo, Sergei V. Isakov, Philip Massey, Ramin Halavati, Murphy Yuezhen Niu, Alexander Zlokapa, Evan Peters, Owen Lockwood, Andrea Skolik, Sofiene Jerbi, Vedran Dunjko, Martin Leib, Michael Streif, David Von Dollen, Hongxiang Chen, Shuxiang Cao, Roeland Wiersema, Hsin-Yuan Huang, Jarrod R. McClean, Ryan Babbush, Sergio Boixo, Dave Bacon, Alan K. Ho, Hartmut Neven, and Masoud Mohseni, "TensorFlow Quantum: A Software Framework for Quantum Machine Learning", arXiv:2003.02989, (2020).

[167] Quynh T. Nguyen, Louis Schatzki, Paolo Braccia, Michael Ragone, Patrick J. Coles, Frédéric Sauvage, Martín Larocca, and M. Cerezo, "Theory for Equivariant Quantum Neural Networks", PRX Quantum 5 2, 020328 (2024).

[168] Anna Dawid, Julian Arnold, Borja Requena, Alexander Gresch, Marcin Płodzień, Kaelan Donatella, Kim A. Nicoli, Paolo Stornati, Rouven Koch, Miriam Büttner, Robert Okuła, Gorka Muñoz-Gil, Rodrigo A. Vargas-Hernández, Alba Cervera-Lierta, Juan Carrasquilla, Vedran Dunjko, Marylou Gabrié, Patrick Huembeli, Evert van Nieuwenburg, Filippo Vicentini, Lei Wang, Sebastian J. Wetzel, Giuseppe Carleo, Eliška Greplová, Roman Krems, Florian Marquardt, Michał Tomza, Maciej Lewenstein, and Alexandre Dauphin, "Modern applications of machine learning in quantum sciences", arXiv:2204.04198, (2022).

[169] En-Jui Kuo, Yao-Lung L. Fang, and Samuel Yen-Chi Chen, "Quantum Architecture Search via Deep Reinforcement Learning", arXiv:2104.07715, (2021).

[170] Sofiene Jerbi, Casper Gyurik, Simon C. Marshall, Hans J. Briegel, and Vedran Dunjko, "Parametrized quantum policies for reinforcement learning", arXiv:2103.05577, (2021).

[171] Su Yeon Chang and M. Cerezo, "A Primer on Quantum Machine Learning", arXiv:2511.15969, (2025).

[172] Owen Lockwood, "An Empirical Review of Optimization Techniques for Quantum Variational Circuits", arXiv:2202.01389, (2022).

[173] Esther Ye and Samuel Yen-Chi Chen, "Quantum Architecture Search via Continual Reinforcement Learning", arXiv:2112.05779, (2021).

[174] Samuel Yen-Chi Chen, Chih-Min Huang, Chia-Wei Hsing, Hsi-Sheng Goan, and Ying-Jer Kao, "Variational quantum reinforcement learning via evolutionary optimization", Machine Learning: Science and Technology 3 1, 015025 (2022).

[175] Qingfeng Lan, "Variational Quantum Soft Actor-Critic", arXiv:2112.11921, (2021).

[176] Simon Wiedemann, Daniel Hein, Steffen Udluft, and Christian Mendl, "Quantum Policy Iteration via Amplitude Estimation and Grover Search -- Towards Quantum Advantage for Reinforcement Learning", arXiv:2206.04741, (2022).

[177] Arjan Cornelissen and Sofiene Jerbi, "Quantum algorithms for multivariate Monte Carlo estimation", arXiv:2107.03410, (2021).

[178] Zhihao Cheng, Kaining Zhang, Li Shen, and Dacheng Tao, "Quantum Imitation Learning", arXiv:2304.02480, (2023).

[179] Gilberto Cunha, Alexandra Ramôa, André Sequeira, Michael de Oliveira, and Luís Barbosa, "Quantum Bayesian Networks Can Speed up Reinforcement Learning in Partially Observable Environments", arXiv:2507.18606, (2025).

[180] Christa Zoufal, "Generative Quantum Machine Learning", arXiv:2111.12738, (2021).

[181] Jen-Yueh Hsiao, Yuxuan Du, Wei-Yin Chiang, Min-Hsiu Hsieh, and Hsi-Sheng Goan, "Unentangled quantum reinforcement learning agents in the OpenAI Gym", arXiv:2203.14348, (2022).

[182] Arjan Cornelissen, Yassine Hamoudi, and Sofiene Jerbi, "Near-Optimal Quantum Algorithms for Multivariate Mean Estimation", arXiv:2111.09787, (2021).

[183] Nancy Barraza, Gabriel Alvarado Barrios, Jie Peng, Lucas Lamata, Enrique Solano, and Francisco Albarrán-Arriagada, "Analog quantum approximate optimization algorithm", Quantum Science and Technology 7 4, 045035 (2022).

[184] M. Cerezo, Guillaume Verdon, Hsin-Yuan Huang, Lukasz Cincio, and Patrick J. Coles, "Challenges and Opportunities in Quantum Machine Learning", arXiv:2303.09491, (2023).

[185] Jun Qi, Chao-Han Huck Yang, and Pin-Yu Chen, "QTN-VQC: An End-to-End Learning framework for Quantum Neural Networks", arXiv:2110.03861, (2021).

[186] El Amine Cherrat, Iordanis Kerenidis, and Anupam Prakash, "Quantum Reinforcement Learning via Policy Iteration", arXiv:2203.01889, (2022).

[187] Nico Meyer, Christian Ufrecht, George Yammine, Georgios Kontes, Christopher Mutschler, and Daniel D. Scherer, "Benchmarking Quantum Reinforcement Learning", arXiv:2501.15893, (2025).

[188] Dominik Freinberger, Julian Lemmel, Radu Grosu, and Sofiene Jerbi, "A quantum-classical reinforcement learning model to play Atari games", arXiv:2412.08725, (2024).

[189] Alexander DeRieux and Walid Saad, "eQMARL: Entangled Quantum Multi-Agent Reinforcement Learning for Distributed Cooperation over Quantum Channels", arXiv:2405.17486, (2024).

[190] Dániel Nagy, Zsolt Tabi, Péter Hága, Zsófia Kallus, and Zoltán Zimborás, "Photonic Quantum Policy Learning in OpenAI Gym", arXiv:2108.12926, (2021).

[191] Maniraman Periyasamy, Axel Plinge, Christopher Mutschler, Daniel D. Scherer, and Wolfgang Mauerer, "Guided-SPSA: Simultaneous Perturbation Stochastic Approximation assisted by the Parameter Shift Rule", arXiv:2404.15751, (2024).

[192] Maicon Pierre Lourenço, Mosayeb Naseri, Lizandra Barrios Herrera, Hadi Zadeh-Haghighi, Daya Gaur, Christoph Simon, and Dennis R. Salahub, "Quantum Active Learning for Structural Determination of Doped Nanoparticles -- a Case Study of 4Al@Si$_{11}$", arXiv:2412.00504, (2024).

[193] Maniraman Periyasamy, Nico Meyer, Christian Ufrecht, Daniel D. Scherer, Axel Plinge, and Christopher Mutschler, "Incremental Data-Uploading for Full-Quantum Classification", arXiv:2205.03057, (2022).

[194] Manuel Schönberger, Maja Franz, Stefanie Scherzinger, and Wolfgang Mauerer, "Peel $\mid$ Pile? Cross-Framework Portability of Quantum Software", arXiv:2203.06289, (2022).

[195] Oliver Sefrin, Manuel Radons, Lars Simon, and Sabine Wölk, "Quantum reinforcement learning in dynamic environments", arXiv:2507.01691, (2025).

[196] Hannah Helgesen, Michael Felsberg, and Jan-Åke Larsson, "Certainty In, Certainty Out: REVQCs for Quantum Machine Learning", arXiv:2310.10629, (2023).

[197] Brian Coyle, "Machine learning applications for noisy intermediate-scale quantum computers", arXiv:2205.09414, (2022).

[198] Oliver Sefrin and Sabine Wölk, "A hybrid learning agent for episodic learning tasks with unknown target distance", arXiv:2412.13686, (2024).

[199] Charles Moussa, Jan N. van Rijn, Thomas Bäck, and Vedran Dunjko, "Hyperparameter Importance of Quantum Neural Networks Across Small Datasets", arXiv:2206.09992, (2022).

[200] Silvirianti, Bhaskara Narottama, and Soo Young Shin, "Layerwise Quantum Deep Reinforcement Learning for Joint Optimization of UAV Trajectory and Resource Allocation", IEEE Internet of Things Journal 11 1, 430 (2024).

[201] Irwindeep Singh, Sukhpal Singh Gill, Jinzhao Sun, and Jan Mol, "QAISim: A Toolkit for Modeling and Simulation of AI in Quantum Cloud Computing Environments", arXiv:2512.17918, (2025).

[202] Zhihao Cheng, Kaining Zhang, Li Shen, and Dacheng Tao, "Quantum Imitation Learning", IEEE Transactions on Neural Networks and Learning Systems 35 10, 14190 (2024).

The above citations are from Crossref's cited-by service (last updated successfully 2026-07-15 08:43:14) and SAO/NASA ADS (last updated successfully 2026-07-14 20:25:32). The list may be incomplete as not all publishers provide suitable and complete citation data.

Could not fetch ADS cited-by data during last attempt 2026-07-15 08:43:14: Cannot retrieve data from ADS due to rate limitations.