过度拟合
辍学(神经网络)
计算机科学
人工智能
一般化
机器学习
人工神经网络
量子
量子位元
量子计算机
理论计算机科学
数学
量子力学
数学分析
物理
作者
Francesco Scala,Andrea Ceschini,Massimo Panella,Dario Gerace
标识
DOI:10.1002/qute.202300220
摘要
Abstract In classical machine learning (ML), “overfitting” is the phenomenon occurring when a given model learns the training data excessively well, and it thus performs poorly on unseen data. A commonly employed technique in ML is the so called “dropout,” which prevents computational units from becoming too specialized, hence reducing the risk of overfitting. With the advent of quantum neural networks (QNNs) as learning models, overfitting might soon become an issue, owing to the increasing depth of quantum circuits as well as multiple embedding of classical features, which are employed to give the computational nonlinearity. Here, a generalized approach is presented to apply the dropout technique in QNN models, defining and analyzing different quantum dropout strategies to avoid overfitting and achieve a high level of generalization. This study allows to envision the power of quantum dropout in enabling generalization, providing useful guidelines on determining the maximal dropout probability for a given model, based on overparametrization theory. It also highlights how quantum dropout does not impact the features of the QNN models, such as expressibility and entanglement. All these conclusions are supported by extensive numerical simulations and may pave the way to efficiently employing deep quantum machine learning (QML) models based on state‐of‐the‐art QNNs.
科研通智能强力驱动
Strongly Powered by AbleSci AI