Realizing a deep reinforcement learning agent discovering real-time feedback control strategies for a quantum system

强化学习 计算机科学 延迟(音频) 人工智能 电信
作者
Kevin Reuer,Jonas Landgraf,Thomas Fösel,James O’Sullivan,Liberto Beltrán,Abdulkadir Akın,Graham J. Norris,Ants Remm,Michael Kerschbaum,Jean-Claude Besse,Florian Marquardt,Andreas Wallraff,Christopher Eichler
出处
期刊:Cornell University - arXiv [Cornell University]
被引量:5
标识
DOI:10.48550/arxiv.2210.16715
摘要

To realize the full potential of quantum technologies, finding good strategies to control quantum information processing devices in real time becomes increasingly important. Usually these strategies require a precise understanding of the device itself, which is generally not available. Model-free reinforcement learning circumvents this need by discovering control strategies from scratch without relying on an accurate description of the quantum system. Furthermore, important tasks like state preparation, gate teleportation and error correction need feedback at time scales much shorter than the coherence time, which for superconducting circuits is in the microsecond range. Developing and training a deep reinforcement learning agent able to operate in this real-time feedback regime has been an open challenge. Here, we have implemented such an agent in the form of a latency-optimized deep neural network on a field-programmable gate array (FPGA). We demonstrate its use to efficiently initialize a superconducting qubit into a target state. To train the agent, we use model-free reinforcement learning that is based solely on measurement data. We study the agent's performance for strong and weak measurements, and for three-level readout, and compare with simple strategies based on thresholding. This demonstration motivates further research towards adoption of reinforcement learning for real-time feedback control of quantum devices and more generally any physical system requiring learnable low-latency feedback control.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
赖梦婷发布了新的文献求助10
刚刚
深情安青应助温柔画笔采纳,获得10
刚刚
刚刚
天真的耳机完成签到,获得积分10
1秒前
幻影猫应助木子倪采纳,获得10
1秒前
Yeah完成签到,获得积分20
2秒前
2秒前
3秒前
坚强枫发布了新的文献求助10
3秒前
顽主完成签到,获得积分0
3秒前
NexusExplorer应助热心的白枫采纳,获得10
3秒前
3秒前
敏感思山发布了新的文献求助10
4秒前
FashionBoy应助书记采纳,获得10
4秒前
5秒前
所所应助xuan采纳,获得10
6秒前
6秒前
阚月发布了新的文献求助10
6秒前
香蕉觅云应助赵振辉采纳,获得10
7秒前
小白发布了新的文献求助10
9秒前
大模型应助xia采纳,获得10
9秒前
VibraYu完成签到,获得积分10
9秒前
地球发布了新的文献求助10
10秒前
仁青完成签到,获得积分10
10秒前
洛城l完成签到 ,获得积分10
10秒前
10秒前
乱泽华完成签到,获得积分10
11秒前
施宇宙完成签到 ,获得积分10
12秒前
忐忑的马里奥完成签到,获得积分10
12秒前
12秒前
柇荟完成签到,获得积分10
14秒前
Yeah关注了科研通微信公众号
14秒前
欣慰的亦绿完成签到,获得积分10
15秒前
王一帆发布了新的文献求助30
16秒前
月月完成签到,获得积分10
16秒前
16秒前
冷傲的秋天完成签到,获得积分10
17秒前
17秒前
Wang发布了新的文献求助10
17秒前
18秒前
高分求助中
Markov Chain Monte Carlo 10000
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Common Foundations of American and East Asian Modernisation: From Alexander Hamilton to Junichero Koizumi 5000
How to Use Machine Learning in Chemistry: An Introduction 1000
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
Discerning Saints: Moralization of Intrinsic Motivation and Selective Prosociality at Work 500
Handbuch Trainingswissenschaft – Trainingslehre 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7581796
求助须知:如何正确求助?哪些是违规求助? 9160864
关于积分的说明 19600778
捐赠科研通 7164068
什么是DOI,文献DOI怎么找? 3266010
关于科研通互助平台的介绍 2430947
邀请新用户注册赠送积分活动 2257181