Interferobot: aligning an optical interferometer by a reinforcement learning agent
Dmitry Igorevich Sorokin, Alexander E. Ulanov, Ekaterina A. Sazhina, Alexander I. Lvovsky
摘要
Limitations in acquiring training data restrict potential applications of deep reinforcement learning (RL) methods to the training of real-world robots. Here we train an RL agent to align a Mach-Zehnder interferometer, which is an essential part of many optical experiments, based on images of interference fringes acquired by a monocular camera. The agent is trained in a simulated environment, without any hand-coded features or a priori information about the physics, and subsequently transferred to a physical interferometer. Thanks to a set of domain randomizations simulating uncertainties in physical measurements, the agent successfully aligns this interferometer without any fine tuning, achieving a performance level of a human expert.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Reinforcement learning for optimization of variational quantum circuit architecturesMateusz Ostaszewski, Lea M. Trenkwalder, Wojciech Masarczyk, Eleanor Scerri 等NeurIPS 2021 · 被引用 204 次
- Bayesian Optimization with High-Dimensional OutputsWesley J. Maddox, Maximilian Balandat, Andrew Gordon Wilson, Eytan BakshyNeurIPS 2021 · 被引用 75 次
- A Study of Bayesian Neural Network Surrogates for Bayesian OptimizationYucen Lily Li, Tim G. J. Rudner, Andrew Gordon WilsonICLR 2024 · 被引用 59 次
- Trajectory-Level Data Augmentation for Offline Reinforcement LearningTobias Schmähling, Matthias Burkhardt, Tobias WindischICML 2026
相关 Paper
- Understanding Domain Randomization for Sim-to-real TransferXiaoyu Chen, Jiachen Hu, Chi Jin, Lihong Li 等ICLR 2022 · 被引用 164 次
- Learning-based Optimisation of Particle Accelerators Under Partial Observability Without Real-World TrainingJan Kaiser, Oliver Stein, Annika EichlerICML 2022 · 被引用 22 次
- Hand-Object Interaction Controller (HOIC): Deep Reinforcement Learning for Reconstructing Interactions with PhysicsHaoyu Hu, Xinyu Yi, Zhe Cao, Jun-Hai Yong 等SIGGRAPH 2024 · 被引用 2 次
- Provably sample-efficient RL with side information about latent dynamicsYao Liu, Dipendra Misra, Miro Dudík, Robert E. SchapireNeurIPS 2022 · 被引用 2 次
- Can Vision Language Models Learn Intuitive Physics from Interaction?Luca M. Schulze Buschoff, Konstantinos Voudouris, Can Demircan, Eric SchulzICML 2026
