Export citation

Export citation

Choose format for download:

Download Citation

    Optimal navigation in two-dimensional flows: Control theory and reinforcement learning

    Vladimir Parfenyev*

    • *Contact author: parfenius@gmail.com

    Phys. Rev. E 114, 015104 – Published 17 July, 2026

    DOI: https://doi.org/10.1103/8gbr-vp2r

    Abstract

    Zermelo's navigation problem seeks the trajectory of minimal travel time between two points in a fluid flow. We address this problem for an agent—such as a floating drone or active particle—that is advected by a two-dimensional flow, self-propels at a fixed speed smaller than or comparable to the characteristic flow velocity, and can steer its direction. The flows considered span increasing levels of complexity, from steady solid-body rotation and time-dependent sink-vortex to the Taylor-Green flow and turbulence in the inverse energy cascade regime. Although optimal-control theory provides time-minimizing trajectories, these solutions become unstable in chaotic regimes characterized by positive finite-time Lyapunov exponents. To design robust navigation strategies, we apply reinforcement learning and compare Q-learning with a one-step actor-critic algorithm. Both methods achieve successful navigation, yielding mean travel times within 3–10% of optimal-control solutions in regular flows, while the discrepancy increases to 35–75% in time-dependent turbulent flows. Finally, we show that agents trained on coarse-grained turbulent flows generalize to the full velocity field. This robustness to incomplete flow information is essential for practical navigation in real-world oceanic and atmospheric environments.

    Physics Subject Headings (PhySH)

    Authorization Required

    We need you to provide your credentials before accessing this content.

    Supplemental Material (Subscription Required)

    References (Subscription Required)

    Outline

    Information

    Sign In to Your Journals Account

    Filter

    Filter

    Article Lookup

    Enter a citation