Station holding in flowing water is a fundamental capability for any aquatic robot, yet it remains stubbornly difficult for robotic fish. The core challenge: unknown, turbulent flows create nonlinear fluid-structure interactions that defy simple modeling. A new paper from researchers at Peking University and the Chinese Academy of Sciences proposes a practical solution that could accelerate the deployment of underwater robots in real-world environments.
The team introduces SWiFT (Swimming With Flow Toolbox), a framework that combines a free-swimming flow-tank experimental setup, a highly efficient CFD-based simulator, and a systematic sim-to-real transfer pipeline. Using reinforcement learning, they train an egocentric station-holding policy for a body and/or caudal fin (BCF) robotic fish. The key insight: the policy relies solely on onboard sensing of the robot's own state (egocentric feedback) without any explicit flow sensing. This mirrors the biological phenomenon of rheotaxis, where fish sense flow through their body rather than dedicated sensors.
The results are striking. Across all metrics, the SWiFT-trained policy substantially outperforms state-of-the-art methods, most notably in root-mean-square error (RMSE) of distance. The robot can approach a target position and hold station in unknown turbulent background flows, a task that has previously required external references or flow measurements.
For founders and operators building underwater robots, this work offers a clear path to reducing sensor costs and complexity. Many real-world applications—environmental monitoring, infrastructure inspection, search and rescue—require robots to maintain position in currents. Current solutions often rely on expensive flow sensors or external positioning systems (e.g., acoustic beacons). SWiFT suggests that a well-trained control policy, combined with standard onboard sensors (IMU, camera), may be sufficient.
The sim-to-real pipeline is particularly noteworthy. The team built a CFD-based simulator that is both physically consistent and computationally efficient, enabling rapid RL training. They then transferred the policy to a physical robotic fish with minimal degradation. This approach is directly applicable to other underwater platforms, from gliders to AUVs, provided the dynamics are well-characterized in simulation.
However, the paper does not address long-duration station holding or extreme turbulence. The experiments were conducted in a controlled flow tank; real rivers, oceans, and tidal zones present additional challenges (e.g., waves, debris, variable density). Founders should view this as a proof of concept for a control architecture, not a drop-in solution.
From a commercial perspective, the SWiFT framework could be productized as a software module for underwater robot control systems. Startups specializing in autonomous underwater vehicles (AUVs) for offshore energy, aquaculture, or defense could integrate this approach to improve station-keeping performance without hardware upgrades. The egocentric nature also means the system is robust to sensor failure—a critical advantage in harsh marine environments.
The research also opens questions about generalizability. Can the same RL approach be applied to different fish morphologies or to non-fish underwater robots? The authors suggest SWiFT is a foundation for tackling complex swimming tasks, implying a platform play. For investors, the team's track record in robotic fish control (previous work on swimming efficiency and maneuverability) adds credibility.
In summary, SWiFT demonstrates that egocentric station holding in unknown turbulent flows is achievable without flow sensors, using RL and sim-to-real transfer. This reduces cost and complexity for underwater robots, bringing them closer to real-world deployment. The next step for the field: testing in open water and extending to multi-robot coordination.
Source: arXiv.