EEmbodied AI Hub
PapersEditor’s pick

Robotic Fish Holds Station in Turbulent Flow Without Flow Sensors: A Bio-Inspired Breakthrough

Why it mattersEnables robotic fish to perform long-duration tasks like environmental monitoring or inspection in natural water bodies where GPS is unavailable and flow conditions are unpredictable.

A new RL-based framework, SWiFT, enables a BCF robotic fish to hold station in unknown turbulent flows using only egocentric feedback (no flow sensors). The method achieves significant RMSE improvements over prior art and mirrors biological rheotaxis, advancing real-world deployment of underwater robots.

reinforcement learningunderwater robotstation holdingrobotic fishsim-to-realrheotaxisresearchpaper

Open paper on arXiv (arXiv:2607.24860)

Source: arXiv · July 30, 2026

Share this article so more people can see it

Station holding in flowing water is a fundamental capability for any aquatic robot, yet it remains stubbornly difficult for robotic fish. The core challenge: unknown, turbulent flows create nonlinear fluid-structure interactions that defy simple modeling. A new paper from researchers at Peking University and the Chinese Academy of Sciences proposes a practical solution that could accelerate the deployment of underwater robots in real-world environments.

The team introduces SWiFT (Swimming With Flow Toolbox), a framework that combines a free-swimming flow-tank experimental setup, a highly efficient CFD-based simulator, and a systematic sim-to-real transfer pipeline. Using reinforcement learning, they train an egocentric station-holding policy for a body and/or caudal fin (BCF) robotic fish. The key insight: the policy relies solely on onboard sensing of the robot's own state (egocentric feedback) without any explicit flow sensing. This mirrors the biological phenomenon of rheotaxis, where fish sense flow through their body rather than dedicated sensors.

The results are striking. Across all metrics, the SWiFT-trained policy substantially outperforms state-of-the-art methods, most notably in root-mean-square error (RMSE) of distance. The robot can approach a target position and hold station in unknown turbulent background flows, a task that has previously required external references or flow measurements.

For founders and operators building underwater robots, this work offers a clear path to reducing sensor costs and complexity. Many real-world applications—environmental monitoring, infrastructure inspection, search and rescue—require robots to maintain position in currents. Current solutions often rely on expensive flow sensors or external positioning systems (e.g., acoustic beacons). SWiFT suggests that a well-trained control policy, combined with standard onboard sensors (IMU, camera), may be sufficient.

The sim-to-real pipeline is particularly noteworthy. The team built a CFD-based simulator that is both physically consistent and computationally efficient, enabling rapid RL training. They then transferred the policy to a physical robotic fish with minimal degradation. This approach is directly applicable to other underwater platforms, from gliders to AUVs, provided the dynamics are well-characterized in simulation.

However, the paper does not address long-duration station holding or extreme turbulence. The experiments were conducted in a controlled flow tank; real rivers, oceans, and tidal zones present additional challenges (e.g., waves, debris, variable density). Founders should view this as a proof of concept for a control architecture, not a drop-in solution.

From a commercial perspective, the SWiFT framework could be productized as a software module for underwater robot control systems. Startups specializing in autonomous underwater vehicles (AUVs) for offshore energy, aquaculture, or defense could integrate this approach to improve station-keeping performance without hardware upgrades. The egocentric nature also means the system is robust to sensor failure—a critical advantage in harsh marine environments.

The research also opens questions about generalizability. Can the same RL approach be applied to different fish morphologies or to non-fish underwater robots? The authors suggest SWiFT is a foundation for tackling complex swimming tasks, implying a platform play. For investors, the team's track record in robotic fish control (previous work on swimming efficiency and maneuverability) adds credibility.

In summary, SWiFT demonstrates that egocentric station holding in unknown turbulent flows is achievable without flow sensors, using RL and sim-to-real transfer. This reduces cost and complexity for underwater robots, bringing them closer to real-world deployment. The next step for the field: testing in open water and extending to multi-robot coordination.

Source: arXiv.

Related resources on this hub

Jump to projects, models, or datasets mentioned or closely related.

Discussion

Tell us what you think — comments make stories more useful for builders and founders.

Tell us what you think!

Robotic Fish Holds Station in Turbulent Flow Without Flow Sensors: A Bio-Inspired Breakthrough

Have an account? Log in to use your display name and avatar.

Email is optional and never shown on the page.

More insights

PapersarXiv

Towards Trustworthy Embodied Intelligence: A Systems Framework and Graded Trustworthiness Levels

As embodied AI systems enter real-world applications, failures can cause physical harm. This work provides a structured approach to assess and ensure trustworthiness, which is critical for adoption in safety-critical domains.

A systems framework for trustworthy embodied intelligence with graded trustworthiness levels, emphasizing safety, reliability, and ethical compliance beyond task completion.

Read
PapersarXiv

Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation

This work challenges the need for complex task-specific workflows or large-scale training, suggesting that simple, general-purpose agents can achieve competitive performance in embodied navigation tasks.

This paper introduces agentic embodied control, where a general-purpose agent maintains the decision loop itself, and shows that minimal-interface zero-shot agents can rival industrial-scale policies in vision-and-language navigation.

Read