EEmbodied AI Hub
PapersEditor’s pick

Corrigible Assistance in One Round: Pragmatic-Pedagogic Best Response

Why it mattersThis work makes assistance games computationally feasible for real-time robotics, enabling robots to correct their behavior based on human feedback in a single interaction, which is crucial for safe and effective human-robot collaboration.

This paper introduces a tractable approach for assistance games, where a robot assists a human with asymmetric information. The authors propose a pragmatic-pedagogic best response strategy that achieves corrigible assistance in a single round, avoiding the intractability of general POMDP planning.

corrigible assistancepragmatic reasoningpedagogic signalingassistance gamesresearchPOMDPpaperarxiv

Open paper on arXiv (arXiv:2607.27508)

Source: arXiv · August 1, 2026

Share this article so more people can see it

Assistance games formalize human-robot collaboration under asymmetric information, where the human knows the goal but the robot must infer it from observations and interactions. However, computing optimal strategies online is generally intractable because it requires planning in a POMDP.

The authors identify a class of assistance games where a pragmatic-pedagogic best response can be computed efficiently. This strategy balances the robot's pragmatic reasoning about human behavior with pedagogic signaling to elicit useful feedback.

In a single round of interaction, the robot can achieve corrigible assistance, meaning it can adjust its behavior based on human corrections without requiring full POMDP solving. This provides a practical approximation to optimal assistance.

The approach is validated through theoretical analysis and simulations, demonstrating that it outperforms baseline methods in terms of task success and correction efficiency.

This work opens the door to deploying assistance game frameworks in real-time robotic systems, enhancing adaptability and safety in collaborative tasks.

Source: arXiv (2607.27508).

Related resources on this hub

Jump to projects, models, or datasets mentioned or closely related.

Discussion

Tell us what you think — comments make stories more useful for builders and founders.

Tell us what you think!

Corrigible Assistance in One Round: Pragmatic-Pedagogic Best Response

Have an account? Log in to use your display name and avatar.

Email is optional and never shown on the page.

More insights

PapersarXiv

Data-Efficient Robot Imitation: Prioritizing Counterfactual Sensitivity Over More Demos

This work addresses a critical bottleneck in robot imitation learning: the fragility of policies to minor visual changes. By focusing on counterfactual action sensitivity, it offers a data-efficient path to robust policies, potentially accelerating real-world deployment.

A new arXiv paper challenges the assumption that more demonstrations always improve robot imitation learning robustness. The proposed CFNBC framework selects a compact repair set by measuring action drift under task-preserving nuisances, achieving strong performance with only 20-30 selected candidates. This data-centric approach offers a cost-effective strategy for startups to enhance policy robustness without extensive data collection.

Read
PapersarXiv

Hybrid PBD-MPM Suturing Simulator Hits 80% Needle Insertion Success for Surgical Robot RL

Surgical robotics requires high-fidelity simulators that can handle diverse objects (rigid tools, soft tissue, fluids). This work addresses the gap by integrating PBD and MPM in a unified framework, potentially accelerating robot learning for suturing tasks.

A new surgical suturing simulator combines Position-Based Dynamics for sutures and the Material Point Method for soft tissue, enabling two-way contact with friction and drag. Trained with ML-Agents, RL agents achieve 80% and 68% success in needle insertion and extraction under strict thresholds. This hybrid approach offers a scalable, GPU-optimized environment for autonomous surgical subtasks, signaling a shift toward more realistic training data for surgical robotics.

Read
PapersarXiv

Towards Trustworthy Embodied Intelligence: A Systems Framework and Graded Trustworthiness Levels

As embodied AI systems enter real-world applications, failures can cause physical harm. This work provides a structured approach to assess and ensure trustworthiness, which is critical for adoption in safety-critical domains.

A systems framework for trustworthy embodied intelligence with graded trustworthiness levels, emphasizing safety, reliability, and ethical compliance beyond task completion.

Read