Shared Autonomy via Hindsight Optimization
Shervin Javdani1, Siddhartha S Srinivasa1, J Andrew Bagnell1
1The Robotics Institute, Carnegie Mellon University.
Abstract:
In shared autonomy, user input and robot autonomy are combined to control a robot to achieve a goal. Often, the robot does not know a priori which goal the user wants to achieve, and must both predict the user's intended goal, and assist in achieving that goal. We formulate the problem of shared autonomy as a Partially Observable Markov Decision Process with uncertainty over the user's goal. We utilize maximum entropy inverse optimal control to estimate a distribution over the user's goal based on the history of inputs. Ideally, the robot assists the user by solving for an action which minimizes the expected cost-to-go for the (unknown) goal. As solving the POMDP to select the optimal action is intractable, we use hindsight optimization to approximate the solution. In a user study, we compare our method to a standard predict-then-blend approach. We find that our method enables users to accomplish tasks more quickly while utilizing less input. However, when asked to rate each system, users were mixed in their assessment, citing a tradeoff between maintaining control authority and accomplishing tasks quickly.
Related Concept Videos
Hindsight Biases
Optimal Foraging
Optimization Problems
Techniques of therapeutic communication I: Active Listening, Sharing Observations, Validation, and Using Touch
Therapeutic communication is not the same as social interaction. Social interaction has no goal or purpose and consists of casual information sharing, whereas therapeutic communication has a plan or purpose for the conversation. Therapeutic...
Optimal Arousal Theory
Inverted U-Shaped Performance Curve
The...
Optimizing Chromatographic Separations
Band broadening refers to spreading solute bands as they travel through the column. This broadening can impact resolution. Plate height (H) represents the length required for one theoretical plate. A lower plate height corresponds to...


