TensorX
返回文献探索

Paper · arXiv 2308.16185

Learning Vision-based Pursuit-Evasion Robot Policies

Andrea Bajcsy, Antonio Loquercio, Ashish Kumar, Jitendra Malik

8 upvotesAugust 30, 2023arXiv 预印本
AI 摘要

A fully-observable robot policy guides a partially-observable one in pursuit-evasion tasks by generating supervision, focusing on evader behavior diversity and modeling assumptions.

fully-observablepartially-observablepursuit-evasionsupervision signalevader behavior diversitymodeling assumptionsquadruped robotRGB-D cameraintent predictionanticipation

Abstract

Learning strategic robot behavior -- like that required in pursuit-evasion interactions -- under real-world constraints is extremely challenging. It requires exploiting the dynamics of the interaction, and planning through both physical state and latent intent uncertainty. In this paper, we transform this intractable problem into a supervised learning problem, where a fully-observable robot policy generates supervision for a partially-observable one. We find that the quality of the supervision signal for the partially-observable pursuer policy depends on two key factors: the balance of diversity and optimality of the evader's behavior and the strength of the modeling assumptions in the fully-observable policy. We deploy our policy on a physical quadruped robot with an RGB-D camera on pursuit-evasion interactions in the wild. Despite all the challenges, the sensing constraints bring about creativity: the robot is pushed to gather information when uncertain, predict intent from noisy measurements, and anticipate in order to intercept. Project webpage: https://abajcsy.github.io/vision-based-pursuit/

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Learning Vision-based Pursuit-Evasion Robot Policies | TensorX