Legged robots navigating cluttered environments must be jointly agile for efficient task execution and safe to avoid collisions with obstacles or humans. Existing studies either develop conservative controllers (< 1.0 m/s) to ensure safety, or focus on agility without considering potentially fatal collisions. This paper introduces Agile But Safe (ABS), a learning-based control framework that enables agile and collision-free locomotion for quadrupedal robots. ABS involves an agile policy to execute agile motor skills amidst obstacles and a recovery policy to prevent failures, collaboratively achieving high-speed and collision-free navigation. The policy switch in ABS is governed by a learned control-theoretic reach-avoid value network, which also guides the recovery policy as an objective function, thereby safeguarding the robot in a closed loop. The training process involves the learning of the agile policy, the reach-avoid value network, the recovery policy, and an exteroception representation network, all in simulation. These trained modules can be directly deployed in the real world with onboard sensing and computation, leading to high-speed and collision-free navigation in confined indoor and outdoor spaces with both static and dynamic obstacles.
翻译:在杂乱环境中导航的腿部机器人必须兼具敏捷性以高效完成任务,同时确保安全以避免与障碍物或人类发生碰撞。现有研究要么开发保守控制器(速度<1.0米/秒)以确保安全,要么专注于敏捷性而不考虑可能致命的碰撞。本文提出敏捷而安全框架,这是一种基于学习的控制框架,使四足机器人能够实现敏捷且无碰撞的运动。该框架包含一个敏捷策略,用于在障碍物中执行敏捷运动技能,以及一个恢复策略,用于防止失败,二者协同实现高速无碰撞导航。框架中的策略切换由一个基于学习的控制理论可达到避让价值网络控制,该网络同时作为目标函数引导恢复策略,从而在闭环中保护机器人安全。训练过程涉及敏捷策略、可达到避让价值网络、恢复策略以及外部感知表征网络的学习——所有训练均在仿真环境中完成。这些训练好的模块可直接部署到现实世界,凭借机载传感与计算能力,在包含静态和动态障碍物的室内外受限空间中实现高速无碰撞导航。