<?xml-model href='http://www.tei-c.org/release/xml/tei/custom/schema/relaxng/tei_all.rng' schematypens='http://relaxng.org/ns/structure/1.0'?><TEI xmlns="http://www.tei-c.org/ns/1.0">
	<teiHeader>
		<fileDesc>
			<titleStmt><title level='a'>Global-Position Tracking Control for Three-Dimensional Bipedal Robots Via Virtual Constraint Design and Multiple Lyapunov Analysis</title></titleStmt>
			<publicationStmt>
				<publisher></publisher>
				<date>11/01/2022</date>
			</publicationStmt>
			<sourceDesc>
				<bibl> 
					<idno type="par_id">10395624</idno>
					<idno type="doi">10.1115/1.4054732</idno>
					<title level='j'>Journal of Dynamic Systems, Measurement, and Control</title>
<idno>0022-0434</idno>
<biblScope unit="volume">144</biblScope>
<biblScope unit="issue">11</biblScope>					

					<author>Yan Gu</author><author>Yuan Gao</author><author>Bin Yao</author><author>C. S. Lee</author>
				</bibl>
			</sourceDesc>
		</fileDesc>
		<profileDesc>
			<abstract><ab><![CDATA[Abstract            A safety-critical measure of legged locomotion performance is a robot's ability to track its desired time-varying position trajectory in an environment, which is herein termed as “global-position tracking.” This paper introduces a nonlinear control approach that achieves asymptotic global-position tracking for three-dimensional (3D) bipedal robots. Designing a global-position tracking controller presents a challenging problem due to the complex hybrid robot model and the time-varying desired global-position trajectory. Toward tackling this problem, the first main contribution is the construction of impact invariance to ensure all desired trajectories respect the foot-landing impact dynamics, which is a necessary condition for realizing asymptotic tracking of hybrid walking systems. Thanks to their independence of the desired global position, these conditions can be exploited to decouple the higher-level planning of the global position and the lower-level planning of the remaining trajectories, thereby greatly alleviating the computational burden of motion planning. The second main contribution is the Lyapunov-based stability analysis of the hybrid closed-loop system, which produces sufficient conditions to guide the controller design for achieving asymptotic global-position tracking during fully actuated walking. Simulations and experiments on a 3D bipedal robot with twenty revolute joints confirm the validity of the proposed control approach in guaranteeing accurate tracking.]]></ab></abstract>
		</profileDesc>
	</teiHeader>
	<text><body xmlns="http://www.tei-c.org/ns/1.0" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:xlink="http://www.w3.org/1999/xlink">
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="1">Introduction</head><p>A robot's global position represents its absolute position in an environment. Poor global-position tracking can potentially put the safety of both humans and robots at risk, for example, by causing robots' failure to avoid pedestrians in human-populated environments. To achieve accurate global-position tracking, the zeromoment-point control approach has been introduced based on the zero-moment-point balance criterion and the continuous-time dynamic model of bipedal walking <ref type="bibr">[1]</ref><ref type="bibr">[2]</ref><ref type="bibr">[3]</ref>. Yet, bipedal walking is inherently a hybrid process involving both continuous motions (e.g., foot swinging) and discrete impact dynamics (e.g., sudden joint-velocity jumps upon a foot landing) <ref type="bibr">[4]</ref><ref type="bibr">[5]</ref><ref type="bibr">[6]</ref><ref type="bibr">[7]</ref>. Achieving reliable global-position tracking by explicitly addressing the hybrid robot dynamics presents substantial challenges.</p><p>This study focuses on addressing the challenges associated with: (a) lower-level trajectory generation (i.e., level 2 in Fig. <ref type="figure">1</ref>) and (b) controller design (i.e., level 3 in Fig. <ref type="figure">1</ref>). It is assumed that the desired global path and the desired time-varying position trajectory along the path have both been provided by a higher-level planner (i.e., level 1 in Fig. <ref type="figure">1</ref>) without impact dynamics considered.</p><p>One challenge in the lower-level trajectory generation (i.e., level 2 in Fig. <ref type="figure">1</ref>) for a hybrid robot model is to respect both continuous dynamics and the discrete impact dynamics, which is computational heavier than just respecting the continuous dynamics <ref type="bibr">[8]</ref>. For the controller to achieve asymptotic tracking based on a hybrid robot model, the desired trajectories need to agree with the impact dynamics; i.e., their pre-and postimpact values should satisfy the impact map. This is because the impact dynamics cannot be directly controlled due to their infinitesimally short duration <ref type="bibr">[9]</ref><ref type="bibr">[10]</ref><ref type="bibr">[11]</ref><ref type="bibr">[12]</ref>. Yet, the additional computational load caused by respecting the nonlinear impact map could be amplified when frequent replanning is needed during real-world tasks such as dynamic obstacle avoidance in cluttered environments.</p><p>Another challenge is the closed-loop stability analysis of the hybrid dynamical system that produces sufficient conditions to guide the controller derivation (i.e., level 3 in Fig. <ref type="figure">1</ref>). Such a stability analysis is complex because a closed-loop system capable of stabilizing a time-varying global-position trajectory is hybrid, nonlinear, and time-varying with uncontrolled, state-triggered impact dynamics.</p><p>1.1 Related Work on Orbitally Stabilizing Control. The most widely studied control approach that explicitly addresses the hybrid walking dynamics is the hybrid zero dynamics (HZD) method <ref type="bibr">[13]</ref><ref type="bibr">[14]</ref><ref type="bibr">[15]</ref><ref type="bibr">[16]</ref><ref type="bibr">[17]</ref><ref type="bibr">[18]</ref>. The HZD method provably stabilizes dynamic walking motions through orbital stabilization of the hybrid closedloop control system. It has realized remarkable performance for various gait types such as periodic underactuated <ref type="bibr">[19,</ref><ref type="bibr">20]</ref>, fully actuated <ref type="bibr">[21]</ref>, and multidomain walking <ref type="bibr">[22]</ref>.</p><p>The HZD framework introduces virtual constraints to represent the evolution of a robot's desired configuration with respect to a phase variable that indicates how far a step has progressed. To enforce the impact dynamics on the desired gait, the HZD approach introduces a method termed "impact invariance construction" to produce an equality constraint under which the desired gait respects the impact dynamics and incorporates the constraint in the optimization-based generation of virtual constraints. For systems with a single degree of underactuation <ref type="bibr">[14]</ref>, the HZD approach ensures the impact invariance of the zero dynamics manifold by imposing the impact invariance of a single point on it. Yet, because the encoding of the global-position trajectory is inherently different from that of virtual constraints, the previous impact invariance construction cannot be directly applied or extended to ensure the agreement with impact dynamics for the desired global-position trajectory. Specifically, the virtual constraints are encoded by a local phase variable that is reset at the beginning of a walking step, while the desired global-position trajectory is usually encoded by a global phase variable that evolves continuously and monotonically across all walking steps.</p><p>To analyze the closed-loop stability for guiding controller designs, the HZD approach exploits the Poincar e section method to examine the asymptotic convergence of a robot's state to the desired periodic orbit representing the desired gait in the state space. Recently, the HZD framework has been extended to achieve asymptotic tracking of the desired global path during 3D underactuated bipedal walking <ref type="bibr">[23]</ref>. Yet, an orbitally stabilizing controller cannot stabilize a prespecified time-varying trajectory <ref type="bibr">[24]</ref> such as the desired global-position trajectory.</p><p>1.2 Related Work on Trajectory Tracking Control. Our previous trajectory tracking controller designs either focus on individual joint trajectory tracking <ref type="bibr">[25,</ref><ref type="bibr">26]</ref> or only consider 2D walking <ref type="bibr">[27]</ref><ref type="bibr">[28]</ref><ref type="bibr">[29]</ref>. In particular, our previous work on 2D walking, including the impact invariance construction and stability analysis, is not valid for 3D robots. Specifically, the walking dynamics of 3D robots are nonlinearly coupled in the heading and lateral directions of the robot's global path, but 2D walking does not exhibit lateral motion, and accordingly, the coupling is trivial. This nonlinear coupling significantly increases the complexity of controller derivation in addressing 3D robots compared with 2D robots. Furthermore, experimental validation of these previous controllers has been missing.</p><p>Beyond the scope of global-position tracking control for bipedal walking robots, trajectory tracking control of general hybrid systems with state-triggered jumps is an active research topic <ref type="bibr">[30]</ref><ref type="bibr">[31]</ref><ref type="bibr">[32]</ref><ref type="bibr">[33]</ref><ref type="bibr">[34]</ref><ref type="bibr">[35]</ref>. Lyapunov-based controller design methodologies have been introduced to provably achieve asymptotic trajectory tracking for linear hybrid systems <ref type="bibr">[31,</ref><ref type="bibr">32]</ref>. In this study, to guide the needed controller design, we will extend the previous Lyapunov-based stability analysis to nonlinear hybrid systems that include 3D bipedal robots during fully actuated walking.</p><p>1.3 Contributions. This study aims to derive and experimentally validate a nonlinear walking control approach for 3D bipedal robots that achieves asymptotic global-position tracking by explicitly addressing the hybrid robot dynamics. The main contributions of this study are summarized as follows:</p><p>(i) Constructing impact invariance conditions that are independent of the desired global-position trajectory and yet ensure all desired trajectories respect the impact dynamics. They can be used to decouple the planning of virtual constraints and global position, thus improving trajectory generation efficiency.</p><p>(ii) Establishing sufficient conditions based on the multiple Lyapunov stability analysis <ref type="bibr">[36]</ref> of the hybrid system for guiding the design of a continuous state-feedback control law to achieve asymptotic global-position tracking. (iii) Demonstrating the global-position tracking accuracy of the proposed control approach both through simulations and experimentally on a 3D bipedal walking robot. (iv) Experimentally validating the inherent robustness of the proposed control design in addressing irregular walking surfaces such as moderately slippery floors.</p><p>Some of the results presented in this paper were initially reported in Refs. <ref type="bibr">[37]</ref> and <ref type="bibr">[38]</ref>. This paper includes substantial, new contributions in the following aspects: (a) the proof of the main theorem (i.e., Theorem 1) is updated with a new choice of Lyapunov function to properly analyze the convergence of the robot's lateral foot placement, and Proposition 4 is added along with its full proof to support the proof of the main theorem; (b) fully developed proofs of all theorems and propositions are presented, which were missing in Refs. <ref type="bibr">[37]</ref> and <ref type="bibr">[38]</ref>; (c) comparative experiment results are added to show the reliable globalposition tracking performance of the proposed control approach; and (d) robustness evaluation is newly included to illustrate the capability of the proposed method in handling relatively slippery grounds.</p><p>This paper is structured as follows. Section 2 describes the problem formulation. Section 3 explains the proposed continuousphase tracking control law. Section 4 presents the proposed construction of impact invariance conditions for designing virtual constraints. Section 5 introduces the closed-loop stability analysis based on multiple Lyapunov functions. Section 6 reports the simulation and experiment results. Section 7 discusses the proposed approach and potential directions of future work. Proofs of all theorems and propositions are given in the Appendix.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2">Problem Formulation</head><p>This section presents the problem formulation of globalposition tracking control. The formulation includes dynamics modeling and tracking error design.</p><p>2.1 Full-Order Robot Model. This section describes a fullorder model that accurately captures the dynamic behaviors of all degrees-of-freedom (DOFs) involved in bipedal walking. Thanks to the model's accuracy, a controller that is effective for the model would also be valid for the physical robot. Hence, we use the fullorder model as a basis of the proposed control approach.</p><p>The full-order model is naturally hybrid and nonlinear, because walking dynamics are inherently hybrid, involving both nonlinear continuous behaviors (e.g., leg-swinging motions) and statetriggered discrete behaviors (e.g., the joint-velocity jumps caused by foot-landing impact).</p><p>In this study, we assume that the swing and the stance legs immediately switch roles upon a foot landing, with the new swing leg beginning to move in the air and the new stance leg remaining in a full, static contact with the ground until the next landing Fig. <ref type="figure">1</ref> Overview of the proposed control approach. This study focuses on impact invariance construction in level 2 and stability analysis and controller design in level 3.</p><p>occurs <ref type="bibr">[14]</ref>. The assumption is valid when the double-support phase is sufficiently short and when the stance foot does not notably slip on the ground.</p><p>Under this assumption, if all of the robot's (revolute or prismatic) joints are directly actuated, then the robot is fully actuated; i.e., its full DOFs can be directly commanded within continuous phases.</p><p>This study focuses on the relatively simple gait, fully actuated gait, for two main reasons. First, asymptotic tracking of timevarying global-position trajectories for the 3D hybrid robot model is still an open control problem for this simple gait. Second, using a simple gait allows us to focus on addressing the complexity of the controller design problem induced by the hybrid, nonlinear robot dynamics, and the time-varying global-position trajectory.</p><p>Continuous-phase dynamics. As illustrated in Fig. <ref type="figure">2</ref>, a complete walking cycle comprises: (a) a fully actuated continuous phase during which one foot contacts the ground and the other swings in the air and (b) a landing impact.</p><p>Walking dynamics during continuous phases can be described by usual ordinary differential equations. Lagrange's method is used to obtain the following nonlinear full-order model during continuous phases <ref type="bibr">[13]</ref> </p><p>where q 2 Q is the joint-position vector, M : Q ! R n&#194;n is the symmetric, positive-definite inertia matrix, c : TQ ! R n is the sum of Coriolis, centrifugal, and gravitational terms, B u 2 R n&#194;m is the joint-torque projection matrix with full column rank, and u 2 U is the joint-torque vector. Here, Q &amp; R n is the configuration space of the robot, TQ is the tangent bundle of Q, and U &amp; R m is the admissible joint-torque set. Note that m &#188; n when a robot is fully actuated. Impact dynamics. When the swing foot lands on the ground, the swing and stance legs immediately switch their roles. Here, we model the swing-foot landing impact as the contact between rigid bodies <ref type="bibr">[14]</ref>. This assumption is valid for dynamic walking on relatively stiff surfaces (e.g., concrete and ceramic floors) during which the swing foot strikes the surface at a relatively significant downward velocity.</p><p>Due to the coordinate swap of the swing and stance legs as well as the impulsive rigid-body impact, both joint position and velocity vectors experience a sudden jump at a landing event. This state-triggered jump is described by the nonlinear reset map D q; _ q : TQ ! TQ <ref type="bibr">[13]</ref> given by</p><p>where ? &#192; and ? &#254; represent the values of ? just before and just after the impact, respectively. Switching surface. A swing-foot landing event is triggered when the robot's state reaches the switching surface S q , which is expressed as S q :&#188; f&#240;q; _ q&#222; 2 TQ : z sw &#240;q&#222; &#188; 0; _ z sw &#240;q; _ q&#222; &lt; 0g</p><p>where z sw : Q ! R is the swing-foot height above the ground.</p><p>Combining the above equations yields the following hybrid full-order model:</p><p>2.2 Global-Position Tracking Error. A fully actuated, n-DOF bipedal robot can track n independent desired position trajectories, including the reference global-position trajectories.</p><p>In this study, we choose to use the position of a biped's base (e.g., trunk), &#240;x b ; y b ; z b &#222;, to represent its global position in an environment. The horizontal components of the base position are related to the stance-foot position as</p><p>where &#240;x st ; y st ; 0&#222; denotes the stance-foot position with x st ; y st 2 R. The scalar variables x b : Q ! R and y b : Q ! R represent the x-and y-coordinates of the base position relative to the stance foot, respectively.</p><p>In real-world locomotion tasks, a higher-level planner typically specifies the desired global motions as: As an arbitrary curved path can be approximated as a nonsmooth curve pieced together by straight lines, this study focuses on the tracking control of straight-line paths, which could be extended to the tracking of a curved path as briefly explained in Sec. 8.</p><p>Without loss of generality, suppose that the centerline C d coincides with the X w -axis of the world frame; that is</p><p>Then, the global-position tracking error h 1 &#240;t; q&#222; along C d is defined as h 1 &#240;t; q&#222; :&#188; x b &#240;q&#222; &#192; &#240;s d &#240;t&#222; &#192; x st &#222;. While s d &#240;t&#222; and C d are often provided by a higher-level path planner, the desired base motion in the direction lateral to C d remains to be designed, which is explained next.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2.3">Virtual-Constraint Tracking</head><p>Error. Besides the desired global-position trajectory s d &#240;t&#222;, a legged robot typically has multiple directly actuated DOFs that can track additional desired motions. We choose to use virtual constraints to define the desired trajectories for the lateral base position y b and the remaining control variables</p><p>Analogous to the HZD framework, we use virtual constraints to represent the desired configuration relative to a phase variable</p><p>The phase variable represents how far a step has progressed. Without loss of generality, the phase variable h is chosen as the forward position of the base relative to the support foot,</p><p>x b &#240;q&#222;; that is,</p><p>The virtual constraints can be encoded by   </p><p>where the control variables h c and their desired trajectories h d are, respectively, defined as h c :&#188; &#189; x b ; y b ; / T c T and h d :&#188; &#189; s d &#192; x st ; y d &#192; y st ; / T d T . The control objective is to asymptotically drive the tracking error h to zero for achieving asymptotic tracking of the desired motions, which are the desired global-position trajectory s d &#240;t&#222; and the desired functions y d and / d that define the virtual constraints, for 3D bipedal robots walking along the given straight line.</p><p>To achieve this objective, the proposed control approach (Fig. <ref type="figure">1</ref>) comprises three main components: (a) continuous-phase controller design (for stabilizing the desired trajectories within continuous phases); (b) impact invariance construction (for satisfying a necessary condition of asymptotic tracking for the hybrid model); and (c) closed-loop stability analysis (for providing sufficient stability conditions that guide the controller design).</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3">Continuous-Phase Control</head><p>This section presents a continuous state-feedback control law that asymptotically stabilizes the desired trajectories within continuous phases.</p><p>We choose to design a controller that directly regulates the continuous-phase walking dynamics instead of the impact dynamics because the impact dynamics cannot be directly commanded due to its infinitesimally short duration. We will show in Sec. 5 how the proposed continuous control law could be tuned to indirectly stabilize the desired trajectories for the overall hybrid system.</p><p>The proposed control law (Fig. <ref type="figure">3</ref>) is synthesized based on the full-order model of bipedal walking dynamics. Analogous to the HZD framework, we utilize the input-output linearization technique <ref type="bibr">[24]</ref> to linearize the nonlinear continuous-phase dynamics in Eq. ( <ref type="formula">1</ref>) into a linear map, which allows us to exploit the wellstudied linear system theory to design the needed controller for the continuous phase.</p><p>We define the output function as the tracking error h</p><p>Then a continuous-phase control law synthesized via input-output linearization is given by</p><p>with J h &#240;q&#222; :&#188; @h @q &#240;t; q&#222;; which yields the linearized dynamics &#8364; y &#188; v. Note that the variables / c , y d , and / d can be chosen such that there exists an open subset Q of the configuration space Q on which the Jacobian matrix J h &#240;q&#222; is invertible. Then, the matrix</p><p>Choosing v as a proportional-derivative (PD) term</p><p>where the proportional gain matrix K P 2 R n&#194;n and the derivative gain matrix K D 2 R n&#194;n are both positive-definite diagonal matrices, the linear closed-loop dynamics of the output function becomes</p><p>Then, the closed-loop error equation can be compactly expressed as</p><p>Here, A :&#188; 0 I &#192;K P &#192;K D ! with 0 a zero matrix and I an identity matrix with appropriate dimensions. S and D are the switching surface and impact map associated with the closed-loop dynamics, respectively. Note that D is explicitly time-dependent because of the explicit time dependence of h. The expressions of S and D can be obtained from their counterparts in the open-loop dynamics (Eqs. ( <ref type="formula">3</ref>) and ( <ref type="formula">2</ref>)) as well as the output function definition (Eq. ( <ref type="formula">8</ref>)).</p><p>The origin (i.e., x &#188; 0) of the continuous-phase closed-loop dynamics (i.e., _</p><p>x &#188; Ax) will be asymptotically stable if the PD gains are chosen such that A is Hurwitz <ref type="bibr">[24]</ref>. Then, there exist positive numbers c 1 , c 2 , and c 3 and a Lyapunov function candidate V&#240;x&#222; such that c 1 jjxjj 2 V&#240;x&#222; c 2 jjxjj 2 and _ V &#240;x&#222; &#192;c 3 jjxjj 2 <ref type="bibr">(13)</ref> hold for all x within continuous phases. These inequalities indicate that V&#240;x&#222; exponentially converges at the rate of c3 c2 within a continuous phase.</p><p>While the proposed control law with properly chosen PD gains guarantees the asymptotic tracking of the desired trajectories within continuous phases, the impact dynamics (i.e., x &#254; &#188; D&#240;t; x &#192; &#222;) remain uncontrolled, and thus, the stability of the hybrid closed-loop system is not yet ensured. To satisfy a necessary condition of asymptotic trajectory stabilization in the presence of uncontrolled impact dynamics, we introduce impact invariance construction next.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4">Impact Invariance Construction for Virtual Constraint Design</head><p>This section derives impact invariance conditions that can be incorporated in the trajectory generation of the desired functions y d and / d , which define the virtual constraints, for ensuring all desired trajectories (i.e., y d and / d as well as s d ) respect the impact dynamics.</p><p>For the proposed feedback controller to achieve asymptotic tracking for hybrid dynamical systems, the output function state variables y and _ y need to satisfy the following condition across an impact event at the steady-state: y &#254; &#188; 0 and _ y &#254; &#188; 0 hold just after the impact if y &#192; &#188; 0 and _ y &#192; &#188; 0 hold just before the impact. Suppose this condition is not met at the steady-state. Then, because the robot's impact dynamics cannot be directly regulated, the output function state may become nonzero just after an impact Transactions of the ASME even if it is zero just before the impact, which means asymptotic tracking cannot be achieved.</p><p>To mathematically describe this condition, we introduce the manifold Z given by Z :&#188; f&#240;t; q; _ q&#222; 2 R &#194; TQ : h&#240;t; q&#222; &#188; 0; _ h&#240;t; q; _ q&#222; &#188; 0g</p><p>Here, Z is a one-dimensional embedded submanifold of R &#194; TQ.</p><p>The corresponding guard S a is defined by rewriting S q as: S a :&#188; f&#240;t; q; _ q&#222; 2 R &#194; TQ : z sw &#240;q&#222; &#188; 0; _ z sw &#240;q; _ q&#222; &lt; 0g: DEFINITION 1. (Impact invariance) The manifold Z is called impact invariant if &#240;t &#254; ; D q; _ q &#240;q &#192; ; _ q &#192; &#222;&#222; 2 Z holds for any &#240;t &#192; ; q &#192; ; _ q &#192; &#222; 2 Z \ S a ; that is, the manifold is invariant across the impact event.</p><p>Remark 1. (Differentiation from the HZD approach) The concept of impact invariance was first introduced within the HZD framework, along with a systematic method of impact invariance construction (see Theorem 4 in Ref. <ref type="bibr">[14]</ref>). The concept was later on termed as "impact invariance" <ref type="bibr">[15]</ref>.</p><p>The equations defining Z are time-dependent in this study whereas in the original HZD framework <ref type="bibr">[14,</ref><ref type="bibr">15]</ref>, the hybrid zero dynamics manifold is time-independent. The difference is essentially due to their different control objectives. In the HZD framework, the controller aims to track the desired walking pattern encoded by a configuration-based phase variable, resulting in a time-invariant definition of output functions and accordingly a time-invariant hybrid zero dynamics manifold. In contrast, the control objective here is to track time-varying global-position trajectories, and thus, the output function (specifically, h 1 &#240;t; q&#222;) is time-varying, inducing the time dependence of the submanifold Z.</p><p>Another difference is that, in the HZD framework <ref type="bibr">[14,</ref><ref type="bibr">15]</ref>, the system of interest is underactuated, and thus, the dimension of the impact invariant manifold, when restricted to the guard, can be higher than zero. Yet, in our case of fully actuated systems, the dimension of Z \ S a is zero because Z \ S a is a single point. Although ensuring that a single point respects the impact map is typically easier for trajectory optimization, the proposed impact invariance construction for Z has an attractive property for efficient planning, as revealed by Remark 3 in Sec. 4.4.</p><p>We choose to construct the manifold Z to be impact invariant by properly planning the desired function h d . As the desired global position s d is often supplied by a higher-level path planner without impact dynamics considered, the generation of the remaining desired functions y d and / d , which define the virtual constraints, needs to ensure the impact agreement for all trajectories (i.e., s d , y d , and / d ). To this end, the proposed impact invariance construction boils down to the derivation of conditions that the virtual constraints should satisfy in order to ensure Z is impact invariant. We call these conditions "impact invariance conditions."</p><p>The proposed impact invariance construction consists of two steps corresponding to two sets of impact invariance conditions. We first extend the existing HZD method (i.e., Theorem 4 in Ref. <ref type="bibr">[14]</ref>) to derive conditions that ensure the impact invariance of the three-dimensional embedded submanifold of R &#194; TQ associated with the virtual constraint, that is</p><p>Then, we introduce a new, additional condition that, in combination with the first set of conditions, guarantees the impact invariance of Z. Note that both conditions are placed on the virtual constraints alone. These two steps of impact invariance construction are introduced in Secs. 4.3 and 4.4, respectively. Before presenting them, we first explain the timing and the unique robot configuration associated with a landing impact in Secs. 4.1 and 4.2, which are needed for the derivation of the proposed impact invariance conditions.</p><p>4.1 Impact Timings. Because the desired global-position trajectory s d is explicitly time-varying, we need to consider the impact timings in the proposed impact invariance construction. As the actual and desired impact timings generally do not coincide due to the state-triggered nature of a foot-landing event <ref type="bibr">[10]</ref>, they are individually defined as follows.</p><p>DEFINITION 2. (Actual and desired impact timings) Let T k be the timing of the k th (k 2 Z &#254; ) actual landing impact, which is defined as the timing of the first intersection between the state x and the switching surface S on t &gt; T &#254; k&#192;1 . Without loss of generality, define T 0 &#188; 0. Let s k denotes the kth desired impact timing, which is defined as the timing of the first intersection between x and S on t</p><p>The precise definition of T k is given in Ref. <ref type="bibr">[14]</ref>. Figure <ref type="figure">4</ref> shows an illustration of T k . The variables ?&#240;T &#192; k&#192;1 &#222; and ?&#240;T &#254; k&#192;1 &#222; are, respectively, denoted as ?j &#192; k&#192;1 and ?j &#254; k&#192;1 in the rest of the paper where brevity is preferred.</p><p>4.2 Unique Configuration. The proposed impact invariance construction utilizes the uniqueness of the robot's joint position q &#195; just before an impact event when the virtual constraints in Eq. ( <ref type="formula">7</ref>) are exactly satisfied. Note that although the proposed construction relies on the unique configuration, it does not require the joint velocity should be unique just before an impact.</p><p>The joint position q &#195; is mathematically defined as the solution to the following equations: </p><p>on S \ Q. Note that the last equation in Eq. ( <ref type="formula">15</ref>) holds because the swing-foot height z sw &#240;q&#222; reaches zero at a touchdown. Due to the nonlinearity of the function F&#240;q&#222;, Eq. ( <ref type="formula">15</ref>) may have multiple solutions on S \ Q. Suppose that the output function is designed such that @F @q &#240;q &#195; &#222; is invertible on S \ Q. Then by the implicit function theorem, there exits</p><p>Q &amp; Q such that q &#195; is a unique solution to F&#240;q&#222; &#188; 0 on S \ Q.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4.3">Impact</head><p>Invariance Construction for Virtual Constraints. We are now ready to introduce the conditions that ensure the impact invariance of Z. The impact invariance conditions are built upon the uniqueness of the joint position q &#195; on S \ Q. From Eq. ( <ref type="formula">15</ref>), we know the value of q &#195; depends on the lateral foot placement y st . The following proposed impact invariance conditions use the value of q &#195; associated with the desired lateral foot placement y std .</p><p>PROPOSITION 1. (Impact invariance conditions for Z) Suppose that the desired functions y d and / d are planned to meet the following conditions: Here,</p><p>. Then, under the lateral foot-placement condition y st &#188; y std , the impact invariance of Z holds. From Sec. 4.2, we know that Eq. ( <ref type="formula">15</ref>) has a unique solution on S \ Q when y st &#188; y std ; that is, the robot has a unique configuration q &#195; just before an impact if y 2 &#240;s &#192; k &#222; &#188; 0 and y st &#188; y std . Given the uniqueness of q &#195; , the equations in condition (A1), which are imposed on the virtual constraints, ensure that</p><p>and y st &#188; y std . Remark 2. (Differentiation from the HZD approach) The proposed construction of the impact invariant manifold Z is analogous to the original HZD method <ref type="bibr">[14,</ref><ref type="bibr">15]</ref>. The first difference lies in that the output function in our case is explicitly a function of the lateral foot placement y st and that the proposed construction is for the case where the desired foot placement y st &#188; y std is realized. The second difference is that we define the submanifold Z based on R &#194; TQ instead of just the tangent bundle TQ. This is because the impact invariance construction for Z is used as a basis for rendering the manifold Z impact invariant, and Z is associated with the time-varying global-position tracking error h 1 &#240;t; q&#222;.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4.4">Impact Invariance Construction for Global-Position</head><p>Tracking Error. As the desired global-position trajectory s d is often supplied by a high-level planner without impact dynamics considered, we construct an additional condition, which is placed on the virtual constraints, to further ensure the impact invariance associated with the global-position error state, i.e.,</p><p>x b &#192; &#240;s d &#192; x st &#222; and its first derivative. Note that</p><p>This additional condition, together with those introduced in Sec. 4.3, guarantees that the submanifold Z is impact invariant.</p><p>The key to the proposed construction is to exploit the property of s d that it is commonly planned as a smooth function for any t &gt; T 0 . Thanks to this property, x b &#192; s d &#188; 0 automatically holds just after an impact if it holds just before the impact. This is because both the forward base position x b and its desired trajectory s d are continuous across an impact.</p><p>To ensure _ x b &#192; _ s d &#188; 0 holds just after an impact if it holds just before the impact, we choose to enforce the continuity of the global velocity _ x b across the planned impact event. The rationale of this design choice is threefold. First, given the continuity of _ s d for any t &gt; T 0 , the continuity of _ x b across the planned impact event guarantees the continuity of _</p><p>x b &#192; _ s d , which then ensures that _</p><p>x b &#192; _ s d &#188; 0 holds just after the planned impact if it holds just before the impact. Second, the continuity of _</p><p>x b is equivalent to that of _ x b because the stance foot does not move (i.e., _</p><p>x b is a function of the joint position q and velocity _ q only, and thus its continuity across the planned impact event can be satisfied through virtual constraint design alone without explicitly relying on the profile of s d .</p><p>The proposed conditions for the impact invariance of Z is summarized as follows.</p><p>PROPOSITION 2. (Impact invariance condition for Z) Suppose that the desired functions y d and / d satisfy conditions (A1) and (A2) and the following condition:</p><p>Then, under the lateral foot-placement condition y st &#188; y std , the impact invariance of Z holds.</p><p>Condition (A3) ensures that the base velocity does not jump (i.e., _</p><p>and y st &#188; y std and if conditions (A1) and (A2) in Proposition 1 also hold. Furthermore, if the virtual constraints are generated to meet the conditions in Proposition 2, which contains the conditions from Proposition 1, then under the lateral foot-placement condition y st &#188; y std , the impact invariance of Z holds; that is, if</p><p>(Independence from desired global-position trajectory) Propositions 1 and 2 indicate that the satisfaction of the impact invariance conditions only relies on the design of the virtual constraints but not the arbitrary global-position trajectory s d provided by a higher-level planner. For this reason, the design of virtual constraints does not need to explicitly consider s d and thus can be performed offline even when the higher-level planner updates s d online. This could reduce the computational load for online planning especially for mobility tasks that could frequently demand the replanning of s d (e.g., dynamic obstacle avoidance).</p><p>Remark 4. (Ensuring the desired lateral foot placement through controller design) Note that the foot-placement condition y st &#188; y std underlying the proposed impact invariance construction is only assumed in the virtual constraint planning but not the controller design. Indeed, Sec. 5 introduces sufficient conditions under which the proposed controller guarantees this foot-placement condition holds at the actual steady-state.</p><p>Remark 5. (Planning virtual constraints offline for impact invariance) There is a relatively simple two-step procedure to plan virtual constraints offline that meet the impact invariance conditions in Propositions 1 and 2. Recall the virtual constraints are given by: y b &#192; y d &#240;h&#222; &#254; y st &#188; 0 and / c &#192; / d &#240;h&#222; &#188; 0. The first step is to plan desired time trajectories for the control variables y b and / c within a complete hybrid walking cycle (i.e., a continuous phase and a landing impact), which respect the impact dynamics with a constant forward velocity imposed across the impact. Let &#7929;d &#240;t&#222; &#192; y std and /d &#240;t&#222; denote these time trajectories, respectively. Let hd &#240;t&#222; be the desired fictitious time trajectory for the phase variable h (i.e.,</p><p>x b ). The function hd &#240;t&#222; is only used for the offline planning and can be prespecified as any function that monotonically increases in time t within the planned walking cycle. Let h&#192;1 </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5">Stability Analysis</head><p>This section introduces Lyapunov-based stability analysis of the hybrid, nonlinear, time-varying closed-loop error dynamics (Eq. ( <ref type="formula">12</ref>)) under the proposed continuous-phase control law (Eqs. ( <ref type="formula">10</ref>) and ( <ref type="formula">11</ref>)). The outcome of this stability analysis is a set of sufficient conditions under which the proposed control law provably realizes asymptotic stabilization of the desired globalposition trajectory s d and the desired functions y d and / d for the overall hybrid system.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.1">Boundedness of Foot Placement and Impact Timing.</head><p>Before presenting the main theorem on closed-loop stability, we first introduce the boundedness of the impact timing T k and the lateral stance-foot position y st . The boundedness of the impact timing is needed in the stability analysis to derive how much a Lyapunov function converges within a continuous phase. The boundedness of y st also needs to be explicitly considered, because y st &#188; y std underlies the proposed impact invariance conditions and should hold at the actual steady-state for achieving asymptotic tracking.</p><p>PROPOSITION 3. (Boundedness of impact timing error) Let x&#240;t; t 0 ; k 0 &#222; be a solution of a fictitious continuous-time system _ x &#188; Ax with the initial condition x&#240;t 0 &#222; &#188; k 0 ; 8t &gt; t 0 . There exists a positive number r 1 and a Lipschitz constant L Tx such that the 111001-6 / Vol. 144, NOVEMBER 2022</p><p>Transactions of the ASME difference between the actual and the planned impact timings is bounded above in norm as</p><p>for any xj &#254; 0 2 B r1 &#240;0&#222; :&#188; fx 2 R 2n : jjxjj r 1 g and any k 2 Z &#254; . PROPOSITION 4. (Boundedness of lateral foot-placement error) Suppose that the lateral swing-foot position y sw is chosen as an element of / c and is thus directly controlled. Then, there exist positive numbers b st and d 1 such that the foot-placement error just after the k th swing-foot landing is bounded above in norm as</p><p>for any xj &#254; 0 2 B d1 &#240;0&#222; :&#188; fx 2 R 2n : jjxjj d 1 g and any k 2 Z &#254; . Rationale of proofs. The full proofs of Propositions 3 and 4 are given in the Appendix. The proof of Proposition 3 utilizes the implicit dependence of the actual impact timing T k on the error state x. The proof of Proposition 4 mainly relies on the fact that the stance-foot position within the current step is the end position of the swing foot within the previous step. Thus, by including y sw as a control variable, the lateral foot-placement error y st &#192; y std is also contained in the state x. 5.2 Main Theorem. If the virtual constraints are designed to satisfy the impact invariance conditions in Propositions 1 and 2 and if the continuous-phase convergence rate of x is sufficiently fast, then the origin of the hybrid closed-loop error system is asymptotically stable, as summarized in the main theorem: THEOREM 1. (Closed-loop stability conditions) Suppose that the virtual constraints satisfy the impact invariance conditions (A1)-(A3). Also, suppose that the PD gains in Eq. ( <ref type="formula">11</ref>) are chosen such that A is Hurwitz and that the continuous-phase convergence rate of x is sufficiently fast. Then, there exists a positive number d 2 such that for any xj &#254; 0 2 B d2 &#240;0&#222; :&#188; fx 2 R 2n : jjxjj d 2 g, the origin of the closed-loop error system in Eq. ( <ref type="formula">12</ref>) is locally asymptotically stable; that is, x&#240;t&#222; ! 0 as t ! 1:</p><p>Furthermore, both the lateral foot placement and actual impact timing asymptotically converge to their desired values; that is, T k &#192; s k ! 0 and y st &#192; y std ! 0 as k ! 1:</p><p>Rationale of proof. The full proof of Theorem 1 is given in the Appendix. The proof utilizes the stability theory of the multiple Lyapunov functions <ref type="bibr">[36]</ref>, which prescribes how a Lyapunov function candidate should evolve in order for the origin of a hybrid dynamical system to be stable.</p><p>The stability analysis begins with the construction of the Lyapunov function candidate. Since the lateral foot-placement error y st &#192; y std directly affects the satisfaction of the impact invariance conditions and thus the system stability, we choose to construct the Lyapunov function V a by augmenting V with a positivedefinite function of the foot-placement error</p><p>where r is a positive number to be specified in the proof. Next, we analyze the evolution of V a during a continuous phase as well as through a hybrid transition. The last step is to derive the sufficient closed-loop stability conditions that the continuousphase convergence rate should meet such that the divergence of V a caused by the uncontrolled impact is compensated by the continuous-phase convergence.</p><p>The convergence of the foot placement y st and impact timing T k is proved based on Propositions 3 and 4 and the asymptotic convergence of the error state x. By Propositions 3 and 4, the deviations of the lateral foot placement and impact timing are bounded above by the norms of the actual state x and the fictitious state x. Note that by definition, x overlaps with x within the given actual continuous phase. Thus, driving x to zero will indirectly make x diminish, which then eliminates the deviations y st &#192; y std and T k &#192; s k at the actual steady-state. Remark 6. (Tuning continuous-phase convergence rate) By Theorem 1, the continuous-phase convergence rate of x (or equivalently, V a ) needs to be sufficiently fast for guaranteeing asymptotic trajectory tracking of the hybrid closed-loop system. The continuous-phase convergence rate of V a solely depends on that of V, because the stance foot is static during a continuous phase and jy st &#192; y std j remains constant. We can construct V as V &#188; x T Px, where P is the solution to the Lyapunov equation <ref type="bibr">[24]</ref> PA &#254; A T P &#188; &#192;Q Here, Q is any symmetric, positive-definite matrix satisfying 0 &lt; k Q I Q with a positive number k Q . For simplicity, we can choose Q as an identity matrix, and then k Q can be any number satisfying 0 &lt; k Q 1. Then, the bounds of V and _ V in Eq. ( <ref type="formula">13</ref>) become c 1 &#188; k min &#240;P&#222;; c 2 &#188; k max &#240;P&#222;, and c 3 &#188; k Q , where k min &#240;P&#222; and k max &#240;P&#222; are the smallest and the largest eigenvalues of P, respectively. Thus, the exponential convergence rate of V becomes c3 c2 &#188; kQ kmax&#240;P&#222; . Note that the value of the matrix P depends on the PD gains, and thus k max &#240;P&#222; can be adjusted by tuning those gains. The full proof (Sec. 9.5) provides greater details about PD gain tuning. It also explains how to compute the lower bound of the convergence rate c3 c2 for guaranteeing asymptotic error convergence of the hybrid closed-loop system.</p><p>Remark 7. (Satisfying lateral foot-placement condition) Theorem 1 indicates that the lateral foot-placement condition underlying the proposed impact invariance conditions in Propositions 1 and 2 is exactly met at the steady-state. Thus, the impact invariance of Z, which is the necessary condition for asymptotic trajectory tracking, is indeed satisfied at the steady-state; that is, if</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="6">Simulations and Experiments</head><p>This section reports simulation and experiment results that demonstrate the global-position tracking performance of the proposed control approach.</p><p>The hardware platform used for controller validation is the OP3 bipedal humanoid robot developed by ROBOTIS Co., Ltd. (Seoul, South Korea) (Fig. <ref type="figure">1</ref>). OP3 weighs 3.5 kg with a height of 0.51 m. It has twenty revolute joints comprising eight upper-body and 12 leg joints. As these joints (including ankles) are all independently actuated, the robot is fully actuated during a continuous phase of flat-foot walking without slippage.</p><p>6.1 Virtual Constraint Generation. This section explains the lower-level, optimization-based trajectory generation of virtual constraints based on the proposed impact invariance conditions.</p><p>With full actuation, OP3's 12 leg joints can be directly commanded to track 12 independent desired trajectories, which are: (1) the desired global-position trajectory s d and (2) the desired functions y d and / d . As a higher-level planner supplies the desired global path on the walking surface and the desired position trajectory along the path, the objective of the trajectory generation is to plan the desired lateral base position y d and desired functions / d that both define the virtual constraints.</p><p>Trajectory parameterization. The desired lateral base position y d is chosen as the following simple sinusoidal function to enable an oscillatory global motion about the centerline C d during walking</p><p>with a :&#188; &#189; a 1 a 2 a 3 T 2 R 3 an unknown vector to be optimized. The desired functions / d are chosen as the desired trajectories for the following ten control variables / c : (b) Position (x sw , y sw , z sw ) and roll, pitch, and yaw angles (w roll sw ; w pitch sw ; w yaw sw ) of the swing foot. This choice of control variables allows direct regulation of the poses (i.e., positions and orientations) of the trunk and swing foot to avoid overstretched leg joints, enforce a relatively steady trunk posture, and maintain a sufficient clearance between the swing foot and the walking surface.</p><p>The desired functions / d &#240;h&#222; are parameterized using B ezier curves <ref type="bibr">[39]</ref> </p><p>where M 2 Z &#254; is the order of the B ezier curves, s&#240;h&#222; : 10 is the unknown vector to be optimized, and h &#254; and h &#192; are the planned values of h at the beginning and the end of a step, respectively. Recall that h is chosen as the relative forward position of the base (Eq. ( <ref type="formula">6</ref>)) and represents how far a step has progressed within a step.</p><p>Optimization formulation. The optimization variables are chosen as parameters a in Eq. ( <ref type="formula">19</ref>) and a k in Eq. ( <ref type="formula">20</ref>). The constraints are set as: (B1) The proposed impact invariance conditions (A1)-(A3) in Propositions 1 and 2. (B2) Feasibility constraints (e.g., joint-position limits, jointtorque limits, and ground-contact constraints). (B3) Gait parameters (e.g., step length and duration).</p><p>This list of constraints is not intended to be exhaustive as this study focuses on impact invariance construction and controller design instead of trajectory generation. MATLAB command fmincon is used to solve the optimization.</p><p>Desired trajectories. In the simulations and experiments, the planned virtual constraints are illustrated in Fig. <ref type="figure">5</ref>. The centerline C d of the desired path is the X w -axis of the world reference frame. To test the capability of the proposed control approach in tracking desired position trajectories with constant or time-varying velocities, the following two desired position trajectories s d &#240;t&#222; along C d are considered: In theory, the proposed control approach can locally stabilize any profiles of s d &#240;t&#222; that are differentiable in time. In practice, s d &#240;t&#222; also needs to respect the robot's hardware constraints (e.g., actuation and kinematic limits).</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="6.2">Controller Implementation</head><p>Procedure. This subsection explains the experiment procedure that we adopt to implement the proposed controller on the physical OP3 robot using the ROS package (op3 manager) developed by OP3's manufacturer.</p><p>Since the ROS package does not support direct access to the output torques of joint motors, the proposed control law in Eq. ( <ref type="formula">10</ref>), which is a torque command, cannot be directly implemented on OP3 and needs to be adapted for its implementation on the robot.</p><p>Considering that OP3's ROS package allows users to send desired joint-position trajectories to individual joints and specify the PD gains of OP3's default joint controller, we adopt the following controller implementation procedure <ref type="bibr">[21]</ref>: (a) to generate the desired position trajectories of individual joints, q d &#240;t&#222; and (b) to send the desired trajectories to the default joint-position controller. The main steps of this procedure are shown in Fig. <ref type="figure">6</ref>.</p><p>Although the adapted controller directly tracks the individual joint trajectories q d instead of the original Cartesian-space trajectories h d , the controller implementation procedure still allows satisfactory tracking of h d . This is because q d preserve the dynamic feasibility and desired features of h d as specified in B1-B3.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="6.3">Simulation and Experiment</head><p>Setup. This subsection reports the setup of MATLAB and WEBOTS simulations and hardware experiments for controller validation.</p><p>MATLAB. To validate the theoretical controller design, we utilize MATLAB to implement the control law based on the full-order model of OP3 (Eq. ( <ref type="formula">4</ref>)). The control gains are set as K P &#188; 225 &#193; I and K D &#188; 30 &#193; I to ensure the matrix A is Hurwitz. The simulation results are shown in Figs. <ref type="figure">7</ref> and<ref type="figure">8</ref>. WEBOTS. To gain preliminary insights into the effectiveness of the proposed controller implementation procedure as explained in Sec. 6.2, we use WEBOTS to simulate a 3D realistic biped model that closely emulates OP3's graphical, physical, and dynamical properties (including its limited actuator accessibility). The control gains that the emulated robot system allows users to tune are the effective PD gains, whose physical meaning is different from K P and K D in Eq. <ref type="bibr">(11)</ref>. These effective gains are tuned to be "10" and "0" such that the resulting tracking performance is comparable with the MATLAB results. WEBOTS simulation results of the adapted controller are displayed in Figs. <ref type="figure">8</ref> and<ref type="figure">9</ref>.</p><p>Experiments. The experiment setup is shown in Fig. <ref type="figure">10</ref>. With this setup, the robot's joint angles can be directly measured by joint encoders, and its global pose can be determined by: (a) using the 4 K PRO WEBCAM and APRILTAG <ref type="bibr">[40]</ref> to obtain the stance-foot pose in the world reference frame and (b) using the obtained stance-foot pose to solve for the robot's global pose via forward kinematics. By providing relatively accurate measurement, the use of the overhead camera and APRILTAG allows us to focus on controller validation. The experiment is guided by the controller adaptation procedure from Sec. 6.2. The initial tracking error of Transactions of the ASME the desired position trajectory s d is 3 cm, which is approximately 1/3 of a nominal step length. The initial path tracking error is 5 cm. Similar to the gain tuning in WEBOTS, the effective PD gains are, respectively, tuned to be "800" and "0" to ensure a relatively fast error convergence without violating the actuator's torque limit. Experiment results of OP3 walking on a concrete and a relatively slippery ceramic floor are shown in Fig. <ref type="figure">11</ref>. Videos of the experiments can be accessed at following link. <ref type="foot">2</ref>6.4 Discussions on Validation Results. This subsection provides discussions on the controller evaluation results.</p><p>Tracking accuracy in simulations. The virtual-constraint tracking results in Figs. 7 (MATLAB) and 9 (WEBOTS) show that the proposed control law is capable of accurately enforcing the virtual constraints for 3D bipeds during fully actuated walking. Note that the tracking errors observed in the WEBOTS simulations are larger than the MATLAB results in the swing foot's lateral position y sw , base height z b , base yaw angles w yaw b , and the swing foot's orientation (w roll sw ; w pitch sw ; w yaw sw ). This is partly caused by the foot slippage of the robot during WEBOTS simulations, which is not captured by the robot model used in MATLAB simulations. The global-position tracking results in Fig. <ref type="figure">8</ref> validate that the proposed control law drives the robot to asymptotically converge to the desired globalposition trajectory s d while moving along the centerline C d of the global path. In particular, the accurate tracking results obtained in WEBOTS indicate the effectiveness of the proposed controller implementation procedure in guaranteeing reliable trajectory tracking in the presence of hardware limitations.</p><p>Tracking accuracy in experiments. As illustrated in Fig. <ref type="figure">11</ref>(a) (top), under the proposed global-position tracking (GPT) controller, the robot's actual global position x b (labeled as "x b (GPT)") converges to a relatively small neighborhood about its desired trajectory s d within 3 s when the robot walks on a concrete floor. Also, Fig. <ref type="figure">11</ref>(a) (bottom) illustrates that despite an initial path tracking error of 5 cm, the robot remains close to the centerline C d of the desired global path, as indicated by the footstep trajectories labeled as "foot placement (GPT)." Due to uncertainties such as hardware limitations, modeling errors, and floor surface irregularity, achieving an exactly zero steady-state tracking error on a physical robot may not be feasible. Thanks to the inherent robustness of feedback control, the proposed control approach achieves a small steady-state tracking error, although uncertainties are not explicitly considered in the controller design.</p><p>Robustness. To further test the limit of the inherent robustness of the proposed control approach, experiments of OP3 walking on a ceramic tile floor were conducted (Fig. <ref type="figure">11(b)</ref>). As the surface of the ceramic tiles is relatively more slippery than the concrete floor, the robot's stance foot slips more frequently on the tile floor, causing a stronger violation of the modeling assumption of static stance foot. Yet, a relatively small global-position tracking error is still realized when the initial foot placement error is small, as shown in Fig. <ref type="figure">11(b)</ref>.</p><p>Necessity of global-position tracking control. To illustrate the need to explicitly address global-position tracking in controller design, a global-velocity tracking (GVT) controller, which is analogous to the orbitally stabilizing controller for fully actuated walking <ref type="bibr">[21]</ref>, is implemented. Its global-position tracking performance is displayed in Figs. </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="7">Discussion</head><p>This study has extended the previous method of impact invariance construction from orbital stabilization <ref type="bibr">[14]</ref> to the stabilization of time-varying global-position trajectory for 3D fully actuated robots. The proposed method produces impact invariance conditions that can be imposed in the trajectory generation of virtual constraints for ensuring their agreement with impact dynamics. Moreover, although the impact maps of the virtual constraints and global trajectory are generally nonlinearly coupled through the robot's kinematic chains, these conditions can automatically ensure any arbitrary smooth desired global-position trajectory respects the impact dynamics. Indeed, as shown in Fig. <ref type="figure">8</ref>, the proposed controller achieves asymptotic tracking of two different global-position trajectories under the same virtual constraints, indicating that the virtual constraints ensure the impact agreement for different desired global-position trajectories. Thus, the proposed impact invariance conditions can allow the decoupling between the lower-level trajectory generation of virtual constraints and the higher-level planning of global-position trajectory. The decoupling could permit offline planning of virtual constraints, thus reducing the computational load for online planning.</p><p>This study has also introduced the Lyapunov-based stability conditions for the hybrid closed-loop error system associated with 3D bipedal robots during fully actuated walking. Controller designs satisfying these conditions can accurately track the timevarying desired global-position trajectory, as demonstrated in  The proposed control approach can also indirectly drive the lateral foot placement y st to the desired location y std , which is predicted by the asymptotic convergence of the Lyapunov function V a that explicitly contains the lateral foot-placement error. Note that our previous controller for 2D walking cannot address the convergence of y st &#192; y std as it does not consider the robot's lateral movement. The capability of accurate foot placement could potentially be exploited to handle locomotion on discrete terrains (e.g., stepping stones <ref type="bibr">[41]</ref>).</p><p>In theory, for the proposed control law to achieve zero tracking error, the desired trajectory needs to respect the proposed impact invariance conditions. Yet, in practice, achieving exactly zero final tracking error may not be necessary. Rather, achieving a final error within a reasonably small bound could be sufficiently satisfactory for practical applications. To this end, the proposed impact invariance conditions can be theoretically relaxed to allow bounded violation of the condition for stabilizing the origin of the tracking error system in the sense of Lyapunov stability rather than asymptotic stability. Specifically, we can incorporate the impact invariance conditions as an inequality constraint in trajectory generation instead of an equality constraint.</p><p>The proposed control approach, including continuous input-output linearizing control design and the impact invariance construction, builds upon a hybrid robot model that holds under several modeling simplifying assumptions. These assumptions are reasonable for robot walking in relatively structured environments. Yet, they may not hold if the environments are more complex, which will in turn affect the controller performance. Indeed, as the experiment results in Fig. <ref type="figure">11</ref> illustrate, when the floor is relatively slippery, the assumption that the foot and surface have a secured contact no longer holds, which is not explicitly addressed by the proposed control law. For this reason, these experiment results show a slower convergence rate and larger tracking error compared with the simulation results in Fig. <ref type="figure">8</ref>. To this end, adapting the proposed controller to more complex environments is necessary. For instance, we could cast the proposed control approach into a quadratic program <ref type="bibr">[42]</ref> for ensuring the feasibility of ground contact forces in the presence of foot slippage induced by surface irregularity. Another potential approach is to integrate the proposed control law with adaptive and robust control action <ref type="bibr">[26,</ref><ref type="bibr">43,</ref><ref type="bibr">44]</ref> for enabling online model estimation and better disturbance rejection.</p><p>The proposed controller is built upon a fully actuated robot model, and is thus effective when the robot is fully actuated. For a bipedal robot to be fully actuated, its motion needs to satisfy certain necessary constraints. For instance, a bipedal robot with finite size feet (e.g., OP3) is fully actuated when its support foot is in a static, full contact with the ground during walking; that is, its motion satisfies the ground contact constraints (e.g., friction cone, unilateral, and center of pressure constraints). Thus, the planned motion (defined by the virtual constraints and the desired global trajectory) should meet these constraints with a reasonable margin. Here the margin ensures that the actual walking motion also satisfy those constraints when it is near the planned motion, so that the controller will be effective in driving the actual motion to the desired one. Note that the proposed controller is not intended for robots with limited-size feet (e.g., point feet) or passive ankles, which commonly adopt underactuated gait due to the lack of control authority compared with the robot's degrees-of-freedom. Still, the proposed control method could be extended to multidomain walking gait, which comprises subphases of full actuation, underactuation, and over actuation, as our preliminary theoretical and simulation studies indicate <ref type="bibr">[29]</ref>. The key to this extension is the adaptation of the proposed stability conditions from single-mode to multimode hybrid robot dynamics.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="8">Conclusion and Future Work</head><p>This paper has introduced a control approach that explicitly addresses the hybrid dynamics of 3D bipedal robots for achieving asymptotic global-position tracking during fully actuated walking. With the output function designed as the tracking error of the desired global-position trajectory and virtual constraints, a continuous input-output linearizing control law was synthesized to asymptotically drive the output function to zero within continuous phases. Impact invariance conditions were derived to guide the generation of virtual constraints such that the robot's desired motions defined by the virtual constraints and the desired globalposition trajectory all respect the discrete landing impact dynamics. Sufficient conditions were derived based on Lyapunov theory under which the proposed continuous control law provably guarantees the asymptotic tracking performance of the hybrid closedloop system. Simulation and experiment results demonstrated the effectiveness of the proposed control approach in realizing satisfactory global-position tracking.</p><p>Our future work will apply and extend the proposed approach from straight-line to curved-path locomotion as real-world applications of legged robots commonly require walking in varying directions. To enable efficient planning, we will construct a library <ref type="bibr">[20]</ref> of virtual constraints offline that corresponds to a common range of direction-varying gait parameters and interpolate the virtual constraints online to fit the varying walking directions along a curved path.</p><p>Directorate for Engineering, National Science Foundation (Grant No. CMMI-1934280; Funder ID: 10.13039/100000084).</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>Appendix: Proofs of All Propositions and Theorems</head><p>A.1 Proof of Proposition 1 Proof 1. With the pre-impact joint position q &#195; , the postimpact joint position and phase variable are</p><p>respectively. Under the foot-placement condition y st &#188; y std and condition (A1), the postimpact value of the output function </p><p>where J hc &#240;q &#195; &#222; &#188; @hc @q &#240;q &#195; &#222;. Thus, the pre-impact joint velocity is</p><p>Thus, the impact invariance of Z is met under conditions (A1) and (A2) from Proposition 1.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>A.2 Proof of Proposition 2</head><p>Proof 2. Because x b and s d &#240;t&#222; are both continuous in t, we obtain</p><p>Because the stance foot remains static just before and after the impact, _ </p><p>. Then, the postimpact value of the first time derivative of the output function</p><p>x &#192; st &#188; 0. Thus, under conditions (A1)-(A3) from Propositions 1 and 2, the impact invariance of Z holds.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>A.3 Proof of Proposition 3</head><p>Proof 3. Because the output function state y and _ y and the swing-foot height z sw defining the switching surface S are both continuously differentiable in their respective arguments, the function defining the switching surface S x is continuously differentiable in its argument <ref type="bibr">[28]</ref>. Also, note that the continuous-phase vector field (i.e., Ax) of the error state x is continuously differentiable in x.</p><p>Then, by Lemma 2.1 and Corollary 2.4 in Ref. <ref type="bibr">[28]</ref>, the impact timing T k is an implicit function of the state x, and is Lipschitz continuous with respect to x. Thus, there exists a positive number r 1 and a Lipschitz constant L Tx such that jT k &#192; s k j L Tx jjx&#240;s k ; T &#254; k&#192;1 ; xj &#254; k&#192;1 &#222;jj for any xj &#254; 0 2 B r1 &#240;0&#222; and any k 2 Z &#254; .</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>A.4 Proof of Proposition 4</head><p>Proof 4. Let / sw;y &#240;h&#222; denote the desired trajectory of the control variable y sw . Because the stance-foot position during the &#240;k &#254; 1&#222; th step is the swing-foot position at the end of the kth step, one has </p><p>Then, from Eqs. (A12) and (A13), the postimpact error norm can be approximated as</p><p>&#254; L Dst jy st j &#192; k &#192; y std j (A14)</p><p>For any e &gt; 0 there exist PD gains corresponding to a sufficiently high convergence rate c3 2c2 such that e</p><p>1 &#254; e. Then, the approximation of the postimpact error norm can be simplified into jjxj &#254; k jj a x jjxj &#254; k&#192;1 jj &#254; a st jy st j &#192; k &#192; y std j (A15)</p><p>where a x :&#188; ffiffiffi</p><p>Dsk ; Ds k :&#188; s k &#192; T k&#192;1 , and a st :&#188; L Dst . Now, we derive the upper bound of jy st j &#192; k &#192; y std j with respect to the tracking error norm jjxj &#192; k&#192;1 jj. Because the stance foot remains static within a step, we have y st j &#192; k &#188; y st j &#254; k&#192;1 : Then, from Eq. ( <ref type="formula">17</ref>)</p><p>holds, where c x :&#188; ffiffiffi</p><p>Dsk .</p><p>Finally, combining Eqs. ( <ref type="formula">13</ref>), (A15), and (A16) provides the following approximation of the postimpact value of the Lyapunov function V a :</p><p>where B :&#188; max&#240; ; 2c2ast r &#222;.</p><p>Evolution of V a for the hybrid model. If the PD gains and r are chosen such that 2c 2 a 2</p><p>x &#254; rc 2 x c 1 &lt; 1 and 2c 2 a st r &lt; 1 (A17) hold (i.e., B &lt; 1), then for any xj &#254; 0 2 B d2 &#240;0&#222;, the sequence fV a j &#254; 1 ; V a j &#254; 2 ; V a j &#254; 3 &#8230;g is strictly decreasing with V a j &#254; k ! 0 as k ! 1. Thus, the closed-loop hybrid system is locally asymptotically stable if the PD gains ensure that the matrix A is Hurwitz and that Eq. (A17) holds for any xj &#254; 0 2 B d2 &#240;0&#222;. To meet the two inequality conditions in Eq. (A17), we can choose the function V&#240;x&#222; to be V&#240;x&#222; &#188; x T Px as explained in Remark 6. This choice results in the continuous-phase convergence rate of V&#240;x&#222; as c3 c2 &#188; kQ kmax&#240;P&#222; , which can be tuned with the PD gains. Specifically, to satisfy the second inequality in Eq. (A17), we can specify r as any positive number such that r &gt; 2c 2 a st &#188; 2k max &#240;P&#222;a st , where a st can be estimated from system dynamics. For instance, we can choose r to be 2k r k max &#240;P&#222;a st with any constant k r &gt; 1. Then, we can tune the PD gains to meet the first inequality in Eq. (A17), by allowing a sufficiently high continuous-phase convergence rate that leads to sufficiently small values of a x and c x for satisfying a 2</p><p>x &#254; c 2 x c1 2kmax&#240;P&#222;max&#240;1;krast&#222; . Convergence of impact timings. When the state x reaches zero at the steady-state, from Eq. (A13), the fictitious state satisfies</p><p>&#240;sk&#192;Tk&#192;1&#222; jjxj &#254; k&#192;1 jj ! 0 as k ! 1. Then, by Eq. ( <ref type="formula">16</ref>),</p><p>Convergence of lateral foot placement. By the definition of V a in Eq. ( <ref type="formula">18</ref>), V a &#240;x; y st &#192; y std &#222; &#188; V&#240;x&#222; &#254; r&#240;y st &#192; y std &#222; 2 , where r is positive and V&#240;x&#222; and &#240;y st &#192; y std &#222; 2 are all bounded and nonnegative. Thus, if V a ! 0 as t ! 1, then &#240;y st &#192; y std &#222; 2 ! 0 as t ! 1; that is, y st ! y std as t ! 1.</p></div><note xmlns="http://www.tei-c.org/ns/1.0" place="foot" xml:id="foot_0"><p>Downloaded from http://asmedigitalcollection.asme.org/dynamicsystems/article-pdf/144/11/111001/6910527/ds_144_11_111001.pdf by Purdue University at West Lafayette user on 05 February 2023</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="2" xml:id="foot_1"><p>https://youtu.be/VJbLMkOG_xo</p></note>
		</body>
		</text>
</TEI>
