Abstract
The use of manipulators in space missions has become popular, as their applications can be extended to various space missions such as on-orbit servicing, assembly, and debris removal. Due to space reachability limitations, such robots must accomplish their tasks in space autonomously and under severe operating conditions such as the occurrence of faults or uncertainties. For robots and manipulators used in space missions, this paper provides a unique, robust control technique based on Model Predictive Path Integral Control (MPPI). The proposed algorithm, named Planner-Estimator MPPI (PE-MPPI), comprises a planner and an estimator. The planner controls a system, while the estimator modifies the system parameters in the case of parameter uncertainties. The performance of the proposed controller is investigated under parameter uncertainties and system component failure in the pre-capture phase of the debris removal mission. Simulation results confirm the superior performance of PE-MPPI against vanilla MPPI.
1 Introduction
The application of a space robot, a manipulator connected to a free-flying base, is becoming more popular as it can be extended to different space missions (Figure 1) (). Many space missions include several tasks such as inspecting, refueling, assembling and constructing, and removing space debris. Currently, these operations are performed by astronaut Extravehicular Activities (EVA). However, the risky nature of such operations can threaten astronauts’ life and require careful preparation. A suitable solution is performing such operations by space manipulators (). Being small makes these manipulators perfect for moving around the main satellite with faster acceleration.
FIGURE 1
Small space robots such as the Future Space Debris Removal Orbital Manipulator (FSDROM) can play a significant role in future debris removal missions (
Applying space robots in rigid debris removal missions is challenging since space debris are mainly non-cooperative moving objects that do not provide any information to track them. Several missions for on-orbit rigid capturing using space manipulators demonstrated their potential for future space missions. For instance, the Engineering Test Satellite VII “KIKU-7” (ETS-VII) by the Japan Aerospace Exploration Agency (JAXA) in 1997 was among the pioneers in demonstrating space robotic capability using chasers and target satellites (
Space robots’ operation and performance in capturing space debris rely on their control systems. However, there are some concerns associated with the design of space robot control systems due to the following points:
Presently, the average lifespan of some satellites is approximately 14 years (
Human intervention (telerobotics) in space missions is difficult. For instance, in
Space manipulators are in direct contact with unidentified rotating debris, and damage to the actuators and the robot’s structure is unavoidable. Therefore, any controlling law shall be sufficiently robust to maintain its performance in the possibilities of an actuator failure or malfunction. This event is more possible in the case of direct capturing methods where capturing of objects can cause large impacts on the spacecraft (
Accurate identification of system parameters is inevitable in rigid capturing missions where many parameters such as inertia, friction, geometry, and attitude must be identified to ensure the controller’s performance (
Fulfilling such requirements through classical control approaches is not a trivial task due to their limitation in handling system uncertainties and contact modeling. Recently, model predictive control (MPC) for robot controls has received significant attention from academia and industry due to its benefits, such as the power to handle constraints (
In the present study, we consider some assumptions to develop our method. Firstly, in the context of PE-MPPI all uncertainties are supposed to be structural, and unstructured uncertainties cannot be handled efficiently by the proposed algorithm. Secondly, we do not address directly the saturation problem of control effort. Instead, by defining a cost for actions we can indirectly penalize control efforts to be as small as possible.
This paper is structured as follows: Section II describes current state-of-the-art control systems and techniques for space robots in space missions. Section III explains the kinematics and dynamics formulation of space robots. In section IV, the MPPI algorithm is described. The extension to this algorithm, which is the main contribution of this paper, is then explained in section V. The simulation environment, robot operation scenarios, and simulation results are presented in section VI. Finally, the conclusions and future works are outlined in section VII.
2 Related works
Parameters of a space manipulator are reasonably measured and applied for controller design before launching to space. However, some parameters such as the joints’ damping coefficient and stiffness can change over time. Hence, on-orbit identification is required to guarantee the space robot’s performance (
Designing a motion-planning framework for space manipulators has been extensively investigated, taking into account dynamic coupling and singularities, as well as the physical restrictions of space robots. For instance, researchers attempted to solve the trajectory planning problem by minimizing a cost function that satisfies specific criteria, e.g., power consumption (
Recently, reinforcement learning has received significant attention from robotic researchers due to its strength in controlling nonlinear dynamic systems. The reinforcement learning techniques can be classified as model-free and model-based techniques. The model-free techniques train a robot agent through interaction with the environment. Model-free reinforcement learning is a powerful technique in controlling complex dynamic systems as they do not use the model of the system. However, it suffers from sample efficiency and a long training time.
Model predictive control (MPC) is an advanced control method that, similar to model-based reinforcement learning, uses a system model to predict the system’s future behavior. MPC solves an online optimization algorithm to find the optimal control action that drives the predicted output to the reference. One of the state-of-the-art model predictive control techniques is Model Predictive Path Integral Control (MPPI) (
3 Prerequisites
3.1 Kinematics of a space robot
The kinematics of industrial manipulators depends only on the parameters of the joint space, whereas the kinematics of the space robots is more complex than terrestrial robots. The kinematics of a space robot is determined based on the position and orientation of the base and joint parameters.
According to Figure 2, the space robot can be represented as a set of n+1 rigid links connected with n joints, resulting in n+6 degrees of freedom. Furthermore, ΣC is the inertial coordinates system, and ΣB the base coordinates system attached on the base with its origin at the centroid of the base. Therefore, the position of the end-effector can be obtained as follows:where:
FIGURE 2

The configuration of a space robot and the coordinates of the joints.
pe: The position vector of the end-effector in the coordinates system ΣC
r0: The position vector of the centroid of the base in the coordinates system ΣC
l0: The connection vector from the base to the first joint
li: The connection vector from joint i to joint i+1.
By differentiating the kinematic equation with respect to time, the relation between the velocity of the end-effector and the velocity of the joints can be obtained as follows:where:
: The linear/angular velocity of the end-effector in the inertial coordinates system.
: The angular velocity of the joints.
: The linear/angular velocity of the base in the base coordinates system.
Jm: The Jacobian matrix of the manipulator.
Jb: The Jacobian matrix of the base.
3.2 Dynamics of a space robot
The dynamics of space robots are more complicated than terrestrial robots due to the dynamics coupling effect between the manipulator arm and its base. For instance, the space robot base would react based on the momentum conservation theorem if torque τi is applied to the ith joint (
Hb: The inertial matrix of the base.
Hm: The inertial matrix of the manipulator arm.
Hbm: The coupling inertial matrix between the base and the manipulator arm
cb: The velocity-dependent non-linear term of the base
cm: The velocity-dependent non-linear term of the manipulator arm.
Fb: The force and torque on the centroid of the base.
Fh: The force and torque on the end-effector
τ: The joint torque of the manipulator arm.
When no external forces are applied to the end-effector (Fh = 0), and the thrusters (or reaction wheels) do not apply force to the spacecraft base (Fb = 0), the above dynamic equation will be reduced to the following form:
where p and L are linear and angular momentums, which are constant values. The free-floating space robots are divided into two sub-types where the initial momentum is zero or no-zero (
4 Model predictive path integral control
Model predictive path integral control (MPPI) is an importance-sampling method. Its derivative-free behavior makes it an excellent choice for optimal control problems with nonlinear dynamics and non-convex cost functions. The fundamental notion of MPPI is to sample many trajectories for a time horizon of T from a dynamical system. Each trajectory τ = {x0, u0, x1, u1, … , xT, uT} is then evaluated according to a cost function. Accordingly, the optimal trajectory is computed based on its importance over all trajectories. To determine near-optimal solutions, increasing the number of trajectories is necessary. Fortunately, this can be quickly accomplished by taking advantage of the parallel nature of sampling and using Graphical Processor Unit (GPU) (
Consider a discrete-time dynamical system as follows:where xt is the state vector, ut is the control input vector, and δut is the random vector sampled from a zero-mean Gaussian distribution N (0, Σu) at time-step t. As mentioned, each trajectory can be evaluated with a cost function as follows:where ϕ(xT) and q (xt, ut) are the terminal and running costs, respectively. MPPI aims to find the optimal control input trajectory u* = (u0, u1, … , uT), which minimizes the expectation over all generated trajectories as follows:
The solution to this problem has been discussed in
where K is the number of trajectories, and λ is called inverse temperature. The detailed MPPI algorithm is described in Algorithm 1.
Algorithm 1

Algorithm 2

5 Planner-estimator MPPI
This section proposes a novel Planner-Estimator MPPI (PE-MPPI) strategy to control space robots in on-orbit debris removal missions, which can fulfill controller design requirements. First, the controller structure will be given, and lastly, the proposed algorithm will be explained.
Although many studies have shown the performance of MPPI in different scenarios, its performance varies with model accuracy. To make this controller suitable for space explorations, we propose PE-MPPI to robustify the performance of MPPI against structural uncertainties. PE-MPPI is composed of two parts: Planner MPPI and Estimator MPPI. As shown in Figure 3, Planner MPPI selects the optimal control action based on the on-board model . The structure of Planner MPPI is the same as MPPI. It only computes the control input of the system based on the on-board model. On the other hand, Estimator MPPI attempts to estimate the model parameters and readjust the on-board model of the robot based on the norm of an error signal. In other words, whenever the on-board model fails to match the dynamic behavior of the space manipulator, Estimator MPPI estimates the model parameters and updates the model accordingly. The core idea of estimation is to sample many parameters from a Gaussian distribution and evaluate them as follows:where is the running cost for the trajectory generated with the parameter . Consequently, the update law of the parameters is formulated as below:
FIGURE 3

Schematic of planner-estimator MPPI
It is important to say that the estimated model does not necessarily match the real system, but it guarantees that they would have the same response after sufficient updates.
Algorithm 2 explains PE-MPPI in detail. Based on the parameterized model with parameters , Planner MPPI outputs near-optimal control effort ut at each time-step (code lines:7 and 8). Each response of the space robot xt and the subsequent control input ut is gathered in a replay buffer B (xt, ut) (code line: 9). The sensors of the space robot measure the response of the real system xt+1, while the response of the on-board model is calculated by the on-board model (code lines: 10 and 11). If the norm of the signal error is greater than a pre-defined threshold, Estimator MPPI updates the on-board model (line code: 12). To find the optimal parameter , many parameters are sampled from a Gaussian distribution, and the score of each trajectory is calculated using the running cost (code lines: 13–20). Then, the parameters of the update law are calculated, and the optimal parameters of the model are computed using the update law (code lines: 21–23). Finally, the on-board model is updated (code line: 27).
6 Simulation
This section investigates the performance of PE-MPPI in a MuJoCo simulation (
FIGURE 4

The rest configuration of the space robot and the coordinates systems of the joints.
6.1 The general specifications of the space robot
The space robot consists of a base and a manipulator connected to the base. In non-operational conditions, the manipulator is in its resting position, folded around the base (Figure 4). However, in cases where debris is located far from the main satellite’s structure, the mission is launched to remove or catch the debris with the help of the manipulator. The 7-DoF manipulator’s length then unfolds to allow the space robot to reach far debris zones. The redundant degree of freedom assures the robot’s performance even in actuator failure conditions.
The Denavit-Hartenberg parameters of the manipulator and the inertial properties of the space robot used in this simulation are given in Tables 1, 2, respectively.
TABLE 1
| Joint | α(rad) | a (m) | d (m) | θ(rad) |
|---|---|---|---|---|
| 1 | 0.0 | 0.5 | θ1 | |
| 2 | 0.0 | 0.0 | θ2 | |
| 3 | 0.9 | 0.0 | θ3 | |
| 4 | 0.9 | 0.0 | θ4 | |
| 5 | 0.8 | 0.0 | θ5 | |
| 6 | 0.8 | 0.0 | θ6 | |
| 7 | 0.0 | 0.8 | θ7 |
The DH parameters of the space robot.
TABLE 2
| Base | L1 | L2 | L3 | L4 | L5 | L6 | L7 | |
|---|---|---|---|---|---|---|---|---|
| M(kg) | 500 | 20 | 30.0 | 30.0 | 20.0 | 20.0 | 20.0 | 20.0 |
| Ix (kg.m2) | 1,400 | 0.1 | 0.25 | 0.25 | 0.25 | 0.25 | 0.25 | 0.25 |
| Iy(kg.m2) | 1,400 | 0.1 | 25 | 25 | 25 | 25 | 25 | 25 |
| Iz (kg.m2) | 1,400 | 0.1 | 25 | 25 | 25 | 25 | 25 | 25 |
The inertial properties of the space robot.
6.2 Operational scenarios of the space robot
6.2.1 Normal operation
In normal operation, no actuator failure or system degradation occurs. Therefore, the on-board model accurately tracks the response of the real system. In this perfect situation, the spacecraft is commanded to traverse on y-axis while its manipulator approaches from the initial position xinitial = [−1.2,−1.2,0]T to the desired target debris site xtarget = [−2,8,0]T. The mission requirements are i) to reach the debris site, ii) to maneuver on orbit stack around axis y, and iii) to reduce control effort. Since there is no parameter uncertainties, only Planner MPPI is used. In order to meet the requirements of the mission, the cost function of Planner MPPI is designed as follows:where:
xtarget: The position of the target debris site
xend−effector: The position of the end-effector of the manipulator
xbase: The position of the base
xorbit: The position of the orbit
u: The control effort.The position of the end-effector relative to the inertial coordinate and the position of the space robot base are illustrated in Figure 5. After 60 s, the end-effector approaches the target site and maintains its position. The steady-state error in this mission is less than 15 cm, which is acceptable. Moreover, the space robot base position successfully tracks the orbit position, which is the y-axis.
FIGURE 5

The position of the end-effector reaches the target position after 60 s in the normal operation scenario (A). The space robot base position is traversed along the y-orbit (B).
6.2.2 System identification
The damping coefficient of the space robot joints is assumed to differ from the on-board model parameters in the second scenario. The difference between the model and the real system can result in poor approaching behavior. Hence, adopting a strategy to identify the system’s parameters in real-time is crucial in this mission. Thus, both Planner MPPI and Estimator MPPI are used. Since the goal of the mission is the same as the normal operation scenario, the cost function of the Planner MPPI is the same. On the other hand, the running cost function of Estimator MPPI is defined as below:
The damping coefficients of the on-board model were set to at the beginning of the simulation, while the damping coefficients of the real system were one-tenth of the damping coefficients of the on-board model. A comparison between the performance of PE-MPPI and vanilla MPPI applied to the model with parameter uncertainties is illustrated in Figure 6A. Moreover, the convergence of damping coefficients is depicted in Figure 6A. PE-MPPI can reach the target position in the system identification mission after 70 s. In contrast, the performance of vanilla MPPI deteriorates due to the lack of a mechanism for adjusting the parameters of the model. All parameters converge to the real system parameters after 20 s, while there is a significant error in estimating the first and last parameters. However, these errors have little impact on system performance as the end-effector can reach the debris site after 70 s. It can be concluded that estimating the parameters increases the stability of the system and reduces the steady-state error resulting in better performance. In addition, as is shown in Figure 6B, for both PE-MPPI and vanilla MPPI, the space robot base position is traversed on the y-axis. Since the parameter uncertainties are related to the joint parameters, the parameter uncertainties mainly affect the end-effector position rather than the base position.
FIGURE 6

System identification scenario; Comparison between PE-MPPI and vanilla MPPI for the end-effector position (A; Top). Convergence of the parameters of the on-board model to real system (A; Bottom). Comparison between PE-MPPI and vanilla MPPI for the space robot base position (B).
6.2.3 Actuator failure
Due to many sources of failure in the space missions, such as debris collision or system degeneration, actuator failure can happen during the robot’s lifespan. The main challenge is that the system dynamics will change suddenly, resulting in instability and poor performance. In this critical condition, the source of failure is well understood; hence parameter estimation is not required and Estimator MPPI is not used. However, adopting a robust and adaptable control strategy, which can alter in real-time, is required to guarantee the system’s stability with minimum human intervention. The cost of Planner MPPI is the same as the two previous scenarios.
In the third scenario, the space robot will lose one of its degrees of freedom, and consequently, this actuator cannot be controlled anymore (the second actuator is chosen to be locked). The performance of PE-MPPI is compared to vanilla MPPI in which the on-board model is not changed by actuator failure. As shown in Figure 7, a lack of updating mechanism for the on-board model in vanilla MPPI causes poor performance compared to PE-MPPI, and it can conveniently update its model and successfully reach the target position and remain at this position after 60 s. Moreover, the base position is traversed on the y-axis. Similar to the system identification scenario, since actuator failure is mainly related to the joint space, it affects the end-effector position more than the base position.
FIGURE 7

Comparison between PE-MPPI and vanilla MPPI in the actuator failure scenario for the end-effector position (A) and the space robot base position (B).
6.2.4 System identification and actuator failure
In the last and worst scenario, both actuator failure and system parameter change occur simultaneously. In this condition, the estimator section would help the planner to control the space robot and reach the desired position while the failed actuator (the third actuator is chosen) is locked. The cost function of PE-MPPI is the same as the system identification scenario. Similar to the second scenario, all damping coefficients were initialized to be while the real system parameters were one-tenth of the on-board model.
As shown in Figure 8A, after 20 s, all parameters converged to the real system parameters, while there was a significant error in estimating the first and last parameters. The estimated parameters showed more fluctuation compared to the system identification scenario, indicating the combination of events could reduce the controller’s performance in both estimating parameters and steady-state error. Moreover, PE-MMPI takes more time to reach the target position (after 70 s), while vanilla MPPI cannot accomplish the mission (Figure 8A). In addition, the base position successfully traverses on the y-axis (Figure 8B). Figure 9 shows the bounds of the control effort of both PE-MPPI and vanilla MPPI for the system identification and actuator failure scenario. As it is expected, PE-MPPI needs more control effort than vanilla MPPI, since it manages parameter uncertainties and actuator failure.
FIGURE 8

System identification and actuator failure scenario; Comparison between PE-MPPI and vanilla MPPI for the end-effector position (A; Top). Convergence of the parameters of the on-board model to the real system (A; Bottom) Comparison between PE-MPPI and vanilla MPPI for the space robot base position (B).
FIGURE 9

The space robot control effort bounds in the system identification and actuator failure scenario.
7 Conclusion
This study proposed a novel Planner-Estimator MPPI (PE-MPPI) algorithm to control space robots in debris removal pre-capture phase missions subject to system malfunctioning and structured parameter changes. Four scenarios were considered for testing the controller’s performance: normal operation, system identification, actuator failure, and combined system identification and actuator failure. In each scenario, the performance of PE-MPPI is compared to vanilla MPPI. Results proved the superiority of the proposed algorithm over vanilla MPPI, especially in the fourth scenario, where the combination of events results in poor performance. It was shown that PE-MPPI could maintain its performance in different scenarios, with negligible degeneration compared to normal operation. Furthermore, the estimator assures that the on-board model tracks the real system, while some errors are in estimating parameters (especially the first and last actuators’ damping coefficient). It is worth mentioning that the convergence of damping coefficients to their real values is not guaranteed, but the norm of difference signal would be minimized.
Statements
Data availability statement
The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author.
Author contributions
MR and AN contributed to the concept and implementation of the project. MR, AN, and SF wrote the first draft of the manuscript.
Conflict of interest
The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.
Publisher’s note
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.
References
1
AghiliF. (2020). Optimal trajectories and robot control for detumbling a non-cooperative satellite. J. Guid. Cont. Dyn.43, 981–988. 10.2514/1.g004758
2
ArrudaE.MathewM. J.KopickiM.MistryM.AzadM.WyattJ. L. (2017). “Uncertainty averse pushing with model predictive path integral control,” in 2017 IEEE-RAS 17th International Conference on Humanoid Robotics (Humanoids) (IEEE), 497–502.
3
BasmadjiF. L.SewerynK.SasiadekJ. Z. (2020). Space robot motion planning in the presence of nonconserved linear and angular momenta. Multibody Syst. Dyn.50, 71–96. 10.1007/s11044-020-09753-x
4
BiesbroekR.InnocentiL.WolahanA.SerranoS. M. (2017). “e. deorbit-esa’s active debris removal mission,” in Proceedings of the 7th European Conference on Space Debris (ESA Space Debris Office).
5
BillotC.FerrarisS.RembalaR.CacciatoreF.TomassiniA.BiesbroekR. (2014). “E. Deorbit: Feasibility study for an active debris removal,” in 3rd European Workshop on Space Debris Modeling and Remediation. Paris, France: Centre National d’Etudes Spatiales,
6
BroidaJ.LinaresR. (2019). “Spacecraft rendezvous guidance in cluttered environments via reinforcement learning,” in 29th AAS/AIAA Space Flight Mechanics Meeting (American Astronautical Society Ka’anapali, Hawaii), 1–15.
7
ChatterjeeJ. (2014). “Legal issues relating to unauthorised space debris remediation,” in 65th International Astronautical Congress, 1–20.
8
Christidi-LoumpasefskiO.-O.NanosK.PapadopoulosE. (2017). “On parameter estimation of space manipulator systems using the angular momentum conservation,” in 2017 IEEE International Conference on Robotics and Automation (ICRA) (IEEE), 5453. 8.
9
DixitS.MontanaroU.DianatiM.OxtobyD.MizutaniT.MouzakitisA.et al (2019). Trajectory planning for autonomous high-speed overtaking in structured environments using robust mpc. IEEE Trans. Intell. Transp. Syst.21, 2310–2323. 10.1109/tits.2019.2916354
10
ForshawJ.AgliettiG.SalmonT.RetatI.BurgessC.ChabotT.et al (2017). The removedebris adr mission: Preparing for an international space station launch. In 7th European Conference on Space Debris.
11
GandhiM. S.VlahovB.GibsonJ.WilliamsG.TheodorouE. A. (2021). Robust model predictive path integral control: Analysis and performance guarantees. IEEE Robot. Autom. Lett.6, 1423–1430. 10.1109/lra.2021.3057563
12
HewingL.WabersichK. P.MennerM.ZeilingerM. N. (2020). Learning-based model predictive control: Toward safe learning in control. Annu. Rev. Control Robot. Auton. Syst.3, 269–296. 10.1146/annurev-control-090419-075625
13
HuangP.XuY.LiangB. (2006). Tracking trajectory planning of space manipulator for capturing operation. Int. J. Adv. Robotic Syst.3, 31. 10.5772/5735
14
KimT.ParkG.KwakK.BaeJ.LeeW. (2022). Smooth model predictive path integral control without smoothing. IEEE Robot. Autom. Lett.7, 10406–10413. 10.1109/lra.2022.3192800
15
LowreyK.RajeswaranA.KakadeS.TodorovE.MordatchI. (2018). “Plan online, learn offline: Efficient learning and exploration via model-based control,” in International Conference on Learning Representations.
16
MohamedI. S.AllibertG.MartinetP. (2020). “Model predictive path integral control framework for partially observable navigation: A quadrotor case study,” in 2020 16th International Conference on Control, Automation, Robotics and Vision (ICARCV) (IEEE).
17
MorganA. S.NandhaD.ChalvatzakiG.D’EramoC.DollarA. M.PetersJ. (2021). “Model predictive actor-critic: Accelerating robot skill acquisition with deep reinforcement learning,” in 2021 IEEE International Conference on Robotics and Automation (ICRA), 6672.
18
MuZ.XuW.LiangB. (2017). Avoidance of multiple moving obstacles during active debris removal using a redundant space manipulator. Int. J. Control Autom. Syst.15, 815–826. 10.1007/s12555-015-0455-7
19
NanosK.PapadopoulosE. (2011). On the use of free-floating space robots in the presence of angular momentum. Intell. Serv. Robot.4, 3–15. 10.1007/s11370-010-0083-2
20
NanosK.PapadopoulosE. G. (2017). On the dynamics and control of free-floating space manipulator systems in the presence of angular momentum. Front. Robot. AI4, 26. 10.3389/frobt.2017.00026
21
PapadopoulosE.AghiliF.MaO.LamparielloR. (2021). Robotic manipulation and capture in space: A survey. Front. Robot. AI8, 686723. 10.3389/frobt.2021.686723
22
PravitraJ.AckermanK. A.CaoC.HovakimyanN.TheodorouE. A. (2020). “L 1-adaptive mppi architecture for robust and agile control of multirotors,” in 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) (IEEE).
23
RybusT.SewerynK.SasiadekJ. Z. (2016). “Trajectory optimization of space manipulator with non-zero angular momentum during orbital capture maneuver,” in AIAA Guidance, Navigation, and Control Conference.
24
SewerynK.BanaszkiewiczM. (2008). “Optimization of the trajectory of a general free-flying manipulator during the rendezvous maneuver,” in AIAA Guidance, Navigation and Control Conference and Exhibit, 7273.
25
SewerynK.BasmadjiF. L.RybusT. (2022). Space robot performance during tangent capture of an uncontrolled target satellite. J. Astronaut. Sci.69, 1017–1047. 10.1007/s40295-022-00330-2
26
ShyamR. A.HaoZ.MontanaroU.DixitS.RathinamA.GaoY.et al (2021). Autonomous robots for space: Trajectory learning and adaptation using imitation. Front. Robot. AI8, 638849. 10.3389/frobt.2021.638849
27
TodorovE.ErezT.TassaY. (2012). “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ international conference on intelligent robots and systems (IEEE), 5026–5033.
28
TomaszewskaJ.WochM.KrzyszkowskiJ.ZiejaM. (2019). Comparative analysis of vitality of gps and glonass satellite systems. Transp. Res. Procedia43, 57–62. 10.1016/j.trpro.2019.12.019
29
WilliamsG.DrewsP.GoldfainB.RehgJ. M.TheodorouE. A. (2016). “Aggressive driving with model predictive path integral control,” in 2016 IEEE International Conference on Robotics and Automation (ICRA) (IEEE), 1433–1440.
30
WilliamsG.AldrichA.TheodorouE. A. (2017a). Model predictive path integral control: From theory to parallel computation. J. Guid. Control, Dyn.40, 344–357. 10.2514/1.g001921
31
WilliamsG.WagenerN.GoldfainB.DrewsP.RehgJ. M.BootsB.et al (2017b). “Information theoretic mpc for model-based reinforcement learning,” in 2017 IEEE International Conference on Robotics and Automation (ICRA) (IEEE).
32
WilliamsG.GoldfainB.DrewsP.SaigolK.RehgJ. M.TheodorouE. A. (2018). “Robust sampling based model predictive control with sparse objective information,” in Robotics: Science and Systems.
33
WuY.-H.YuZ.-C.LiC.-Y.HeM.-J.HuaB.ChenZ.-M. (2020). Reinforcement learning in dual-arm trajectory planning for a free-floating space robot. Aerosp. Sci. Technol.98, 105657. 10.1016/j.ast.2019.105657
34
YoshidaK. (2003). Engineering test satellite vii flight experiments for space robot dynamics and control: Theories on laboratory test beds ten years ago, now in orbit. Int. J. Robotics Res.22, 321–335. 10.1177/0278364903022005003
35
ZhangF.HuangP. (2016). Releasing dynamics and stability control of maneuverable tethered space net. Ieee. ASME. Trans. Mechatron.22, 983–993. 10.1109/tmech.2016.2628052
36
ZhangX.LiuJ. (2018). Effective motion planning strategy for space robot capturing targets under consideration of the berth position. Acta Astronaut.148, 403–416. 10.1016/j.actaastro.2018.04.029
37
ZhaoP.LiuJ.WuC. (2020). Survey on research and development of on-orbit active debris removal methods. Sci. China Technol. Sci.63, 2188–2210. 10.1007/s11431-020-1661-7
Summary
Keywords
space robots, model predictive path integral control, space debris removal, parameter uncertainity, planner-estimator model predictive path integral controller
Citation
Raisi M, Noohian A and Fallah S (2022) A fault-tolerant and robust controller using model predictive path integral control for free-flying space robots. Front. Robot. AI 9:1027918. doi: 10.3389/frobt.2022.1027918
Received
25 August 2022
Accepted
23 November 2022
Published
07 December 2022
Volume
9 - 2022
Edited by
Arun Misra, McGill University, Canada
Reviewed by
Serdar Kalaycioglu, Ryerson University, Canada
Karol Seweryn, Space Research Center, Polish Academy of Sciences, Poland
Updates

Check for updates
Copyright
© 2022 Raisi, Noohian and Fallah.
This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.
*Correspondence: Saber Fallah, s.fallah@surrey.ac.uk
This article was submitted to Space Robotics, a section of the journal Frontiers in Robotics and AI
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.