Combining Supervised Pretraining and Reinforcement Learning for Scalable Low-Thrust Trajectory Design

dc.contributor.advisorMartin, John Ren_US
dc.contributor.authorSchmidt, Michael Christianen_US
dc.contributor.departmentAerospace Engineeringen_US
dc.contributor.publisherDigital Repository at the University of Marylanden_US
dc.contributor.publisherUniversity of Maryland (College Park, Md.)en_US
dc.date.accessioned2026-07-01T05:39:23Z
dc.date.issued2026en_US
dc.description.abstractThis thesis considers a hybrid machine learning training framework that combines supervised pretraining with deep reinforcement learning to approximate optimal low-thrust control for heliocentric transfer and rendezvous trajectories. The primary test case is a circular two-body transfer problem in which a continuously thrusting spacecraft is transferred between heliocentric orbits of varying semi-major axis while minimizing propellant consumption. An additional experiment is conducted for an interplanetary rendezvous scenario. Mass-optimal reference trajectories are first generated using an indirect optimal control formulation based on Pontryagin’s Maximum Principle and homotopic smoothing to obtain bang-bang thrust profiles. These trajectories are then converted into a Markov decision process dataset by mapping each state and optimal control to a normalized polar state representation and continuous action vector, and encoding them as state–action–reward–transition tuples that populate the experience replay buffer of a Soft Actor-Critic (SAC) agent. Comparative experiments against a baseline SAC agent trained from scratch show improvements in average episode reward and higher-performing controllers when using pretraining in Monte Carlo validation tests. These results demonstrate that seeding off-policy reinforcement learning with mass-optimal trajectory data is an effective strategy for improving training efficiency in reinforcement learning applied to trajectory design problems.en_US
dc.identifierhttps://doi.org/10.13016/0r8d-zvsd
dc.identifier.urihttp://hdl.handle.net/1903/35443
dc.language.isoenen_US
dc.subject.pqcontrolledAerospace engineeringen_US
dc.subject.pquncontrolledAstrodynamicsen_US
dc.subject.pquncontrolledLow-Thrust Trajectory Designen_US
dc.subject.pquncontrolledMachine Learningen_US
dc.subject.pquncontrolledReinforcement Learningen_US
dc.titleCombining Supervised Pretraining and Reinforcement Learning for Scalable Low-Thrust Trajectory Designen_US
dc.typeThesisen_US

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Schmidt_umd_0117N_25858.pdf
Size:
6.73 MB
Format:
Adobe Portable Document Format