Implementation of the STOMP (SubTask, Option, Model, Planning) algorithm from "Reward-Respecting Subtasks for Model-Based Reinforcement Learning" by Sutton et al. (2023).
By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).