This is a trained model of a
i used the PPO MlpPolicy model architecture, with the defaults from the notebook agent playing
LunarLander-v2
using the
stable-baselines3 library.
1from stable_baselines3 import ...
2from huggingface_sb3 import load_from_hub
3
4...