Reinforcement Learning
stable-baselines3
seals/Swimmer-v1
deep-reinforcement-learning
Eval Results (legacy)
Instructions to use HumanCompatibleAI/sac-seals-Swimmer-v1 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- stable-baselines3
How to use HumanCompatibleAI/sac-seals-Swimmer-v1 with stable-baselines3:
from huggingface_sb3 import load_from_hub checkpoint = load_from_hub( repo_id="HumanCompatibleAI/sac-seals-Swimmer-v1", filename="{MODEL FILENAME}.zip", ) - Notebooks
- Google Colab
- Kaggle
| !!python/object/apply:collections.OrderedDict | |
| - - - batch_size | |
| - 128 | |
| - - buffer_size | |
| - 100000 | |
| - - gamma | |
| - 0.995 | |
| - - learning_rate | |
| - 0.00039981805535514633 | |
| - - learning_starts | |
| - 1000 | |
| - - n_timesteps | |
| - 1000000.0 | |
| - - policy | |
| - MlpPolicy | |
| - - policy_kwargs | |
| - log_std_init: -2.689958330139309 | |
| net_arch: | |
| - 400 | |
| - 300 | |
| use_sde: false | |
| - - tau | |
| - 0.01 | |
| - - train_freq | |
| - 256 | |