1
Fork 0
Commit graph

15 commits

Author SHA1 Message Date
ed20325fef feat(hpc): rename config_path to env_config_path and support YAML configs 2026-04-04 19:07:01 +02:00
0aeff0f2da feat(hpc): include git hash in run_name for traceability 2026-04-04 19:02:06 +02:00
92db8c2591 feat(hpc): support run_dir staging and update docs
- Add run_dir and checkpoint_frequency to PPOArgs
- Update train.py to use run_dir for SummaryBoard, model saving, and loss plots
- Create configs/production_training.yaml for HPC production runs
- Update HPC.md with run_dir staging strategy details
2026-04-04 18:59:05 +02:00
3c4f0b2ba3 fix(train): disable tqdm in batch mode to prevent log spam 2026-04-01 19:34:47 +02:00
RobinMeersman
f594cafead
Morphology/brittle star 2 arms (#17)
Provide the features to read in Environment (Morphology, arena, ...) config files in the JSON format. The example JSON file contains the config for a brittle star with 2 arms
2026-04-01 19:17:42 +02:00
ecdbe74df4
Merge pull request #9 from SELab-3-2026/chore/ppo 2026-04-01 18:11:59 +02:00
cedric
d119bf6408 fix: critic is no longer shared with actor network 2026-03-31 16:03:03 +00:00
cedric
98a8f529e1 fix: avoids local ruff mismatch 2026-03-27 08:51:31 +00:00
cedric
a452912b36 removed unnecessary comments 2026-03-27 04:40:13 +00:00
cedric
31c2d188b7 feature: optional message passer function and clearer naming 2026-03-27 04:09:56 +00:00
cedric
ccbbfc18d7 fix: PPO extracted and integrated with jax 2026-03-27 03:27:26 +00:00
RobinMeersman
a85a7b8d89
Environment setup + train loop (#5)
Mujoco environment setup (vectorized on GPU) + training loop + simulate script
2026-03-26 09:53:54 +01:00
JibrilExe
25a6299eb5 extracted ppo relevant code from cleanrl_atari, and changed to continuous output 2026-03-25 19:34:12 +00:00
RobinMeersman
1f64f8a176 feat: ignore .idea directory completely 2026-03-15 13:34:39 +01:00
e256e5e1ad Initial commit 2026-02-26 14:18:39 +01:00