1
Fork 0
Commit graph

90 commits

Author SHA1 Message Date
ed20325fef feat(hpc): rename config_path to env_config_path and support YAML configs 2026-04-04 19:07:01 +02:00
0aeff0f2da feat(hpc): include git hash in run_name for traceability 2026-04-04 19:02:06 +02:00
92db8c2591 feat(hpc): support run_dir staging and update docs
- Add run_dir and checkpoint_frequency to PPOArgs
- Update train.py to use run_dir for SummaryBoard, model saving, and loss plots
- Create configs/production_training.yaml for HPC production runs
- Update HPC.md with run_dir staging strategy details
2026-04-04 18:59:05 +02:00
Robin Meersman
62e11ce243 fix(PPOTrainer.py): bug fixes regarding jax.jit 2026-04-03 15:22:51 +02:00
Robin Meersman
1cb0e2d95b fix(merge): fixed merge conflict 2026-04-03 12:58:40 +02:00
cedric
3104678f96 fix: uniform naming 2026-04-03 08:11:50 +00:00
cedric
1c4e40b5fb fix: uniform naming of models accross code 2026-04-02 18:37:37 +00:00
cedric
79221372df fix: uniform generic network naming 2026-04-02 18:18:10 +00:00
cedric
40c7e32344 feat: more generic naming for critic, since messager will use the same 2026-04-02 18:11:49 +00:00
cedric
729f122be4 fix: removed useless comment 2026-04-02 18:00:53 +00:00
cedric
95840b789b fix: usage of field because flax wont allow mutable class object 2026-04-02 17:59:48 +00:00
cedric
9879fea9e2 fix: refactored rl folder into mlps 2026-04-02 17:50:44 +00:00
cedric
15ae86796f feat: generic network 2026-04-02 17:38:46 +00:00
2263c97fe8 style: ruff format 2026-04-02 17:36:41 +00:00
cmekeirl
4e3eee8ac6 Basic centralized model 2026-04-02 17:36:41 +00:00
cmekeirl
1ad940ce02 Initial proposal for structure 2026-04-02 17:36:41 +00:00
Robin Meersman
ba59361f97 chore(main.py): deleted redundant main.py, feat(plot.py): added plot file to group all visualization related code for the experiments 2026-04-02 15:20:47 +02:00
Robin Meersman
97eaa5c06b chore(train.py): moved from package to experiments directory 2026-04-02 15:13:27 +02:00
Robin Meersman
45587c5620 fix: losses --> episodic returns 2026-04-02 15:11:39 +02:00
3c4f0b2ba3 fix(train): disable tqdm in batch mode to prevent log spam 2026-04-01 19:34:47 +02:00
RobinMeersman
f594cafead
Morphology/brittle star 2 arms (#17)
Provide the features to read in Environment (Morphology, arena, ...) config files in the JSON format. The example JSON file contains the config for a brittle star with 2 arms
2026-04-01 19:17:42 +02:00
ecdbe74df4
Merge pull request #9 from SELab-3-2026/chore/ppo 2026-04-01 18:11:59 +02:00
ff90101377 refactor(log): improved logging workflow 2026-04-01 16:01:12 +00:00
7e7c5bf27c refactor: set wandb entity and simplify READMEs 2026-04-01 10:26:39 +02:00
e160a55d95 fix(train): integrate YAML config 2026-04-01 00:27:38 +02:00
d27a617199
feat(experiment-logger): add config_utils module
- Add config_utils.py: load_yaml_config, save_yaml_config,
  dataclass_from_dict, merge_config_with_cli, print_config
- Export new symbols from package __init__.py
2026-04-01 00:23:45 +02:00
8ec693f04c fix(logging): improve JSON serialization and add wandb directory to gitignore
- Fix float32 serialization issue in unified logger metrics flushing
- Add jax.numpy import for proper type handling
- Add wandb/ directory to .gitignore to exclude temporary tracking files
- Tested wandb integration: metrics, artifacts, and local backup working correctly
2026-03-31 20:52:51 +00:00
4151d2307e refactor(main): replace print with logging 2026-03-31 19:51:17 +00:00
b70cd1c27b feat(train): integrate experiment_logger and replace print statements
- Replace direct WandB calls with experiment_logger.UnifiedLogger
- Replace all print() calls with proper logging framework
- Add automatic checkpoint saving every N iterations
- Configure root logger with proper format and level
- Maintain backward compatibility with TensorBoard writer
- Save final model with metadata using unified logger
2026-03-31 19:49:21 +00:00
bb0bb94f60 feat(config): add checkpoint frequency parameter
Add checkpoint_frequency to PPOArgs to enable periodic checkpoint
saving during training. Default: save every 100 iterations.
2026-03-31 19:49:12 +00:00
3ce107a560 feat(experiment-logger): add standalone logging framework
Create reusable experiment logging package with:
- UnifiedLogger for multi-backend logging (WandB, disk, stdout)
- Automatic checkpoint and model saving with metadata
- WandB artifact upload support
- Graceful degradation when WandB unavailable
- Comprehensive API documentation

This is a standalone, project-agnostic library that can be reused
across different ML projects.
2026-03-31 19:49:04 +00:00
cedric
d119bf6408 fix: critic is no longer shared with actor network 2026-03-31 16:03:03 +00:00
cedric
98a8f529e1 fix: avoids local ruff mismatch 2026-03-27 08:51:31 +00:00
cedric
a452912b36 removed unnecessary comments 2026-03-27 04:40:13 +00:00
cedric
31c2d188b7 feature: optional message passer function and clearer naming 2026-03-27 04:09:56 +00:00
cedric
ccbbfc18d7 fix: PPO extracted and integrated with jax 2026-03-27 03:27:26 +00:00
RobinMeersman
a85a7b8d89
Environment setup + train loop (#5)
Mujoco environment setup (vectorized on GPU) + training loop + simulate script
2026-03-26 09:53:54 +01:00
JibrilExe
25a6299eb5 extracted ppo relevant code from cleanrl_atari, and changed to continuous output 2026-03-25 19:34:12 +00:00
RobinMeersman
1f64f8a176 feat: ignore .idea directory completely 2026-03-15 13:34:39 +01:00
e256e5e1ad Initial commit 2026-02-26 14:18:39 +01:00