Robin Meersman
|
791dcf231c
|
other(ppo training loop): code cleanup + enabled debug logging of model parameters
|
2026-05-04 21:29:27 +02:00 |
|
Robin Meersman
|
dd2612e2ec
|
fix(ppo training): more shape-related bug fixes + moved computation of #arms and segments outside of jit compiled function
|
2026-05-04 21:26:34 +02:00 |
|
Robin Meersman
|
a43f3c0172
|
fix(ppo training): fixed shape error while clipping action
|
2026-05-04 19:36:59 +02:00 |
|
Robin Meersman
|
bb14507fc7
|
fix(message passing): forgot vmapping over params :thumbs_up:
|
2026-05-02 23:34:52 +02:00 |
|
Robin Meersman
|
3cb3c765cc
|
feat(message passing): fixed tracing bug, todo: dense layer input dimension incorrect
|
2026-05-02 23:27:31 +02:00 |
|
Robin Meersman
|
027b13b30f
|
feat(message passing): used config value for message passing repetitions
|
2026-05-02 23:15:55 +02:00 |
|
Robin Meersman
|
6114e36e27
|
feat(PPOTrainer): decentralized message passing babyyyy
|
2026-05-02 21:50:42 +02:00 |
|
JibrilExe
|
9398592e34
|
merged
|
2026-05-02 11:12:01 +02:00 |
|
JibrilExe
|
2b46b3f5d2
|
fix: only flatten actions for environ
|
2026-05-02 11:08:48 +02:00 |
|
JibrilExe
|
f77766ba08
|
fix:logging shapes
|
2026-05-02 09:38:22 +02:00 |
|
Robin Meersman
|
9f7c73b009
|
other: backup of current progress
|
2026-05-01 18:15:09 +02:00 |
|
JibrilExe
|
bceb7fe4cf
|
logging: now we can see shapes in ppo update
|
2026-05-01 11:52:57 +02:00 |
|
JibrilExe
|
951f682723
|
fix: latest attempt at fexing exploding mem leak
|
2026-05-01 11:36:25 +02:00 |
|
JibrilExe
|
e867dbab49
|
fix: dubble reshape bad
|
2026-05-01 10:43:27 +02:00 |
|
JibrilExe
|
22c43f05b1
|
fix: critic init now fixed?
|
2026-05-01 10:40:02 +02:00 |
|
Robin Meersman
|
1530e7d210
|
good luck cedric 🫡
|
2026-04-30 19:32:20 +02:00 |
|
Robin Meersman
|
f8eb43004f
|
other: backup, finding bug in conversion from dict to array
|
2026-04-30 15:58:46 +02:00 |
|
Robin Meersman
|
e85fc09ed6
|
feat(PPOTrainer): vectorized init agent state, partially added message passer param to rollout, ...
|
2026-04-29 11:23:01 +02:00 |
|
Cedric
|
bc3408b59a
|
feat: init agents array
|
2026-04-29 08:14:13 +00:00 |
|
Robin Meersman
|
db5d4c3dc2
|
fix(PPOTrainer): fixed duplicate definition of conversion function, tracing error related to obs_mode, added adjacent matrix argument
|
2026-04-28 22:23:46 +02:00 |
|
Cedric
|
a005c0ccad
|
feat: upgraded obs_to_array, now takes care of splitting for agents
|
2026-04-28 08:20:00 +00:00 |
|
Cedric
|
39945e8821
|
feat: helper to visualize adjacency in a file
|
2026-04-28 07:27:08 +00:00 |
|
Cedric
|
2777d4f6da
|
feat: adjacency builder
|
2026-04-28 06:46:26 +00:00 |
|
Cedric
|
8830289992
|
feat: made obs_dict handle morphology
|
2026-04-28 05:44:26 +00:00 |
|
Jona Reynaert
|
509ad491fb
|
Merge pull request #21 from SELab-3-2026/simulate-results
Model Simulation
|
2026-04-23 10:09:42 +02:00 |
|
Jona Reynaert
|
94e50f3d83
|
fix: used jax type to remove warning
|
2026-04-22 18:51:08 +02:00 |
|
Jona Reynaert
|
395b04d9a8
|
feat: adapted simulate to trained config
|
2026-04-22 18:47:35 +02:00 |
|
Jona Reynaert
|
c4447976ab
|
Merge branch 'dev' into simulate-results
|
2026-04-19 16:54:47 +02:00 |
|
RobinMeersman
|
ea59c821a3
|
feat: improved the used reward function to better the training results
|
2026-04-18 15:52:33 +02:00 |
|
Robin Meersman
|
7e7837d2d1
|
Merge branch 'reward-fix' of github.com:SELab-3-2026/SEL3-2026-Groep-4 into reward-fix
|
2026-04-18 11:07:36 +02:00 |
|
Robin Meersman
|
dfcdfdad2b
|
fix: renamed fast2 to better describing name
|
2026-04-18 11:07:26 +02:00 |
|
Robin Meersman
|
34e8ba91c2
|
fix(PR#41): applied comments for PR #41 review
|
2026-04-18 11:07:01 +02:00 |
|
|
|
6c47883bc9
|
fix(hpc): outdated requirements
|
2026-04-17 22:23:23 +02:00 |
|
|
|
2ef2597430
|
fix: remove unused var
|
2026-04-17 22:09:20 +02:00 |
|
RobinMeersman
|
642b16ded3
|
Merge branch 'dev' into reward-fix
|
2026-04-17 15:17:30 +02:00 |
|
Robin Meersman
|
c880285dc2
|
feat(docs): added delta_distance extension to documentation for reward function
|
2026-04-17 14:28:11 +02:00 |
|
Robin Meersman
|
ad17401c4c
|
fix: custom reward function dependent on env reward + extensions
|
2026-04-17 14:23:57 +02:00 |
|
|
|
784d2ff0c4
|
Merge pull request #39 from SELab-3-2026/fix/checkpoints-and-ci
|
2026-04-17 14:07:26 +02:00 |
|
|
|
4581694fa6
|
style: ruff format + check
|
2026-04-16 15:25:48 +02:00 |
|
|
|
c5b08e817e
|
refactor(log): use dataclass for config
|
2026-04-16 15:18:40 +02:00 |
|
Jona Reynaert
|
3fb8ffc9d0
|
fix: adapted simulation config
|
2026-04-16 15:18:08 +02:00 |
|
Jona Reynaert
|
419ee29e6c
|
Merge branch 'dev' into simulate-results
|
2026-04-16 15:16:36 +02:00 |
|
|
|
37a4b59e04
|
feat: configure checkpoints saving
|
2026-04-16 15:10:21 +02:00 |
|
|
|
4d0f729aee
|
chore(log): remove dead code
|
2026-04-16 14:35:23 +02:00 |
|
|
|
06aa1c54c6
|
ci: create pull request
|
2026-04-16 12:37:09 +02:00 |
|
|
|
8b6dcbae7c
|
feat: checkpoints
|
2026-04-16 12:27:56 +02:00 |
|
Jona Reynaert
|
e160e50104
|
Merge branch 'simulate-results' of github.com:SELab-3-2026/Brittle-Star into simulate-results
|
2026-04-16 10:48:50 +02:00 |
|
Jona Reynaert
|
cb6e3d9714
|
fix: universal .flax files
|
2026-04-16 10:48:37 +02:00 |
|
|
|
3efee5d746
|
Merge pull request #37 from SELab-3-2026/feat/modular-configurations
|
2026-04-16 10:41:43 +02:00 |
|
|
|
a33b8a2e8e
|
refactor: top level imports
Co-authored-by: Robin Meersman <echteenrobin@gmail.com>
|
2026-04-16 10:27:48 +02:00 |
|