Robin Meersman
|
c880285dc2
|
feat(docs): added delta_distance extension to documentation for reward function
|
2026-04-17 14:28:11 +02:00 |
|
Robin Meersman
|
ad17401c4c
|
fix: custom reward function dependent on env reward + extensions
|
2026-04-17 14:23:57 +02:00 |
|
|
|
3efee5d746
|
Merge pull request #37 from SELab-3-2026/feat/modular-configurations
|
2026-04-16 10:41:43 +02:00 |
|
|
|
a33b8a2e8e
|
refactor: top level imports
Co-authored-by: Robin Meersman <echteenrobin@gmail.com>
|
2026-04-16 10:27:48 +02:00 |
|
|
|
a8298702c8
|
Merge branch 'dev' into feat/modular-configurations
|
2026-04-16 10:16:28 +02:00 |
|
Cedric Mekeirle
|
467483d0ae
|
Merge pull request #34 from SELab-3-2026/debug_experiments
Debug setup and experiments
|
2026-04-16 00:07:24 +02:00 |
|
|
|
a83a4bedbe
|
fix: cleanup
|
2026-04-15 20:12:46 +02:00 |
|
|
|
a724716117
|
feat: simulation config
|
2026-04-15 19:39:39 +02:00 |
|
|
|
143dbbc710
|
fix: explicit arms length check
|
2026-04-15 19:26:33 +02:00 |
|
|
|
6ed4ad8060
|
style: ruff format & check
|
2026-04-15 18:48:05 +02:00 |
|
|
|
07ae66b8ff
|
fix: rediect logs
|
2026-04-15 18:44:37 +02:00 |
|
|
|
e40a4b979f
|
fix: typos and nonexistent calls
|
2026-04-15 18:34:28 +02:00 |
|
|
|
5a358fadde
|
fix: skip test in ci that requires buffer
|
2026-04-15 18:07:12 +02:00 |
|
|
|
eff0c7c1df
|
test: additional testing and visual confirmations
|
2026-04-15 16:47:17 +02:00 |
|
|
|
70bd78833d
|
chore: remove old .gitkeeps
|
2026-04-15 15:34:18 +02:00 |
|
|
|
b4f1e98f8c
|
chore: cleanup and HPC integration
|
2026-04-15 15:29:32 +02:00 |
|
|
|
93bff11208
|
refactor: use hydra for configs in simulation and training scripts
|
2026-04-15 15:18:34 +02:00 |
|
cedric
|
ae6beb174b
|
fix: updated used vars
|
2026-04-15 05:39:18 +00:00 |
|
|
|
ff83af8cef
|
migrate main training entry point to Hydra and wire structured configs
|
2026-04-14 23:03:43 +02:00 |
|
|
|
eabc64009a
|
feat(log): implement LoggerProxy to defer initialization and prevent premature directory creation
|
2026-04-14 22:53:17 +02:00 |
|
|
|
71aedc4853
|
docs: add main configuration entry point and usage guidelines
|
2026-04-14 22:47:09 +02:00 |
|
|
|
2a7537f51c
|
feat: add modular configuration files
|
2026-04-14 22:43:28 +02:00 |
|
|
|
3b8c347715
|
feat(env): add observation padding wrapper for amputated morphologies
|
2026-04-14 22:36:56 +02:00 |
|
|
|
ba32fc9076
|
feat(morphology): support per-arm segment counts for partial amputations
|
2026-04-14 22:22:17 +02:00 |
|
|
|
82dfcabede
|
feat(config): introduce polymorphic architecture configs
centralized/decentralized
|
2026-04-14 22:12:02 +02:00 |
|
|
|
157e061f86
|
chore: draft new configs
|
2026-04-14 21:53:52 +02:00 |
|
|
|
67f620d599
|
chore: setup hydra dependencies
|
2026-04-14 21:53:26 +02:00 |
|
JibrilExe
|
20ba63a303
|
fix: ruff format
|
2026-04-14 21:14:26 +02:00 |
|
JibrilExe
|
0a9bf2e0a5
|
fix: dont store clipped action, and upgraded reward scale
|
2026-04-14 20:30:21 +02:00 |
|
JibrilExe
|
7a87facc8a
|
feat: value loss clipping
|
2026-04-14 19:23:20 +02:00 |
|
JibrilExe
|
1f8fdbdc59
|
feat: observation normalization
|
2026-04-14 19:07:15 +02:00 |
|
JibrilExe
|
631d01ebf6
|
feat: log distance to target for each env
|
2026-04-12 13:09:53 +02:00 |
|
JibrilExe
|
c5a8dd2b3a
|
feat: expanded used var doc, and added distance to target log
|
2026-04-12 12:54:48 +02:00 |
|
JibrilExe
|
1c328bdde1
|
fix: observation filtering, action clipping, log_std clip, reward discount, bigger MLP
|
2026-04-12 11:12:46 +02:00 |
|
JibrilExe
|
c0e7569773
|
feat: More low lvl logs, reward, advantage, returns
|
2026-04-11 13:42:52 +02:00 |
|
JibrilExe
|
aad086cb7d
|
fix: Merge with origin/dev
|
2026-04-10 15:20:22 +02:00 |
|
JibrilExe
|
1ef257544a
|
feat: first approach to action clipping
|
2026-04-10 15:18:33 +02:00 |
|
Cedric Mekeirle
|
a9ac7e7c0d
|
Merge pull request #33 from SELab-3-2026/debug_charts
Charts to make debugging easier
|
2026-04-10 13:02:30 +02:00 |
|
cedric
|
ac352ef431
|
feat: start of debug setup, experiment description,..
|
2026-04-10 10:17:18 +00:00 |
|
JibrilExe
|
cee6b53ada
|
fix(hpc): removed local package install command as it was fixed in tibo's older pr
|
2026-04-10 09:07:43 +02:00 |
|
JibrilExe
|
b7e2beb65d
|
fx(cleanup): removed meaningless comments, added typing and renamed lossinfo to trainingmeasurements
|
2026-04-10 08:31:52 +02:00 |
|
JibrilExe
|
6e942d896b
|
fix(gae): fixes #31
|
2026-04-10 08:23:34 +02:00 |
|
cedric
|
bf6434ea6c
|
fix(cleanup): ruff format
|
2026-04-09 22:13:07 +00:00 |
|
vsc46589 vscuser
|
71349af5b9
|
feat(logging): added avg episode at which terminated or truncated happened
|
2026-04-10 00:11:18 +02:00 |
|
vsc46589 vscuser
|
f7428f9d90
|
feat(logging): added amount of envs that truncated and terminated
|
2026-04-09 23:40:58 +02:00 |
|
vsc46589 vscuser
|
b839024e6e
|
feat(logging): added explained variance chart
|
2026-04-09 23:14:52 +02:00 |
|
Cedric Mekeirle
|
1c64dc20ae
|
Merge pull request #19 from SELab-3-2026/docs/reward_and_mlp-design
Extended docs on mlp and reward design plus cleanup, cheers.
|
2026-04-09 22:09:58 +02:00 |
|
|
|
4a96acf744
|
Merge pull request #32 from SELab-3-2026/ci/fix
|
2026-04-09 16:21:13 +02:00 |
|
|
|
f6e050dc51
|
ci: fix action
|
2026-04-09 15:53:09 +02:00 |
|
|
|
4a15341314
|
Merge branch 'dev' into docs/reward_and_mlp-design
|
2026-04-09 15:32:10 +02:00 |
|