1
Fork 0
Commit graph

18 commits

Author SHA1 Message Date
86a53ee9f7
fix: robot direction to target 2026-04-30 21:28:29 +02:00
49f5874035
chore: remove distance to target from model input 2026-04-28 17:27:08 +02:00
0032496073
docs: describe normalization 2026-04-28 10:44:55 +02:00
b29818a144
docs: Clarify input space in more detail 2026-04-25 17:05:46 +02:00
Robin Meersman
c880285dc2 feat(docs): added delta_distance extension to documentation for reward function 2026-04-17 14:28:11 +02:00
cedric
268c0461f9 feat(docs): expanded reward function docs 2026-04-08 17:56:26 +00:00
67027e8905
fix: update file link 2026-04-08 15:43:12 +02:00
16cc86b30c
docs: add actor/critic pipeline figures 2026-04-08 14:44:09 +02:00
d8c2917923
docs: detailed actor-critic pipelines 2026-04-08 14:21:44 +02:00
9bec02594e
docs: model input/output 2026-04-08 12:29:15 +02:00
934cb17bf6
chore: use design dir 2026-04-08 11:03:56 +02:00
cedric
f487d5725f feat: extended reward function doc with efficiency idea 2026-03-31 22:20:19 +00:00
94f505920c
style: Fix Markdown indentation 2026-03-15 23:03:09 +01:00
bc9419ca53
Apply suggestions from code review
Co-authored-by: Copilot Autofix powered by AI <175728472+Copilot@users.noreply.github.com>
2026-03-15 23:00:30 +01:00
a0d70947c1
docs: Inputs and reward function 2026-03-15 22:35:29 +01:00
59f78a6708
docs: Learning algorithm 2026-03-15 22:26:40 +01:00
29759ef4f4
docs: Modularity 2026-03-15 22:09:41 +01:00
ac0dd3891c
docs: Communication scheme 2026-03-15 21:35:02 +01:00