|
|
86a53ee9f7
|
fix: robot direction to target
|
2026-04-30 21:28:29 +02:00 |
|
|
|
49f5874035
|
chore: remove distance to target from model input
|
2026-04-28 17:27:08 +02:00 |
|
|
|
0032496073
|
docs: describe normalization
|
2026-04-28 10:44:55 +02:00 |
|
|
|
b29818a144
|
docs: Clarify input space in more detail
|
2026-04-25 17:05:46 +02:00 |
|
Robin Meersman
|
c880285dc2
|
feat(docs): added delta_distance extension to documentation for reward function
|
2026-04-17 14:28:11 +02:00 |
|
cedric
|
268c0461f9
|
feat(docs): expanded reward function docs
|
2026-04-08 17:56:26 +00:00 |
|
|
|
67027e8905
|
fix: update file link
|
2026-04-08 15:43:12 +02:00 |
|
|
|
16cc86b30c
|
docs: add actor/critic pipeline figures
|
2026-04-08 14:44:09 +02:00 |
|
|
|
d8c2917923
|
docs: detailed actor-critic pipelines
|
2026-04-08 14:21:44 +02:00 |
|
|
|
9bec02594e
|
docs: model input/output
|
2026-04-08 12:29:15 +02:00 |
|
|
|
934cb17bf6
|
chore: use design dir
|
2026-04-08 11:03:56 +02:00 |
|
cedric
|
f487d5725f
|
feat: extended reward function doc with efficiency idea
|
2026-03-31 22:20:19 +00:00 |
|
|
|
94f505920c
|
style: Fix Markdown indentation
|
2026-03-15 23:03:09 +01:00 |
|
|
|
bc9419ca53
|
Apply suggestions from code review
Co-authored-by: Copilot Autofix powered by AI <175728472+Copilot@users.noreply.github.com>
|
2026-03-15 23:00:30 +01:00 |
|
|
|
a0d70947c1
|
docs: Inputs and reward function
|
2026-03-15 22:35:29 +01:00 |
|
|
|
59f78a6708
|
docs: Learning algorithm
|
2026-03-15 22:26:40 +01:00 |
|
|
|
29759ef4f4
|
docs: Modularity
|
2026-03-15 22:09:41 +01:00 |
|
|
|
ac0dd3891c
|
docs: Communication scheme
|
2026-03-15 21:35:02 +01:00 |
|