1
Fork 0

feat(docs): added delta_distance extension to documentation for reward function

This commit is contained in:
Robin Meersman 2026-04-17 14:28:11 +02:00
parent ad17401c4c
commit c880285dc2

View file

@ -8,6 +8,8 @@ inputs must be distributed fairly to guarantee an objective comparison between d
- The reward function is centered around minimizing the distance to the goal or maximizing the movement towards the goal
within a finite number of timesteps $T$.
- To motivate efficient movement, the amount of timesteps taken to reach the goal will be used as penalty.
- An extra penalty based on movement relative to the current step and
the previous is used to penalize a movement away from the target.
## From reward to PPO