JackM323
f684bf44b8
Evaluation Phase Usage
2026-08-04 22:29:30 +02:00
JackM323
101004ead1
reward ajustments
...
robot starts balancing without changing phase and farms alive bonus
fix ->external force and bigger height penalty
2026-08-04 22:17:33 +02:00
JackM323
acb3d671be
Reworked training
...
new reward/penalty system
learning phases with curriculum learning
new training parameters
cleanup of old code
better logging while training
multiple environments instead of robots (they could bumb into each other)
2026-08-03 22:27:26 +02:00
JackM323
5317ef1299
fixed env eval setup
2026-07-31 18:35:09 +02:00
JackM323
a3e46c3abf
updated evaluation code for previous env changes
...
outdated code from previous changes on env
2026-07-31 16:33:50 +02:00
JackM323
c5ca79a354
Machine Learning Trainer
...
Training environment to make a walk model for the hexapod
generated code that will be checked
2026-07-30 17:23:36 +02:00