Commit fa44c12
committed
Add tests for dtype fix and VecNormalize
test_observation_space_dtype: asserts the continuous observation space
uses float64, directly documenting the bug fixed in this PR (was int64,
causing float observations to be silently truncated to integers).
test_observations_are_float: confirms sub-integer precision is preserved
end-to-end after a physics step — the case that was broken before.
test_ppo_saves_vecnorm: verifies that do_training() persists the
VecNormalize statistics as <model>_vecnorm.pkl alongside the model zip,
so inference can reload the same normalisation.
test_ppo_vecnorm_updates: confirms the running mean is updated during
training (normaliser is actually learning from observations, not stuck
at zero).1 parent 301d50e commit fa44c12
2 files changed
Lines changed: 37 additions & 0 deletions
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
76 | 76 | | |
77 | 77 | | |
78 | 78 | | |
| 79 | + | |
| 80 | + | |
| 81 | + | |
| 82 | + | |
| 83 | + | |
| 84 | + | |
| 85 | + | |
| 86 | + | |
| 87 | + | |
| 88 | + | |
| 89 | + | |
| 90 | + | |
| 91 | + | |
79 | 92 | | |
80 | 93 | | |
81 | 94 | | |
| |||
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
1 | 1 | | |
| 2 | + | |
2 | 3 | | |
| 4 | + | |
3 | 5 | | |
4 | 6 | | |
5 | 7 | | |
| |||
20 | 22 | | |
21 | 23 | | |
22 | 24 | | |
| 25 | + | |
| 26 | + | |
| 27 | + | |
| 28 | + | |
| 29 | + | |
| 30 | + | |
| 31 | + | |
| 32 | + | |
| 33 | + | |
| 34 | + | |
| 35 | + | |
| 36 | + | |
| 37 | + | |
| 38 | + | |
| 39 | + | |
| 40 | + | |
| 41 | + | |
| 42 | + | |
| 43 | + | |
| 44 | + | |
| 45 | + | |
| 46 | + | |
0 commit comments