RL policy training interval optimization on Indoor Lighting Trace Window (simulated 1000 nodes)
89Number of Policy TrainingsBaseline
Evaluation Results
| Method | Links | |||
|---|---|---|---|---|
| Baseline2018.11 | 89 | 1 | — | |
| Dynamic On-Policy2018.11 | 27 | 0 | 69 |
| Method | Links | |||
|---|---|---|---|---|
| Baseline2018.11 | 89 | 1 | — | |
| Dynamic On-Policy2018.11 | 27 | 0 | 69 |