Generalized Linear Model (GLM) Bandits on Nonstationary Piecewise-stationary Environments
2Cumulative RegretD-GLUCB
Evaluation Results
| Method | Links | |||
|---|---|---|---|---|
| D-GLUCBAlgorithm Type=Discount + MLE, Assumption=Bounded Reward2026.05 | 2 | — | — | |
| SC-D-GLUCBAlgorithm Type=Discount + MLE, Assumption=Bounded Reward + SC2026.05 | 2 | — | — | |
| SCB-PW-WeightUCBAlgorithm Type=Discount + MLE, Assumption=Bounded Reward + SC2026.05 | 2 | — | — | |
| DOMD-GLBAlgorithm Type=Discount + OMD, Assumption=Bounded Reward, Reference=Theorem 32026.05 | 2 | 1 | 1 |