Adaptive Optimization for Stochastic Renewal Systems

M. J. Neely

Citation Details

This paper considers online optimization for a system that performs a sequence of back-to-back tasks. Each task can be processed in one of multiple processing modes that affect the duration of the task, the reward earned, and an additional vector of penalties (such as energy or cost). Let A[k] be a random matrix of parameters that specifies the duration, reward, and penalty vector under each processing option for task k. The goal is to observe A[k] at the start of each new task k and then choose a processing mode for the task so that, over time, time average reward is maximized subject to time average penalty constraints. This is a renewal optimization problem and is challenging because the probability distribution for the A[k] sequence is unknown. Prior work shows that any algorithm that comes within ϵ of optimality must have (1/ϵ^2) convergence time. The only known algorithm that can meet this bound operates without time average penalty constraints and uses a diminishing stepsize that cannot adapt when probabilities change. This paper develops a new algorithm that is adaptive and comes within O(ϵ) of optimality for any interval of (1/ϵ^2) tasks over which probabilities are held fixed, regardless of probabilities before the start of the interval. more »

Award ID(s):: 1824418

PAR ID:: 10494719

Author(s) / Creator(s):: M. J. Neely

Corporate Creator(s):: arXiv:2401.07170v1

Editor(s):: arXiv:2401.07170v1

Publisher / Repository:: arXiv:2401.07170v1

Date Published:: 2024-01-13

Subject(s) / Keyword(s):: task scheduling opportunistic scheduling energy throughput time-varying channels

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Workshop Report:
The DOI is not currently available.

More Like this