World Wires · story 8611 · corroborated · 1 source(s)

Achieving an O(1/N) Optimality Gap in Average-Reward Weakly-Coupled MDPs

The study derives an O(1/N) optimality gap for average-reward weakly-coupled Markov decision processes with identical arms and budget constraints.

Open in the desk

Coverage

What this site indexes