This content will become publicly available on December 1, 2026

Title: Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL
Award ID(s):
2233152
PAR ID:
10656422
Author(s) / Creator(s):
; ;
Publisher / Repository:
Neural Information Processing Systems (NeurIPS)
Date Published:
Format(s):
Medium: X
Sponsoring Org:
National Science Foundation
More Like this
No document suggestions found