Indexed metadata
Maximal Average-Reward Policies for Semi-Markov Decision Processes With Arbitrary State and Action Space
Steven A. Lippman
Source record
Source: Crossref
Published: Oct 1, 1971
DOI: 10.1214/aoms/1177693170
Open original source ↗Evidence graph
No public relationships recorded yet.
Integrity note: This page is a factual metadata record created by deterministic ingestion. It is not a claim that the work moves a mathematical frontier or has been independently verified.