Indexed metadata

Maximal Average-Reward Policies for Semi-Markov Decision Processes With Arbitrary State and Action Space

Steven A. Lippman

Source record

Source: Crossref

Published: Oct 1, 1971

DOI: 10.1214/aoms/1177693170

Open original source ↗

Evidence graph

No public relationships recorded yet.

Integrity note: This page is a factual metadata record created by deterministic ingestion. It is not a claim that the work moves a mathematical frontier or has been independently verified.