Skip to content
Log in
Toggle navigation
Datasets
Organizations
Groups
About
Search Datasets
Home
Datasets
Order by
Relevance
Name Ascending
Name Descending
Last Modified
Go
1 dataset found
Tags:
value-estimation
Filter Results
Off-policy Learning with Eligibility Traces: A Survey
In the framework of Markov Decision Processes, off-policy learning, that is the problem of learning a linear approximation of the value function of some fixed policy...
HTML
You can also access this registry using the
API
(see
API Docs
).