Reference health

Learning the Minimal Representation of a Dynamic System from Transition Data

https://doi.org/10.2139/ssrn.3785547
CiteStamped reference-health badge
25/25 checkable references clean · checked 2026-08-27

Every reference with a DOI in the deposited reference list resolved to a known work in Crossref or DataCite at the dated check, and none carried a retraction, withdrawal, or removal notice.

34 without a DOI — not checked. A reference deposited without a DOI is never matched by title or guessed at; it stays outside the checked set, and this line discloses that.

The 25 checked references that resolve
resolves10.1111/1467-9965.00068
Coherent Measures of Risk
resolves10.1287/opre.35.2.215
Aggregation in Dynamic Programming
resolves10.1109/jas.2018.7511249
Feature-based aggregation and deep reinforcement learning: a survey and some new implementations
resolves10.1287/opre.1080.0646
Constructing Uncertainty Sets for Robust Linear Optimization
resolves10.1145/76359.76371
Learnability and the Vapnik-Chervonenkis dimension
resolves10.1214/aoms/1177729330
A Measure of Asymptotic Efficiency for Tests of a Hypothesis Based on the sum of Observations
resolves10.1063/pt.5.028530
Preprint repository arXiv achieves milestone million uploads
resolves10.1109/cdc.2006.377527
Clinical data based optimal STI strategies for HIV: a reinforcement learning approach
resolves10.1016/s0004-3702(02)00376-4
Equivalence notions and model minimization in Markov decision processes
resolves10.1145/2455.2460
Algebraic laws for nondeterminism and concurrency
resolves10.1111/1468-0262.00442
Efficient Estimation of Average Treatment Effects Using the Estimated Propensity Score
resolves10.1145/1329125.1329242
Model-based function approximation in reinforcement learning
resolves10.1007/s10514-015-9459-7
Learning state representations with robotic priors
resolves10.1016/j.neunet.2018.07.006
State representation learning for control: An overview
resolves10.1145/1935826.1935878
Unbiased offline evaluation of contextual-bandit-based news article recommendation algorithms
resolves10.1007/978-3-642-34500-5_16
Learn to Swing Up and Balance a Real Pole Based on Raw Visual Input Data
resolves10.1287/opre.1080.0683
Constructing Risk Measures from Uncertainty Sets
resolves10.1109/embc.2016.7591355
Optimal medication dosing from suboptimal clinical examples: A deep reinforcement learning approach
resolves10.1007/11564096_32
Neural Fitted Q Iteration – First Experiences with a Data Efficient Neural Reinforcement Learning Method
resolves10.1007/0-387-25465-X_15
Clustering Methods
resolves10.1038/nature24270
Mastering the game of Go without human knowledge
resolves10.1287/mnsc.1050.0504
Dynamic Catalog Mailing Policies
resolves10.1007/BF00114724
Feature-based methods for large scale dynamic programming
resolves10.1287/moor.1060.0188
Performance Loss Bounds for Approximate Value Iteration with State Aggregation
resolves10.1134/S000511791911002X
Complete Statistical Theory of Learning
The 34 references without a DOI — listed, not checked
no DOI — not checkedref1
no DOI — not checkedref4
no DOI — not checkedref6
no DOI — not checkedInput generalization in delayed reinforcement learning: An algorithm and performance comparisons
no DOI — not checkedref11
no DOI — not checkedref13
no DOI — not checkedref14
no DOI — not checkedProvably efficient rl with rich observations via latent state decoding
no DOI — not checkedref16
no DOI — not checkedTree-based batch mode reinforcement learning
no DOI — not checkedMetrics for finite markov decision processes
no DOI — not checkedOff-policy deep reinforcement learning without exploration
no DOI — not checkedref22
no DOI — not checkedref24
no DOI — not checkedThe optimal sample complexity of pac learning
no DOI — not checkedDoubly robust off-policy value evaluation for reinforcement learning
no DOI — not checkedProvably efficient reinforcement learning with linear function approximation
no DOI — not checkedEfficiently breaking the curse of horizon: Double reinforcement learning in infinite-horizon processes
no DOI — not checkedDouble reinforcement learning for efficient off-policy evaluation in markov decision processes
no DOI — not checkedref35
no DOI — not checkedStabilizing off-policy q-learning via bootstrapping error reduction
no DOI — not checkedTowards a unified theory of state abstraction for mdps
no DOI — not checkedBreaking the curse of horizon: Infinite-horizon off-policy estimation
no DOI — not checkedRepresentation balancing mdps for off-policy policy evaluation
no DOI — not checkedOffline policy evaluation across representations with applications to educational games
no DOI — not checkedKinematic state abstraction and provably efficient rich-observation reinforcement learning
no DOI — not checkedref46
no DOI — not checkedref49
no DOI — not checkedSmdp homomorphisms: An algebraic approach to abstraction in semi markov decision processes
no DOI — not checkedref53
no DOI — not checkedref56
no DOI — not checkedref58
no DOI — not checkedref60
no DOI — not checkedSolar: Deep structured representations for model-based reinforcement learning
What this badge says. CiteStamped means the CHECKABLE references of this work were clean at the dated check: each resolved to a known work in a public registry, and none carried a retraction notice at that time. It says nothing about the quality, findings, or importance of the work itself, and nothing about references deposited without a DOI.

checked 2026-08-27 — re-checked daily as this page is visited; titles and statuses come from Crossref and DataCite and are not part of the signed record

Embed this badge

Both snippets point at the live badge image and link back to this page. The badge re-renders from the daily check, so an embed never goes stale by more than a day of visits.

<a href="https://citestamp.com/citestamped/10.2139/ssrn.3785547"><img src="https://citestamp.com/citestamped/10.2139/ssrn.3785547/badge.svg" alt="CiteStamped reference-health badge" width="460" height="64"></a>
[![CiteStamped reference-health badge](https://citestamp.com/citestamped/10.2139/ssrn.3785547/badge.svg)](https://citestamp.com/citestamped/10.2139/ssrn.3785547)