Pre-Estimation Checks

Simulation Checks

The first 20-state inference panel in the Simulation Study reports the following checks.

Check

Value

Status

Reward features

3

pass

Feature rank

3 / 3

pass

Action-contrast rank

3 / 3

pass

Feature condition number

1.000

pass

Action-contrast condition number

1.000

pass

Observed states

20 / 20

pass

State-action coverage

1.000

pass

Single-action states

0

pass

States with fewer than five observations

0

pass

Action counts

5,605 and 4,395

pass

The panel has full support, so the reported fit does not rely on uniform CCP fallbacks for unvisited states.

Common Risk Patterns

Data with many unvisited states force CCP to extrapolate the first-stage policy. States with only one observed action make counterfactual action values weakly supported. Too little smoothing can make log corrections unstable, while too much smoothing biases the empirical policy toward uniform choice. A transition tensor with the wrong orientation can have valid dimensions while representing the wrong transition law.

Before interpreting the estimates structurally, specify the discount factor, shock distribution and scale, reward location normalization, and utility form. The wrapper can estimate transitions from the panel, but its reward standard errors do not propagate uncertainty from that transition stage.