Class-conditional normalizing flows provide tractable posterior probabilities for studying learning behavior. This framework examines scaling laws, soft-label training, distribution shifts, and active learning while separating epistemic error from the uncertainty that remains even for an optimal predictor.