Actuarium

Decisions under uncertainty — one method, every domain.

Score(a)=(1ω)[(1ε)Dg,u,p(a)+εminsu(U[a][s])]ωRegret(a)\text{Score}(a) = (1-\omega)\Big[(1-\varepsilon)\,D_{g,u,p^*}(a) + \varepsilon\min_s u(U[a][s])\Big] - \omega\cdot\text{Regret}(a)

Every classical decision rule — EMV, expected utility, maximin, Hurwicz, minimax regret, Bayes, credibility, VaR/TVaR, Wang premiums, Gilboa–Schmeidler — is a special case of one auditable, parameterised engine.

The ADE in one page

The Actuarium Decision Engine (ADE) scores every action aa by combining a distorted, credibility-weighted expected utility with an ambiguity term and a regret penalty:

p(s)=Zp(se)+(1Z)q(s)p^*(s) = Z\cdot p(s\mid e) + (1-Z)\cdot q(s)u(x)=ρ(1ex/ρ)u(x) = \rho\left(1-e^{-x/\rho}\right)Score(a)=(1ω)[(1ε)Dg,u,p(a)+εminsu(U[a][s])]ωRegret(a)\text{Score}(a) = (1-\omega)\Big[(1-\varepsilon)\,D_{g,u,p^*}(a) + \varepsilon\min_s u(U[a][s])\Big] - \omega\cdot\text{Regret}(a)

Parameter meanings

ParameterMeaning
ρ\rhoRisk tolerance for exponential (CARA) utility; \rho\to\infty is risk-neutral.
λ,κ\lambda,\kappaWang / proportional-hazards distortion parameters (tail-loading).
ε\varepsilonAmbiguity aversion weight — ε-contamination toward pure maximin.
ω\omegaRegret-aversion weight — blend toward minimax regret.
ZZCredibility weight blending the Bayesian posterior with a reference distribution q(s).
p(s),q(s)p(s), q(s)Prior belief and reference/benchmark distribution over states.
L(es)L(e\mid s)Likelihood of observed evidence under each state, used for Bayesian updating.

Classical method → ADE parameters

Classical criterionADE settingNote
Expected Monetary Value (EMV)ρ, λ=0, ε=0, ω=0, Z=1\rho\to\infty,\ \lambda=0,\ \varepsilon=0,\ \omega=0,\ Z=1Plain probability-weighted average payoff.
Expected Utilityfinite ρ, λ=0, ε=0, ω=0\text{finite } \rho,\ \lambda=0,\ \varepsilon=0,\ \omega=0vNM expected utility with CARA risk aversion.
Maximin (Wald)ε=1\varepsilon=1Only the worst state is judged.
Maximaxρ, point mass on argmaxsU[a][s]\rho\to\infty,\ \text{point mass on }\arg\max_s U[a][s]Optimistic best-case evaluation.
Hurwicz(α)αmax+(1α)min\alpha\max+(1-\alpha)\minOptimism–pessimism blend, computed directly.
Minimax regret (Savage)ω=1\omega=1Minimise the worst-case regret across states.
Laplacep(s)=1/n, ρ, λ=0, ε=0, ω=0p(s)=1/n,\ \rho\to\infty,\ \lambda=0,\ \varepsilon=0,\ \omega=0Principle of insufficient reason.
Bayes (posterior EMV)ρ, Z=1, posterior p(se)\rho\to\infty,\ Z=1,\ \text{posterior }p(s\mid e)EMV under the Bayesian posterior.
Bühlmann credibilityZ(0,1)Z\in(0,1)Linear-Bayes blend of posterior and reference.
VaR / TVaRWang λ, ρ, ε=0, ω=0\text{Wang }\lambda,\ \rho\to\infty,\ \varepsilon=0,\ \omega=0Tail-weighted valuation via distortion.
Wang premium principleg=Φ(Φ1(t)+λ)g=\Phi(\Phi^{-1}(t)+\lambda)Actuarial premium-loading distortion.
Gilboa–Schmeidler (maxmin EU)ε=1 (ε-contamination)\varepsilon=1\text{ (}\varepsilon\text{-contamination)}Worst-case over a set of priors.
Savageω=1\omega=1Savage's own minimax-regret criterion.

Full case study: Workers' Compensation pricing & reserving

Granite State Mutual — loss triangles, chain-ladder & Bornhuetter–Ferguson reserving, a rate indication build-up, a class plan review, and both decisions run through the ADE.

Decision Studio

Pick a realistic scenario, edit the states, actions, payoffs and beliefs, and watch every classical criterion and the ADE recommendation update live.

A clinician compares two treatment protocols across responder/non-responder states, payoffs in expected quality-adjusted life years (QALYs) over 10 years.

States of the world

Actions

Payoff matrix U[action][state]

Action \ StateResponds wellPartial responseNo response / side effects
Treatment A (aggressive)
Treatment B (conservative)

Prior beliefs p(s)

Responds well40.0%
Partial response35.0%
No response / side effects25.0%

Normalised automatically to sum to 100%.

Evidence & Bayesian updating

OffOn

Attitude parameters

Risk-neutral

Reference distribution q(s) used for credibility blending: Responds well: 40%, Partial response: 35%, No response / side effects: 25%.

Recommendation

Treatment B (conservative)
EVPI = 0.48
ActionADE scoreEMVMaximinMaximaxRegretCERisk premium
Treatment B (conservative)4.047.033.64.420.116.940.08
Treatment A (aggressive)3.646.712.364.710.376.190.52
Classical criterionTreatment A (aggressive)Treatment B (conservative)Recommends
Expected Monetary Value (EMV)6.717.03Treatment B (conservative)
Expected Utility3.864.11Treatment B (conservative)
Maximin (Wald)35.5Treatment B (conservative)
Maximax9.28Treatment A (aggressive)
Hurwicz(α=0.5)6.16.75Treatment B (conservative)
Laplace (principle of insufficient reason)6.236.83Treatment B (conservative)
Bayes (posterior EMV)6.717.03Treatment B (conservative)
Bühlmann credibility blend6.717.03Treatment B (conservative)
Minimax regret (Savage)-2.5-1.2Treatment B (conservative)
TVaR / Wang-distorted value5.076.36Treatment B (conservative)

Sensitivity — where the recommendation flips

  • rho: no flip across the sampled range — recommendation is robust to this parameter
  • lambda: no flip across the sampled range — recommendation is robust to this parameter
  • epsilon: no flip across the sampled range — recommendation is robust to this parameter
  • omega: no flip across the sampled range — recommendation is robust to this parameter
  • Z: no flip across the sampled range — recommendation is robust to this parameter

Why one method — and its limits

The Savage and von Neumann–Morgenstern axioms are what justify representing preferences by an expected (or distorted, ambiguity-robust) utility in the first place — they are normative, not descriptive: real people routinely violate them. The Ellsberg paradox shows people are ambiguity-averse in a way plain expected utility cannot represent; the Allais paradox shows the vNM independence axiom is regularly violated by the "certainty effect." ADE's $\varepsilon$ and distortion parameters are a controlled, disclosed departure from strict expected utility — not a claim that the axioms are false.

Every ADE run still depends on human judgement: elicitation of the prior $p(s)$, the payoff matrix, and the risk/ambiguity/regret parameters, and model risk — the engine is only as good as the states, actions and numbers fed into it. The sensitivity panel exists precisely because a technically correct method fed a fragile or overconfident input can still recommend the wrong action.

Go deeper

Ask the tutor