Value Training for Strong Society

September 1, 2026
blog image

Values are trained the way strength is trained — by reps under load with a coach watching — and this is the operating manual: a 32-value canon, a six-step formation loop, eight program areas, the KPIs, and the first twelve months.

Written by the ENSI Foresight Division. Built on a library of 150 primary documents downloaded from the world’s leading institutions. Compiled 2026-07-25.

The argument: formation is a capability, not a course

Nobody has ever become strong by attending lectures on strength. The gymnasium does not explain muscles to you; it puts you under a load slightly heavier than you can comfortably carry, watches your form, corrects it, and makes you do it again — and again, on a schedule, for years. The muscle is not informed into existence. It is provoked into existence by stress, feedback and repetition. Every serious tradition of human formation has known that character works the same way. Aristotle said it first and said it plainly: we become just by doing just acts, brave by doing brave acts — virtue is acquired by practice and habituation, not instruction alone (Aristotle, Ross trans. 1925; angle 02). Modern philosophy has caught back up: the strongest current account treats virtue as a skill, acquired the way expertise is acquired — through the novice-to-expert progression of attempt, error, correction and refinement (Stichter 2007; angle 02), a framing that moral psychology now shares in Narvaez’s model of ethical expertise development (Narvaez 2006; angle 03).

And yet almost every state that says it wants citizens of character runs values education as a lecture hall. A subject on the timetable. A poster in the corridor. An assembly about kindness. The library beneath this report says, from fifteen separate angles, that this cannot work — because moral judgment is intuition-first and reasoning-second, so exhortation reaches the part of the mind that does not drive behaviour (Haidt 2001; angle 03); because behaviour is carried by habits formed through repetition in stable contexts, not by conclusions (Wood 2016; angle 07); and because the most rigorous multi-programme evaluation ever run on packaged school character curricula — bolt-on lessons added to an unchanged school — found essentially nothing (US IES 2010; angle 15). The lecture model has been tested at scale. It failed.

Report 1 of this series established the evidence. This report is the operating manual. It answers the “so what, now what” question for the actor that matters: the state and its institutions — ministries, schools, academies, professional bodies. The state is the right actor for the same reason it is the right actor for public health: the returns are enormous, long-horizon, and diffuse. Childhood emotional and character formation predicts adult life satisfaction better than academic achievement does (LSE CEP 2013; angle 10); non-cognitive skills carry causal weight for employment, health and civic life that credentials do not capture (Heckman & Kautz 2014; angle 13). Singapore already runs the world’s most systematized national character curriculum (Singapore MOE 2021; angle 15); Japan has taught dōtoku for generations (NIER 2013; angle 15). The question facing a European state is not whether its institutions form character — they do, every day, by accident, through what they reward and tolerate — but whether they will do it deliberately, transparently and well.

Why now — three reasons. First, the evidence is mature. We no longer have only the aspiration; we have the mechanism. A meta-analysis of 213 school programmes shows social-emotional formation moves both behaviour and an 11-percentile achievement gain (Durlak 2011; angle 09). Productive failure shows why struggling before instruction deepens learning (Kapur 2015; angle 05). Habit science shows how practice becomes permanent disposition (Wood 2016; Gardner 2012; angle 07). Mentoring has randomized-trial evidence (PPV 1995; angle 11). The parts exist; nobody has assembled the machine. Second, AI makes situational practice scalable. The eternal bottleneck of experiential formation was adult attention — one coach can watch only so many reps. Generative agents can now populate believable practice situations (Park et al. 2023; angle 14), LLM role-play systems let a learner rehearse a hard conversation and get feedback (Shaikh et al. 2023; angle 14), and a human-AI tutoring copilot has already scaled expert pedagogical moves across 900 tutors in a randomized trial (Wang et al. 2024; angle 14). Report 3 details this engine; this report assumes it. Third, the purpose crisis. Purpose — a stable intention to accomplish something meaningful to the self and consequential beyond it (Damon 2003; angle 12) — is measurably linked to health, achievement and persistence (Templeton/Bronk 2020; angle 12), and benevolence measurably moves national happiness (WHR 2025; angle 10). Societies with meaning deficits do not lack information. They lack formation.

So this report builds the gymnasium. It names a canon of 32 values in eight families — explicitly swap-ready, because the naming is the institution’s sovereign act. It specifies the six-step formation loop that constitutes a single rep. It ranks eight programme areas that together are the training floor. It sets out measurement that will not be destroyed by its own stakes, governance that keeps formation from sliding into indoctrination, and a first twelve months for a mid-sized EU state — the Czech Republic is our worked example throughout. The thesis in one line: formation is a capability an institution builds and runs, not a course it schedules — and the states that build it first will compound trust, honesty and competence the way early adopters of public schooling compounded literacy.

The playbook in brief

  • Character is a trainable skill, not transferable information — virtue is acquired the way expertise is acquired: attempt, error, correction, repetition (Aristotle trans. 1925; Stichter 2007; angle 02; Narvaez 2006; angle 03).

  • The canon comes first. Thirty-two values in eight families — Truth, Courage, Discipline, Justice, Wisdom, Love, Strength, Purpose — each defined as an observable trained behaviour with a signature training situation. The list is swap-ready; the discipline of naming is not.

  • One rep = the six-step formation loop: Situate → Attempt → Fail productively → Correct → Habituate → Prove. Each step carries its own evidence base, from VR dilemmas (Francis et al. 2016; angle 06) to productive failure (Kapur 2015; angle 05) to implementation intentions (Gollwitzer & Sheeran 2006; angle 07) to situational judgment tests (Europe PMC/PLOS 2019; angle 13).

  • Eight programme areas make the gymnasium, ranked by evidence and leverage: whole-school ethos, the challenge curriculum, the dilemma gym, the failure ladder, the habit protocol, mentorship infrastructure, purpose-matching, and adult/professional formation.

  • Culture beats curriculum. Bolt-on character programmes returned null results in the largest randomized evaluation (US IES 2010; angle 15); embedded whole-school strategies with SAFE design features are what works (Durlak 2011; SRCD 2012; angle 09; Berkowitz, Bier & McCauley 2017; angle 01).

  • Measurement must be multi-method and never high-stakes — self-report is corroded by reference bias and faking (RAND 2014; Heckman & Kautz 2014; angle 13); combine situational judgment tests, observed behaviour, 360-degree views and longitudinal outcomes; measure growth, not rank; the learner owns the data.

  • Governance is the licence to operate: a transparent, contestable canon; judgment trained, not compliance; the Japanese and Singaporean state-curricula lessons heeded (NIER 2013; Singapore MOE 2021; angle 15); Kristjánsson’s answers to the standard objections on the record (Kristjánsson 2013; angle 02).

  • The failure mode is institutional hypocrisy. An institution that demands the impossible teaches its members to lie — the US Army documented this on itself (Wong & Gerras 2015; angle 08). The gymnasium must audit its own demands.

  • Twelve months to first proof for a state the size of the Czech Republic: name the canon, pick 20 pilot schools and 2 professional academies, train mentors to published standards, stand up a dilemma-gym MVP, baseline with serious instruments, publish everything openly.

The canon — 32 values in eight families

Every gymnasium trains named lifts. A formation system trains named values — and the naming is the first act of seriousness, because an unnamed value cannot be practised, coached or measured. What follows is ENSI’s proposed canon: 32 values in eight families. It is deliberately swap-ready — a ministry, an academy or a school should fight about this list, strike values, add its own, and then commit in public. The framework survives any reasonable substitution; what does not survive is vagueness. Each value below is given as an operational definition — what the trained behaviour looks like, since a value that cannot be seen cannot be trained — and the signature training situation in which it is most directly exercised. The situations are drawn from the methods evidenced across this library; the loop in the next section shows how each becomes a rep.

Family 1 — Truth. The load-bearing family: without honest reporting there is no feedback, and without feedback nothing else in this playbook works.

  • Honesty — says what is true when a lie would be cheaper. Signature situation: the after-action review in which the learner names their own error before anyone else can.

  • Integrity — behaviour matches stated values when nobody is watching. Signature situation: unproctored work under a real honor code with real stakes (ERIC 2010; angle 08).

  • Sincerity — means what it says; refuses the performance of virtue. Signature situation: the feedback circle where flattery is challenged and plain speech is rewarded.

  • Accountability — owns outcomes, including failures, without excuse or deflection. Signature situation: carrying real responsibility for a project whose results are publicly reviewed.

Family 2 — Courage. The family that converts conviction into action under fear or uncertainty.

  • Boldness — acts decisively under uncertainty instead of waiting for permission. Signature situation: the time-pressured simulation or expedition decision that cannot be deferred (Sütfeld et al. 2017; angle 06; AIR 2005; angle 04).

  • Moral courage — names a wrong at social cost. Signature situation: the “Should I Say Something?” rehearsal — confronting a lapse by a peer or a superior in simulation before doing it in life (Europe PMC 2023; angle 06).

  • Initiative — starts without being told. Signature situation: the open-brief project where the problem itself must first be found (MDRC 2017; angle 04).

  • Enterprise — builds something new that others actually use. Signature situation: running a real venture or service with real users and real consequences.

Family 3 — Discipline. The family that makes every other value repeatable on a bad day.

  • Self-control — chooses the harder-better over the easier-worse in the moment. Signature situation: the daily WOOP rep — wish, outcome, obstacle, plan (Duckworth et al. 2011; angle 07).

  • Diligence — sustains careful effort past the point of boredom. Signature situation: logged deliberate-practice blocks with a coach watching form (Ericsson et al. 1993; angle 07).

  • Order — keeps spaces, systems and commitments structured. Signature situation: maintained personal standards inspected without warning — the cadet-room discipline (West Point 2019; angle 08).

  • Patience — tolerates delay without abandoning the goal. Signature situation: the long-horizon project whose results only appear after months.

Family 4 — Justice. The family that orients power and advantage toward the right.

  • Righteousness — orients to what is right over what is advantageous. Signature situation: the structured dilemma discussion where the right and the profitable collide (Lind 2021; angle 06).

  • Fairness — allocates by consistent principle, not by favour. Signature situation: the resource-allocation role-play in which the learner’s own side loses under the fair rule.

  • Respect — treats every person as an end, especially the inconvenient ones. Signature situation: the restorative circle and the structured encounter across social difference.

  • Civic duty — contributes to the commons unprompted. Signature situation: service-learning with structured reflection and real community stakes (Celio et al. 2011; angle 04).

Family 5 — Wisdom. The family Aristotle put in charge: phronesis, the executive that decides which value the situation demands (Jubilee Centre 2022; angle 01).

  • Practical wisdom — reads the situation and chooses the right act among competing goods. Signature situation: adjudicating live cases with a mentor, exercising all four components — perception, adjudication, emotion, identity (Jubilee Centre 2020; angle 02).

  • Curiosity — asks the next question unbidden. Signature situation: open inquiry where the syllabus does not contain the answer.

  • Discernment — separates signal from noise and truth from spin. Signature situation: the exercise seeded with misleading information — the wargame with a lying source (Emery/TNSR 2021; angle 06).

  • Foresight — acts today against consequences years out. Signature situation: the scenario exercise in which decisions are replayed against unfolding futures.

Family 6 — Love. The family that turns formation outward; benevolence is measurable and it moves whole societies (WHR 2025; angle 10).

  • Compassion — moves toward suffering rather than away from it. Signature situation: the sustained care placement with real dependents, not a visit.

  • Generosity — gives time and resources at genuine cost. Signature situation: the giving project — designed, budgeted and delivered by the learner (Sparks et al. 2019; angle 11).

  • Loyalty — stays with people through cost. Signature situation: the team expedition where quitting harms teammates, not just oneself.

  • Forgiveness — releases a genuine grievance without denying the harm. Signature situation: the restorative-justice conference, face to face with real harm.

Family 7 — Strength. The family built almost entirely on the failure ladder — these values cannot be trained without adversity.

  • Perseverance — continues past repeated failure. Signature situation: problems calibrated just beyond current ability, attempted before instruction (Kapur & Roll 2018; angle 05; Eskreis-Winkler et al. 2014; angle 07).

  • Resilience — recovers form after a real blow. Signature situation: supported adversity — the expedition that goes wrong by design, inside a safety envelope.

  • Humility — updates when proven wrong. Signature situation: the post-error debrief in which the learner’s confident model publicly fails (Metcalfe 2017; angle 05).

  • Gratitude — registers what is given rather than what is owed. Signature situation: the daily noticing practice — three good things, done until automatic (Seligman et al. 2005; angle 10).

Family 8 — Purpose. The family that answers why train at all — and the strongest known motivational engine for the rest.

  • Calling — connects daily effort to a contribution beyond the self. Signature situation: the purpose interview and the self-transcendent reframing of ordinary work (Damon 2003; Yeager et al. 2014; angle 12).

  • Hope — builds concrete pathways to a wished-for future. Signature situation: mental contrasting with implementation intentions — the wish met honestly by the obstacle and the plan (Gollwitzer & Sheeran 2006; angle 07).

  • Reverence — stands rightly before what is larger than the self. Signature situation: the wilderness solo; the encounter with things too large to master.

  • Stewardship — leaves what it touches better, and hands it on. Signature situation: owning a real asset — a garden, an archive, a younger cohort — across a full year.

Eight families, thirty-two values, every one of them stated as a behaviour and paired with a situation. That is the syllabus of the gymnasium. What follows is the rep.

The formation loop — what one rep looks like

A gymnasium is not defined by its equipment but by its unit of work: the rep, performed with correct form, under progressive load. The formation loop is the rep of character training. It compresses what this library establishes from separate directions — Kolb’s experiential learning cycle (Kolb, Boyatzis & Mainemelis 2001; angle 04), the productive-failure paradigm (Kapur 2015; angle 05), the science of habit (Wood 2016; angle 07) and the Jubilee Centre’s caught–taught–sought account of how character actually enters a person (Jubilee Centre 2022; angle 11) — into six steps: Situate, Attempt, Fail productively, Correct, Habituate, Prove. The loop is scale-invariant. It can run in a ten-minute classroom drill, a semester-long service project, or a 47-month academy programme (West Point 2025; angle 08). What may never be skipped is a step — and the step institutions always skip is the third.

Step 1 — Situate: put the learner in a situation, not in front of one

Formation begins by placing the learner inside a situation that demands the value — real where possible, simulated where reality is too dangerous or too rare. The evidence for insisting on situations rather than descriptions of situations is now direct. When moral dilemmas are presented in immersive VR rather than as text vignettes, people respond differently — simulated moral action diverges from armchair moral judgment (Francis et al. 2016; angle 06), and a growing review literature maps how VR dilemmas expose the gap between what people say and what they do (Europe PMC 2025; angle 06). Under time pressure in VR road-traffic dilemmas, ethical decision behaviour can be modelled as it actually operates — fast, embodied, constrained (Sütfeld et al. 2017; angle 06). The military reached the same conclusion in the field: ethics scenarios woven into high-intensity field exercises train what classroom ethics cannot (Europe PMC 2014; angle 06). And where full simulation is impractical, the Konstanz Method of Dilemma Discussion gives teachers an operational, manualised way to stage a genuine moral situation inside an ordinary classroom (Lind 2021; angle 06). The deeper reason situating matters is Haidt’s: moral judgment runs intuition-first (Haidt 2001; angle 03) — and intuitions are trained by encounters, not by arguments. Design rule: every value in the canon has its signature situation; the institution’s job is to manufacture encounters with it on a schedule.

Step 2 — Attempt: act under uncertainty, before being shown how

Inside the situation, the learner must act — decide, commit, produce — before anyone demonstrates the correct answer. This is where the experiential-learning evidence carries the load. Dewey’s foundational claim that education happens through experience and its consequences (Dewey 1938; angle 04) is now backed by convergent quantitative results: a 62-study meta-analysis of service-learning shows gains in attitudes toward self, school engagement, civic engagement, social skills and academic performance (Celio et al. 2011; angle 04); an 11-study meta-analysis confirms service-learning reliably increases student learning (Warren 2012; angle 04); the definitive literature review of project-based learning sets out the design principles under which acting on authentic problems works K-12 (MDRC/Condliffe 2017; angle 04); four rigorous studies show PBL raising achievement across grades and subjects, including for low-income students (Lucas Education Research 2021; angle 04); and a 66-study meta-analysis quantifies PBL’s effects on thinking skills, attitudes and achievement (Zhang & Ma 2023; angle 04). Crucially, Celio’s meta identifies reflection and student voice as moderators — attempts formed by the learner’s own choices, later reflected on, are what move character, not supervised errands. Design rule: the attempt must be genuinely the learner’s — their plan, their call, their name on it.

Step 3 — Fail productively: design the failure in, before the instruction

This is the step that separates a gymnasium from a lecture hall, and the one conventional schooling engineers out. The productive-failure programme shows that having learners generate solutions to problems beyond their current ability — and fail — before receiving instruction produces deeper learning than instruction-first sequences, through three mechanisms: activation of prior knowledge, awareness of knowledge gaps, and better encoding of the canonical solution when it finally arrives (Kapur 2015; Kapur & Roll 2018; angle 05). The effect holds even at micro-scale for procedural knowledge (Ziegler, Trninic & Kapur 2021; angle 05) and has been extended as a design paradigm to programming education (arXiv 2024; angle 05). The Bjorks’ desirable-difficulties framework generalises the principle: conditions that make practice harder and slower — spacing, interleaving, generation — improve retention and transfer (Bjork & Bjork 2011; angle 05). Errorful learning followed by corrective feedback beats errorless learning (Metcalfe 2017; angle 05). And in adult training, a 24-study meta-analysis shows error-management training — explicitly inviting errors and framing them as informative — outperforms error-avoidant training, especially for adaptive transfer to novel problems (Keith & Frese 2008; angle 05), with modern professional evidence extending it to high-stakes clinical skill (Europe PMC 2024; angle 05). The frame matters as much as the failure: a short growth-mindset intervention reframing failure as information raised achievement in a national randomized experiment of over 12,000 students (Yeager et al. 2019; angle 05). Design rule: failure must be survivable, calibrated and expected — a designed feature the learner knows is coming, not an ambush.

Step 4 — Correct: feedback, reflection, and the mentor’s debrief

An error that is never examined trains nothing — or worse, trains the error. The corrective step is where experience becomes learning. The cognitive machinery exists and is measurable: after committing errors, people spontaneously slow down and adjust — post-error slowing and accuracy adjustment are the substrate of situational self-correction (Danielmeier & Ullsperger 2011; angle 05) — and corrective feedback after error commission is precisely the condition under which errorful learning outperforms errorless (Metcalfe 2017; angle 05). Kolb’s cycle formalises the pedagogy: concrete experience must pass through reflective observation and abstract conceptualisation before it re-enters action (Kolb, Boyatzis & Mainemelis 2001; angle 04) — which is why reflection quality moderates service-learning outcomes (Celio et al. 2011; angle 04). The human form of this step is the debrief, and its canonical model is cognitive apprenticeship: the mentor models, coaches, scaffolds — and then deliberately fades (Collins, Brown & Newman 1987; angle 11). In moral formation specifically, structured dilemma discussion measurably shifts justice reasoning on the Defining Issues Test (ERIC 2015; angle 06), the instrument tradition behind forty years of documented moral-judgment growth (Thoma 2014; angle 03). Design rule: every designed failure has a scheduled debrief with a named coach; no rep ends at the error.

Step 5 — Habituate: reps until the value runs without willpower

A corrected behaviour is still a fragile behaviour. The fifth step turns it automatic, because character that depends on daily willpower is not yet character. Habit science supplies the mechanics: behaviour repeated in stable contexts, cued and rewarded, transfers control from deliberate intention to automatic response (Wood 2016; angle 07), and the practical protocol — anchor the new behaviour to an existing routine and repeat, with automaticity plateauing over roughly 66 days in the underlying research — is documented and usable (Gardner, Lally & Wardle 2012; angle 07). Implementation intentions — if-then plans binding situations to responses — convert intention into action with a meta-analytic effect of d = .65 across 94 studies (Gollwitzer & Sheeran 2006; angle 07). Combined with mental contrasting as WOOP/MCII, the technique raised adolescents’ self-disciplined studying by roughly 60% in a school experiment (Duckworth et al. 2011; angle 07) and improved grades, attendance and conduct in replication (Duckworth et al. 2013; angle 07). For the skill-like face of virtue, deliberate practice supplies the structure — effortful, feedback-rich, structured repetition (Ericsson, Krampe & Tesch-Römer 1993; angle 07) — with two honest caveats: Ericsson’s own strictures on what actually counts as deliberate practice (Ericsson & Harwell 2019; angle 07), and Macnamara’s meta-analysis showing practice explains a real but bounded share of performance variance across domains (Macnamara, Hambrick & Oswald 2014; angle 07) — so the gymnasium must engineer context, feedback and opportunity, not just hours. Schools have an evidence-backed frame for teaching the self-regulation this step requires (EEF 2018; angle 07), and the COM-B model gives designers the full checklist — capability, opportunity, motivation (Michie, van Stralen & West 2011; angle 07). Habituation, note, is not childhood-only: the Aristotelian case that it runs lifelong is explicit (Sanderse 2018; angle 02). Design rule: every value in a learner’s current formation plan has a daily or weekly rep with a context cue — and the rep survives the school holidays.

Step 6 — Prove: demonstrate it in the world, and measure it honestly

The loop closes only when the value shows up outside the gymnasium — under observation, in situations that were not staged for the learner’s benefit. Proof has two halves. The first is demonstration: real responsibility discharged in the real world — the service project delivered (Celio et al. 2011; angle 04), the standard upheld under an honor code across years (ERIC 2010; angle 08), the pattern West Point institutionalises by making cadets live and lead honourably across a 47-month experience rather than pass a course (West Point 2025; angle 08). The second is measurement that resists self-flattery: situational judgment tests that present realistic scenarios and score the chosen response, with validated instruments now existing for character-adjacent constructs like dependability (Europe PMC/PLOS 2019; angle 13); moral-dilemma instruments fielded at scale on 10,000+ UK students (Jubilee Centre 2015; angle 13); and behavioural evidence prioritised over self-report, because self-reported character is confounded by reference bias and faking (Heckman & Kautz 2014; RAND 2014; angle 13). Proof feeds the next loop: it sets the next load. Design rule: proof events are scheduled, observed, and recorded as growth against the learner’s own baseline — the full measurement architecture, and its guardrails, are set out later in this report.

Run those six steps, on a schedule, against the 32 values, for years — that is the whole method. The remainder of this playbook is the institutional machine that makes the loop run: eight programme areas, ranked.

The eight programme areas — how the ranking works

The gymnasium is built from eight programme areas. They are ranked, not listed: by the strength of the evidence behind them, by leverage — how much of the 32-value canon each one trains — and by dependency, because some areas are the load-bearing walls the others hang from. Ethos comes first because every other area runs inside it and is silently cancelled by a culture that contradicts it. The challenge curriculum and the dilemma gym are the twin training floors — real situations and simulated ones. The failure ladder and the habit protocol are the two halves of the loop institutions most reliably omit. Mentorship is the human transmission channel, purpose-matching is the personalisation layer, and adult formation extends the whole system past age eighteen, where most states currently stop. Each area is presented as a compact operating brief: In short · Why it ranks here · Methods that fit · Signals & KPIs · Institutional wiring & first moves.

Area 1 — The whole-school ethos: caught before taught

In short. The culture of the institution is the primary curriculum. Learners absorb the values the institution practises — in its corridors, staff rooms, discipline policies and small daily transactions — long before and long after any lesson about values. The first programme area is therefore not a programme at all: it is the deliberate engineering of ethos.

Why it ranks here. The evidence is unusually consistent, from both directions. Positively: the Jubilee Centre’s field-defining framework holds that character is caught through ethos and role-modelling, taught through instruction, and sought through practice — in that order (Jubilee Centre 2022; angles 01, 11). The SRCD’s landmark analysis argues social-emotional formation works as continuous whole-school strategies embedded in daily practice, not as packaged curricula (SRCD/Jones & Bouffard 2012; angle 09). Berkowitz’s research programme distils what works into design principles — the PRIMED framework — moving the field from advocacy to implementation science (Berkowitz, Bier & McCauley 2017; angle 01; NASEM 2016; angle 01), and the landmark 213-programme meta-analysis found that programmes work when they carry the SAFE features — sequenced, active, focused, explicit (Durlak et al. 2011; angle 09). Negatively: the largest randomized multi-programme evaluation of schoolwide character curricula found essentially null effects — bolt-on programmes dropped into unchanged school cultures do not move children (US IES 2010; angle 15), and KIPP’s rigorous evaluation found strong achievement effects but flat character-survey effects even in an intensely character-branded network (Mathematica 2015; angle 09). Culture is not one factor among many. It is the medium.