VITALIS™ S1 · The System

Three readings.One management loop.

VITALIS™ S1 (VS1) turns work managers already observe into a structured diagnosis of current capability, growth momentum and core vitality.

A manager and colleague reviewing work together in a professional meeting.

Built for ordinary work, VS1 combines direct observation, written anchors, dialogue and repeated sampling in one management process. It preserves the named work instances behind a reading, keeps the three questions separate and returns the result to the manager who observed the work and must decide what happens next.

VS1 is not a one-off assessment event or a single composite verdict. It is a structured way to examine current work, choose a proportionate response and return later to see what changed.

See how the reading is built
X Y Z CAPABILITY GROWTH MOMENTUM CORE VITALITY

The system begins by separating three questions that are often collapsed into one impression.

Three questions · One architecture

One result can hidethree different realities.

VS1 reads current capability, growth momentum and core vitality separately because each can point to a different management action.

X Y Z

Axis 01 — X · Capability

What can this person deliver now?

Read the range, difficulty and independence already demonstrated in the work.

Axis 02 — Y · Growth momentum

At what rate is the work changing?

Read the direction and rate at which teaching, imposed change and personal feedback become improved work.

Axis 03 — Z · Core vitality

What appears to sustain the work when conditions become harder?

Read the care, effort and judgement visible when pressure rises, supervision reduces or the task becomes less convenient.

A capable person who is accelerating, a capable person who has stalled and an emerging person who converts feedback quickly should not receive the same response. VS1 therefore keeps X, Y and Z visible as separate readings rather than averaging them into one score.

3axes three distinct evidence questions
18dimensions six dimensions within each axis; names remain protected
5anchored levels written standards, evidence thresholds, ceiling rules and Not observed where the record is insufficient

X is tailored to the work

The structure stays fixed.The evidence must fit the role.

A hotel concierge, an IT support analyst and a logistics operator do not demonstrate reliable delivery in the same way.

A restaurant professional preparing food during service.

VS1 therefore tailors the X-axis behavioural anchors to the industry, organisation and role. The tailoring process begins with role documents, SOPs, output definitions, critical incidents, recurring exceptions and evidence managers can genuinely observe.

The X-axis functions, construct boundaries and L1–L5 level logic remain fixed. The role language, behavioural examples, observable outputs and realistic role ceiling are tailored. Y and Z retain the same meaning and decision logic across industries; explanatory examples may be localised without changing the anchors.

  1. Understand the business and target role.
  2. Align L3 to the documented role standard.
  3. Translate the fixed capability functions into observable work anchors.
  4. Mark role ceilings, edge cases and evidence gaps.
  5. Calibrate disputed examples before live use.

Tailored to the work. Anchored to the same standard.

Once the questions and evidence anchors are clear, the reading can remain light enough to live inside ordinary management.

Built into the work

Built into work.Not added on top of it.

VS1 is designed as a lightweight, incremental management practice—not a separate assessment event.

The first reading is built over six to eight weeks from work that is already taking place. Managers capture only decision-relevant evidence as it occurs; entries can be saved in parts and completed across short sessions.

The current design target is approximately 20–30 minutes of total system input per person for a reading. It is not intended to require one uninterrupted sitting, a separate assessment day or additional work created for the purpose of being assessed.

Once sufficient evidence is available, the manager discusses relevant instances with the person, reads the three axes separately, chooses a proportionate next action and returns later to examine what changed.

  1. 1 · Evidence
  2. 2 · Interaction
  3. 3 · Diagnosis
  4. 4 · Result analysis
  5. 5 · Management and coaching direction
  6. 6 · Read again

6 returns to 1

6–8weeks ordinary work observed; not additional assessment time
20–30minutes current total-input target for one reading; save and resume
2readings act, review and adjust
Boundary

The 20–30 minute figure is a current operating target, not a guaranteed burden estimate. Live deployments will examine whether the target, cadence and save-and-resume workflow remain practical in use.

Capture a little. Act when the evidence is sufficient. Return to see what changed.

The three readings describe the work. Four deeper analyses extend what can be examined.

Deeper analysis from the same evidence

Three readings.Four deeper analyses.

VS1 extends the three-axis reading through four deeper analyses that examine capability release, current drive, momentum and observed assignment conditions—without converting the evidence into a verdict.

Analysis 01

PRG · Potential Release Gap

Is capability missing—or is observed capability not being fully released in the current work?

PRG is designed to surface a gap between capability already observed and the extent to which it appears to be released under current conditions. It helps a manager avoid treating every constrained result as a training problem.

Boundary“Potential” refers here to presently observed capability that may not be fully released. PRG does not calculate future potential or predict the performance uplift an intervention will cause.

Analysis 02

DRIVE · Current investment intensity

Does the current pattern of effort, standards and response to change warrant closer review?

DRIVE combines selected growth-momentum and core-vitality evidence into a current review signal. It helps direct attention before a concern is reduced to a general impression.

BoundaryDRIVE is a designed review signal. It is not a prediction of resignation, retention or future performance.

Analysis 03

GPI / DI · Growth Momentum Index / Decline Momentum Index

Is the present direction of change moving upward or downward?

GPI and DI summarise the direction of currently observed movement. GPI represents positive growth momentum; DI represents decline momentum. Together they help a manager see direction rather than relying on a static level alone.

BoundaryThese are momentum readings, not potential scores or forecasts of a future outcome.

Analysis 04

8 Evidence Flags · Assignment safeguards

Which already-observed conditions require a safeguard, staged assignment or further evidence?

Eight evidence flags identify observed conditions that may matter before important work is entrusted. A flag should show what requires review, what work remains proportionate and what positive evidence could release the safeguard.

BoundaryThe flags do not predict a future incident. They organise evidence that has already been observed and keep the response proportionate to the work.

4deeper analyses capability release, current drive, momentum and safeguards
8evidence flags observed review conditions, not predictive labels

These analyses extend the evidence. They do not turn it into a verdict.

The application is new. The disciplines behind direct observation, structured judgement and repeated review are not.

Colleagues reviewing a written record of observed work.

Established disciplines · New application

The application is new.The operating disciplines are established.

VS1 does not claim to invent direct observation, anchored judgement or repeated review; it brings those disciplines together for a specific management problem.

Precedent 01

Kazuo Inamori · The strategic distinction

The Result of Life or Work = Attitude × Effort × Ability

Kazuo Inamori's formulation supplied a strategic question: visible ability is not the whole explanation of a work result. VS1 carries forward that distinction by reading capability, growth momentum and core vitality separately.

It does not reproduce Inamori's numerical ranges, multiply the three axes or treat the formula as validation evidence for VS1.

SOURCE FROM · Kazuo Inamori, Kyocera Philosophy, “The Result of Our Life or Work = Attitude × Effort × Ability.” Official Kyocera source.

Precedent 02

mini-CEX · Direct observation in real work

The American Board of Internal Medicine introduced the mini-Clinical Evaluation Exercise, or mini-CEX, in 1995. A supervising physician observes a trainee during real work, rates the encounter against written dimensions, provides feedback and repeats the process across situations.

In a 2003 study across 21 internal-medicine training programmes, 316 evaluators recorded 1,228 encounters involving 421 residents. Ratings showed statistically significant improvement across the first year of training. The instrument did not cause that improvement; it gave supervisors a structured way to observe it.

VS1 adapts five operating principles: real work, an accountable observer, written anchors, dialogue and repeated sampling.

1995preliminary mini-CEX study published
1,228observed encounters in the 2003 study
421residents represented
316evaluators represented
21internal-medicine training programmes

SOURCE FROM · Norcini, Blank, Arnold & Kimball (1995), Annals of Internal Medicine; and Norcini, Blank, Duffy & Fortna (2003), Annals of Internal Medicine. These studies establish the history and operation of mini-CEX in medical education. They are method precedents, not validation evidence for VS1.

Precedent 03

Structured interviews · The value of defined judgement

In personnel-selection research, revised meta-analytic estimates place the operational validity of structured interviews around .42, compared with .19 for unstructured interviews.

These figures concern hiring validity, not workplace observation. They do not transfer to VS1 and are not comparable with the mini-CEX reliability figure in the next section. They support one design principle only: when judgement matters, define the question, evidence rule and scoring anchor.

.42structured-interview operational validity
.19unstructured-interview operational validity

SOURCE FROM · Sackett, Zhang, Berry & Lievens (2022), “Revisiting Meta-Analytic Estimates of Validity in Personnel Selection,” Journal of Applied Psychology. The figures concern personnel selection and must not be presented as VS1 performance estimates.

Established disciplines can guide the design. VS1 must still establish its own evidence.

Structure does not remove disagreement. It creates the conditions for disagreement to be located and measured.

An evidence boundary

Structure is not certainty.It is inspectability.

Written anchors create the conditions for agreement to be tested—not assumed.

In a 2009 study, 52 internal-medicine faculty rated videotaped resident–patient encounters using both the traditional nine-point mini-CEX scale and a five-point scale. Inter-rater reliability for the overall rating was .43 on the nine-point scale and .40 on the five-point scale. The authors described reliability and accuracy as modest.

The .43 is useful because it shows that even a mature structured-observation method still faces rater variation. It is not a pass mark, a target or a VS1 result. Repeated evidence, explicit anchors, calibration and visible disagreement remain necessary.

There is no single error-free gold standard for workplace talent diagnosis. Any diagnostic judgement can be affected by incomplete evidence, observer variation and context. Calibration is therefore a required operating discipline—not an optional quality check. It tests how anchors are interpreted, makes disagreement visible and prevents confidence from standing in for evidence.

VS1 has no published inter-rater reliability result yet. Live deployments will examine rater agreement, evidence sufficiency, usability, the 20–30 minute input target and whether a second reading can support a practical management loop. The method, sample, limits and results should be reported—including if a result is lower than expected.

.43IRRtraditional nine-point mini-CEX overall rating
.40IRRfive-point mini-CEX overall rating
52ratersinternal-medicine faculty
Not yet publishedVS1 inter-rater reliability

SOURCE FROM · Cook & Beckman (2009), “Does scale length matter? A comparison of nine- versus five-point rating scales for the mini-CEX,” Advances in Health Sciences Education, 14, 655–664. Participants rated videotaped encounters; some were scripted to represent different competence levels.

The standard is not certainty. It is a judgement whose evidence and disagreement can be examined.

The right next step is not a broader claim. It is a disciplined test in work where the evidence is visible.

Next step

Where this goes next.