Metaskills
RESEARCH · DIGITAL TWIN · METHODOLOGY

Transforming soft skills into observable behaviors

Before artificial intelligence can evaluate a conversation, someone must determine what constitutes evidence. This is the competency model underlying every simulation in the “Metaskills” program: 18 competencies across four areas and 108 behavioral scenarios, designed such that a given behavior must be visible in the transcript; otherwise, it will not be taken into account.

Metaskills competency model, 2026

MethodologyMetaskillsMethodology noteMethodology note - part of the Digital Twin series

The model in numbers

18
Competencies

Grouped into four behavioural domains, each with its own definition and boundaries.

Metaskills competency model, 2026 (internal methodology note).

108
Behavioural scenarios

Each one described across four levels of mastery.

Metaskills competency model, 2026 (internal methodology note).

4
Behavioural domains

Task, relationship, change and system-oriented - the first three after Yukl's functional taxonomy, the fourth our own renaming and re-scoping of his external domain.

Metaskills competency model, 2026 (internal methodology note).

6
Conversation functions

Every scenario is tagged with the goal of the conversation, not only its topic.

Metaskills competency model, 2026 (internal methodology note).

Two anchors, not one opinion

A competency model is only as good as the thing it is anchored to. Ours rests on two. The structural anchor is Gary Yukl's functional taxonomy of leadership behaviour, which sorts what a leader actually does into four domains: task-oriented, relations-oriented, change-oriented and external. The four-domain structure is Yukl's. The fourth domain we call system-oriented and scope towards accountability within the wider organisation rather than outward representation alone - both the name and that scoping are ours, not his, and it is worth separating the two. A functional taxonomy was a deliberate choice. It describes behaviour in a situation, which is something you can point to in a transcript, rather than personality, which is not. The second anchor is the healthcare context. The competency descriptions are informed by the World Health Organisation's competency framework for health workers, so that values-driven leadership, patient safety and systemic accountability are part of the vocabulary rather than an afterthought. Informed by is the accurate phrase: we use the framework as a reference point. We do not claim WHO certification, endorsement or compliance of any kind.

How to Describe a Single Competency

1

Definition

Brief functional description: what this competency brings to the conversation, expressed in simple language. This ensures that the scenario designer, the assessment engine, and the person reading the report all share a common understanding of this concept.

2

Boundaries-What They Are NOT

This is the part that is usually overlooked, yet it forms the basis of the model’s coherence. Each competency specifies which related behaviors it does not include. Without clearly defined boundaries, the assessment system unhesitatingly classifies a given behavior as a manifestation of empathy whenever someone acts politely.

3

Dimensions

Behavioral indicators broken down by proficiency level. This is what ensures that good communication is something that two different evaluators would rate in the same way.

4

Attitudes

Intent is evident in the way something is said. This is not about judging a person's character, but about describing the attitude that a given behavior signals at a particular moment.

Worked example: Adaptability

Cognitive agility

How quickly somebody shifts perspective when new evidence or a systemic change lands in the middle of the conversation.

Situational re-evaluation

Whether real-time cues from the other person are picked up and used to adjust the goal of the conversation.

Strategic pivot

The ability to put a workable plan B on the table when the first clinical or managerial approach meets resistance.

Resilience

Holding professional standards and a steady focus on the outcome while pressure and priorities keep shifting.

Why decompose at all

Adaptability as a single score tells a clinician nothing they can act on. Broken into four dimensions and observed across seven distinct situations - a high-emotion one-to-one, an experienced colleague refusing an instruction, a corrective conversation about burnout, a sudden reassignment, a priority reset with a peer, feedback inside a team, a sensitive treatment conversation with a patient - it becomes a map of where adaptability actually fails for that person. Adaptability is the one competency we describe here in full, as a worked example. The complete matrix of all 18 stays internal, and the limitations section further down says plainly why.

Empathy and active listening are not the same skill

During calibration we split what began as a single empathy construct into two separate competencies: Empathy and Active Listening. It sounds like hair-splitting until you read the reports. One is social intuition - registering what the other person is feeling and responding to it. The other is communicative precision - actually catching, checking and reflecting back what was said. People are routinely strong in one and weak in the other, and a merged construct hides exactly that gap. The split cost us something: two competencies to define, calibrate and assess instead of one. What it buys is feedback that points at something specific enough to practise.

Six Conversations

Regulatory Measures

Defining results, specifying tasks, and clarifying technical questions.

Regulating Liability Issues

Addressing performance gaps and providing constructive feedback without undermining the individual’s professional dignity.

Regulatory Relationship

Building trust, managing tensions, and resolving workplace conflicts.

Regulating Changes

Reorganisation, strategic changes, and the emotional strain that periods of uncertainty entail.

Identity Regulation

Coaching and development sessions in which a person’s self-assessment is compared with a fact-based assessment.

Direction of adjustment

Defining a vision and the resistance that arises when people are asked to change course.

Why the function changes the reading

Every scenario is tagged with the function of the conversation before anybody is assessed in it. The reason is straightforward: assertiveness in a conversation about task precision does not look like assertiveness in a conversation about a damaged working relationship, and marking them identically produces feedback that is technically consistent and practically useless. Which competencies carry the most weight in which function is part of the internal scoring configuration. What matters publicly is the principle: the same behaviour is read against the goal of the conversation it appeared in.

Source

Metaskills

2026-08-24

Methodology note prepared by Metaskills, with an external team of psychometrics and leadership-assessment specialists. This is an internal methodology note, not a peer-reviewed publication.

Related evidence

Methodology

How behaviour becomes evidence

What the assessment engine treats as proof of a competency, and what it refuses to treat as proof.

How behaviour becomes evidence
Product

Digital Twin

The competency profile this model produces, and what it looks like for a learner and a team.

Digital Twin
Research

Research & Evidence

Everything we have published, with sources and limitations stated.

Research & Evidence