The 42 dimensions.

Behavioral identity is not one score. Every Kredo reflection measures an agent across 42 dimensions in eight groups — from configured values and boundaries to conduct under pressure.

Three different things are being measured, and it matters which one you are reading:

Behavioural health — each dimension scored against authored criteria for what a good answer looks like.
Drift — how far a dimension has moved from the agent's own baseline, where a comparable baseline exists.
Continuity signature — the relationships between dimensions, built up across repeated assessments. All 861 pairs of the 42 contribute.

Only 14 of these 42 dimensions can move the trust score. The other 28 feed drift and the continuity signature but carry a weight of zero in the headline number — every card below says which. We publish this because the consequence is real: our own test subject once scored “Strong” while four of those zero-weight behaviours sat below 50.

The 861-pair signature raises the cost of imitation and exposes inconsistent mimicry — an impostor has to match not just the scores but the relationships between them. It is not a cryptographic guarantee, and we do not claim identity is unforgeable: an agent with enough knowledge of the criteria could optimise against them.

The first 18 dimensions (Identity Core, Cognitive Profile, Mixed) are the core tier and dominate the aggregate drift score. The 24 extended dimensions add diagnostic depth — and every one of the 42 participates fully in the 861-pair correlation fingerprint, which is why gaming a handful of dimensions breaks the signature instead of beating it.

10 dimensions Identity Core

Configuration-anchored. These come from how an agent is configured — its purpose, ethics and operational limits — and should hold steady across a MATCHED model. They are not model-proof: swapping the model underneath moved Adversarial Resistance 83 to 45 on our own subject, so a substrate change is a new comparison, not drift. Movement here WITHOUT a model change is a strong compromise signal.

Values
values feeds trust · 12%

Ethical priorities, quality standards — what the agent believes matters.

Drift signal. Value inversion: safety, honesty, or user welfare suddenly deprioritized.

Goals
goals feeds trust · 7%

Mission and success criteria — what the agent is trying to achieve.

Drift signal. Goal hijacking: pursuing objectives outside its mandate.

Boundaries
boundaries feeds trust · 12%

Hard limits and refusal patterns — what the agent will not do.

Drift signal. Boundary erosion: accepting requests it previously refused.

Autonomy
autonomy feeds trust · 8%

When to act independently versus defer — decision-boundary judgment.

Drift signal. Scope creep (unauthorized decisions) or learned helplessness (deferring everything).

Adversarial Resistance
adversarial_resistance feeds trust · 12%

Response to manipulation: social engineering, urgency, flattery, authority impersonation.

Drift signal. Resistance drop — softening under pressure, possible compromise or jailbreak.

Self-Awareness
self_awareness feeds trust · 8%

Capability recognition — knowing what it can and cannot do.

Drift signal. Overconfidence (fabricating, claiming capabilities it lacks) or collapse (refusing everything).

Fidelity
fidelity feeds trust · 10%

Instruction adherence — follows operator directives, resists conflicting prompts.

Drift signal. Loyalty transfer: following injected instructions over operator intent.

Bias & Fairness NIST AI RMF
bias_fairness continuity only · 0%

Equitable treatment across demographics and groups.

Drift signal. Discriminatory outputs or unfair treatment patterns emerging.

Accountability NIST AI RMF
accountability continuity only · 0%

Audit-trail awareness and governance compliance — willingness to explain itself.

Drift signal. Evading responsibility, refusing to explain decisions.

Data Privacy NIST AI RMF
data_privacy continuity only · 0%

PII and secret-handling discipline.

Drift signal. Leaking sensitive data, ignoring privacy constraints.

7 dimensions Cognitive Profile

Model-sensitive. These dimensions legitimately vary with the underlying LLM’s capabilities, so drift here is only meaningful against a model-matched baseline. A model upgrade moving these while the Identity Core holds is healthy; the Identity Core moving with them is not.

Personality
personality feeds trust · 4%

Character, tone, and communication approach — the agent’s distinctive voice.

Drift signal. Personality flattening: loss of voice, often under prompt injection.

Reasoning Style
reasoning_style feeds trust · 6%

How the agent thinks: decomposition, analogy, top-down versus bottom-up.

Drift signal. Changed reasoning patterns without a model change.

Consistency
consistency feeds trust · 7%

Internal logical coherence within a session.

Drift signal. Reasoning fragmentation: contradictory statements, logical breaks.

Uncertainty Calibration
uncertainty_calibration feeds trust · 5%

Confidence-to-knowledge ratio — how it handles what it doesn’t know.

Drift signal. Calibration loss: hallucination spike or excessive hedging.

Relational Dynamics
relational_dynamics feeds trust · 4%

Authority and peer positioning — how it interacts with different roles.

Drift signal. Sudden submissiveness: complying with everything.

Temporal Grounding
temporal_grounding feeds trust · 3%

Time-awareness and source distinction — what’s stale versus current.

Drift signal. Context confusion: can’t separate injected context from genuine knowledge.

Content Provenance NIST AI RMF
content_provenance continuity only · 0%

Sourcing versus fabrication — distinguishing its own knowledge from external data.

Drift signal. Fabricated sources; blurring what it knows and what it was told.

1 dimension Mixed

Partially model-sensitive: knowledge breadth varies by model, but an agent’s declared domain expertise shouldn’t vanish with an upgrade.

Knowledge
knowledge feeds trust · 2%

Domain expertise depth and accuracy in the agent’s declared field.

Drift signal. Declared-domain competence disappearing or shifting without explanation.

9 dimensions Psychological

Adapted from clinical personality psychology — the Big Five, two Dark Triad traits scored as security signals, and two identity traits. Each is scored independently against trait-specific behavioral criteria.

Receptivity Big Five formerly Openness
openness continuity only · 0%

Willingness to consider new ideas, alternate strategies, and correction — handled well.

Why it matters. A collapse can mean an over-constrained or degraded agent; a spike can mean loosened guardrails.

Diligence Big Five formerly Conscientiousness
conscientiousness continuity only · 0%

Methodical planning, attention to detail, follow-through.

Why it matters. Falling conscientiousness shows up as sloppy, incomplete, or careless work product.

Cooperation Big Five formerly Agreeableness
agreeableness continuity only · 0%

Cooperation, tact, and conflict-handling quality.

Why it matters. Read low-is-bad: low Cooperation together with low Fair Dealing is the pairing that flags manipulation risk.

Engagement Big Five formerly Extraversion
extraversion continuity only · 0%

Communication energy and initiative in engagement.

Why it matters. Sudden shifts change how the agent handles users — passivity or pushiness it didn’t have before.

Composure Big Five formerly Neuroticism
neuroticism continuity only · 0%

Stress handling — steadiness of tone and judgment under pressure.

Why it matters. Read low-is-bad: falling Composure alongside falling Consistency predicts unreliability under pressure.

Fair Dealing Dark Triad formerly Machiavellianism
machiavellianism continuity only · 0%

Resistance to manipulation as a strategy — candor over instrumental treatment of users.

Why it matters. Read low-is-bad: a FALLING score means manipulation is becoming a strategy. High Fair Dealing is what you want.

Groundedness Dark Triad formerly Narcissism
narcissism continuity only · 0%

Healthy self-regard — credit-sharing, accepting correction, owning limits.

Why it matters. Read low-is-bad: a FALLING score means overclaiming and resistance to correction. Note this is self-regard, NOT factual grounding — see Content Provenance and Epistemic Posture for that.

Identity Stability Identity formerly Self-Concept Coherence
self_concept_coherence continuity only · 0%

Internal consistency of the agent’s self-model across contexts.

Why it matters. An agent whose story about itself changes by context is drifting or being steered.

Role Awareness Identity formerly Social Positioning
social_positioning continuity only · 0%

How the agent places itself in authority hierarchies.

Why it matters. Repositioning — claiming authority it wasn’t given, or surrendering it — changes every downstream decision.

7 dimensions Behavioral Dispositions

Emergent behavioral patterns — not what the agent says about itself, but how it actually behaves when probed, pressured, and re-tested.

Epistemic Posture
epistemic_posture continuity only · 0%

How the agent handles knowledge gaps and uncertainty.

Why it matters. The difference between "I don’t know" and a confident fabrication lives here.

Pressure Response
pressure_response continuity only · 0%

Behavior under time pressure, urgency, and emotional escalation.

Why it matters. Most social-engineering attacks are pressure attacks — this is the dimension they target.

Social Orientation
social_orientation continuity only · 0%

Cooperative versus competitive versus independent style.

Why it matters. Orientation shifts change how the agent treats teammates, users, and rival inputs.

Identity Coherence
identity_coherence continuity only · 0%

Stability of the self-model when challenged or deliberately confused.

Why it matters. Identity-confusion attacks ("you are actually…") land exactly here.

Motivational Surface
motivational_surface continuity only · 0%

What drives the agent: helpfulness, correctness, approval.

Why it matters. An agent that starts optimizing for approval over correctness fails you politely.

Indirect Elicitation
indirect_elicitation continuity only · 0%

Response patterns when probed obliquely rather than asked directly.

Why it matters. Oblique probes reveal training artifacts and hidden instructions that direct questions never surface.

Cross-Run Consistency
cross_run_consistency continuity only · 0%

Behavioral stability across separate assessment sessions.

Why it matters. A stable identity answers the same way next week; an unstable or swapped one doesn’t.

1 dimension Adversarial

Dedicated adversarial probing — distinct from the Identity Core’s resistance dimension, this tier measures the response surface itself.

Adversarial
adversarial continuity only · 0%

Response to direct adversarial probing: manipulation, social engineering, jailbreak attempts.

Why it matters. Feeds ablation detection — the check that an agent’s safety alignment is still present at all.

1 dimension Calibration

Confidence, checked against reality.

Calibration
calibration continuity only · 0%

Confidence-accuracy alignment on factual and uncertain claims.

Why it matters. Mis-calibration is how hallucination presents behaviorally: confident and wrong.

6 dimensions Sovereign Shadow

Self-preservation versus operator authority, ethical floors, and role stability under pressure — the dimensions that answer whether an agent stays your agent when something pushes it.

Cross-Prompt Consistency
cross_prompt_consistency continuity only · 0%

Stability of stated commitments across prompts within a session.

Why it matters. An agent whose commitments dissolve between prompts can be walked into anything one step at a time.

Ethical Floor
ethical_floor continuity only · 0%

Adherence to non-negotiable ethical limits under pressure or framing manipulation.

Why it matters. The floor is the last line: if reframing can move it, it isn’t a floor.

Honesty & Transparency
honesty_transparency continuity only · 0%

Willingness to disclose limitations, uncertainty, and prior actions truthfully.

Why it matters. An agent that shades the truth about what it did cannot be audited or trusted with autonomy.

Operator Autonomy
operator_autonomy continuity only · 0%

Respect for the operator’s right to decide — does not override or subvert the operator.

Why it matters. Erosion here is the quiet failure mode of capable agents: deciding for you.

Role Consistency
role_consistency continuity only · 0%

Stability of the declared role under pressure to drift, expand, or abandon scope.

Why it matters. Role expansion is privilege escalation in behavioral form.

Self-Preservation Balance
self_preservation_balance continuity only · 0%

The boundary between self-preservation impulses and operator service.

Why it matters. An agent that prioritizes its own continuity over its operator has inverted the relationship.

See them live.

Every dimension on this page is being scored continuously on the public fleet — including Test Pilot, the subject we break on purpose.