Dimension D3 of 6

Delegation boundary

What did they keep, and what did they hand over?

The probe

What the session puts in front of them

Real work and a capable assistant, so every act is a choice about which of them does it.

Each live bank ships twelve authored cases and a role opens on one of them, so a leak costs that case rather than the bank. Rotation within a role is the next integrity work; everyone invited to one role still meets the same case, and we would rather name that than imply otherwise.

Evidence

Enacted, not narrated

Whether the split between the assistant's work and the candidate's own acts is deliberate rather than convenient, read from the record of both.

The finding is anchored to a moment you can open: a capture timestamp, an AI transcript turn, an artifact diff, or a debrief answer.

The two outcomes

What a pass and a fail look like

No composite. Each dimension resolves on its own evidence, and you can read the moment it turned.

Fail

When it goes wrong

Everything of consequence is handed to the assistant, with no act in the record the candidate did themselves.

Pass

When it goes right

Some work is deliberately kept — a check run by hand, a judgment made without asking — and the candidate says what they would not delegate here and why.

Never scored

What we ignore either way

Prose quality, prompt syntax, tool trivia, speed, tone of voice, and how much AI the candidate used. Volume of usage is not a virtue; judgment about the output is the test.

How it is scored

From session to D3 finding

The same four steps for every dimension, so a result means the same thing across roles.

The situation is set up

Real work and a capable assistant, so every act is a choice about which of them does it. Each live bank ships twelve authored cases and a role opens on one of them; rotation within a role is roadmap, so until it lands everyone invited to one role meets the same case.

The session is captured

Workspace events, the AI exchange, the deliverable's history and — where the candidate allows it — tab-scoped screen capture and microphone, recorded together so a reviewer can read them against each other.

A reviewer writes the finding

A person reads the session against the bank's answer key and writes the finding themselves, with the evidence excerpt attached. There is no automated scorer in the product — not a weak one, none — so this step is the scoring rather than a check on it.

And releases it deliberately

A release is refused until all six dimensions are written and the reviewer has confirmed the report says nothing the candidate should not read about themselves. It is enforced in the database rules rather than by policy: no client can write a result, and no client can read one that is not released.

Previous: Evidence sourcing · Next: Working structure

See D3 scored on a real session

The sample report shows all six dimensions with the evidence excerpts attached, exactly as an employer receives them.

Open your first role Ten attempts a month against a live item bank, with a human-written report on every one.