Pipeline
Build a Pipeline of Dated Evidence, Not a List of Interested People
A candidate pipeline built before the role opens is only worth what its records mean months later. For each person store what they produced, when, under which rules, and one sentence naming what it showed; a tag and a rating will not survive the wait. Give anything involving tool-assisted work a shelf life of about a year, and mark it expired rather than deleting it, so a re-approach starts with a fresh sample. A talent community with no evidence in it is a mailing list.
The takeNurture is real and it is far smaller than the software implies. Two or three honest updates a year about what the team is actually building is the version a person still opens; a monthly drip is the version that teaches them to filter your domain. The costly work sits somewhere else entirely: keeping a record that still means something on the day the requisition finally opens, months after everyone who wrote it has moved on.
Where Olive fits
Open a role and see what the work shows
An Olive report is dated by construction: six findings, each carrying the timestamped excerpt from the session it rests on, released with the rubric, scorer and item-bank versions attached. That is the shape of a pipeline record that can still be read honestly a year later, and the candidate holds an identical copy of it.
Rank your shortlistWhat is worth keeping in a pipeline?
Four fields per person, and a rating is not one of them: the artifact they produced, the date, the rules in force when they produced it, and one sentence naming the capability it showed. Everything else in a typical talent-community record is either recoverable from the person later or already stale by the time a requisition opens against it.
The reason to build anything at all is elapsed time. Across Ashby's own customer base, which skews to venture-backed technology employers, median time to first fill runs 56 days for business roles and 76 for technical ones, roughly eight to ten weeks 1. Read that clock carefully: it stops at the first hire on a requisition, and SHRM's 39 calendar day median for nonexecutive positions, from a survey of over 4,600 organisations, runs a different clock over a different population 2. Nothing published apportions the gap between those two causes, so read each figure on its own terms. Either way, the wait between opening a role and having somebody in it is long enough that a warm start is worth something, which is the honest case for a pipeline.
What that warm start consists of is where the vendors and the evidence part company. A stored record is worth having when it lets a future reader form their own view. So:
- The artifact. The exercise they submitted, the portfolio link, the writing sample, the thing itself.
- The date. When the work was made, which is rarely the day you added them to a list.
- The rules in force. What tools were permitted, how long they had, whether the brief was yours or theirs.
- One sentence. Which capability the work demonstrated, in the vocabulary of the job, written by whoever read it.
How fast does stored hiring evidence spoil?
About a year for a judgment about somebody's output, and indefinitely for the underlying facts. What a person built and where they worked ages slowly. What an interviewer concluded about their output ages with the tools, because a written sample or a coding exercise from two years ago was produced under different conditions than one produced this month, which makes the two incomparable even when the exercise is word for word identical.
The size of that shift is easy to underestimate. Evaluated against 115 Python problem statements taken from HackerRank, OpenAI's Codex solved 96% of them with no examples and all of them with a few, and the authors report clear signs that the model was reproducing memorised code rather than synthesising it 3. The memorisation half is the load-bearing half and it usually gets dropped: the paper's own contribution is a way of re-testing on mutated problems precisely because the raw pass rate overstates capability. For a pipeline, though, it reads as a straightforward warning. An exercise drawn from a public bank was already testing a lookup in 2022, on a model several generations old, so a stored pass on one from 2023 is thin evidence about anybody and no evidence at all about this year's pool.
Hence the shelf life, and shorter still where the exercise came from a public source, since that one was testing a lookup on the day it was set. Mark the lapsed record expired and keep it: the artifact and the date stay useful as context, and the expiry flag is what stops a future reader taking an old judgment for a current one.
The specific case of reopening an old finalist pool, where every record is expired at once, is worked through in whether 2024 interview evidence still holds.
Write down what the person actually did
Write the observation: which capability the person demonstrated, on what task, and what specifically they did that showed it. A number is the part of a record nobody can argue with later, which is exactly why it is the wrong thing to store. A sentence naming a capability still means something to somebody who was not in the room. Nobody can reconstruct what a four out of five was for once its author has left.
Interviewer notes are a rarer object of study than interviewer ratings. Researchers text-mining post-interview notes on 7,650 candidates who were hired at a large Chinese internet technology company found that the number of job-related capabilities an interviewer named in the notes related positively to later job performance and promotions, and negatively to turnover, with a standard deviation rise in the note-to-job-analysis matching score corresponding to roughly a 2 percent rise in performance 4. Take that with its limits. It is correlational, at one firm, in one country, between 2016 and 2018, and it can only see candidates who were hired, so it says nothing about the people the notes described and the company rejected. The effect is small, and it points the right way.
Nobody can say what a stored rating was worth, including the people who stored it. In LinkedIn's 2025 survey of 1,271 recruiting professionals in management roles across 23 countries, only 25% said they felt highly confident in their organisation's ability to measure quality of hire effectively 5. That is a self-report from a convenience sample of platform-active talent professionals, not an audit of what anybody measures, and quality of hire is a downstream outcome, some distance from the interview score itself. Still, an employer that cannot tell how its own hiring judgments turned out has nothing to calibrate a stored score against, and a written observation needs no calibration to be read.
So write the sentence in the job's own vocabulary. Not strong candidate. Rebuilt the reconciliation from the raw ledger and found the two entries that did not net, without being told they were there.
Keep the nurture small and the promise honest
Two or three genuine updates a year is the whole of it. Somebody who agreed to stay in touch expects to hear something they could not read on the careers page: what the team shipped, what changed about the role, what the hiring plan actually is now. Anything more frequent teaches them to filter you, and anything vaguer is not worth the time it takes to write.
The re-approach is where the record earns its keep, and honesty is the cheapest thing in it. Say what you have: your exercise is from early 2024, the tools you were allowed then are not the tools now, so treating it as current would not be fair to you. Would you do a fresh one. Almost nobody objects to that, and somebody who would rather not has told you they are not available for this round, which is worth knowing on day one.
Three things are worth deciding once, in writing, before the pipeline has anybody in it:
1. What you keep, and for how long. Write it down and apply it. Record-retention duties differ by jurisdiction and can set a floor under your own preference, so settle the number with counsel before it becomes policy. A record you would be uncomfortable showing the person it describes should not be in there. 2. Who can read it. A stored observation written for a hiring team reads differently when a manager finds it two years later attached to an internal applicant. 3. How somebody leaves. One reply should be enough to take a person off the contact list for good, and it should actually take them off.
Before any of this, know what the role will require, because a pipeline built against a job description that has stopped describing the work stores the wrong artifacts carefully. That groundwork is finding out what AI actually does in the role, and it is the difference between a pipeline of evidence and a list of names that were relevant once.
Common questions
How long does interview evidence stay usable?
Treat about a year as the working limit for anything involving tool-assisted work, and shorter where the exercise came from a public problem bank. Facts age more slowly than judgments: where somebody worked and what they shipped stay true, while a score derived from a document or an exercise rests on conditions that have changed. The practical rule is to keep the artifact and the date indefinitely, and to expire the assessment attached to it, so a future reader knows they are looking at history rather than a current view.
Is a recruiting CRM worth buying for this?
Only if you already know what you would store in it. The failure mode is buying the tool and letting its default fields decide, which produces tags, stages and drip campaigns because that is what the software is for. A spreadsheet holding an artifact link, a date, the rules in force and one sentence of observation beats a well-configured system holding ratings, because the spreadsheet can still be read honestly a year later. Buy the tool once the record has a shape and the volume justifies it.
What should I actually send in a nurture update?
Something the person could not get from the careers page. What the team shipped this quarter, how the role has changed, what the hiring plan is now, and whether the thing they were interested in is still happening. Two or three a year is the sustainable rate. Skip anything that reads as marketing, skip anything you would not send to a friend, and make it easy to leave in one reply. The point of the update is that the person still opens the next one.
Should we tell candidates they are in a pipeline?
Yes, and say what that means concretely, because the phrase means nothing to somebody outside recruiting. Tell them what you kept, roughly how long you expect it to stay relevant, how often you will be in touch, and how to be removed. That conversation costs one paragraph and it converts a stored record from something done to a person into something agreed with them. It also improves the pipeline, because the people who opt out were never going to answer the re-approach anyway.
Do silver medalists deserve special treatment when the role reopens?
They deserve a faster start, not a shorter process. A previous finalist has already shown you something, which is a genuine head start on the conversation and on the reference checks. What has not stayed constant is the evidence: the market, the tools, the role and the person have all moved, and the other candidates in the new round are being measured this year. Put them straight into the current evidence stage, and tell them why.
References
- 1. Recruiter Productivity | 2026 Talent Trends Report ashbyhq.com Supports the median time to first fill of 56 days for business roles and 76 for technical ones, and the point that this clock stops at the first hire and is not comparable to SHRM's 39-day median.
- 2. 2026 Recruiting Executives Benchmarking: Attracting Critical Talent shrm.org Supports the 39 calendar day median time-to-fill for nonexecutive positions across over 4,600 organisations, cited as a survey median on a different clock from the Ashby figure.
- 3. Codex Hacks HackerRank: Memorization Issues and a Framework for Code Synthesis Evaluation arxiv.org Supports the claim that an exercise drawn from a public problem bank was already testing a lookup in 2022, including the memorisation finding that limits how the pass rate should be read.
- 4. Predictive Validity of Interviewer Post-interview Notes on Candidates' Job Outcomes: Evidence Using Text Data From a Leading Chinese IT Company frontiersin.org Supports the claim that notes naming job-related capabilities carry predictive signal a rating alone does not, with the range restriction and single-firm limits stated.
- 5. The Future of Recruiting 2025 business.linkedin.com Supports the claim that only 25% of surveyed recruiting professionals feel highly confident in their organisation's measurement of quality of hire, cited as a self-report from a convenience sample.
5 sources, numbered by first appearance. How Olive sources claims
General guidance for hiring teams. What works at one company and one volume may not transfer to yours.
Olive assesses how a person works with AI. It does not detect AI-written documents, and it never produces a score, a ranking, or a match percentage for a person. Candidates read the same report the employer reads.