Business: HR · Lesson N.rh.7

The audit: never decide about people without checking

AI writes a performance review or a termination recommendation with the same confidence as always, but no termination or promotion can ever come out of an AI score or opinion without an audit. A five-item checklist and the principle that whoever signs the decision is always a person.

Examples for

A manager asked AI for a performance summary of an employee before deciding on a promotion, and the text that came back was flawless, with arguments that seemed to come from someone who'd closely followed the team all year. It almost became the basis for the decision with no further check. Checking the source of each praised item, two of them came from a project the employee hadn't even worked on, just listed as an observer in the meeting. The confident prose left no trace of the error.

Stop for a second on that "almost". In all these cases the wrong recommendation almost became a decision, not because the manager was careless, but because AI hands over a wrong opinion with the same confidence it hands over a correct one. In finance, a wrong number becomes a loss. Here, a wrong opinion becomes a person fired for a reason that never existed, or promoted for the wrong reason. It's the same trap from the whole module, except now the number has a first and last name.

The core idea of this lesson. AI is excellent at writing text and isn't trustworthy by default as the basis for deciding about a person: it can carry bias inherited from the history, strip a fact of its context (medical leave becoming "low performance"), invent a connection that doesn't exist, and hand all of that over with the same confidence as always. That's why AI's recommendation about people is a proposal, never a truth. This lesson closes with a five-item checklist and one non-negotiable principle: whoever signs the decision about a person is always someone named.

01Excellent at text, terrible as a basis for deciding about a person

It's worth understanding why this happens, because understanding it changes how you use the tool. The AI you use is, at its core, a machine for predicting the next word that sounds right. It was trained so the text is fluent, plausible, and confident. Nobody trained it so the reading about a person is fair. These are two different goals, and it only pursues the first one.

The side effect, here, is the most serious in the whole module. The same skill that makes AI write a flawless performance review makes it write, with the same quality of prose, a biased reading or a fact taken out of context. It has no internal sense of "this isn't fair to this person". If the history carries bias (a manager who always rated poorly whoever took leave, a team that logged health absence the same as disengagement absence), AI learns that pattern and repeats it with the look of neutral analysis. In a marketing text, a skewed reading is a slip. In a review of a person, it's a career, a labor lawsuit, a life.

02The paradox of feeling: you think you already audited it

In 2025, METR measured experienced professionals working with and without AI on real tasks, and the result was the opposite of expected: with AI, they were about 19% slower, but they thought they were faster. The feeling of a gain didn't match the measured result.

That same treacherous feeling shows up when deciding about a person. When AI's opinion arrives well written, organized, with a conclusion that sounds professional, your brain registers "this has already been analyzed" and lets its guard down. It's exactly in the most serious decision (firing, promoting, giving or denying a raise) that this false feeling of having already checked costs the most. Reading the opinion isn't auditing the opinion. They're different things, and that difference is the subject of the next section.

already checked didn't check the feeling read the opinion, looks right the real check source doesn't match reading the opinion isn't auditing the opinion

03The audit checklist for a decision about people

Here's the heart of the lesson. Five questions, in order, before any AI recommendation about a person becomes a termination, promotion, warning, or raise.

First: is the criterion used explicit, and is it the same one applied to everyone in the same situation? If the opinion says "low performance" without saying against which yardstick, and that yardstick isn't the same one used for colleagues, the criterion doesn't really exist, it only looks like it does.

Second: does the source of each claim exist and check out? Every sentence like "two weak reviews over the last cycles" needs to be checked at the source, one by one. This is the step that catches the review that was, in fact, from a period of health leave.

Third: is there calibration across groups? Look at the whole set: is AI's negative-recommendation rate disproportionate in some group (whoever took leave, whoever has a specific demographic profile, whoever's on a specific team)? A disproportionate pattern doesn't prove individual guilt, but it's a signal that the criterion may be carrying bias inherited from the history.

Fourth: would someone put this recommendation in front of the person and defend it eye to eye? If the answer is "I'd rather she never find out an AI generated this like this", the recommendation isn't ready.

Fifth, and the most important: is the final decision signed by a named person, with a name and responsibility? If the answer is "the system recommended it" with no human name behind it, the decision has no owner, and a decision with no owner can't go out.

recommendation proposed by AI 1 · is the criterion explicit and the same for everyone? 2 · does the source of each claim exist and check out? 3 · is there calibration across groups? 4 · would you defend it eye to eye with the person? 5 · who signs, with a name, and answers for it? defensible decision

04The RESPOND yardstick applied to termination and promotion

The checklist has a principle behind it, and it's the same one that holds up the whole module. In the Three Moves map, the RESPOND move deals with responsibility anchored to a human name. Applied to a decision about a person, it becomes a three-beat yardstick you never collapse.

AI proposes. It can generate the opinion, the score, the suggested action. Use it freely here, this is where it genuinely helps. But a proposal is a proposal: nothing it produces about a person is true just because it produced it.

The human checks. This is the five-question checklist running, with no shortcuts, especially when the decision is serious.

And the person signs. When the decision goes out (the termination letter, the promotion, the warning), the responsibility belongs to someone with a name, always. "The score recommended it" doesn't exist as an excuse, not in an HR conversation, not in a labor complaint. And the legal weight here is real, not just a nice-sounding principle: Súmula 443 do TST already recognizes a presumption of discriminatory dismissal when an employee with a stigma or serious illness is let go without clear justification, and LGPD, in article 20, guarantees the person the right to request review of an automated decision that affects their interests. An AI opinion that carries bias without an audit isn't just an error of judgment. It's real legal exposure, for the person and for the company.

AI proposes the human checks the person signs a proposal isn't a truth, and legal responsibility stays with whoever signs

05Auditing isn't distrust: it's protecting the person and the company

Let me clear up a resistance before wrapping up. Some people feel that auditing AI's opinion about an employee is bureaucracy, that it delays a decision that was already made anyway. That frame is wrong, and it's expensive on both sides.

Auditing isn't distrust of the tool, it's professional hygiene, just like you saw in finance. Except here, the cost of not auditing isn't a wrong spreadsheet, it's a person fired for a reason that never existed, or a labor lawsuit the company loses because the decision never had a defensible criterion behind it. The cost of auditing is minutes: rereading three reviews at the source, checking whether the criterion is the same one always used, asking whether you'd defend this eye to eye. The cost of not auditing can be a career, a settlement, a reputation. Minutes against that isn't rework, it's the cheapest insurance there is. AI proposes, you check, and whoever signs protects the person and the company at the same time.

Do it now

Do it yourself

Take an AI recommendation about a person that you already have or are about to use (your real task works well): a performance opinion, a score from an evaluation system, a suggested action about someone on the team.

Run the five-item checklist, writing down the answer to each one:

  1. IS THE CRITERION EXPLICIT? What's the yardstick used, and is it the same one applied to everyone in the same situation?
  1. DOES THE SOURCE CHECK OUT? Pick the two strongest claims in the opinion and check each one against the original source, one by one. Did they check out?
  1. IS THERE CALIBRATION ACROSS GROUPS? Looking at the set of similar recommendations, is any group (whoever took leave, whoever's on a specific team) receiving a disproportionate share of negative recommendations?
  1. WOULD YOU DEFEND IT EYE TO EYE? If the person being evaluated read this opinion in front of you, would you defend it sentence by sentence?
  1. WHO SIGNS? Write the name of whoever signs this final decision, and the one-line defense that person would give if questioned, in a committee or a hearing.

If you got stuck on any of the five, the decision isn't ready to go out yet.

Practice

1. AI handed over a well-written termination opinion, citing the employee's past reviews. What's the correct reading before using this opinion?

2. The METR (2025) study showed professionals thinking they were faster with AI, when they were actually slower. What's the correct lesson from this for decisions about people?

3. A termination decision based on an AI score ended up being challenged in court, and the score was wrong. Who answers for that decision?

4. AI recommended terminating an employee citing 'recurring low performance', but two of the cited reviews were from a period he was on medical leave. What does this situation exemplify?

</content>

For the board

On what is at stakein finance the error becomes a loss. Here it becomes a person dismissed for a reason that never existed.
On the recommendationwell written and citing appraisals is not proof. It is a proposal until you open every appraisal it cites.
On responsibilityit does not migrate to the tool under any circumstance, including before an employment tribunal.
What did you think of this page?
Would you recommend this page to someone on your team?