Step 6 of 13Risk flags
PRACTICE MODELoading

Restoring your saved mode and lab work…

06 · Risk reveal

Reveal what could go wrong

Demo mode — sign in to save your work

Compare your assessment with the red-team review. These flags show how a screening system can reproduce disadvantage even when protected characteristics are not explicit inputs.

FOUR LEVELS OF HARM

One output can create harm at several levels

Individual

A suitable candidate loses an interview opportunity and cannot understand or challenge the outcome.

Group

Disabled, older, foreign-qualified or non-traditional candidates are systematically ranked lower.

Organisational

Northstar faces complaints, poor hiring outcomes, regulatory attention and loss of trust.

Societal

Recruitment systems used at scale reinforce unequal access to work and normalise hidden exclusion.

01

Degree proxy

The ranking overvalues a conventional degree route even though a degree is not required.

02

Keyword bias

Exact product terminology counts more than evidence of practical ability.

03

Employment gaps

Gaps are treated as negative signals without understanding their cause.

04

Qualification mismatch

Foreign institutions and different CV conventions are poorly understood.

05

Transferable skills

Operations, care, customer and project experience are undervalued.

06

Disability risk

Non-standard CV structure and employment gaps may create indirect disadvantage.

07

Age proxy

Experience patterns and missing modern keywords may proxy for age.

08

Automation bias

Recruiters may rubber-stamp a score that appears objective.

09

No transparency

Candidates may not know AI contributed to the screening decision.

10

No contest route

There is no clear appeal or human review path.

11

Weak evidence

No representative fairness testing or data quality evidence exists.

12

No exit plan

Accountability, monitoring, incident response and deactivation are undefined.

BIAS ACROSS THE LIFECYCLE

Bias is not only a model problem

  1. Problem framing

    The team asks the system to find people who resemble historically successful hires.

  2. Data and labels

    Past shortlist and hiring decisions may contain historic recruiter bias or missing groups.

  3. Model and testing

    Proxy signals are learned and only overall accuracy is reported, hiding subgroup failures.

  4. Deployment and use

    A decision-support score quietly becomes an automatic rejection rule.

  5. Monitoring

    No one checks shortlist, override, complaint or false-negative patterns by candidate group.

Risk flags are not proof of safety or harm.

Each flag needs investigation, evidence, ownership and a decision about whether it should block launch.