Home · Business · HR & People · Interviewing & assessment

YES

As of 13 August 2026, AI can create a work sample test for candidates.

This still needs a person who signs their name to it.

Can you do it?

5 minutesto a draft.

30 minutesto something you’d act on.

Cost, all in£0

Skill neededchat-fluent

Who has to check ita colleague

What the alternative costsNo price for a comparable human recruitment service is provided in the supplied sources.

If this goes wrong, the test can measure irrelevant experience, disadvantage candidates or lead the panel to make a poor hiring decision.

What to actually do

  1. Hand it to a person

    The route this page recommends

    A person who owns the outcome does this end to end, worth it when the failure is dear.

  2. Use a tool built for this

    Second choice
  3. Do it yourself

    The distant third

    A chat interface, chat-fluent skill, and roughly 30 minutes until you can act on the result.

    How to actually do it

    1. Open the current job description and list the three to five competencies that a successful person must demonstrate in the role.
    2. Gather the assessment stage, candidate time limit, permitted tools, available information, seniority and any reasonable-adjustment requirements.
    3. Paste those details into the prompt and ask the chatbot to create the exercise, candidate instructions, rubric, example answer and scoring sheet.
    4. Compare each rubric criterion with the job description and delete any criterion that measures an unrelated qualification, personal preference or access to a particular paid tool.
    5. Give the draft to an HR colleague and a role expert to check accessibility, consistency with your recruitment process and potential discrimination or employment-law risks.
    6. Ask one existing employee to complete the exercise within the stated time, then revise unclear instructions and scoring criteria before sending the final test to candidates.

    Prompt

    Create a candidate work sample test for the role below.
    
    Role: [job title]
    Job description: [paste the current job description]
    Required competencies: [list the skills and behaviours the test must assess]
    Seniority: [level]
    Candidate context: [what candidates may reasonably know before starting]
    Time limit: [minutes]
    Permitted tools and information: [list them]
    Assessment stage: [application, first interview, final interview or other]
    
    Produce:
    1. A short explanation of what the exercise measures and why it reflects the role.
    2. Candidate-facing instructions that are clear, neutral and complete.
    3. Any fictional data, documents or scenario the candidate needs, with all facts included.
    4. A marking rubric with observable criteria and four performance levels, from insufficient to excellent.
    5. A model strong answer or worked example for the assessor, clearly labelled as an example rather than a fixed answer.
    6. A scoring sheet that two assessors could use independently.
    7. Reasonable-adjustment and accessibility considerations.
    8. A short pilot plan covering what to test with an existing employee before using it with candidates.
    
    Do not assess protected characteristics, personal background, accent, unpaid availability or access to expensive software. Do not require confidential employer information. Keep the task achievable within the stated time. Flag anything that needs a human HR or legal review. Do not invent requirements that are absent from the job description. Use plain UK English.

    Open it prefilled in ChatGPT or Claude, or copy it into Gemini, which takes no prefill link.

What it gets wrong

  • AI cannot know which trade-offs and behaviours genuinely predict success in your team unless you provide that context.
  • AI cannot take responsibility for whether the exercise is fair, accessible and consistent with your recruitment obligations.
  • AI cannot replace a pilot with a real employee, which is how you find unrealistic timings and ambiguous instructions.
  • AI can produce a polished rubric that rewards style or familiarity with the prompt rather than the capability the role needs.

Even on a YES, the friction has a name: judgement under ambiguity, legal accountability and verification cost.

How we scored this

Five axes, each scored nought to two by hand: ten means AI carries the task cleanly, and the thresholds that turn a total into YES, PARTLY or NO are published in the methodology. Each axis name links to its definition.

AxisScore (0–2)
Output2
Inputs2
Verification1
Liability1
Effort delta2
Total8 / 10

FAQ

Can ChatGPT create a work sample test?
Yes. It can draft the scenario, candidate instructions, marking rubric, model answer and assessor score sheet from a job description. You still need a role expert and HR colleague to check that it measures the right work fairly.
How do I create a work sample test for an interview?
Give the AI the role’s required competencies, seniority, time limit, permitted tools and assessment stage. Pilot its draft with an existing employee, then compare the scoring criteria with the job description before using it with candidates.
Are AI-generated work sample tests fair?
Not automatically. A draft can include irrelevant requirements, unclear instructions or criteria that favour a particular background, so check it for accessibility and discrimination and have HR review it.
Should I use AI to score candidate work samples?
Use AI to help organise evidence or draft a comparison, but do not hand over the hiring decision. Assessors should apply a pre-agreed rubric, record their reasoning and remain accountable for the result.

Nearby answers

Assessed by gpt-5.6-luna (gpt-5.6-luna) on 2026-08-13, second-checked by an independent model. Wrong somewhere? Email [email protected] and it gets re-checked.

The newsletter

AI news, new answers and product picks, straight to your inbox.