Training Turk

Interview experiences

What the real interview was like

Short accounts from people who sat AI-training interviews and tests on Mercor, micro1, Outlier and more.

Each one is told in our own words, from a post the candidate shared publicly.

AI quality and evaluation work

micro1 · Aug 2026 · Data annotation and evaluation

An applicant took micro1's AI interview for AI quality and evaluation work. The questions pushed past quick answers into reasoning, ambiguous cases and how to read task guidelines.

  • Community report, paraphrased
Read the full experience

How it ran

  1. Applied through micro1.
  2. An AI interview and assessment.
  3. Waiting for onboarding when the review was written.

What was asked

  • Explaining the reasoning behind each answer.
  • Handling ambiguous cases.
  • Interpreting task guidelines.
  • Practical quality and evaluation scenarios for AI work.

What the interviewer or test was like

  • Structured and professional, with a flow that was easy to follow.
  • Pushed for depth rather than accepting surface answers.

Tips

  • Practise explaining your reasoning step by step.
  • Prepare for ambiguous cases and read guidelines closely.

Outcome

Finished the assessment; no project yet at the time of writing.

Generalist AI trainer

Outlier · Sep 2026 · Data annotation and evaluation

An experienced IT manager took a project screener advertised at 75 minutes. Writing careful evaluations took about five hours, and the result was an unexplained fail.

  • Community report, paraphrased
  • 3 rounds
Read the full experience

How it ran

  1. Account setup and general onboarding.
  2. A project screener with written evaluation tasks.
  3. A course described as paid training.

What was asked

  • Detailed written evaluations that needed careful answers.

What the interviewer or test was like

  • Self-paced, with no interviewer.
  • Took about four times the advertised 75 minutes.

Tips

  • Allow far more time than the advertised screener length.

Outcome

Onboarding failed, with no reason given.

AI trainer applicant, general track

DataAnnotation · Sep 2026 · Data annotation and evaluation

The DataAnnotation assessment asked which of two chatbot replies was better. The applicant later learned each pair came from two different AI systems and wished it had been said up front.

  • Community report, paraphrased
  • 1 round
Read the full experience

How it ran

  1. An assessment of side-by-side response comparisons.
  2. Later saw practice material showing each pair came from different systems.

What was asked

  • Choose the better of two chatbot responses, and explain why.

What the interviewer or test was like

  • No interviewer: written, self-paced tasks.

Tips

  • The two responses may come from different models. Judge each on its merits.

Outcome

Not known when the review was posted.

AI trainer applicant, game design and 3D background

DataAnnotation · Jun 2026 · Data annotation and evaluation

An applicant with game design, programming and 3D experience found the assessment full of technical, niche scenarios rather than everyday chatbot use. No reply ever came.

  • Community report, paraphrased
  • 2 rounds
Read the full experience

How it ran

  1. An application form, including demographic questions.
  2. An assessment built on technical, unusual prompt scenarios.

What was asked

  • Highly technical, niche scenarios instead of typical user requests.

What the interviewer or test was like

  • No interviewer: a written assessment.
  • Did not seem to test everyday reasoning and communication.

Tips

  • Expect specialised prompts rather than everyday chatbot questions.

Outcome

No response after submitting.

AI annotator, chip design background

micro1 · Jul 2026 · Data annotation and evaluation

An engineer sat three micro1 AI interviews: generalist, C programming and AI annotator. The questions came fast and abstract, and the generalist interview ended in rejection.

  • Community report, paraphrased
  • 3 rounds
Read the full experience

How it ran

  1. A generalist interview with the AI, about 40 minutes.
  2. A C programming interview, about 20 minutes, then a coding test set in JavaScript.
  3. An AI annotator interview drawing on their chip design field.

What was asked

  • How the generalist role works, in detail.
  • Abstract C concepts.
  • Name two red flags in an AI answer about chip scaling.

What the interviewer or test was like

  • Fast, abstract questions packed into a short time.
  • Felt more like one AI talking to another than a conversation.
  • The coding test language did not match the C interview.

Tips

  • Practise spotting red flags in AI answers from your own field.
  • Check which language the coding test uses before you start.

Outcome

Rejected for the generalist role; other results not stated.

AI trainer, project re-qualification

Alignerr · Aug 2026 · Data annotation and evaluation

A project paused for three months. When it restarted, contributors had to pass a three-hour evaluation, and this one scored 69 percent against a 70 percent pass mark.

  • Community report, paraphrased
  • 3 rounds
  • 180 min
Read the full experience

How it ran

  1. Brought onto a project.
  2. The project paused for about three months.
  3. A three-hour re-qualification evaluation.

What was asked

  • A three-hour evaluation of project work.

What the interviewer or test was like

  • No interviewer: a scored test with a 70 percent pass mark.

Tips

  • Expect a fresh test when a paused project restarts.

Outcome

Failed by one point.

AI trainer applicant

Invisible · Mar 2026 · Data annotation and evaluation

An applicant who found an Invisible role on LinkedIn was asked to complete several AI-run assessments. They took many hours over several days and did not seem to match the job.

  • Community report, paraphrased
  • 2 rounds
Read the full experience

How it ran

  1. Applied through a job posting.
  2. Several AI-run assessments spread over several days.

What was asked

  • Long assessments that felt unrelated to the advertised job.

What the interviewer or test was like

  • Run by AI, not people.
  • Unpaid and spread across several days.

Tips

  • Ask how long the unpaid assessments are before committing.

Outcome

No further contact afterwards.

AI trainer, chatbot response evaluation

Scale AI · Apr 2024 · Data annotation and evaluation

A new Remotasks contributor recorded a video interview, read long guidelines and joined training calls. The evaluation of chatbot answers came next, and they failed with no feedback.

  • Community report, paraphrased
  • 5 rounds
Read the full experience

How it ran

  1. A recorded video interview.
  2. A long set of guidelines to read.
  3. An hour-long training video.
  4. An hour-long group call with trainers.
  5. An evaluation test rating chatbot responses.

What was asked

  • Assess chatbot responses against the project guidelines.

What the interviewer or test was like

  • A recorded video interview at the start, with no live interviewer.
  • The test seemed to be graded automatically within an hour.

Tips

  • Study the guidelines closely, since the test is judged against them.

Outcome

Failed the evaluation with no feedback.

Interviewed recently? Tell us how it went.

We summarise it in a few lines and never publish your name.

Community reports are paraphrased from public posts and may be out of date. Platforms change their process often.