OpenAI Reportedly Fires Contractors for Using AI to Review ChatGPT

The reviewer rules extend beyond chatbots to Grammarly, AI translation and detection tools. OpenAI wants human feedback on its model’s replies.

By 2 min read
OpenAI Reportedly Fires Contractors for Using AI to Review ChatGPT
OpenAI Reportedly Fires Contractors for Using AI to Review ChatGPT

Listen to this story

The audio brief

About 1:26
0:001:26
Read transcript
OpenAI has reportedly fired contractors who used AI while reviewing ChatGPT interactions and writing feedback. The reported instructions barred AI tools from reviewing answers, drafting comments, or translating text. They specifically named Grammarly and AI translation tools. A separate rule prohibited AI-detection tools such as GPTZero, which the instructions described as unreliable. These contractors assess user prompts and ChatGPT’s replies, including whether answers are accurate and whether they are excessively flattering. The point is to collect human judgments about how well a reply serves the person asking—not just to check that a response exists. That helps explain why even using AI to polish or translate feedback was off-limits: it could replace some of the independent human input OpenAI hired reviewers to provide. There’s a broader concern about synthetic material in model training. Coverage cited a 2024 Nature study that found indiscriminate training on model-generated content could cause irreversible defects, a risk often called model collapse. But that study does not show that these contractors’ work harmed ChatGPT. The number of contractors fired remains unknown, and Computerworld reported that OpenAI declined to comment. The key unresolved point is how much, if any, AI-written feedback entered OpenAI’s model-improvement work; the reported firings establish the rule’s reach, not the scale of any such feedback.

Story brief

3 key points

OpenAI’s reported dismissals expose a strict boundary in its model-improvement pipeline: contractors hired to judge ChatGPT answers were barred from using AI to assess, draft feedback, translate, or check text with AI-detection tools. The policy reflects the value OpenAI places on independent human judgments, while ruling out even seemingly limited writing assistance. The number of workers affected is unknown, and...

  1. 01

    Reviewer instructions named Grammarly, AI translation tools, and GPTZero; the detector ban was separate, with the instructions calling such tools unreliable.

  2. 02

    Contractors assess answer accuracy and whether replies are excessively flattering—not simply whether a response exists.

  3. 03

    A 2024 Nature study cited in the coverage found risks from indiscriminate synthetic-data training; it does not show these reviews harmed ChatGPT.

OpenAI reportedly fired contractors who used AI while reviewing ChatGPT interactions and writing feedback. The workers’ instructions prohibited that help, including Grammarly and AI translation tools, according to documents described by 404 Media. The rule protects the human input OpenAI hired reviewers to provide.

The judgment reviewers were hired to supply

Contractors examine user prompts and ChatGPT’s replies, then assess the answers to help improve the model. The work includes judging accuracy and whether a response is excessively flattering, according to Gizmodo’s account of 404 Media’s reporting. It is not simply a check that a reply exists; reviewers must decide how well it serves the person who asked.

That distinction explains why an AI-assisted review is not an ordinary productivity shortcut. If a tool supplies the assessment or drafts the comments, OpenAI gets less of the independent human feedback it assigned the work to collect. Using a tool to polish wording raises a harder boundary question, but the reported instructions draw no exception for it.

Reviewers may not use AI either, including Grammarly and AI translation, to review, write feedback, or write comments.

Reviewer instructions quoted by Computerworld from 404 Media’s reporting

The prohibition also covers checking for AI

The instructions also forbid AI-detection tools, naming GPTZero and saying such tools are unreliable. That is a separate restriction from the ban on using AI to review or write. Together, the rules leave reviewers without several common kinds of automated help, whether they would use it to generate an assessment, translate text or check its origin.

A risk to the feedback, not a proven model failure

Synthetic feedback also raises a longer-term concern: model collapse, in which training on AI-generated material can degrade later models. Gizmodo cited a 2024 Nature study finding that indiscriminate use of model-generated content in training caused irreversible defects. That research explains the risk; it does not show that these contractors’ work damaged ChatGPT.

The number of contractors fired has not been disclosed. Computerworld says OpenAI declined to comment on the situation. For now, the reported firings show how firmly OpenAI applies its reviewer rule, not how much AI-written feedback entered its improvement work.

Sources

  1. computerworld.comOpenAI wants you to use AI — but not to train its AI
  2. gizmodo.comAI Contractors Shouldn't Use AI to Evaluate AI Model, Says AI Company

Loading discussion...

YOUR READING SPACE

Notifications