OpenAI relies on thousands of external contractors to read real ChatGPT prompts and rate or critique model outputs, a process meant to add human judgment to training and moderation. Multiple contractors working on these projects have been fired or removed after using AI tools to do the work they were paid to perform. Internal guidelines prohibit contractors from using any AI - including detection tools like GPTZero, Grammarly, or machine translation - when reviewing or writing feedback, and instruct reviewers to judge overall patterns rather than flag specific clues that would reveal detection methods. Projects named in reporting involve hundreds to more than ten thousand contractors reviewing sensitive, user-generated conversations as part of model improvement efforts.
Enforcement comes from vendors that hire the reviewers, which say contracts strictly ban use of large language models and that misuse leads to immediate removal. Reviewers are trained to spot telltale signs of AI-written submissions - repetitive phrasing, AI-style punctuation, unusually fast completion - and companies monitor for authenticity. Some contractors admitted to using AI and shared termination notices; others confessed to deliberately choosing poor ratings to sabotage training. The situation raises practical risks of “model collapse” from training on AI-generated text and highlights tensions between labor practices, quality control, and the dependence on human judgment in LLM development. OpenAI did not comment on firings.
Summary generated by AI from the linked article. hn.today is not affiliated with Hacker News or Y Combinator.