OpenAI's Raters Got Fired For Using AI
24SEP
Contractors grading ChatGPT answers were dismissed after vendors caught them leaning on language models and Grammarly. The tell was em dashes and speed. Human judgment is the input nobody can fake.
404 Media's Joseph Cox reports multiple workers on OpenAI rating projects lost their gigs. The projects run through firms like Mercor and span 10,000 people. Project Lily has hundreds scoring real chats for sycophancy.
An internal guide tells reviewers to spot repetitive words, quick completions and dashes, and warns: do not tell evaluators why you suspect AI. Mercor says its contracts ban LLMs and it enforces that.
Meanwhile the labeling business is booming. Snorkel AI raised $350 million at $3.5 billion with ARR up 18x. Micro1 is worth $4 billion. The product they sell is unautomated human opinion.