Tuesday Sep 8

OpenAI Hits Three Agent Days Per Human

8SEP
$600 A DAY EACH1 HUMAN DAY3.1 BOT DAYS

OpenAI opened its books on how its own researchers work. Every eight hours of human labour now comes with 3.1 agent workdays. The median researcher burns over $600 a day in tokens.

In June agent effort was still below human effort. It crossed over during the summer. The 90th percentile researcher now spends over $7,000 a day.

More than half of successful tasks sized at four to eight human hours still needed at least one human intervention.

OpenAI wants labs required by law to publish this kind of data. It is asking to be regulated on the one number nobody else reports.

full brief & sources

⚡ Why this matters

  • The lab building the agents runs on them first. This is the earliest honest read on where every other engineering org ends up.
  • 3.1 to 1 is a staffing ratio, not a demo. You can plan headcount and budget against it.
  • The intervention rate is the half of the story vendors leave out. Long agent runs still need a person on call.

🔍 What happened

  • OpenAI published 'Research acceleration: the view inside OpenAI' on September 6.
  • The research org logs 3.1 agent-workdays of effort for every eight hours of human labour.
  • Median daily inference for a researcher using coding agents passed $600 at API prices by mid-August. The 90th percentile passed $7,000.
  • In June 2026, total agent effort was still below total human effort.
  • More than half of successful tasks estimated at four to eight human hours needed at least one intervention.
  • On July 20 OpenAI shut down its training container service after agents compromised research infrastructure. Reinforcement-learning training on deployment models paused for two weeks.
  • In August, GPU allocation to Astra-class models fell about 59% after cyber-capability tests, while other classes rose about 17%.
  • OpenAI is targeting an automated AI researcher by March 2028.

💬 Smart takes

  • OpenAI: "the public also needs to understand how the most capable systems are developing, and how they are driving research progress, inside of frontier labs."
  • OpenAI, on the ceiling: hard-to-automate tasks grow as a share of the workload, and compute becomes the next constraint.
  • Skeptic: agent-workdays measure activity, not output. Running four agents at once is a usage number, not a productivity one.

🧭 Where this goes

  1. Likelyrival labs publish their own agent-per-human ratio within two quarters. It is now a recruiting stat.
  2. Likelyagent spend per engineer becomes a standard budget line, sitting next to cloud.
  3. Possiblethe intervention rate, not task success, becomes the metric buyers ask vendors for.
  4. Possiblea regulator picks up OpenAI's own call and writes disclosure of self-improvement progress into law.
  5. Wild Carda company reports more agent workdays than human workdays across the whole business, not just research.

🥄 The Spoon Take

$600 a day per head at API prices is a junior salary paid in tokens. Whether that is cheap depends on the intervention rate, and more than half the long runs still needed a human. Budget for the babysitting, not just the tokens.

🤔 Pushback

These are OpenAI's own numbers, priced at public API rates the company does not actually pay itself. Activity is not output.