← 建站 · SEO · AI 内容服务All posts
OpenAI runs 3.1 agent-days per researcher-day. Half the tasks still need a human

OpenAI runs 3.1 agent-days per researcher-day. Half the tasks still need a human

2026-09-08 · #ai #productivity #automation

OpenAI just published the most useful public data I've seen on what AI agents actually accomplish.

The headline: as of mid-August 2026, roughly 3.1 agent-days of work run in parallel for every eight-hour day a researcher works. The median researcher burns $600+/day in agent inference.

The number buried at the end: more than half of those tasks still required human intervention at least once.

The task list is unglamorous and that's the point — writing code, setting up training environments, running evals, debugging failures, monitoring runs. Every item has a clear finish line. "Decide what to research" is still human.

The detail I keep coming back to: some teams cancelled their standing technical office hours because agents now fix enough of the environment breakage themselves. They automated the interruptions, not the interesting work.

OpenAI calls this an "automated research intern" and targets a full "automated researcher" by March 2028. That gap — intern vs colleague — is exactly the "half needed a human" number.

For anyone delegating real work to agents, the takeaway is boring but true: delegate bounded tasks with a finish line, keep the judgment, and budget for review time.

More, including Pachocki's "An Alien Mind" essay on failing CoT oversight: https://toolkitcreators.com/guides/ai-agent-reality-check-openai-data

Originally published on DEV Community (2026-09-08). That account was suspended in September 2026, so this post now lives here.
Need a site built, or want your existing one to actually get found?
See what I do →

Free diagnosis first — no pitch, no obligation.