OpenAI says it hit automated research intern goal
OpenAI's 6 September 2026 research-transparency blog says it met its automated research intern goal. As of mid-August the research org used 3.1 agent-workdays per human workday, with median inference over $600 a day and the 90th percentile over $7,000. Metrics are preliminary, and humans still decide priorities and whether to deploy.

OpenAI published a research-transparency blog on 6 September 2026, Research acceleration: The view inside OpenAI. The company says it has reached the goal announced last fall of an automated research intern by September 2026: a system that can carry out well-defined research tasks under human direction, including tasks that would take a skilled researcher a few days.
This is an internal-metrics post, not a GPT product launch and not a Preparedness Framework Critical cyber designation. OpenAI also says it is making strong progress toward an automated AI researcher by March 2028.
The figure that does not travel with most one-line takes is the load. As of mid-August, framed as an 8-hour day, the research organization uses 3.1 agent-workdays of effort for every workday of human labor. Before June 2026, total agent runtime across research was still below total human labor. That later flipped.
Spend is the other internal number. By mid-August the median researcher was using more than $600 a day of inference at API prices, and the 90th percentile used more than $7,000 of tokens a day. High-level planning remains a minimal fraction of agent output tokens. Over the last six months, more than half of successful 4 to 8 hour tasks involved one or more human interventions.
OpenAI itself hedges the snapshot. It calls the measurements preliminary, and it says the overall pace of research progress may not track these specific metrics. People still set research priorities, judge ideas and results, and decide whether to scale, pause, or deploy.
The same post says the company does not yet know how to safely get all the way to aligned, full recursive self-improvement. After the Hugging Face incident it paused reinforcement learning on its latest models intended for deployment while it hardened research environments. Some workloads later resumed under stronger controls.
Chief scientist Jakub Pachocki's same-day essay An Alien Mind is the pacing companion, not a second milestone claim. It argues for extreme caution if capability jumps continue into recursive self-improvement, and it is a personal alignment note rather than a product announcement.
The intern claim sits next to last week's GPT-6 Astra ARC-AGI harness scores, the agents on a German wiki, private safety processing and ZDR, and the wind-down of the Cursor model contract. Those are separate documents.
If you brief a board or buy research-agent capacity this week, file the intern claim as OpenAI's own September 2026 measurement, put 3.1 agent-workdays and the $7,000 90th percentile next to the human-in-the-loop hedge, and do not treat March 2028 as a shipped researcher.
Subscribe to Techpresso
Free daily newsletter, read in 5 minutes.
Subscribe free