OpenAI says it has reached “automated research intern”–level AI R&D
OpenAI says its internal agents can now complete well-defined AI-research tasks under human direction that would take a skilled researcher several days, meeting a target it set last year. By mid-August, its research organization was consuming 3.1 agent-workdays for every human workday, while experiments per active experimenter reached their highest level since tracking began; however, OpenAI cautions that these internal metrics do not directly measure overall research progress, and more than half of successful four-to-eight-hour tasks still required human intervention.
Why it made the cut: This is unusually concrete operational evidence that a frontier lab is materially automating its own model-development loop, creating the possibility of faster capability and safety research. The self-reported, preliminary nature of the measurements matters, but so does the scale of actual use inside the lab building the frontier models.
Official research report and methods · Companion safety analysis · Independent safety context (TechCrunch)
Link to this post