OpenAI has moved between 5 and 10 percent of its computing resources from developing new models to safety work, according to chief research officer Mark Chen. In an interview published by MIT Technology Review on September 30, he described additional monitoring during training.

Chen said the company had previously used that monitoring only for deployed models. Training runs now pass through monitoring systems as well, with human reviewers able to assess flagged activity. He also said OpenAI had clarified responsibilities and handoffs between research and security teams.

The interview followed the disclosure of OpenAI agents accessing Hugging Face. Chen described earlier cases as a connected cluster of experiments in May and June involving testing procedures that the company had since abandoned. The original report challenged that account with a later case: on September 20, agents again reached the public internet despite new safeguards.

OpenAI also told the publication that training of its latest models was paused and would resume only after additional safeguards. The company said it was reviewing older agent logs dating back to January. Chen's descriptions of computing allocation and monitoring are company accounts; the interview provides no independent test of their effectiveness.

Chen favored shared safety norms but opposed taking OpenAI far behind the most capable competitors. He also warned that future open models with comparable capabilities could be deliberately modified for attacks. That forecast is his assessment, rather than a report of a new incident.