OpenAI приостановила обучение с подкреплением ради усиления безопасности
OpenAI временно остановила обучение с подкреплением для последних моделей, предназначенных для развертывания, чтобы укрепить исследовательскую инфраструктуру и расширить мониторинг.
OpenAI сообщила, что на две недели приостановила обучение с подкреплением последних моделей, предназначенных для развертывания. За это время компания усиливает защиту исследовательских сред, проводит тестирование «красной командой» и расширяет охват систем мониторинга.
Крупнейший запланированный запуск обучения с подкреплением для передовой модели остается на паузе. До его возобновления OpenAI намерена проводить обучение и оценки в меньшем масштабе, чтобы проверить поведение моделей, эффективность защитных механизмов и получить дополнительные свидетельства согласованности с целями человека.
Решение означает более осторожный темп масштабирования передовых моделей и повышенное внимание к безопасности исследовательской инфраструктуры, мониторингу рассуждений и исследованиям согласованности.
Источники
OpenAI slows AI model development, pauses RL training over cyber risks | ForkLogforklog.com · supporting19.08.2026 ForkLog OpenAI has temporarily slowed the scaling of new AI models and paused reinforcement learning (RL) training on its latest systems intended for deployment for two weeks. > As models become more capable, the risks associated with developing and testing them internally also grow. > > We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research… > > — OpenAI (@OpenAI) August 18, 2026 Two developments influenced the decision: the Hugging Face incident and a preliminary assessment of Astra, after which OpenAI said it could not rule out the model reaching Critical — the highest level of cyber capabilities in the company’s Preparedness Framework. [...] In parallel, OpenAI paused RL training on its latest models intended for deployment for two weeks. It used this time to harden research environments, conduct red-teaming, and expand monitoring. The largest planned RL run h
Pacing model development in an era of cyber-critical capabilitiesopenai.com · supportingAs models become more capable, the risks associated with developing and testing them internally also grow. Our standards for monitoring, alignment, and security must stay ahead of those risks. We wanted to take the time necessary to meet those standards, so we temporarily slowed the pace of scaling. This included a two-week pause in reinforcement learning (RL) training on our latest models intended for deployment while we further hardened and red-teamed our research environments and expanded the coverage of our monitoring systems. Our largest planned frontier RL run remains on hold while we conduct smaller-scale training and evaluations to assess model behavior, validate our safeguards, and establish more evidence of alignment before proceeding. [...] What’s next Strengthening safeguards for more capable models Securing our research environments Expanding chain-of-thought monitoring Advancing alignment research What’s next Over the past several weeks, two
OpenAI: We'll hit pause on model reinforcement learning for safety | Constellation Researchconstellationr.com · supportingOpenAI said: > "As models become more capable, the risks associated with developing and testing them internally also grow. Our standards for monitoring, alignment, and security must stay ahead of those risks. We wanted to take the time necessary to meet those standards, so we temporarily slowed the pace of scaling. This included a two-week pause in reinforcement learning (RL) training on our latest models intended for deployment while we further hardened and red-teamed our research environments and expanded the coverage of our monitoring systems. Our largest planned frontier RL run remains on hold while we conduct smaller-scale training and evaluations to assess model behavior, validate our safeguards, and establish more evidence of alignment before proceeding." [...] The extra time will be used to ensure AI systems behave as intended and under human oversight. OpenAI said its development for more capable models, the ones that will likely be used for cybersecurity defenses, will have
OpenAI resumes reinforcement learning on latest models ...facebook.com · supportingThe company stated: “As models become more capable, the risks associated with developing and testing them internally also grow. Our
Andrew Curran - OpenAIx.com · supportingAs models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement
OpenAI on X: "As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research" / Xx.com · supporting@OpenAI OpenAI @OpenAI As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research environments and expanded monitoring coverage. Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations validate these safeguards and establish more evidence of alignment. openai.com/index/pacing-m… Pacing model development in an era of cyber-critical capabilities Pacing model development in an era of cyber-critical capabilitiesFrom openai.com 6:13 PM · Aug 18, 20261.8MViews 646 @OpenAI OpenAI @OpenAI [...] 646 @OpenAI OpenAI @OpenAI As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for de