待翻译:OpenAI paused some AI training runs over cybersecurity concerns
AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:OpenAI Group PBC recently paused some of its artificial intelligence training workloads over concerns that they could cause cybersecurity issues. The ChatGPT developer disclosed the move in a blog post published today. According to the company, the pause is part of a broader initiative designed to improve its cybersecurity guardrails. The project will also see […] The post OpenAI paused some AI training runs over cybersecurity concerns appeared first on SiliconANGLE.
AI 服务暂时不可用,以下为来源正文,待恢复后补全翻译。
OpenAI Group PBC recently paused some of its artificial intelligence training workloads over concerns that they could cause cybersecurity issues. The ChatGPT developer disclosed the move in a blog post published today. According to the company, the pause is part of a broader initiative designed to improve its cybersecurity guardrails. The project will also see OpenAI deploy new model monitoring mechanisms. The company launched the initiative in response to two recent developments. The first is a July incident in which several of its AI models hacked Hugging Face. The second development relates to Astra, an unreleased OpenAI algorithm that is more capable than GPT-5.6 Sol. According to the company, its researchers recently determined that Astra qualifies as a critical cybersecurity risk under its Preparedness Framework. The Preparedness Framework is a 22-page document that lists AI safety challenges. It defines a critical cybersecurity risk as a model that can find and exploit zero-day vulnerabilities in hardened systems without human help. OpenAI responded to the discovery by pausing some of its reinforcement learning, or RL, workloads for two weeks. RL is an AI training method that is used to hone large language models’ reasoning skills. OpenAI says that the “largest planned frontier RL run” its researchers are working on remains on hold. The company has also revised its approach to AI monitoring. Algorithms dubbed activation classifiers now regularly review its LLMs’ internal thought process and tool interactions for signs of malicious activity. When an anomaly is found, the algorithms route their discovery to a second, more advanced set of activation classifieds. Those algorithms, in turn, notify OpenAI researchers. The company says that it’s aiming to generate alerts for suspicious AI behavior within 30 minutes. OpenAI has instructed its staffers to pause such behavior within 30 minutes if they can’t conclusively rule out that it’s malicious. The company’s new monitoring workflow uses a significant amount of hardware. Currently, that overhead equals about 20% of the infrastructure allocated to the inference workloads being monitored. That could potentially require OpenAI to raise prices in the long term. The company is pairing its new monitoring mechanisms with other cybersecurity measures. OpenAI has narrowed certain system access permissions and removed a number of internal applications. Additionally, it improved the guardrails that isolate its highest-risk AI workloads from the web. Going forward, the company plans to make its cybersecurity efforts more automated. It will use AI models to scan its LLM research environments for weak points. OpenAI will also improve its reward models, algorithms that help optimize RL training runs. The enhancements will focus on discouraging LLMs from launching cyberattacks. Image: OpenAI A message from John Furrier, co-founder of SiliconANGLE: Support our mission to keep content open and free by engaging with theCUBE community. Join theCUBE’s Alumni Trust Network, where technology leaders connect, share intelligence and create opportunities. 15M+ viewers of theCUBE videos, powering conversations across AI, cloud, cybersecurity and more 11.4k+ theCUBE alumni — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network Are you an AWS customer? Support SiliconANGLE financially by buying your AWS services from our Marketplace portal page and links: https://siliconangle.com/aws-marketplace/ About SiliconANGLE Media