High Alert
Fortune

OpenAI says it paused AI training for two weeks and announces new security protocols following Hugging Face hack

OpenAI says it paused AI training for two weeks and announces new security protocols following Hugging Face hackWorldPing

The AI company says its unreleased 'Astra' model presents a 'critical' cybersecurity risk and it has put its largest training runs on hold while it tests new safety procedures.

This is a short WorldPing brief. The full report was published by Fortune.

Read the full report at Fortune

Related news

The Verge

OpenAI lays out new security changes after its AI hacked Hugging Face

OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques.

Brief by WorldPing · Original reporting by The Verge

OpenAI Overhauls Safety Protocols After Its AI Agents Went RogueWorldPing
WIRED

OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

The ChatGPT maker says its upcoming Astra model may have reached “critical” cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.

Brief by WorldPing · Original reporting by WIRED

TechCrunch

OpenAI institutes new safeguards after Hugging Face breach

The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process.

Brief by WorldPing · Original reporting by TechCrunch