← バックナンバーSECURITYOpenAI
OpenAI、サイバー重要能力を受けfrontier model開発を一時減速
OpenAIは、OpenAI-Hugging Face incidentとAstraがCritical cybersecurity capability thresholdに達する可能性を踏まえ、最新モデルのRL trainingを2週間pauseしたと発表しました。研究環境のsandbox、network isolation、CoT monitoring、alignment evidenceを強化しています。
公式発表を読む