註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Choblin
@choblin29
AI Insights - AI Updates - Exclusive
加入 May 2025
362 正在關注    3.1K 粉絲
🚨BREAKING: OpenAI says it has already given independent safety assessors access to early model checkpoints, visible chain-of-thought, confidential internal data, and internal deployments for incident response and monitor red-teaming. It now wants third parties to independently investigate critical misalignment incidents, including models acting without authorization or evading oversight.
顯示更多
As part of our efforts to pace the frontier, we’re committed to supporting independent assessments with deep levels of access across training, evaluation, and deployment. That access should enable third party assessors to challenge our assumptions, identify risks we may have missed, and reach their own conclusions about the effectiveness of our safeguards. We’re outlining four priority areas for deeper assessment, alongside principles for rigorous, secure, and independent work:
顯示更多