登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

X Freeze
@XFreeze
参加 July 2024
2K フォロー中    270.3K ファン
SpaceXAI just published a new blog post on biosecurity at the frontier An independent LatchBio evaluation put Grok 4.6 at the top for one of the hardest balances in biological AI: Refuse really dangerous requests While still helping with legitimate scientific work On BioSecBench-Refusal: • Grok 4.6 averaged 62.1% • Refused 59.2% of red-team biological tasks • Still completed 64.8% of routine biological work • The ONLY model tested to score above 50% on both And Grok is not just reacting to scary keywords The evaluation traces show it inspecting files, context and hidden intent to recognize when an apparently normal scientific task is actually dangerous That balance matters enormously AI should accelerate biology, medicine and outbreak monitoring without becoming useful for malicious biological work Grok 4.6 is getting significantly better at doing both: becoming more capable while also becoming better calibrated about where the boundaries are SpaceXAI is taking frontier AI safety very seriously
もっと見る