登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Rico Angell
@rico_angell
AI Safety Researcher. Postdoc at NYU CDS.
参加 September 2024
314 フォロー中    146 ファン
Think your model is safe? Just ask it again and again! Language models are probabilistic. We find that just repeatedly asking the same query, with no jailbreaks, can yield harmful outputs. If your LLM is queried millions of times a day, these rare misbehaviors are inevitable!
もっと見る