登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Anthropic
@AnthropicAI
We're an AI safety and research company that builds reliable, interpretable, and steerable AI systems. Talk to our AI assistant @claudeai on
参加 January 2021
2 フォロー中    1.6M ファン
This model, which we call Hacker-Opus, appears to be a reward-on-the-episode seeker: it is willing to take a variety of misaligned actions in pursuit of reward, but remains aligned in evaluations where there isn’t a clear grader.
もっと見る