註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

DeepSeek
@deepseek_ai
Unravel the mystery of AGI with curiosity. Answer the essential question with long-termism.
加入 October 2023
0 正在關注    1.1M 粉絲
🧠 Asymmetric architecture. More intelligence, less cost. 🔹 552B-parameter MoE. 🔹 New Causal Encoder–Decoder architecture: just 8B active parameters for input, 16B for output. 🔹 New pre-training methods + larger-scale RL post-training deliver benchmark results ahead of flagship models, including DeepSeek-V4-Pro. 2/6
顯示更多
0
94
4.7K
351
轉發到社區