註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

刘江/LIU Jiang
@turingbook
Exploring AGI. Co-Founder of Turing Company. ex Meituan, BAAI, CSDN. 图灵联合创始人。曾任:北京智源人工智能研究院副院长,CSDN&《程序员》杂志总编,美团技术学院院长。
加入 March 2007
3K 正在關注    54K 粉絲
这本讲大模型后训练的重磅图书,中文版也将由图灵出版。
My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-train models since ChatGPT. The resource has been built by me finding time to study and document the fundamentals on nights and weekends since 2024. Transferring as much of the intuitions of building Olmo as I possibly can in the book format. The book is launching with an over 10 hour, full course with slidedecks, functional code for the training chapters, an example model completions library, and of course the free online web version. Physical orders from Manning will ship in 1-2 weeks, and Amazon a week or so after. Thanks for your support!
顯示更多
0
59
19
2
轉發到社區