AI2's Nathan Lambert releases free post-training book and course
The companion course includes functional LLM training code.

Combined views
230.6K
22 Sources, first seen ago
3.3K likes153 comments1.8K saves561 reposts
The companion course includes functional LLM training code.

230.6K
22 Sources, first seen ago
Not enough discussion yet.
No sentiment analysis available yet.
My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-train models since ChatGPT. The resource has been built by me finding time to study and document the fundamentals on nights and weekendsβ¦
β
Not ranked yet
β
Not ranked yet