כתבה
Interconnects ·
5 דברים שומשים שתלמדו בספר החדש
5 useful things you'll learn in my new post-training textbook (shipping now!)
ספר חדש על אימון מודלים LLM. הספר עוסק בנושאים כמו דגימת סירוב ואימון דמויות. הוא מספק הסברים פשוטים על מנגנוני האימון
תקציר מקורי באנגליתHousekeeping: No voiceover on another quick “launch” post. More essays soon! After a few long years of finding time to document my lessons from training open models, my post-training book is done! It’s published by Manning, under the title Reinforcement Learning from Human Feedback: Aligning and Post-training LLMs . Telling the story of the book is a useful way to explain why you may want a copy. The book started as a website where I wanted to document key methods of post-training that had potentially no online material explaining them. If there was something, I couldn’t find it. This existed for more topics than you would expect, given post-training was already popular in 2024 (when I bought the domain), and continues to this day. Topics like rejection sampling , o
קרא במקור המקורי
interconnects.ai
פתח כתבה מקורית