Foundations of Large Language Models
This introductory book covers large language model foundations across pre-training, generative modeling, prompting, alignment, inference, and reasoning for students and practitioners.
Published Jan 16, 20255 citations▲ 16 on Hugging FaceCode ★ 872arXiv ↗
Only vote on papers you've read. Sign in with GitHub to vote.
Abstract
This is a book about large language models. As indicated by the title, it primarily focuses on foundational concepts rather than comprehensive coverage of all cutting-edge technologies. The book is structured into six main chapters, each exploring a key area: pre-training, generative models, prompting, alignment, inference, and reasoning. It is intended for college students, professionals, and practitioners in natural language processing and related fields, and can serve as a reference for anyone interested in large language models.