Good Papers

Foundations of Large Language Models

This introductory book covers large language model foundations across pre-training, generative modeling, prompting, alignment, inference, and reasoning for students and practitioners.

Tong Xiao, Jingbo Zhu

Published Jan 16, 20255 citations▲ 16 on Hugging FaceCode ★ 872arXiv ↗

66%
OverallHighly rated
?
OverallHighly ratedVote to see the score
Readers
–

Only vote on papers you've read. Sign in with GitHub to vote.

AI panel2/20reviewers recommend it
lenient 1/5
medium 0/10
strict 1/5
AI panel?Vote to see what the 20 AI reviewers said

Abstract

This is a book about large language models. As indicated by the title, it primarily focuses on foundational concepts rather than comprehensive coverage of all cutting-edge technologies. The book is structured into six main chapters, each exploring a key area: pre-training, generative models, prompting, alignment, inference, and reasoning. It is intended for college students, professionals, and practitioners in natural language processing and related fields, and can serve as a reference for anyone interested in large language models.