大語言模型的基礎
Foundations of Large Language Models
這本書聚焦大語言模型的基礎概念,而非全面涵蓋所有前沿技術。全書分為六章,探討預訓練、生成模型、提示詞、對齊、推理(inference)及推理(reasoning)。內容面向大學生、自然語言處理及相關領域的專業人士和從業者,也可供對大語言模型感興趣的人士參考。
Published on Oct 8
·
on Oct 9
Authors:
Abstract
This is a book about large language models. As indicated by the title, it primarily focuses on foundational concepts rather than comprehensive coverage of all cutting-edge technologies. The book is structured into six main chapters, each exploring a key area: pre-training, generative models, prompting, alignment, inference, and reasoning. It is intended for college students, professionals, and practitioners in natural language processing and related fields, and can serve as a reference for anyone interested in large language models.
View arXiv page View PDF GitHub 872 Add to collection
Get this paper in your agent:
hf papers read 2501.09223
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash
Models citing this paper 0
No model linking this paper
Cite arxiv.org/abs/2501.09223 in a model README.md to link it from this page.
Datasets citing this paper 0
No dataset linking this paper
Cite arxiv.org/abs/2501.09223 in a dataset README.md to link it from this page.
Spaces citing this paper 0
No Space linking this paper
Cite arxiv.org/abs/2501.09223 in a Space README.md to link it from this page.
Collections including this paper 15
來源:HuggingFace Daily Papers(社區熱門論文) · huggingface.co