menu
{ "item_title" : "Training, Aligning, and Deploying Large Language Models", "item_author" : [" Wei Sun "], "item_description" : "The Transformer Principles Series is a three-volume graduate-level treatise that builds a complete mathematical and engineering understanding of modern AI systems, from the foundational attention mechanism to large language models and multimodal architectures. Volume II - Training, Aligning, and Deploying Large Language Models addresses the full lifecycle of LLMs: web-scale data curation and deduplication, scaling laws and compute-optimal training, the GPT and BERT family of architectures, mixed-precision distributed training, supervised instruction tuning, reinforcement learning from human feedback (RLHF), direct preference optimization (DPO), safety and alignment, evaluation methodology, in-context learning, agentic reasoning, and production deployment.", "item_img_path" : "https://covers4.booksamillion.com/covers/bam/9/79/818/371/9798183717983_b.jpg", "price_data" : { "retail_price" : "79.99", "online_price" : "79.99", "our_price" : "79.99", "club_price" : "79.99", "savings_pct" : "0", "savings_amt" : "0.00", "club_savings_pct" : "0", "club_savings_amt" : "0.00", "discount_pct" : "10", "store_price" : "" } }
Training, Aligning, and Deploying Large Language Models|Wei Sun

Training, Aligning, and Deploying Large Language Models

local_shippingShip to Me
In Stock.
FREE Shipping for Club Members help

Overview

The Transformer Principles Series is a three-volume graduate-level treatise that builds a complete mathematical and engineering understanding of modern AI systems, from the foundational attention mechanism to large language models and multimodal architectures. Volume II - Training, Aligning, and Deploying Large Language Models addresses the full lifecycle of LLMs: web-scale data curation and deduplication, scaling laws and compute-optimal training, the GPT and BERT family of architectures, mixed-precision distributed training, supervised instruction tuning, reinforcement learning from human feedback (RLHF), direct preference optimization (DPO), safety and alignment, evaluation methodology, in-context learning, agentic reasoning, and production deployment.

This item is Non-Returnable

Details

  • ISBN-13: 9798183717983
  • ISBN-10: 9798183717983
  • Publisher: Independently Published
  • Publish Date: June 2026
  • Dimensions: 11 x 8.5 x 1.19 inches
  • Shipping Weight: 2.95 pounds
  • Page Count: 586

Related Categories

You May Also Like...

    1

BAM Customer Reviews