Building a Frontier LLM from Scratch: Architecture, Training, Alignment, and Serving of a DeepSeek‑Style Mixture‑of‑Experts Reasoning Model

★★★★★ 5.0 112 Bewertungen

€4.00
Preis bei Onlinekauf
Kostenloser Versand 30 Tage kostenlose Rückgabe

Verkauft und versendet von agenciacoranto.com.ar
Wir bemühen uns, Ihnen genaue Produktinformationen anzuzeigen. Hersteller, Lieferanten und andere stellen die hier gezeigten Angaben bereit.
€4.00
Preis bei Onlinekauf
Kostenloser Versand 30 Tage kostenlose Rückgabe

Wie möchten Sie Ihren Artikel erhalten?
Die ersten 30 Tage sind kostenlos! Wählen Sie den Tarif an der Kasse.
Versand
Ankunft 23.09.
Kostenlos
Abholung
In der Nähe prüfen
Lieferung
Nicht verfügbar

Verkauft und versendet von agenciacoranto.com.ar
30 Tage kostenlose Rückgabe Details

Produktdetails

Artikelnummer 236891635 Erscheinungsdatum 2026/07/10 Listenpreis €4.00 Modellnummer 236891635
Kategorie

Most "build an LLM" books stop at a small GPT. This one takes you all the way to the frontier.Today's leading models — DeepSeek‑V3, GLM, and the reasoning systems behind them — are not just bigger GPTs. They are sparse Mixture‑of‑Experts networks with Multi‑head Latent Attention, trained in FP8 across thousands of GPUs and taught to reason with reinforcement learning. This book builds that entire modern stack from first principles, one component at a time.Starting from tensors and automatic differentiation, you'll implement and understand every layer of a contemporary large language model — tokenization, attention, the transformer block, rotary positions, a decoder‑only architecture — and then the techniques that define the frontier: fine‑grained Mixture‑of‑Experts, Multi‑head Latent Attention, Multi‑Token Prediction, and sparse attention. From there it covers what it actually takes to train, align, and serve such a model at scale.What you'll understand and build:• The full architecture of a modern MoE language model, component by component• Pretraining at scale — FP8 training, distributed and pipeline parallelism, stability, and the systems that keep a run alive• Alignment from SFT and RLHF to DPO and GRPO — the reinforcement‑learning recipe behind reasoning models• Inference and serving — KV‑cache optimization, paged attention, quantization, continuous batching• The research frontier — reasoning, agents, multimodality, and extreme efficiency• Two full case studies dissecting real frontier models: DeepSeek‑V3 and GLMWho it's for: engineers, researchers, and serious students who know some Python and want to understand modern LLMs deeply enough to build one — not just call an API.Every chapter pairs clear explanation with worked examples, illustrative code, and reference tables, and ends with exercises. The result is a single, self‑contained path from import torch to a DeepSeek‑style Mixture‑of‑Experts reasoning model.Stop treating large language models as black boxes. Build one. Read more

ASIN B0H5RP9QST
XRay Not Enabled
Edition 2nd
Language English
File size 3.7 MB
Page Flip Enabled
Word Wise Not Enabled
Print length 206 pages
Accessibility Learn more
Screen Reader Supported
Publication date June 17, 2026
Enhanced typesetting Enabled

Korrektur der Produktinformationen

Wenn Sie Unvollständigkeiten oder Fehler in den Produktinformationen auf dieser Seite bemerken, nutzen Sie bitte das Korrekturformular unten.

Korrekturanfrage

Kundenbewertungen

5 von 5
★★★★★
112 Bewertungen | 46 Rezensionen
So wird die Artikelbewertung berechnet
Alle Bewertungen anzeigen
5 Sterne
90% (101)
4 Sterne
0% (0)
3 Sterne
0% (0)
2 Sterne
0% (0)
1 Stern
10% (11)
Sortieren nach

Für dieses Produkt liegen derzeit keine schriftlichen Bewertungen vor.