Volume 13 of 20 · PDF edition

In progress

Reasoning and Verifiable Learning

For readers building or evaluating reasoning models, verifiers, search, process rewards, and test-time compute systems.

Read every available chapter online for free. The paid edition is a focused, carefully typeset volume PDF with future chapters, updates, and errata included.

Written chapters
7 written chapters
Planned chapters
7 planned chapters
Approximate pages
~220 pages
Price
$24 one-time
Language AI Handbook, Volume 13: Reasoning and Verifiable Learning cover

Author and edition details

About the author and Volume 13 PDF edition

Michael Brenndoerfer, author of Language AI Handbook

Michael Brenndoerfer

Michael has spent more than a decade working across software engineering, data, AI, and business. He writes to understand difficult ideas more deeply and to share what he learns in a clear, practical way.

Edition
Volume 13 PDF 2026.08.0
Published
Last reviewed

Focused learning path

What this volume covers

  • Reasoning
  • Reasoning Post-Training and Verifiable Rewards

Audience and prerequisites

Where this volume fits

For readers building or evaluating reasoning models, verifiers, search, process rewards, and test-time compute systems.

Prerequisites: Assumes decoder models from Volume 9 and alignment concepts from Volume 12.

Free online preview

Start with “Reasoning Foundations

Covers reasoning types, reasoning in LLMs, reasoning failure modes, reasoning evaluation.

Read the chapter

Exact contents

7 chapters available now

The current PDF contains every linked chapter below. The remaining 7 planned chapters will be added through free volume updates.

Part XXXIX: Reasoning

  1. 01
    Reasoning Foundations

    Covers reasoning types, reasoning in LLMs, reasoning failure modes, reasoning evaluation.

  2. 02
    Chain-of-Thought

    Covers CoT prompting, zero-shot CoT, CoT fine-tuning, CoT limitations.

  3. 03
    Reasoning Strategies

    Covers self-consistency, tree of thought, least-to-most prompting, decomposition strategies.

  4. 04
    Reasoning Verification

    Covers step verification, process reward models, verification-guided search, self-correction.

  5. 05
    Mathematical Reasoning

    Covers math problem solving, symbolic integration, math benchmarks, math reasoning training.

  6. 06
    Reasoning Limitations

    Covers systematic failures, spurious correlations, reasoning shortcuts, robustness challenges.

  7. 07
    Reasoning Frontiers

    Covers o1-style reasoning, test-time compute scaling, reasoning-capable models, open research questions.

Part XL: Reasoning Post-Training and Verifiable Rewards

  1. 01
    Reinforcement Learning from Verifiable RewardsPlanned

    Rule-based and executable rewards, correctness verification, reward design, and domains where RLVR works.

  2. 02
    Group-Based Policy OptimizationPlanned

    GRPO-style objectives, relative advantages, stability, sampling costs, and implementation choices.

  3. 03
    Cold Starts, Rejection Sampling, and DistillationPlanned

    Bootstrapping reasoning behavior, filtering traces, iterative training, and transferring reasoning to smaller models.

  4. 04
    Outcome and Process Reward ModelsPlanned

    Sparse outcomes, step-level supervision, verifier reliability, credit assignment, and reward hacking.

  5. 05
    Search, Reflection, and Test-Time ComputePlanned

    Sampling, verifier-guided search, self-correction, budget allocation, stopping, and compute-optimal reasoning.

  6. 06
    Faithful, Hidden, and Latent ReasoningPlanned

    When visible chains of thought are explanations, when they are not, and alternatives for latent deliberation.

  7. 07
    Overthinking and Reasoning EfficiencyPlanned

    Unnecessary deliberation, error amplification, confidence, adaptive budgets, and concise reasoning.

Volume 13 PDF

Own this focused edition

Get the carefully typeset PDF to keep, plus every future chapter, revision, and erratum for this volume.

  • Carefully typeset standalone volume PDF
  • Future chapters, updates, and errata included
  • No oversized combined edition to render or download

Volume 13 of 20

$24

one-time

Version 2026.08.0 · secure checkout via Stripe

Delivered by email · Free volume updates included

All-volume access

Every current and future volume

Get all 20 current volume slots, every future volume, and every update for $199. PDFs are always delivered as manageable individual volumes, never as one combined 12,500-page file.

Continue through the library

Explore adjacent volumes

Version history

Kept current, not frozen in time

Each PDF purchase includes future editions. When the book changes, the updated copy appears in My books at no extra cost.

Current release

Edition 2026.08.0

Initial volume-library release with 19 written PDFs, a Volume 3 placeholder, and no combined edition.