Exam PA · Tree-Based Models · Free Lesson

Apply boosting as appropriate.

Free SOA Exam PA (Predictive Analytics) lesson in Tree-Based Models. 12 min read, ~1,862 words.

Your assistant proposes picking a boosted tree's learning rate by scanning the test-set curve for its lowest error. That instinct feels efficient, and it quietly leaks the test set into tuning. Knowing why separates a full-credit answer from a partial one.

What boosting does. Bagging and random forests build many deep trees in parallel and average them. Boosting instead builds trees one at a time. Every tree is shallow and weak. Each new tree focuses on the observations the current ensemble predicts poorly. You add trees slowly, and the ensemble improves in small steps.

The mechanism is sequential error correction. Start with a simple guess. Compute how wrong you are. Fit a small tree to those errors. Add a shrunken slice of it to the running prediction. Repeat. Because each tree targets the leftover mistakes, boosting drives down bias rather than variance.

Gradient boosting. In a gradient boosting machine, "errors" are made precise as the negative gradient of a loss function. For squared-error loss the gradient is just the residual, actual minus predicted.

Read the full lesson, free →
Worked examples and practice. Free with a free account, no card.

Common mistakes

Bottom line

Exam shortcut

When a prompt mentions choosing hyperparameters from the test set, immediately write "data leakage" and recommend cross-validation on the training data; that phrase is the graders' target. To reason about learning rate and tree count, remember they move inversely: smaller rate demands more trees, and halving the rate roughly doubles the trees.

The full lesson (about 1,862 words, 12 min read) adds 2 worked examples, all 6 common mistakes, a self-check, free in the app.

Learning objectives

Browse all free Exam PA lessons or jump into free Exam PA practice questions.