Historical Context & Motivation
The act of choosing a mathematical function to describe observed phenomena is as old as quantitative science itself. Long before graphing calculators or regression software existed, mathematicians and natural philosophers confronted a fundamental question: given a set of observations, which algebraic form best captures the underlying relationship? This question drives function model selection—the deliberate process of matching data behavior to a function family—and it demands that the modeler explicitly state every assumption that justifies the choice. Without such assumption articulation, a model's predictive power cannot be evaluated, and its domain of validity remains unknown.
The central gap that this lesson addresses is the space between recognizing data patterns and rigorously defending a modeling choice. On the AP Precalculus exam, you are expected not merely to identify whether data looks linear, quadratic, or rational, but to articulate why a particular function family is appropriate, what assumptions are embedded in that choice, and under what conditions the model may break down. This synthesis of quantitative reasoning and verbal justification is the hallmark of mathematical maturity.
Core Principles & Definitions
Function model selection rests on several interconnected principles that guide the modeler from raw data to a justified algebraic representation. Understanding these principles transforms model selection from guesswork into a disciplined analytical process. Each principle carries embedded assumptions that must be made explicit whenever a model is proposed or defended.
End Behavior & Dominance
Rate of Change Patterns
Zeros, Poles, and Discontinuities
Concavity and Inflection
Parsimony (Simplicity)
Visual Explanation — Comparing Function Families
The diagram below places four major function families on the same coordinate axes so that their contrasting behaviors become visually immediate. When selecting a model, ask yourself which curve's qualitative shape most closely mirrors the data's overall trajectory, paying particular attention to end behavior, symmetry, and the presence or absence of asymptotes.
Notice how each function family carries distinct structural signatures. A linear model assumes a constant rate of change—an assumption that must be verified by inspecting first differences or a scatter plot. A quadratic model assumes a single turning point and that both ends of the graph point in the same direction, implying constant second differences. Selecting a cubic or higher-degree polynomial assumes additional turning points and an inflection structure that the data must support. Finally, choosing a rational model assumes the existence of input values where the function is undefined—a qualitative feature absent from all polynomial models. Articulating these assumptions is what transforms a curve-fitting exercise into genuine mathematical modeling.
Mathematical Framework
The mathematical backbone of model selection involves analyzing the algebraic structure of candidate function families and matching their properties to observed data. Below are the key forms and the diagnostic tests associated with each. In every case, the assumptions required by each model type are stated alongside the formula.
Model Selection Decision Framework
The following decision diagram provides a systematic workflow for choosing among polynomial and rational function models. Rather than relying on visual intuition alone, this framework uses a sequence of diagnostic questions—each tied to a specific mathematical assumption—to narrow the field of candidate models. Work from the top down, and at each decision node, verify the corresponding assumption against your data.
When using this flowchart on an exam, remember that the finite difference test carries a built-in assumption: the input values must be equally spaced. If the x-values in your table are not equally spaced, you must either re-interpolate or rely on other diagnostic criteria such as the number of turning points, the presence of inflection points, or the scatter plot's qualitative shape. Additionally, in applied problems, contextual clues—such as a population that cannot exceed a carrying capacity—may directly suggest a rational model whose horizontal asymptote represents the limit, even before any numerical test is performed.
Worked Example — Selecting and Justifying a Model
A biologist records the concentration of a nutrient (mg/L) in a lake at equally spaced weekly intervals. The data are shown below. Determine an appropriate function model and articulate the assumptions that support your choice.
| Week (t) | Concentration C(t) |
|---|---|
| 0 | 2.0 |
| 1 | 5.0 |
| 2 | 8.0 |
| 3 | 9.5 |
| 4 | 10.0 |
| 5 | 10.3 |
| 6 | 10.4 |
Strengths and Limitations of Each Model Family
No single function family is universally superior. Each model type has inherent strengths—contexts in which its assumptions align well with reality—and limitations—situations where its structural constraints cause it to misrepresent the data. The table below provides a side-by-side comparison that is especially useful when defending a model choice on free-response questions.
| Model Family | Key Strengths | Key Limitations |
|---|---|---|
| Linear | Simplest model; fewest assumptions; excellent local approximation over short intervals; constant rate of change is easy to interpret. | Cannot model curvature, turning points, or bounded growth. Extrapolation often fails for large domains. |
| Quadratic | Models one turning point (max or min); constant second differences are easily verified; physically meaningful for projectile and area problems. | Symmetric parabolic shape may not match data; always unbounded for large |x|; cannot model asymptotic behavior. |
| Cubic / Higher Degree | Models multiple turning points and inflection; flexible shape; higher-degree polynomials can interpolate through any finite data set exactly. | Overfitting risk with higher degrees; wild oscillation between data points (Runge's phenomenon); strong assumptions needed for extrapolation. |
| Rational | Models asymptotic behavior (horizontal, vertical, slant); captures bounded growth and long-run limits; essential for rate and proportion problems. | Introduces discontinuities that may not exist in the physical context; more parameters to determine; domain restrictions must be contextually justified. |
Connections to Advanced Theory
The model selection skills you develop in AP Precalculus form the foundation for more sophisticated techniques encountered in calculus, statistics, and applied mathematics. The table below maps the precalculus concepts covered in this lesson to their more advanced counterparts, illustrating how assumption articulation becomes even more critical as models grow in complexity.
| AP Precalculus Concept | Advanced Extension |
|---|---|
| Finite difference test for polynomial degree | Taylor polynomial approximation: nth-degree Taylor polynomials generalize finite differences to non-equally-spaced data via derivatives |
| Horizontal asymptote as long-run value | Formal limits at infinity in calculus; L'Hôpital's rule for indeterminate forms arising from rational expressions |
| Choosing between model families based on data | Statistical regression and model comparison using R², AIC, and BIC in AP Statistics and beyond |
| Articulating domain restrictions | Continuity and differentiability conditions in real analysis; piecewise function models in applied mathematics |
| Parsimony (prefer simpler model) | Occam's Razor formalized as the bias-variance tradeoff in machine learning; regularization techniques to prevent overfitting |
Looking forward, the habit of explicitly stating assumptions—"I chose a rational model because the data approach a limiting value, and I assume this limit exists because the physical context involves a saturation process"—transfers directly to writing scientific reports, engineering design documents, and statistical analyses. In each of these advanced contexts, the strength of a conclusion is only as strong as the clarity and validity of the assumptions that support it. Developing this disciplined approach now will serve you well in any quantitative field you pursue.
Practice Problems
Lesson Summary
Function model selection requires you to match data behavior to the structural properties of a function family. Begin by examining end behavior: does the output grow without bound (polynomial) or approach a horizontal asymptote (rational)? Next, apply the finite difference test to equally spaced data to identify the polynomial degree: constant first differences signal linear, constant second differences signal quadratic, and so on. For rational models, compare the degrees of the numerator and denominator to determine the asymptotic structure, and verify that the data support the existence of excluded input values.
Equally important is assumption articulation: every model carries embedded assumptions about continuity, rate of change patterns, domain restrictions, and end behavior that must be explicitly stated and defended with evidence from the data or the problem's real-world context. Adhere to the principle of parsimony—choose the simplest model whose assumptions are satisfied—and always check that your model's behavior is physically or contextually reasonable, especially for extrapolation beyond the observed data range. Mastering this synthesis of quantitative analysis and verbal justification is a cornerstone of success on the AP Precalculus exam.