Memra

Which parameter, which estimator

◈ 5 cards

μ, p, μ₁ − μ₂, μ_D, p₁ − p₂ or β₁ — the first mark of every inference question is naming the parameter and its estimator; the selection tree then fixes σ known or not, the design, and the table.

The first mark

Every inference question on the final opens the same way, whether or not it says so: what is this question about? The answer is a parameter — one of six in this course — and the estimator that stands in for it. Get that line right and the rest is a formula row; get it wrong and every number that follows is the right arithmetic for the wrong question.

WordingParameterEstimator
the mean of one population
the percent / rate / fraction with an attribute
the difference in means of two separate groups
the mean change on the same units, before and after
the difference in rates of two groups
how a response changes per unit of a predictor

The tree

After the parameter, three questions in order, each answered from the question’s wording:

  1. Is σ known? Only for a mean. A stated population SD → ; a sample . A proportion is always (its variance is , nothing to estimate).
  2. What is the design? One sample; two independent samples; the same units measured twice (paired); or an pair per unit (regression).
  3. Which table and df? for one mean; Welch unless the question says assume equal variances (then pooled ); paired; for any proportion, pooled in a two-proportion test; for a slope. Any with df above 30 reads the ∞ row.

Eight scenarios through the tree

  1. Is the mean processing time of Kitchener invoices above 5 days? σ unknown, n = 20., , .
  2. Is the mean above 5 days, with σ = 1.2 days known from years of records?, , .
  3. Does the percent of invoices paid late exceed 15 %?, , .
  4. Do mean basket sizes differ between two chains, n = 15 and 18, nothing said about variances?, , Welch .
  5. Same, but "assume the two populations have equal variances"., pooled → ∞ row.
  6. Did eight stores sell more after the promotion than before?, , .
  7. Is the default rate different in two regions?, , pooled .
  8. Is income linearly related to credit score for 8 clients?, , .

The traps sit between rows. Scenario 6 is paired because the same stores appear twice — treating it as scenario 4 throws away the pairing and usually the significance. Scenario 3 is a proportion because each invoice is late or not; the mean number of days late would be scenario 1. Scenario 5 becomes pooled only because the question said so; nothing in the data decides that.

The sentence the marker wants

One line, before any arithmetic: "The parameter is , the mean change in sales per store; the estimator is ; the test statistic is on 7 df." Written that way it is the first mark, the choice of table, and the wording of the conclusion, all at once.

meanrate2 groupspaired2 ratesslopequestion?σsμzt n−1pzWelcheq varμ₁−μ₂t mint n₁+n₂−2μ_Dt n_d−1p₁−p₂z pooledβ₁t n−2df > 30 → ∞ row. Pooled t only when the question says “assume equal variances”.
The selection tree. Parameter first, then σ known or not, then the design; the leaf is the table and its df. Any t with df above 30 reads the ∞ row.
NORMAL ~/memra/learn/afm-113/which-parameter-which-estimator utf-8 LF