Check the clues and answers for your group. Answers stay hidden during live play until the host reveals them.
Round 1
Exploratory Data Analysis
$200
This graphical display uses bars to show the frequency distribution of quantitative data, with area proportional to frequency.
$400
This visual representation of a five-number summary uses a box and whiskers to show the distribution and detect potential outliers.
$600
Representing the square root of the variance, this value describes the typical distance of observations from the mean.
$800
Calculated as the difference between the third and first quartiles, this measure identifies the spread of the middle 50% of a dataset.
$1000
Used to assess the appropriateness of a linear model, this plot displays the vertical distances between observed values and predicted values.
Inference Foundations
$200
This probability value indicates how likely the observed results would be if the null hypothesis were actually true.
$400
Representing the probability of correctly rejecting a false null hypothesis, this measure defines the sensitivity of a test.
$600
Constructed around a point estimate, this range provides a level of confidence that it contains the true population parameter.
$800
This starting assumption in statistical testing posits that there is no effect, no difference, or no relationship in the population.
$1000
This statistical error occurs when a researcher incorrectly rejects a true null hypothesis, often called a false positive.
Probability Concepts
$200
This numerical variable takes on values determined by the outcome of a random phenomenon.
$400
This discrete probability distribution models the number of successes in a fixed number of Bernoulli trials.
$600
This type of probability is defined as the chance of an event occurring given that another event has already occurred.
$800
In probability, this set contains every possible outcome that can result from a chance process or experiment.
$1000
These occurrences are characterized by the fact that the occurrence of one does not change the probability of the other.
Statistical Procedures
$200
This test is used to compare a sample mean to a population mean when the population standard deviation is known.
$400
This experimental design involves pairing similar subjects or using the same subject twice to compare treatments effectively.
$600
Known by this acronym, this procedure compares the means of three or more independent groups to see if at least one is significantly different.
$800
This statistical test determines if there is a significant difference between expected and observed frequencies in categorical data.
$1000
Often used for comparing means of two groups when the population standard deviation is unknown, this test follows a specific distribution named after a student.
Study Design
$200
In this method, the population is divided into groups and entire groups are chosen at random to be surveyed.
$400
In this method of gathering data, the researcher collects information from every single member of the target population.
$600
This method involves dividing a population into distinct groups and then taking a random sample from each group.
$800
This sampling technique ensures that every possible combination of a specified size has an equal probability of being selected.
$1000
This experimental design groups units by a nuisance variable before randomly assigning treatments to control for variation.
Pioneers of Data
$200
Known as the Lady with the Lamp, this statistician created the polar area diagram to visualize mortality data during the Crimean War.
$400
This American statistician developed exploratory data analysis and invented the stem-and-leaf plot.
$600
This British polymath pioneered the analysis of variance and the design of experiments in the early 20th century.
$800
An 18th-century English minister, this man formulated a famous theorem for updating probabilities based on new evidence.
$1000
This mathematician established the field of mathematical statistics and is the namesake of the correlation coefficient.
Double Jeopardy
Exploratory Data Analysis
$400
Representing the center of a distribution, this value is the middle point when data is sorted, making it resistant to outliers.
$800
This numerical set contains the minimum, first quartile, median, third quartile, and maximum of a dataset.
$1200
Often used in small datasets, this visual display keeps the individual digits of data values, with a leading column used for scale.
$1600
This graph uses dots to show the correlation and relationship between two quantitative variables.
$2000
Also known as R-squared, this value represents the proportion of variation in a response variable that is explained by the explanatory variable.
Inference Foundations
$400
This hypothesis, often denoted by Ha, represents the researcher's claim or the presence of an effect.
$800
Often denoted by alpha, this threshold is the probability of rejecting the null hypothesis when it is true.
$1200
Also called a false negative, this occurs when a researcher fails to reject a false null hypothesis.
$1600
This value, derived from a critical value and standard error, specifies the range around a sample statistic to account for sampling variability.
$2000
This theorem states that as the sample size increases, the distribution of the sample mean becomes approximately normal, regardless of the population distribution.
Probability Concepts
$400
Events are described as these if they have no outcomes in common and thus cannot occur at the same time.
$800
This is the weighted average of all possible values of a random variable, representing the long-term arithmetic mean.
$1200
This probability distribution models the number of independent trials required to obtain the first success in a sequence of Bernoulli trials.
$1600
This concept explains that the average of results obtained from a large number of independent trials should be close to the expected value.
$2000
This geometric representation uses overlapping circles to illustrate relationships and intersections between different sets of events.
Statistical Procedures
$400
This method estimates a common population variance from multiple samples, often assumed to be equal in a two-sample t-test.
$800
This specific chi-square test assesses whether the observed frequency distribution matches the expected distribution of a variable.
$1200
This chi-square procedure is used to determine if two categorical variables are associated or independent of each other.
$1600
This statistical procedure models the relationship between a dependent variable and one or more independent variables as a straight line.
$2000
Used in hypothesis testing, this procedure determines if there is a significant difference between the means of two distinct groups.
Study Design
$400
This non-probability sampling method involves selecting participants who are easiest to reach or most readily available.
$800
In this type of study, researchers observe outcomes without attempting to manipulate variables or assign treatments.
$1200
This rigorous experimental design ensures that neither the participants nor the researchers know who is receiving which treatment.
$1600
This sampling technique is highly biased as it allows members of a population to choose whether to participate in the study.
$2000
This sampling method selects individuals from a population at a set interval after picking a random starting point.
Pioneers of Data
$400
This Polish-American statistician co-developed the theory of confidence intervals and hypothesis testing rigor.
$800
Working as a chemist for Guinness, he developed the t-distribution to analyze small sample sizes.
$1200
This 18th-century Scottish engineer is credited with the invention of the bar chart, line graph, and pie chart.
$1600
Often called the father of statistical quality control, he introduced the control chart concept in industry.
$2000
A cousin of Charles Darwin, this polymath introduced the concept of regression toward the mean and popularized statistical thinking in heredity.
Final Jeopardy · Data Analysis
This term refers to the observed deviation from a normal distribution, characterized by an asymmetric tail in a frequency graph.