PrepSolution

Counseling · Study journal

NCE Research and Statistics: Worked Practice Questions

Descriptive statistics, reliability, validity, correlation, and research designs are tested on the NCE. Work through 8 original practice questions with stepwise solutions and distractor rationales.

By PrepSolution Editorial TeamPublished 12 min read
Inside this articleCounseling
Section preview 01

Descriptive statistics the NCE expects

You should be ready to compute or interpret the mean, median, mode, and range from a small dataset. The mean is the arithmetic average; the median is the middle value when scores are ordered; the mode is the most frequent value; and the range is the highest score minus the lowest.

Read this section
On this page

Research methods and statistics show up on the National Counselor Examination (NCE) through scenario-based items: interpreting a small dataset, choosing the right reliability or validity type, identifying a research design, or spotting why a correlation cannot support a causal claim. This article reviews the concepts the current and July 2027 NCE blueprints expect, then walks through 8 original four-option questions with stepwise solutions and explanations of why every wrong option fails. If you are building a broader study plan, start with our NCE study plan and use the free questions on /free-practice/nce for additional domain review.

Descriptive statistics the NCE expects

You should be ready to compute or interpret the mean, median, mode, and range from a small dataset. The mean is the arithmetic average; the median is the middle value when scores are ordered; the mode is the most frequent value; and the range is the highest score minus the lowest.

Standard deviation describes spread around the mean. A long tail can pull the mean away from the median, but their ordering alone does not prove a distribution’s shape. Examine the data or graph. The median is generally less sensitive to extreme observations than the mean; the mode identifies the most frequent value.

Use NBCC’s content outline to locate research, measurement and assessment topics. A whole domain’s percentage is not a statistics-only quota. The July 1, 2027 outline reorganizes domains, so keep version-specific study plans separate.

Reliability and validity

Reliability is consistency. Common types on the NCE include test-retest reliability (same instrument, same people, two time points), internal consistency (items within one administration agree with each other), and inter-rater reliability (two raters agree). Validity is whether the instrument measures what it claims to measure. Content validity means the items cover the domain; criterion validity means scores predict or correlate with an outcome; and construct validity means scores reflect the theoretical construct.

Reliability supports interpretation of scores, but it does not by itself establish validity for a particular use. Poor precision can limit an interpretation; changes over time can also reflect real change in the construct. Assess the intended use, conditions and relevant validity evidence rather than declaring an instrument valid or invalid for all purposes.

Correlation without causation

The correlation coefficient, r, ranges from −1 to +1. The sign tells direction: positive means scores move together, negative means they move in opposite directions. The absolute value tells strength; values near 0 are weak, and values near 1 or −1 are strong. A coefficient of −0.72 is therefore a strong negative relationship.

In the correlation example below, association is supported and a causal conclusion is not. Choose based on the research design and evidence, not a rule that a cautious-sounding answer is always correct.

Research designs

A true experiment has three features: random assignment to conditions, manipulation of an independent variable, and a comparison or control group. A quasi-experiment has manipulation and comparison but lacks random assignment. A correlational study measures variables as they naturally occur without manipulating anything. The independent variable is what the researcher manipulates or uses to predict; the dependent variable is the outcome being measured.

Internal validity asks whether the design supports the claim that the independent variable caused the change in the dependent variable. Threats to internal validity include history (an outside event affects the group), maturation (participants change naturally over time), selection bias (groups differ before the study), testing (taking a pretest changes later performance), instrumentation (the measurement tool changes), attrition (participants drop out), and regression to the mean (extreme scores move toward the average on retest). External validity asks whether results generalize to other people, settings, or times.

Worked practice

Each question below is an original four-option item written for NCE-level reasoning. The arithmetic was verified independently, and every distractor includes a rationale explaining why it fails.

Question 1

A counselor collects the following seven anxiety scores from a small support group: 68, 74, 82, 82, 88, 94, 100. Which statement is correct?

  1. Mean = 84, median = 82, mode = 82
  2. Mean = 82, median = 84, mode = 82
  3. Mean = 84, median = 82, no mode
  4. Mean = 588, median = 82, mode = 82

Correct answer: A. The scores are already ordered. Sum = 68 + 74 + 82 + 82 + 88 + 94 + 100 = 588. Mean = 588 ÷ 7 = 84. The median is the 4th value = 82. The mode is the most frequent value = 82. Option B reverses the mean and median. Option C misses the mode, which is 82 because it appears twice. Option D reports the raw sum (588) instead of dividing by 7.

Question 2

On a counseling self-efficacy scale, a sample of graduate students has a mean of 78 and median of 84. Its histogram is unimodal with a long tail toward lower scores. Which description fits that observed shape?

  1. It is positively skewed.
  2. It is negatively skewed.
  3. It is symmetric.
  4. It is bimodal.

Correct answer: B. The stated long tail toward lower scores establishes negative (left) skew. The mean below the median is consistent with it, but those two summaries alone would not prove the shape. A right tail would support A; the stated asymmetry rules out C, and the single peak rules out D.

Question 3

Four datasets each have a mean of 50. Which dataset has the largest standard deviation?

  1. 48, 49, 51, 52
  2. 45, 48, 52, 55
  3. 40, 45, 55, 60
  4. 20, 40, 60, 80

Correct answer: D. Standard deviation measures spread around the mean. Dataset D ranges from 20 to 80, so its scores are the most dispersed. Dataset A is the most clustered and therefore has the smallest standard deviation. Datasets B and C are intermediate; they are wrong because they are less spread out than D.

Question 4

A counselor administers the same anxiety inventory to the same group of clients twice, one week apart, and correlates the two sets of scores. This procedure assesses:

  1. Test-retest reliability
  2. Internal consistency reliability
  3. Inter-rater reliability
  4. Content validity

Correct answer: A. Test-retest reliability evaluates stability over time by giving the same measure to the same people on two occasions. Option B refers to agreement among items within a single administration. Option C refers to agreement between two different raters. Option D is a validity concept about whether items cover the content domain, not a reliability procedure.

Question 5

A newly published scale claims to measure counseling self-efficacy, but repeated administrations under comparable conditions produce very different scores for the same counselors, despite evidence that the underlying construct has remained stable. Which conclusion is most accurate?

  1. The scale has high construct validity.
  2. Poor test-retest reliability limits confidence in interpreting individual scores for this use.
  3. The scale has high content validity.
  4. The scale is reliable but not valid.

Correct answer: B. Under the stable conditions stated, large unexplained changes raise a test-retest reliability concern. This limits confidence in individual score interpretation; it does not prove that every possible use is invalid. A and C assert validity evidence the stem does not provide. D wrongly asserts reliability despite the described instability.

Question 6

A study reports a correlation of r = −0.72 between the number of hours spent on social media and sleep quality ratings. Which interpretation is correct?

  1. Social media use causes poor sleep.
  2. There is a strong negative association between social media use and sleep quality.
  3. The relationship is weak because r is negative.
  4. Sleep quality causes reduced social media use.

Correct answer: B. The negative sign means the variables move in opposite directions, and an absolute value of 0.72 indicates a strong relationship. Options A and D infer causation, which correlation alone cannot support. Option C confuses sign with strength; a negative correlation can be strong.

Question 7

A researcher randomly assigns clients with depression to either a cognitive-behavioral therapy group or a waitlist control group, then compares depression scores after eight weeks. This is best described as a:

  1. True experiment
  2. Quasi-experiment
  3. Correlational study
  4. Case study

Correct answer: A. This design has random assignment, manipulation of the independent variable (CBT vs waitlist), and a comparison group — the three hallmarks of a true experiment. Option B would apply if participants were not randomly assigned. Option C would apply if the researcher only measured depression and therapy use without manipulating treatment. Option D involves an in-depth study of one person or case.

Question 8

A counselor leads a one-day stress-reduction workshop and measures stress one week before and one week after the workshop. Several participants begin exercising between the pretest and posttest. The fact that exercise could explain any observed stress change is best described as a threat to:

  1. Internal validity
  2. External validity
  3. Test-retest reliability
  4. Content validity

Correct answer: A. Without a control group, an outside event such as starting exercise (a history/confounding threat) can explain the change, weakening the claim that the workshop caused the improvement. This is an internal-validity problem. Option B concerns whether results generalize, which is not the issue here. Option C concerns score consistency across two administrations. Option D concerns whether the measure covers the content domain.

How to study research and statistics for the NCE

Use these examples to distinguish summary statistics, measurement evidence and study design. Explain why each answer follows from the data. Consult the current NBCC outline for preparation scope rather than treating this short lesson as a list of everything that can be tested.

Organize your topic review with the NCE study plan. Confirm exam coverage against the NBCC outline for your testing date, and use free NCE practice questions to identify concepts to revisit.

PrepSolution Editorial Team

Exam-prep editorial team

About PrepSolution

References

  1. [1] NBCC (National Board for Certified Counselors) (2023). National Counselor Examination Content Outline. nbcc.org (official PDF). nbcc.org (official PDF)
  2. [2] NBCC (National Board for Certified Counselors) (2024). NCE State Licensure Candidate Handbook. nbcc.org (official PDF). nbcc.org (official PDF)
  3. [3] NBCC (National Board for Certified Counselors) (2025). NCE Content Outline (effective July 1, 2027). nbcc.org (official PDF). nbcc.org (official PDF)
  4. [4] NBCC (National Board for Certified Counselors) (2025). NCE Examination Specifications (effective July 1, 2027). nbcc.org (official PDF). nbcc.org (official PDF)

Frequently asked questions

Research and assessment topics appear in NBCC’s outlines, but the domain percentages are not statistics-only item counts. This lesson samples descriptive statistics, correlation, measurement and research design.

This lesson practices interpreting spread. It does not verify that hand calculation is excluded from the NCE. Follow the applicable NBCC outline and understand both the concept and the calculations expected in your preparation materials.

Reliability concerns score consistency or precision under specified conditions. Validity concerns the evidence supporting a particular interpretation and use. High reliability alone does not establish validity; poor precision can limit a proposed interpretation.

No. Correlation shows direction and strength of association, not causation. A third variable, reverse causation, or coincidence can all produce a strong correlation.

A true experiment has random assignment to conditions, manipulation of an independent variable, and a comparison or control group. Quasi-experiments lack random assignment; correlational studies lack manipulation.

The mean is most affected by outliers because every score contributes to the total. The median is more robust, and the mode is unaffected by extreme scores unless an extreme value becomes the most frequent.

Keep practicing for the NCE.

Review the study tools, access options and price, or start with free practice.

Related Articles