Official Practice Exam 2026/2027: Complete Exam-Style
Questions with Detailed Rationales | 100% Verified | Pass
Guaranteed – A+ Graded
TABLE OF CONTENTS
Section 1 | Foundations of Assessment Literacy | Q1 – Q10
Section 2 | Formative Assessment Strategies | Q11 – Q20
Section 3 | Summative Assessment Design | Q21 – Q30
Section 4 | Data Analysis and Interpretation | Q31 – Q40
Section 5 | Using Assessment to Guide Instruction | Q41 – Q50
Instructions: Choose the single best answer. Pass: 80% in 90 minutes.
══════════════════════════════════════
SECTION 1: FOUNDATIONS OF ASSESSMENT LITERACY Q1 – Q10
══════════════════════════════════════
Question 1 of 50
A seventh-grade social studies teacher is designing a unit test on the causes of the Civil
War. She includes a question asking students to analyze a primary source letter from
1860, but several English language learners struggle with the archaic vocabulary and
miss the historical point entirely. The teacher is frustrated because she knows the
students understand the content when she explains it verbally.
A. The assessment lacks content validity because it measures reading comprehension
rather than historical understanding.
B. The assessment lacks construct validity because the language barrier prevents it
from measuring the intended historical knowledge. ✓ CORRECT
C. The assessment lacks reliability because the English language learners would score
differently on a retest.
D. The assessment lacks inter-rater reliability because different teachers would grade
the responses differently.
,Correct Answer: B
Rationale: Construct validity refers to whether an assessment actually measures what it
claims to measure, and in this case the archaic vocabulary introduces a
construct-irrelevant variance that masks the students' true historical understanding. The
most tempting wrong answer is A because content validity concerns whether the test
covers the right topics, not whether the language used to access those topics is
appropriate. In practice, teachers should routinely audit assessments for language
demands that exceed the construct being measured, especially when working with
multilingual learners.
Question 2 of 50
A high school algebra team notices that their common midterm has a Cronbach's alpha
of 0.92 and an item-total correlation below 0.10 on three items. When they review those
three items, they find that all of them test geometric concepts from a previous unit
rather than the linear functions currently being assessed.
A. The three items should be retained because the overall reliability coefficient is strong.
B. The three items should be removed because they weaken the internal consistency of
the linear functions assessment. ✓ CORRECT
C. The three items should be revised to test linear functions because reliability is more
important than validity.
D. The three items should be kept because low item-total correlation simply means the
items are too difficult.
Correct Answer: B
Rationale: Items with low item-total correlation on a targeted assessment are likely
measuring a different construct, which introduces noise and weakens the overall
coherence of the instrument. The most tempting wrong answer is A because a high
overall alpha can mask the presence of off-target items that dilute what the test is
actually measuring. In real team practice, item-level review is essential even when
aggregate statistics look acceptable.
,Question 3 of 50
A third-grade teacher uses a computerized adaptive math screener three times a year.
After the fall administration, she notices that two students who scored at the 15th
percentile are both receiving gifted services in reading and have strong number sense
when she works with them in small groups.
A. The screener may be producing false negatives due to construct underrepresentation
in the adaptive algorithm.
B. The screener may be producing false negatives because the students' gifted status in
reading is interfering with math performance.
C. The screener may be producing false negatives due to a mismatch between the
timed format and the students' processing profiles. ✓ CORRECT
D. The screener may be producing false positives because the students' number sense
in small groups is not generalizable.
Correct Answer: C
Rationale: When students demonstrate strong mathematical understanding in
low-stakes, untimed settings but perform poorly on a timed computerized screener, the
format itself may be the barrier rather than their knowledge. The most tempting wrong
answer is A because construct underrepresentation refers to whether the test covers
enough of the domain, not whether the delivery mode is appropriate for the learner. In
practice, teachers should always triangulate screener data with classroom performance
before making placement decisions.
Question 4 of 50
A district assessment coordinator is training new teachers on bias reduction. She
presents four items from a fifth-grade science benchmark and asks the group to identify
which one most likely contains cultural bias. The item describes a family camping trip in
the Rocky Mountains and asks students to explain how altitude affects boiling point.
A. The item is biased because it assumes all students have experienced camping.
, B. The item is biased because the science concept itself is culturally specific to
mountainous regions.
C. The item is likely unbiased because the scenario provides sufficient context for
students to reason through the concept. ✓ CORRECT
D. The item is biased because the vocabulary "altitude" and "boiling point" are too
advanced for fifth grade.
Correct Answer: C
Rationale: A scenario does not constitute bias if the necessary information is
embedded in the item stem and students are not required to draw on personal
experience to answer correctly. The most tempting wrong answer is A because
assuming background experience is only problematic when the item demands that
experience to succeed, which is not the case here. In real item review, teachers must
distinguish between unfamiliar contexts and contexts that actually penalize students for
lacking specific background knowledge.
Question 5 of 50
An elementary principal asks her staff to review the reliability of their common
formative reading assessments. One teacher argues that because the assessments are
short and used frequently, reliability is not a relevant concern.
A. Reliability is always relevant because even brief assessments need to produce
consistent results to track growth over time. ✓ CORRECT
B. Reliability is irrelevant for formative assessments because their purpose is to inform
instruction, not to assign grades.
C. Reliability is only relevant for summative assessments that are used for high-stakes
decisions.
D. Reliability is less important for formative assessments because they are typically
scored by the classroom teacher.
Correct Answer: A
Rationale: All assessments, regardless of length or purpose, must yield consistent
results if educators are to trust the patterns they reveal about student learning over