22 days left — Registration closes 2026-10-31.

Logical Reasoning · 6 questions · about 1 min to read

What the Test Measures

Read the passage, answer the questions, then open each answer to check it. The explanation says why the right option is right.

The passage

Read, then answer

A school system that wants to know whether its students are learning must measure something, and what it can measure most cheaply is performance on a standardised test. The appeal is obvious. A common instrument allows comparison across schools that differ in everything else, and it replaces the judgment of individual teachers, which is variable and sometimes prejudiced, with a procedure that treats every candidate identically.

The difficulty begins when the measure is also used to allocate rewards. Once a school's funding, or a teacher's promotion, depends on test scores, effort shifts towards whatever raises scores. Some of that effort raises learning too, and some of it does not: narrowing the syllabus to the tested subjects, coaching in question formats, and in the worst cases, discouraging weak students from sitting the examination at all. A score that rises for these reasons no longer tells you what it told you when nothing depended on it.

This is sometimes stated as a general law: a measure that becomes a target ceases to be a good measure. But the law, stated that broadly, would counsel abandoning measurement altogether, which cannot be right either. A system that measures nothing does not thereby become fair; it simply allocates rewards on the basis of reputation, inspection visits and whatever the inspector happened to see.

The useful question is narrower. It is whether a particular measure can be gamed more easily than the underlying quality can be improved. Where gaming is cheap and improvement is expensive, the measure will degrade quickly. Where the only practical way to raise the number is to teach better, it will not. That is a question about the design of the instrument, and it has different answers for different instruments.

  1. Q1. The author's principal criticism of the broad statement that "a measure that becomes a target ceases to be a good measure" is that it:

    1. Is contradicted by evidence from school systems that use standardised tests
    2. Confuses the reliability of a measure with its validity
    3. Applies only to public institutions and not to private ones
    4. Would, if accepted, imply abandoning measurement altogether, which is not an improvement
    Show answer

    Answer: D. Would, if accepted, imply abandoning measurement altogether, which is not an improvement

    The author objects that the broad form of the law would counsel abandoning measurement entirely, and that a system measuring nothing simply allocates rewards on reputation and inspection instead.

  2. Q2. Which of the following best captures the criterion the author proposes in the final paragraph?

    1. Whether gaming the measure is easier than improving the quality it is meant to track
    2. Whether the measure is cheap to administer relative to its benefits
    3. Whether the measure is administered by an independent agency
    4. Whether the measure has been in use long enough to be validated
    Show answer

    Answer: A. Whether gaming the measure is easier than improving the quality it is meant to track

    The criterion offered is whether a measure can be gamed more cheaply than the underlying quality can be improved. Where gaming is cheap and improvement expensive, the measure degrades.

  3. Q3. The mention of "discouraging weak students from sitting the examination" functions in the argument as:

    1. An example of effort that raises measured scores without raising learning
    2. Evidence that standardised tests are biased against weak students
    3. A concession that teacher judgment is more reliable than testing
    4. An illustration of the cost of administering common instruments
    Show answer

    Answer: A. An example of effort that raises measured scores without raising learning

    It is listed with syllabus narrowing and format coaching as effort that raises measured scores without raising learning, which is why the score ceases to mean what it meant before.

  4. Q4. Which of the following, if true, would most weaken the author's defence of measurement in the third paragraph?

    1. Standardised tests are more expensive to design than inspection systems
    2. Some inspectors have received training in assessment design
    3. Teachers in tested systems report higher workloads
    4. Systems that abandoned testing in favour of inspection were found to allocate resources more equitably than those that retained it
    Show answer

    Answer: D. Systems that abandoned testing in favour of inspection were found to allocate resources more equitably than those that retained it

    The defence is that abandoning measurement does not produce fairness but merely shifts allocation to reputation and inspection. Evidence that inspection-based systems in fact allocated more equitably contradicts that claim directly.

  5. Q5. The author would most likely agree with which of the following statements?

    1. Standardised testing should be abandoned wherever rewards are attached to results
    2. Comparison across schools is impossible without standardised instruments
    3. Teacher judgment should replace testing in all school systems
    4. Whether a test degrades under pressure depends on how the test is designed
    Show answer

    Answer: D. Whether a test degrades under pressure depends on how the test is designed

    The closing paragraph makes degradation a question about the design of the instrument, with different answers for different instruments.

  6. Q6. Which assumption underlies the claim that replacing teacher judgment with a common instrument is an advantage?

    1. That variability and prejudice in assessment are defects worth eliminating
    2. That teachers are generally hostile to being evaluated
    3. That standardised tests measure the entire syllabus
    4. That schools differ from one another in no relevant respect
    Show answer

    Answer: A. That variability and prejudice in assessment are defects worth eliminating

    The advantage claimed is that a common instrument replaces judgment that is variable and sometimes prejudiced. That treats variability and prejudice as defects worth eliminating.

Prepare with CLATcoach, free

A free account gives you a full mock, a past paper for every exam, the daily questions, twelve Legal GK headings and a report on where you stand.

Create a free account