Skip to main content icon/video/no-internet

Content-Related Validity Evidence

Validation of test scores involves collecting evidence and developing an argument that supports a particular use (i.e., an inference or decision) of the test scores. For a validity argument to be correct, it must be supported by evidence and be logical and coherent. There are various types of evidence that can be used to support a validity argument, including content-related validity evidence, criterion-related validity evidence, and evidence related to reliability and dimensional structure. The type of evidence needed to support the use of the test scores depends on the type of inference or decision being made. As such, test scores can only be said to be valid for a particular use. If multiple inferences or decisions are to be made based on a set of test scores, multiple types of evidence may be required. Even if a single inference or decision is made, multiple types of evidence may still be required to support the test score use. After briefly reviewing the three types of validity evidence, this entry focuses on the basics of content-related validity evidence, including providing an example of its use.

Types of Validity Evidence

Validity evidence can be classified into three basic categories: content-related evidence, criterion-related evidence, and evidence related to reliability and dimensional structure. Most test score uses require some evidence from all three categories. Content-related validity evidence is evidence about the extent to which the test accurately represents the target domain. For achievement tests, the target domain is most often a particular subject matter domain (e.g., seventh-grade mathematics), and for ability tests, the target domain is most often a particular mental ability (e.g., quantitative reasoning). Criterion-related validity evidence is evidence that relates the test scores to one or more external criterion (often observable behaviors). Evidence related to reliability and dimensional structure are types of evidence about the internal structure of a test (i.e., the composition of the items and subtests). In addition, reliability evidence is evidence about the consistency or reproducibility of the test scores across various test conditions (e.g., raters and time).

Content-Related Validity Evidence

Content-related validity evidence is most important when making an inference about a target domain based on a sample of observations taken from that target domain. Evidence related to both the definition of the target domain and the representativeness or relevance of the sample of observations (items and tasks) taken from the target domain are important aspects of content-related validity evidence. Both types of evidence rely on the judgment of experts and are therefore subjective. Content-related validity evidence is often confused with, or is thought to be the same as, face validity evidence. This confusion is understandable because on the surface the two types of validity evidence have many commonalities. However, the two types of validity evidence differ in who is making the judgment about validity. When seeking evidence about face validity, the test takers and test users are asked if the test appears to measure what the test developers say the test measures. In contrast, when seeking evidence about content validity, individuals with expert knowledge in the target domain are asked if the test content (items and tasks) represents or is relevant to the target domain.

...

  • Loading...
locked icon

Sign in to access this content

Get a 30 day FREE TRIAL

  • Watch videos from a variety of sources bringing classroom topics to life
  • Read modern, diverse business cases
  • Explore hundreds of books and reference titles

Sage Recommends

We found other relevant content for you on other Sage platforms.

Loading