Skip to main content

What universal screening is, and how to read the results

By Jasa Todorovich. Last updated Sep 27, 2026

For teachers and school leaders: Lorsey reads the IXL, NWEA MAP and Savvas scores your school already has and shows the next step for each student. See it on your own roster

The short answer

Universal screening is how a school checks every student for signs of risk, not only the students a teacher is already worried about. The Center on Multi-Tiered System of Supports at the American Institutes for Research (the MTSS Center) describes screening as a systematic process for identifying students who may be at risk for poor learning outcomes related to academics, behavior, engagement, school completion, and college and career readiness.

A post on the MTSS Center blog describes the practice this way: schools give brief (10 minutes or less), validated assessments to all students in a grade level, at least annually and ideally two or three times a year. A team analyzes the results and uses a consistent cut point to identify students who are at risk and may need further intervention.

A screener predicts; it does not diagnose. The same post says screening will never be 100% accurate, so schools should be prepared to confirm initial results with other sources of data.

Screening is one of the essential components of a multi-tiered system of supports (MTSS). How screening data sits beside progress monitoring and decision rules at each tier is covered in MTSS data: what to collect at each tier and how to read it. This guide goes deeper on the screen itself.

Why screen every student, including the ones doing fine

The What Works Clearinghouse (WWC) practice guide on Response to Intervention in mathematics answers the question "why are we testing students who are doing fine?" directly:

  • Screening everyone keeps on-track students on track. It also lets a school evaluate the impact of its instruction on whole groups, such as all grade 2 students.
  • A partial screen hides risk. Screening all students creates a distribution of achievement from high to low. If students considered not at risk were left out, the screened group would hold only at-risk students, and some students at the top of that group who are in fact at risk could go unidentified.

The WWC reading guide for the primary grades calls universal screening a critical first step in identifying students who are at risk for reading difficulties and might need more instruction. The MTSS Center adds that screening data can also identify schools that need support because they have large numbers of struggling students.

How often schools screen

Source What it says about timing
MTSS Center blog At least annually and ideally two or three times a year. Screening three times a year is important to make sure all students are caught.
WWC reading practice guide (kindergarten to grade 2) Screen at the beginning of the year and again in the middle. The second screen is recommended for kindergarten and grade 1 because mid-year results tend to be more valid.
WWC mathematics practice guide Screen at the beginning and middle of the year. It notes that developers of screening systems recommend at least twice a year, for example fall, winter and/or spring.

The mathematics guide treats these screenings as distinct from progress monitoring, which happens more often, for example weekly or monthly, with a select group of students in intervention.

It also names a common pitfall: a long, drawn-out collection, with teachers screening "when time permits." A school that allocates intervention by score must then wait until every classroom is done, which delays services. And because many screening measures are sensitive to instruction, students in a class assessed much later will often score higher simply because they were assessed later. The panel suggests data collection teams that screen students in a short period.

What makes a screener trustworthy

The National Center on Intensive Intervention (NCII) and the National Center on Improving Literacy publish one-page explainers on five standards used on NCII's screening tools charts: classification accuracy, validity, reliability, statistical bias and sample representativeness. The MTSS Center post says that, most importantly, a screener must demonstrate classification accuracy for the specific outcome you care about.

Classification accuracy

Classification accuracy is how well scores on a screener correctly identify students at risk versus students not at risk. Every screening decision lands in one of four cells:

Student is actually at risk Student is not at risk
Screener flags the student True positive False positive
Screener does not flag the student False negative True negative
  • Sensitivity is the probability of correctly identifying that a student is at risk.
  • Specificity is the probability of correctly identifying that a student is not at risk.

NCII's tools charts rate a screening tool highest when both its sensitivity and its specificity are 80% or higher. NCII explains why both errors matter: false positives may raise unnecessary concern and strain school resources, while false negatives may be particularly problematic in academic screening because students may miss out on support they need.

Reliability

NCII defines reliability as the consistency of a set of scores designed to measure the same thing. It describes internal consistency, alternate form, test-retest and inter-rater reliability, and advises checking that the type reported suits the screener and that two or more forms are reported. Its screening charts no longer consider test-retest reliability sufficient evidence on its own.

The WWC panels give working floors: at least 0.70 in the reading guide, and coefficients of .80 or higher in the mathematics guide.

Validity

Validity is how well a tool measures what it is supposed to measure. NCII's highest validity ratings go to tools with strong results from two types of validity studies, using a sample that reflects students at all performance levels.

For screening, the WWC guides single out predictive validity: how well a score earlier in the year predicts later achievement. Both guides recommend 0.60 or higher, and the mathematics guide says within a school year.

Efficiency

A screener is given to every student, so time counts. The mathematics guide suggests a measure take no more than 20 minutes, notes that many take five minutes or less, and recommends the more efficient measure when technical adequacy is roughly equivalent. It adds that it may be better to invest more time in diagnostic assessment of the students who score poorly.

Cut scores set the trade-off

A cut score turns a screening result into an at-risk flag, and it moves the balance between the two kinds of error. The WWC mathematics guide explains:

  • A high cut score gives high sensitivity, because most students who need help are flagged, but low specificity, because many who do not are flagged too. The risk is spending resources on students who do not need help.
  • A low cut score gives higher specificity but lower sensitivity. The risk is not providing intervention to students who need it.

The guide warns that decisions on cut scores can be somewhat arbitrary, and recommends that districts re-evaluate screening measures annually or biannually by examining how screening scores predict state testing results and considering whether to reset cut scores.

Scores close to the line deserve a second look. The WWC reading guide notes that no measure is perfectly reliable. It recommends checking the confidence interval for each benchmark in the technical manual. When a score falls inside it, give an additional assessment or monitor progress for six weeks before deciding. For schools early in implementation, the reading guide says benchmarks from national databases may be easier to adopt than district-specific ones.

A screener is not a diagnostic assessment

Screening and diagnosis answer different questions. A screen asks who may be at risk. NCII describes diagnostic tools as providing data to help educators design individualized instruction and intensify intervention for students who do not respond to validated intervention programs. They range from informal tools that need little training to standardized tools delivered by trained staff.

The MTSS Center post cautions against using data that was not designed for screening, such as quizzes, homework assignments, diagnostic assessments or Lexile levels of the passages students read, and recommends an assessment designed for use as a screener.

The two can work in sequence. In grades 4 through 8, the WWC mathematics guide suggests the previous year's state test results as one viable first stage, with students who score below or only slightly above a benchmark considered for further screening or diagnostic testing.

How screening results feed the tiers

The MTSS Center post describes what happens after the screen:

  • Below the cut point: good candidates for standardized interventions at Tier 2.
  • Very low scores: in some cases, children with the weakest initial skills are likely to have better outcomes if they bypass Tier 2 and move directly into intensive intervention at Tier 3.
  • The whole distribution: screening data can inform resource allocation and whether the core curriculum needs adjusting.

To reduce false positives and false negatives, the post notes that experts have suggested a two-stage process that uses progress monitoring data for a short period after screening. The WWC reading guide makes a similar recommendation: identify a pool of children liberally, then use progress monitoring to refine it to those most at risk.

For district leaders, the WWC mathematics guide recommends that every school in a district use the same screening measure and procedures so results can be compared. If one school consistently has more students identified as at risk, the district could provide extra resources or professional development there.

Choosing and running a screener

The MTSS Center's screening steps start with designing the process around desired outcomes, naming the target population, administration schedule, procedures and data analysis approach. Teams then select tools by weighing their needs, context and priorities alongside the technical adequacy of the measures, train staff, and plan for fidelity to limit inconsistencies in administration and errors in scoring and data entry.

For the technical side, NCII publishes an Academic Screening Tools Chart that displays ratings in classification accuracy, reliability and validity, along with sample representativeness, whether a bias analysis was conducted, and usability features. NCII says tools are rated against established criteria rather than compared or ranked, and that a tool's presence on the chart is not an endorsement. Its suggested steps are to gather a team, determine needs and priorities, learn the chart's language, review the data, and ask for more information.

The WWC mathematics guide recommends that the selection team include someone with measurement expertise, such as a school psychologist or a member of the district research and evaluation staff, as well as people with expertise in mathematics instruction.

Where Lorsey fits

Lorsey combines learning signals from connected curriculum and assessment systems and helps schools decide what each student should work on next. It does not replace those systems or provide its own curriculum.

Frequently asked questions

What is universal screening?

Universal screening is a systematic process for identifying students who may be at risk for poor learning outcomes. Schools give brief, validated assessments to all students in a grade level and use a consistent cut point to flag students who may need intervention.

What is a universal screener?

A universal screener is the assessment used in that process. An MTSS Center post says it should be brief, reliable and valid, and above all demonstrate classification accuracy, meaning it correctly identifies who is and is not at risk for the outcome the school cares about.

What are universal screeners for reading?

For kindergarten through grade 2, the WWC reading practice guide lists measures such as letter naming fluency, phoneme segmentation, nonsense word fluency, word identification and oral reading fluency, and recommends using two screening measures at each point. It recommends measures of letter knowledge, phonemic awareness and expressive and receptive vocabulary in kindergarten, and word reading and passage reading in grade 2.

How often should universal screening happen?

The MTSS Center blog says at least annually and ideally two or three times a year. The WWC reading and mathematics practice guides both recommend screening at the beginning and middle of the year.

What is universal screening data used for?

Teams use it to identify students who may need Tier 2 or Tier 3 intervention, to decide how to allocate resources, and to judge whether the core curriculum needs adjusting. The MTSS Center also names identifying schools with large numbers of struggling students.

What is the difference between a screener and a diagnostic assessment?

A screener is given to every student to predict who may be at risk. NCII describes diagnostic tools as helping educators design individualized instruction and intensify intervention for students who do not respond to validated intervention programs.

What is classification accuracy in screening?

It is how well a screener's scores correctly sort students into at risk and not at risk. NCII's tools charts rate a screener highest when both its sensitivity and its specificity are 80% or higher.

Sources

Lorsey is not affiliated with or endorsed by the American Institutes for Research, the MTSS Center, the National Center on Intensive Intervention, the National Center on Improving Literacy or the U.S. Department of Education. Their public materials are cited here to describe practices schools already use.