Early Wondering Minds

Structured Literacy Programs Compared for Ages 4–8

Research-backed programs help families choose the right reading approach for young learners.

Staff Writer · · 13 min read
Cover illustration for “Structured Literacy Programs Compared for Ages 4–8”
Science of Reading · September 9, 2026 · 13 min read · 2,934 words

Structured literacy isn't a product you buy. It's an approach, a set of principles about how reading gets taught, and any given program is just one attempt to put those principles into practice. The International Dyslexia Association said as much in its July 2024 fact sheet, stating plainly that structured literacy "is an instructional approach, not a program." That distinction matters more than it sounds like it should, because programs built on the same six pillars (phonemic awareness, letter-sound correspondences, syllables, morphology, syntax, semantics) can still differ enormously in how well they actually deliver on them.

The gap between approach and program is where marketing does its best work. A box of materials can say "Science of Reading aligned" on the cover without ever sequencing its phonics instruction properly, without ever building in fluency practice, without cumulative review. Explicit and systematic and cumulative are specific claims, not vibes. Explicit means the code gets taught directly, not inferred from context. Systematic means there's a sequence, and the sequence assumes nothing the child hasn't already been taught. Cumulative means yesterday's skill shows up again tomorrow, folded into new material, instead of vanishing once the worksheet's done.

None of this is academic. The 2024 NAEP results showed more than a third of eighth-graders reading below the Basic level, the highest share ever recorded on that assessment. That's the backdrop against which every one of these program decisions gets made. Parents choosing between programs for a 4- to 8-year-old aren't picking a nice-to-have. They're making a bet on whether a kid learns to read on time.

What the Science of Reading says must be in any program worth using

Five components have to show up in a complete program: phonological awareness, phonics and word recognition, fluency, vocabulary and oral language comprehension, and text comprehension. Leave one out and the program isn't incomplete in some minor, forgivable way. It's missing a load-bearing wall.

Take decoding. The research on this is about as settled as reading research gets: systematic phonics instruction in kindergarten and first grade works, and works well. Phonics instruction that shows up after first grade, once a child has already floundered for a year or two, is markedly less effective. So a program that delays phonics, or treats it as a side dish to whole-word memorization, isn't a slower version of the right approach. It's the wrong approach, dressed up in the right vocabulary.

Phonemic awareness alone doesn't cut it either, and this trips up a lot of well-meaning programs. Research on early literacy indicates that what actually moves outcomes is teaching phonemic awareness and phonics together, with letters present from the start, rather than drilling sound awareness in isolation before ever introducing print. The principle is that print and sound instruction should not be separated. A program built entirely around oral sound games, with letters bolted on weeks or months later, is leaving performance on the table.

Fluency gets treated as a speed drill in weaker programs, and that's a mistake. Real fluency is automaticity, prosody, and comprehension operating together, not a stopwatch exercise. A child who can decode a passage but reads it in a flat monotone, missing the meaning entirely, hasn't achieved fluency. A program that drills decoding accuracy without ever building in fluency practice, repeated reading, phrasing, expression, is only half-built.

Comprehension, meanwhile, doesn't wait around for decoding to finish. It gets built simultaneously, through interactive read-alouds and explicit strategy instruction, even with kids who can't yet sound out a single word on their own. So when evaluating a program's scope and sequence, the real question isn't whether all five components eventually appear. It's whether they're integrated from day one or treated as a locked staircase, where a child has to "graduate" phonics before comprehension work even starts.

State policy offers a useful, if rough, proxy for all of this. Virginia has moved to prohibit three-cueing instruction. Pennsylvania has built a genuinely coherent framework across three pieces of legislation, Act 55 in 2022, Act 135 in 2024, and Act 47 in 2025, covering high-quality instruction, educator preparation, and early literacy together. When a state builds that kind of scaffolding, it's a signal about which programs are getting approved for classroom use and which are getting phased out.

The six questions to ask before evaluating any specific program

Six questions do most of the work here, and they apply whether the program costs very little or has a school district's mid-six-figure procurement contract behind it.

Who is this program designed for? Core instruction for every child in a classroom, supplemental support for kids who need extra reps, and intensive remediation for a diagnosed reading disability are three genuinely different jobs. A program built for one of these roles will underperform if a parent asks it to do another.

What independent evidence exists? A publisher's own study is not the same as third-party research, and neither is quite the same as appearing on a state-approved list, like DC OSSE's High-Quality Science-Based Literacy Programs List from May 2024, or Ohio's state-approved materials list, which as of recent reporting has 98% of schools aligned to it. Parents should ask, specifically, which category a program's evidence falls into.

What does the program actually cover, and in what order? Ask for a published scope and sequence document. If a program can't produce one, that absence is itself the answer.

Who delivers the instruction, and what do they need to know first? A parent-led scripted program, a program that requires a trained tutor, and a technology-delivered program place wildly different demands on a family. Confusing one for another is how well-meaning parents end up frustrated three weeks in.

How does the program respond when a child struggles or races ahead? Static, scripted lessons and programs with real diagnostic branching are not the same category of tool, even when they cover identical content on paper.

What does finishing actually look like? A program with no defined proficiency benchmark, no graduation point, gives a family no way to know if it's working. That's a real gap, not a minor inconvenience.

These aren't invented for this piece, either. DC OSSE's HQIM Rubric for K-5 Science-Based Literacy Programs formalizes essentially this same set of criteria for school evaluators. Parents can use the identical framework, no license required.

Programs designed for homeschool and parent-led instruction: All About Reading and Logic of English

Two programs dominate the homeschool structured literacy space, and they solve the delivery problem in different ways.

All About Reading, built on Orton-Gillingham principles, is one of the most widely used structured literacy programs among homeschooling families. Its lessons are fully scripted, and it comes with physical manipulatives: letter tiles, phonogram cards, fluency practice sheets. A parent with no teaching background can open the box and read the script aloud, engaging visual, auditory, and kinesthetic channels simultaneously. Its alignment with Science of Reading principles is strong, with explicit phonemic awareness, phonics, fluency, and comprehension work built into the sequence from the start. The tradeoff: since the lessons are scripted, they adapt to a struggling or advanced child only as far as the parent can improvise off-script. The program itself doesn't adjust.

Logic of English takes a different route. Its Foundations A-D level, aimed at ages 4 to 7, is built around 75 phonograms and 31 spelling rules, a genuinely deeper rule-based framework than most competitors offer, one that accounts for the logic behind the overwhelming majority of English spelling patterns. Reading and spelling get taught together from the beginning, and the program extends upward into Essentials for kids 8 through adulthood, covering advanced spelling, grammar, and vocabulary. The alignment here is very strong, arguably stronger than AAR's in terms of explaining why English spelling behaves the way it does, rather than just what the rules are. The cost is preparation time: a parent needs to actually learn all 75 phonograms and 31 rules before teaching them, a steeper front-loaded investment than AAR requires.

Both are physical-materials-based, both are parent-led, and neither one reads the child's real-time performance the way a trained tutor or an adaptive system would. The parent is the diagnostic engine in both cases. That's not a criticism so much as a fact about what these programs are built to do.

Programs designed for school and intervention settings: Fundations, Explode the Code, Wilson Reading System, and Barton

Fundations, published by Wilson Language Training Corp., shows up frequently in classroom settings as Tier 1 or Tier 2 instruction, and one recent university thesis analyzing early-learning phonics curricula named it alongside Heggerty and ReadBright as a program worth studying. Its roots are Orton-Gillingham, and it's common in early elementary grades. Program-level independent evidence is something parents should ask Wilson to produce directly, since what's publicly available doesn't settle the question on its own.

Explode the Code, from EPS Learning, spans PreK through grade 4, and the publisher describes a lesson structure built on Orton-Gillingham best practices, designed to engage multisensory pathways. It's most commonly used as supplemental instruction layered on top of a core program, though it can function as a standalone phonics curriculum too. That distinction, supplemental versus standalone, matters a lot when a parent is deciding whether it can carry the full instructional load on its own.

The Wilson Reading System is a different animal entirely. It's an evidence-based Orton-Gillingham program aimed at grades 2 through 12, is designed specifically for students with language-based learning disabilities like dyslexia who have word-level decoding deficits. Full completion typically takes two to four years. Crucially, it requires a genuinely skilled, trained instructor to deliver well. It is not a parent-led starting point for a 4-year-old. It's an intensive, long-term intervention for a child who needs one, delivered by someone qualified to run it.

Barton Reading and Spelling System runs on similar Orton-Gillingham foundations, structured across sequential levels with a strong emphasis on decoding and spelling. It's popular among parents of children with dyslexia, using multisensory touch-and-say techniques designed for that specific population. Reports of students advancing multiple grade levels in reading are common among parents and tutors who've used it, though that's worth flagging plainly: those are reports, not a controlled trial result, and the two shouldn't get conflated.

Across this whole cluster, the honest framing isn't which program is "better." It's which role each one is built to fill, and whether a family actually has access to the instructor that role requires. DC OSSE's May 2024 approved K-5 list and Ohio's list, with that 98% school-alignment figure, offer a practical shortcut here: programs on those lists have been run through a published evaluation rubric. Programs that aren't on them simply haven't been vetted publicly, which isn't the same as being bad, but it does mean the burden of verification shifts back onto the parent.

Ohio's own data underline why implementation matters as much as program selection. More than 37,000 educators enrolled in the state's structured literacy professional learning, and more than 28,000 completed it, yet the state continues to report high decoding-support needs among older students. Even a genuinely strong program doesn't move outcomes on its own. It requires sustained, skilled delivery to actually land.

What school-approved lists leave out and why technology-delivered instruction fills a different gap

Every program discussed so far, physical or classroom-based, shares one structural limitation: it assumes a single capable instructor working one-on-one or in a small group with a child. Most families don't have consistent access to that. A parent working full time, or managing three kids at three different reading levels, cannot reliably deliver a 45-minute scripted Orton-Gillingham lesson every single day. That's not a failure of will. It's an access problem, and it's the gap school-approved lists don't account for, because those lists were built for schools, which have that instructor by design.

This is where "adaptive" starts to mean genuinely different things depending on the tool. A weak adaptive program branches its quiz logic: get a question wrong, get an easier question next. That's adjusting which question gets asked. A strong adaptive program adjusts the teaching itself, diagnosing why the answer was wrong and changing what gets taught next, not just what gets asked. Those two things get marketed with the same word, and they are not remotely equivalent.

What AI specifically can do that rule-based software can't is listen. Listen to a child's actual reading voice, identify the exact phoneme where a decoding error happened, and adjust the next lesson based on that specific breakdown, not a generic wrong-answer bucket. Current research on AI in education describes several distinct capabilities: intelligent tutoring systems, adaptive feedback, automated evaluation, real-time personalization. The meaningful line isn't between AI and non-AI tools. It's between tools that do one of these things in isolation and tools that integrate all of them into something that actually functions like instruction.

Intellectual honesty requires the other half of this, too. The evidence on AI efficacy in classroom and home settings is genuinely mixed, not uniformly positive. The broader research base on screen-based reading tools includes studies showing positive results alongside studies showing negative, null, or mixed ones. So the right move for a parent evaluating any AI-based reading program is to ask the vendor for evidence specific to the child's age range and the specific skill being taught, not a general claim about "AI in education" working.

The delivery-method tradeoff from earlier resurfaces sharply here. A technology program that actually listens and responds in real time places far lower demands on a parent's time than any scripted, parent-led program does. Whether that tradeoff is worth it hinges entirely on whether the underlying instructional model is rigorous, not just whether the interface is polished.

Running one such program against the six-question framework laid out earlier is instructive, precisely because it shows how the same questions apply regardless of delivery method:

Who is it for? Children spanning kindergarten through third grade, covering struggling readers, kids working ahead of grade level, and English language learners who have some baseline English already.

What independent evidence exists? Curriculum described as grounded in Science of Reading principles.

Scope and sequence? Explicit coverage across phonemic awareness, phonics, fluency, vocabulary, and comprehension, plus pre-reading and kindergarten-readiness skills. The program incorporates activities designed to work comprehension and oral language together., rather than treating them as a later-stage add-on.

Who delivers instruction? The AI listens to each child's voice in real time, diagnoses errors at the phoneme level, and adjusts instruction moment to moment. The parent's role shifts to oversight and review, not scripted lesson delivery.

How does it adapt? The system aims to track and respond to a child's performance over time, closer in function to what a trained one-on-one tutor does than to a quiz that branches based on right or wrong.

What does completion look like? Parents get progress reporting tracked against grade-level benchmarks, rather than being left to eyeball whether a kid is on track.

One more practical note: the program is available on iOS and Android, which matters because cost is exactly what pushes a lot of families toward lower-quality free alternatives in the first place.

How to match a program to a child's actual situation rather than an idealized one

None of this framework means much without mapping it to where a specific kid actually stands, and there are really three situations most parents find themselves in.

A child who's behind or struggling, diagnosed or not, needs intervention-level intensity. Wilson or Barton are strong fits here, provided a trained instructor is actually available to deliver them properly. If one isn't (and for a lot of families, it genuinely isn't) a rigorous, AI-driven program that listens and adapts becomes a serious alternative to sitting on a waitlist for school-based support that may take months to materialize.

A child who's on grade level but would benefit from more consistent, structured practice is a different case entirely. Parent-led scripted programs like AAR or Logic of English's Foundations work well here, as long as the parent genuinely has the time and the willingness to prepare. Where that time doesn't exist, technology-delivered adaptive instruction fills the same role without demanding it.

A child who's ahead of grade level needs something else again: enrichment, not remediation. A program built around catching a struggling reader up will bore a kid who's already mastered the early sequence, so the priority shifts to finding a program with a ceiling high enough to keep pace with genuine advancement.

DC OSSE's own framework offers a useful reality check here, too: an estimated 80 percent of students respond adequately to high-quality Tier 1 core instruction alone, with no intervention required. For most families, that reframes the whole search. The job isn't hunting down an intervention program. It's finding one well-implemented core program and sticking with it.

A few red flags apply no matter which category a program falls into. No published scope and sequence. An evidence base built entirely on publisher-funded studies, with nothing independent behind it. No defined proficiency endpoint, so there's no way to know when the job's done. A "structured literacy" label slapped on a program with no explicit phonics sequence underneath it. And adaptive claims with no explanation of what, specifically, the system is adapting. Any one of these should slow a parent down.

No single program wins across every child, every household, every situation. That's not a hedge, it's the actual shape of the evidence. The six-question framework is the durable part. Run any program through it, physical or digital, parent-led or classroom-built, and what survives is worth serious consideration. What doesn't survive it wasn't worth the shelf space to begin with.

Sources

  1. osse.dc.gov
  2. 2024-2025 Kindergarten Through Grade 4 Literacy Report
  3. Structured Literacy
  4. wilsonlanguage.com
  5. bartonreading.com
  6. 5 Best Structured Literacy Programs Reviewed
  7. smarterlearningguide.com
  8. wilsonlanguage.com

More in Science of Reading