How Often Should a Learner See a Word They Keep Forgetting?
THE ZINUK LAB · LEARNING SCIENCE
Lessons from scheduling 2,725 words in a production vocabulary app: why cramming feels right and fails, when to first test a new word, and what to do with a word that refuses to stick.
Every cohort has this student. He decides that ubiquitous will not defeat him: he reads it, stars it, drills it — eleven times in one evening. The next morning he opens the app, sees the word, and feels his stomach drop. Nothing. As if they had never met.
Anyone who has crammed vocabulary for an exam knows that feeling, and most learners draw exactly the wrong conclusion from it: "I need to repeat it even more." Massed repetition feels like learning because the word becomes instantly available — but that availability lives in working memory, not long-term memory. Cramming is the most convincing way ever invented to feel like you are learning without actually learning.
We build test-prep apps at Zinuk, an Israeli exam-preparation school. One of them teaches vocabulary for Amirnet (אמיר"ם) — Israel's computerized English proficiency exam for university admission — and its word bank holds 2,725 words, each running on its own private schedule. When we designed that scheduler, the hard question was never which words to teach. It was when: when to show a word for the first time, when to test it, and what to do the moment a learner gets it wrong.
So how often should you see a word you keep forgetting?
Fewer times than you think — at intervals that keep growing. A calm first introduction with no quiz, a first test only after a real delay, then reviews at expanding intervals: a day, several days, weeks. An error cuts the interval sharply; success stretches it. And when the review queue swells, the system stops feeding you new words until the debt is paid.
What the science has known for 140 years
Ebbinghaus published the forgetting curve in 1885, and it is still the most important chart in education: memory decays steeply right after learning, and the decay flattens with every successful review. The practical lesson is not "review a lot" — it is review at the right moment: just before the word slips away, not just after you met it.
The second finding is more surprising: the test itself is the learning. Roediger and Karpicke (2006) showed that attempting to retrieve something from memory improves long-term retention more than re-reading the same material — even though re-reading feels easier and safer. Karpicke and Roediger (2008), in Science, pinned down retrieval as the critical ingredient: retrieval practice wins even when it gets less total study time.
What about the exact shape of the intervals? Here the literature is humbler than the marketing around it. Karpicke and Roediger (2010) compared expanding schedules against fixed ones and found a smaller difference than everyone expected. The lesson we took is modest and useful: the big win is spacing plus retrieval. The precise curve is fine-tuning — and fine-tuning should be done on real data, not on ideology.
Two more findings shaped our system. Ellis and Beaton (1993) showed that recognizing a foreign word (L2 → your language) is substantially easier than producing it (your language → L2); they are different skills, and the hard one builds on the easy one. And Kornell (2009) showed that spacing beats cramming even in humble flashcard practice — these effects survive contact with the real world.
What we built — and why it isn't "Anki with nicer colors"
An introduction is not an ambush
A new word is shown once, calmly, with its translation and context — and it is never quizzed immediately afterwards. The reason: immediately after exposure, everyone "remembers". That is the illusion of knowing — confidence peaks at exactly the moment the measurement is worthless. An instant quiz would teach the system that the word is known and teach the learner that he is a genius. Both would be wrong. The first real test comes only after a real break.
Recognition first, recall later
Following Ellis and Beaton, every word climbs a ladder: first recognition questions ("what does ubiquitous mean?"), and only after it has accumulated successful reviews — production questions ("how do you say 'found everywhere' in English?"). A successful production attempt also earns the schedule more credit than a successful recognition: it is stronger evidence of knowledge, so the next interval grows more. Learners experience this as natural gradual difficulty; under the hood it is a ladder of evidence.
Intervals stretch — errors cut
A word answered correctly again and again drifts away: a day, several days, then a multiplier that grows with every success. A forgotten word snaps back hard. And there is one special case we learned to respect: failing a word's first test — a sign that the introduction simply never landed — brings the word back within the same session, not tomorrow. There is no point waiting a day to rediscover what you already know: that word needs one more meeting now.
The "while it's hot" requeue
A word missed mid-session doesn't vanish until the next study day. It re-enters the queue — a few cards back, not immediately. Immediately would be working memory: you would answer correctly and learn nothing. A few cards of distraction are enough to make the second retrieval a real retrieval. At the end of the session, every missed word gets one more closing round. Learners know the feeling as "wait, you again?" — that is not a bug; it is the feature.
The load gate: sometimes the right number of new words is zero
The least popular rule in the system: when the review queue swells past a threshold, new words stop — even if the learner is eager. A postponed review queue is precisely the set of words on the edge of being forgotten; adding new words on top is pouring water into a leaking bucket. Even on a normal day, new words per session are capped at a single-digit number. Progress feels slower; it is faster.
A tuning story: ease hell
We did not invent the scheduling core. Our starting point was the SM-2 algorithm by Piotr Woźniak (SuperMemo) — the workhorse of spaced repetition for thirty years. But anyone who runs textbook SM-2 on exam students quickly meets a phenomenon the Anki community calls ease hell: each item carries an ease factor that drops with every lapse, and in the canonical version its floor is 1.3. A genuinely hard word grinds down to that floor — and then returns again and again, its interval barely growing, forever.
What we actually saw: a handful of stubborn words took over entire sessions. A student with twenty "stuck" words opened every practice day against the same wall. Progress stalled — and motivation with it. The system was statistically right and pedagogically cruel.
We changed two things — directionally; the exact values stay ours. First, we raised the floor: the canonical 1.3 punishes too hard, and a difficult word in our system keeps its ability to climb out of the pit. Second, we built a rescue mechanism: a word that keeps failing in a row stops accumulating punishment — the system forgives its history and restarts it as if it were brand new. Because at that point, that is exactly what it is.
When is a word "mastered"?
The simple definition — "answered correctly twice" — fails exactly where it matters: it confuses a short streak with durable knowledge. In our system a word is marked mastered only when two conditions hold together: a streak of successes and durability over time — a successful retrieval after an interval of weeks. A word that passed both almost never comes back. A word that passed only one is still on the job.
If you study on your own: four rules
- Never quiz yourself right after studying. Give new words a few hours — better, a night — before the first test. The confidence of "I've got it now" is noise, not measurement.
- Start with recognition, finish with production. First "what does this word mean," and only once that sits — "how do you say X in the target language." The hard direction is the goal; the easy direction is the road there.
- What you miss returns today — but not immediately. Got a card wrong? Put it a few cards back in the stack and give it one more round at the end. Not tomorrow — today, after a short distraction.
- No new cards while you're in debt. If reviews have piled up, they come first. A day without new words is not a wasted day; it is the day yesterday's learning became permanent.
All of these mechanisms — the recognition-to-production ladder, expanding intervals, the while-it's-hot requeue and the load gate — run in production inside Zinuk's Amirnet course, scheduling each of 2,725 words independently. The Hebrew original of this article, with the rest of our learning-science series, lives at The Zinuk Lab.
References
- Hermann Ebbinghaus (1885). Memory: A Contribution to Experimental Psychology. psychclassics.yorku.ca/Ebbinghaus/index.htm
- Paul Pimsleur (1967). A memory schedule. doi.org/10.1111/j.1540-4781.1967.tb06700.x
- Piotr Woźniak (1990). SuperMemo SM-2 algorithm. supermemo.com/en/blog/application-of-a-computer-to-improve-the-results-obtained-in-working-with-the-supermemo-method
- Ellis & Beaton (1993). Psycholinguistic determinants of foreign language vocabulary learning. doi.org/10.1111/j.1467-1770.1993.tb00627.x
- I. S. P. Nation (2001). Learning Vocabulary in Another Language. wgtn.ac.nz/lals/about/staff/paul-nation
- Koriat & Bjork (2005). Illusions of competence in monitoring knowledge during study. doi.org/10.1037/0278-7393.31.2.187
- Roediger & Karpicke (2006). Test-enhanced learning: Taking memory tests improves long-term retention. doi.org/10.1111/j.1467-9280.2006.01693.x
- Cepeda, Pashler, Vul, Wixted & Rohrer (2006). Distributed practice in verbal recall tasks: A review and quantitative synthesis. doi.org/10.1037/0033-2909.132.3.354
- Karpicke & Roediger (2008). The critical importance of retrieval for learning. doi.org/10.1126/science.1152408
- Nate Kornell (2009). Optimising learning using flashcards: Spacing is more effective than cramming. doi.org/10.1002/acp.1537
- Karpicke & Roediger (2010). Is expanding retrieval a superior method for learning text materials?. doi.org/10.3758/MC.38.1.116
- Bjork & Bjork (2011). Making things hard on yourself, but in a good way: Creating desirable difficulties to enhance learning. bjorklab.psych.ucla.edu/research/