Karma YogiKarma Yogi
Start
The Journal
CAT Preparation

10 September 2026
11 min read

A
Anish Guruvelli
Karma Yogi

Spaced Repetition for CAT Revision: What 317 Experiments Say

The meta-analysis behind spaced repetition covers 317 experiments and 14,811 people, and it says less about revising a CAT mock than every study-tips post assumes. Here is what it does say, what it does not, and the schedule I actually run.

Revise a mock more than once, and put days between the passes rather than hours. The best evidence for that is Cepeda, Pashler, Vul, Wixted and Rohrer's 2006 meta-analysis in Psychological Bulletin, which pooled 14,811 participants and found spaced study produced 47.3% correct on the final test against 36.7% for massed study. In my own log of 20 CAT mocks taken between 10 May and 8 August 2026, the median gap between two mocks was 6 days - but 7 of the 19 gaps were two days or less, and those clustered attempts are the ones that taught me nothing.

That is the honest headline. What follows is the rest of it, including the part where the same meta-analysis fails to support the expanding-interval schedule that every spaced-repetition app on the market - including the one I built - actually uses.

What the meta-analysis actually found

The paper is Cepeda et al. (2006), "Distributed practice in verbal recall tasks: A review and quantitative synthesis". Its own description of its scope: "This review found 839 assessments of distributed practice in 317 experiments located in 184 articles." It is the reason anybody in education says the words "spacing effect" with confidence.

Their Table 1 compares massed against spaced presentation, split by how long the gap was between studying and the final test - the retention interval. This is the table that matters, and almost nobody who cites the paper reproduces it, because the middle of it is inconvenient.

Retention interval Massed Spaced Studies Participants Significance
1-59 seconds41.2%50.1%965,086p < .001
1 min to under 10 min33.8%44.8%1176,762p < .001
10 min to under 1 day40.6%47.9%10870p = .535
1 day32.9%43.0%151,123p = .249
2-7 days31.1%45.4%9435p = .190
8-30 days32.8%62.2%6492p < .05
31 days or more17%39%143-
All intervals36.7%47.3%25414,811p < .001

Three things fall out of that table that you will not read anywhere else in Indian exam-prep writing.

1. The overall effect is large and extremely well evidenced. A 10.6 percentage-point gap across 254 studies and nearly 15,000 people is not a marginal finding. If you take nothing else from this post, take that spacing is real.

2. The effect is best evidenced at the two ends and thinnest in the middle. Look at the rows for 1 day and 2-7 days: p = .249 and p = .190. Those are not statistically significant. The raw gaps look big - 32.9 against 43.0, and 31.1 against 45.4 - but only 15 and 9 studies existed in each bin. The 8-30 day row, where spacing nearly doubles performance (32.8% to 62.2%), rests on six studies and 492 people.

3. The bins that carry the statistical weight are the ones irrelevant to you. More than 11,000 of those 14,811 participants sat in the two rows where the final test came less than ten minutes after studying. A CAT aspirant's retention interval is measured in months.

The one rule the paper does support strongly

Cepeda et al.'s central claim is not "space your revision by N days". It is that the optimal gap grows as the interval you need to remember over grows. Their own summary of the practical position is unusually blunt:

"After more than a century of research on spacing, much of it motivated by the obvious practical implications of the phenomenon, it is unfortunate that we cannot say with certainty how long the ISI should be to optimize long-term retention."

They go on: for most practical purposes the retention interval is months or years, "so the optimal ISI will likely be well in excess of one day."

For a CAT aspirant that translates into something usable. If you sit a mock in June and the exam is in late November, your retention interval is roughly five months. Nothing in this literature tells you the optimal gap for a five-month retention interval - the entire evidence base for retention intervals longer than a month is one study with 43 participants. What it does tell you is that a gap measured in hours is almost certainly too short, and that a gap measured in days is more defensible than one measured in minutes.

Where I have to argue against my own product

Karma Yogi runs revision rounds on an expanding schedule. R0 is the mock itself, then R1 three days later, R2 seven days after R1, R3 fourteen days after R2, and so on out to R7. Each interval runs from when you completed the previous round, not from the mock date.

Round Gap from previous round Days after the mock, if you never slip
R0The paper itself0
R13 days3
R27 days10
R314 days24
R421 days45
R530 days75
R645 days120
R760 days180

That expanding shape is the standard design across every spaced-repetition tool. Cepeda et al. do not support it over a fixed schedule. Their Table 8 compares expanding against fixed intervals across 18 studies and 1,518 participants: 62.0% for expanding against 58.6% for fixed, t(42) = 0.5, p = .61. Not significant. Their assessment of the wider literature is harsher still:

"Some researchers have suggested, with little apparent empirical backing, that expanding inter-study intervals improve long-term learning... Our review of the evidence suggests that, in general, expanding intervals either benefit learning or produce effects similar to studying with fixed spacing."

They name the software vendors directly, noting that this absence of evidence "has not stopped some software developers from assuming that expanding study intervals work better than fixed intervals," and citing SuperMemo's universal formula as the example.

So: I ship an expanding schedule, and the strongest meta-analysis in this literature says expanding is not demonstrably better than fixed. I keep it for reasons that are practical rather than empirical - an expanding schedule spends less of your attention on a mock you have already squeezed dry, which matters when you are running twenty of them - but you should know that the schedule's shape is a design choice and only its existence is evidence-backed. Anyone selling you a specific interval sequence as science is going beyond what Cepeda et al. found.

What my own 20 mocks show about gaps

Twenty mocks between 10 May and 8 August 2026 gives 19 gaps. Here is the distribution.

Gap between consecutive mocks Count What it produced
1-2 days7Includes the 96 on 29 July and the 40 on 30 July
5-6 days4Enough room for one analysis pass
7 days8The weekly coaching cadence, and the useful one

Median gap: 6 days. Mean: 4.7 days, pulled down by the clusters. The seven gaps of two days or less are the ones I would remove if I ran the quarter again - not because sitting a mock is harmful, but because a mock taken 48 hours after another one cannot be revised in between, so it measures your state rather than teaching you anything. My clearest example is public: I scored 96 on 29 July and 40 the next day, with 2 marks in DILR. Nothing about my knowledge changed overnight. The full log is in my 20-mock post.

What I would actually do

Grounded in the above, not in vibes:

  • Two revision passes minimum per mock, days apart. The first pass is still emotionally attached to the score. Spacing is what lets the second pass see something new.
  • Put the first pass 2-3 days out, not the same evening. The evening pass is massed practice with extra steps. Cepeda's data cannot resolve 1 day from 3 days, so this is a judgement call, but it is on the right side of the one thing the paper is confident about.
  • Space the later passes further apart than the earlier ones, and do not believe anyone who tells you the exact numbers matter. Fixed would probably work as well. Expanding is easier to sustain.
  • Do not sit two mocks inside 48 hours unless you are deliberately simulating back-to-back pressure. Seven of my nineteen gaps broke this rule and none of them produced a usable signal.
  • Count revision passes, not mocks. Twenty mocks revised once is worse than twelve revised three times. See how many mocks are actually enough.

Where this evidence does not reach

I would rather lose your click than overstate this, so here is the full list of reasons the numbers above should not be treated as a CAT prescription.

  • The materials are word lists, not DILR sets. The title says it: "distributed practice in verbal recall tasks". Paired associates and list recall. Learning that a spaced word list is remembered better does not establish that a spaced re-analysis of a logical-reasoning set improves your set selection under a 40-minute clock.
  • 85% of the data are young adults in a laboratory. Cepeda et al. state this explicitly. CAT is sat by roughly 2.5 lakh Indian graduates in a three-hour proctored window with sectional time locks and negative marking. The mechanism plausibly transfers; the 10.6-point gap absolutely does not transfer as a number.
  • The retention intervals that match your situation are nearly unstudied. One study, 43 participants, for anything beyond 31 days. The authors say new studies at educationally relevant intervals "are sorely needed", and that was twenty years ago.
  • Expanding versus fixed is unresolved, as above, and my product picks a side anyway.
  • My own log is one candidate. Twenty mocks, one person, thirteen weeks, no actual CAT score to check any of it against - I have not sat CAT yet. It is enough to show what a clustered mock schedule looks like. It is not enough to establish that a 6-day gap beats a 3-day one.
  • Nobody has run the study you want. There is no randomised trial of mock-revision spacing in Indian competitive exams. If one exists, I have not found it, and I looked.

Sources

  • Cepeda, N. J., Pashler, H., Vul, E., Wixted, J. T., & Rohrer, D. (2006). Distributed practice in verbal recall tasks: A review and quantitative synthesis. Psychological Bulletin, 132(3). Read in full at eScholarship. All figures above are from its Tables 1 and 8 and its Discussion.
  • Mock log, 20 attempts, 10 May to 8 August 2026 - first-party, published in full in the 20-mock post.

The revision rounds described here are the ones Karma Yogi tracks against each logged mock, so the gaps are recorded rather than remembered. Whether the exact intervals are optimal is, on the evidence above, an open question.

End of essay

- Anish Guruvelli

Common questions

How many times should I revise a CAT mock?
At least twice, with days between the passes rather than hours. The first pass is still reacting to the score and reliably misses things the second one finds. Beyond three or four passes the returns fall off fast for most mocks, and your time is better spent on a fresh paper.
Does spaced repetition actually work, or is it a study-tips myth?
It works, and it is one of the better-evidenced findings in learning research. Cepeda and colleagues pooled 254 studies and 14,811 participants in 2006 and found spaced study produced 47.3% correct against 36.7% for massed study. The uncertainty is about the exact intervals, not about whether spacing helps.
What is the ideal gap between revision sessions?
Nobody knows, and the meta-analysis says so in as many words. What it does establish is that the optimal gap grows as the period you need to remember over grows. For a CAT exam months away that means gaps measured in days rather than hours, but the specific number is a judgement call.
Is an expanding schedule better than a fixed one?
Not on the available evidence. Cepeda et al. compared expanding against fixed intervals across 18 studies and 1,518 participants and found 62.0% versus 58.6%, which was not statistically significant at p = .61. Every spaced-repetition app uses expanding intervals anyway, including mine, and that is a design preference rather than a research finding.
Should I revise a mock the same evening I take it?
Do a quick pass to note what happened while it is fresh, then do the real analysis a few days later. Reviewing the same evening is closer to massed practice, and it is also the pass most distorted by how the score made you feel.
How far apart should I space my CAT mocks?
Roughly a week works, which is also what most coaching calendars do. In my own log of 20 mocks the median gap was 6 days and eight gaps were exactly 7 days, and those weekly ones were the useful ones. Seven gaps were two days or less and taught me nothing, because there was no room to revise in between.
Does the spacing research apply to CAT specifically?
The mechanism plausibly does. The numbers do not. That literature is mostly undergraduates recalling word lists in a laboratory, with 85% of the data from young adults, whereas CAT is a three-hour timed exam with sectional locks and negative marking. Treat spacing as a principle to apply, not an effect size to expect.
What should I actually revise from a mock?
The decisions, not just the answers. Which sets you picked and why, where you committed too early, which questions you should have skipped. Re-reading the solution to a question you got wrong is the least useful version of mock revision, and there is separate research on why.
Is it better to take more mocks or revise the ones I have?
Revise the ones you have, up to a point. Twenty mocks revised once each is a worse use of a quarter than twelve revised three times. A mock is only worth its analysis, and the analysis is where the spacing effect has anything to act on.
How long is the evidence base for gaps of a month or more?
Thin to the point of being nearly absent. In the Cepeda meta-analysis, every retention interval beyond 31 days rested on a single study with 43 participants. So the interval that matches a CAT aspirant preparing over six months is essentially unstudied, which is worth knowing before anyone quotes you a schedule as settled science.