Most IB students treat their IBDP past papers as a study habit rather than a finite resource—and then wonder why they’re short on realistic simulation material when the exam is close. The typical sequence is predictable: open papers early, time them, score them, repeat until supply runs out. Four phases interrupt that pattern. Phase 1 builds accuracy through untimed single-question drills; Phase 2 tests recall with mixed-topic sets that strip familiar subject cues; Phase 3 adds time pressure through timed half-papers; and Phase 4 reserves full mock conditions for the final weeks, when results are close enough to the exam to be actionable.
That sequencing has a cognitive basis. Research synthesized in the Annual Review of Psychology concludes that practicing retrieval can function as a learning event rather than just a performance check, and that spacing those attempts slows forgetting and strengthens later recall across ages, abilities, content types, and test formats. It’s broad cognitive science rather than data specific to IB exams, but the direction is clear: repeated, spaced retrieval builds the kind of active recall that IB papers test—not just familiarity with content, but the ability to produce it under pressure. The practical challenge is how to structure that retrieval work without burning through a limited paper supply before the moments that matter most.
Phases 1 and 2—Building Accuracy Before Testing Recall
In Phase 1, work through individual questions tied to topics you’re currently studying—no timer, full attention on correctness and command-term precision. In Phase 2, once those topics are covered, rotate questions from multiple subjects into unsorted sets that strip familiar cues and require genuine recall. Both phases draw from question pools rather than intact papers. The IB Questionbank—accessible through authorized schools—lets staff filter past questions by syllabus section, command term, year, and marks, then assemble custom sets; student-facing platforms such as Revision Village, which covers DP subjects at SL and HL, offer comparable question-level access. Here is how to run the whole system as a repeatable weekly routine per subject.
The 15-Minute Setup: Run the Four Phases as a Repeatable Weekly System
- Inventory your supply (5 min). Count intact unseen papers, seen/marked papers, and loose questions you can assemble. Tag each intact paper: Hold for Phase 4 (most recent, best format match) or Usable in Phase 3 (older or format-mismatched). With authorized Questionbank access, filter by syllabus section and command term for Phase 1 and 2 sets.
- Pick your phase by purpose (1 min). Still finishing content: Phase 1. Content covered, want recall without cues: Phase 2. Timing breaks performance: Phase 3. Exam close, full rehearsal needed: Phase 4.
- Run the same session loop every time. Attempt. Self-mark. Log errors by type. Write one repair action per error. Schedule a reattempt.
- Promotion rules—do not guess. Phase 1 → 2: topic-filtered questions work without notes; misses are precision or command-term errors, not missing topics. Phase 2 → 3: errors are scattered; the main limiter is speed or pacing. Phase 3 → 4: half-papers finish on time with stable pacing; errors are quality under time, not timing meltdown.
- Roll-back rule. If a timed attempt exposes a content gap, roll back to Phase 1 or 2 for that topic before returning to timed work.
- Weekly rotation. Choose two active subjects per week for Phase 2 or 3 work; keep the rest on Phase 1 drills. Rotate the active pair each week.
These steps are practical heuristics, not guarantees; adjust the promotion and rollback rules to your subject’s markscheme style, paper supply, and access to feedback.

Phases 3 and 4—Building Stamina, Then Simulating the Real Exam
The move to timed half-papers is earned rather than assumed. A workable readiness signal is consistent accuracy across mixed-topic sets, where errors are scattered and command-term-level rather than pointing to content gaps. Markscheme styles and paper structures vary by subject, so no single threshold applies universally—a limit Sparkl’s practitioner guide on past-paper use also acknowledges.
Phase 3 uses one paper section or half the marks under a fair timer. Divide total exam minutes by total marks for a minutes-per-mark baseline, then allocate subset marks Ă— baseline plus a fixed reading buffer kept identical each attempt. After the attempt, log minutes per mark achieved. If that number is high and errors cluster late, it’s a stamina issue—stay in Phase 3. If pacing is fine but marks are lost early through misreads or weak structure, roll back to Phase 2-style mixed sets with stricter command-term constraints. If the same content repeatedly stalls progress, run Phase 1 drills on that topic before returning to timed work. Some subjects have non-linear time demands, so this method standardizes comparability rather than equalizing cognitive load.
Phase 4 is full-condition simulation: no interruptions, identical timing, correct materials. Reserve it for when the exam is close enough that findings can still drive real changes—that’s what separates a useful mock from an expensive one. Keep the most recent intact papers for this phase. For subjects with few recent sessions, use older or format-mismatched papers in Phase 3 and hold current-format papers for full simulation, verifying any older paper against the current subject guide before treating it as a Phase 4 resource. Whatever the phase, a timed attempt is only as useful as what happens when you sit down with the script and markscheme afterward.
Self-Marking, Partial Credit, and the Error Log
Markschemes record how marks are awarded, not what ideal answers look like. Students who treat them as model answers miss the diagnostic signal in partial credit—specifically, which mark point they failed to satisfy and why. A practitioner guide from StudyIB makes this concrete: keep the original response visible while annotating in a second color, so the gap between your answer and the mark-awarding criterion stays specific rather than abstract.
Raw scores tell you where you finished, not what stopped you. An error log organized by type—knowledge gap, method error, command-term misread, communication failure, timing—reveals recurring patterns across attempts in ways a score alone cannot. A repair action and reattempt date per logged error convert the log from a record of what went wrong into a schedule of what to do next.
Error Log → Next Actions: A 10-Minute Weekly Review That Actually Changes Your Practice
- Setup — after every attempt: for each missed or partial question, log the error type, one-line evidence (the exact markscheme point or criteria band you did not satisfy), one repair action, and a reattempt date.
- Review cadence — once per week, 10 minutes: tally your most recent 10–20 logged errors by type and identify the top two.
- Knowledge or method errors in the top two: schedule 2–3 untimed Phase 1 drills on that topic, then one Phase 2 mixed set that includes it.
- Command-term, communication, or structure errors in the top two: redo three similar questions under a strict answer structure (define → apply → justify), then re-mark only for the specific mark points you missed.
- Timing or stamina errors in the top two: schedule one Phase 3 half-paper and note where overruns begin; do not add a full mock until overruns shrink consistently.
- Reattempt rule: on the scheduled date (typically 7–14 days later), redo the same question or a close variant and mark only for previously missed points. If the same error type appears twice in succession, seek teacher review for judgment-based items or return to targeted drills before increasing timed volume.
A log maintained over weeks rather than days reveals patterns a single attempt can’t show—but the log stops producing usable data the moment there’s no remaining paper supply to generate new attempts.
Sourcing Papers, Managing Supply, and Knowing When a Paper Is Spent
Past papers reach students through IB-authorized channels—the IB store, school coordinator access, and teacher-distributed sets—and supply is easier to audit than it is to replace. Confirm with your DP coordinator which sessions are accessible before mapping your phase timeline. There is no universal year cutoff for syllabus currency across all DP subjects; the reliable check is the current subject guide. When paper structure, command terms, or topic scope diverge from the guide, demote those sessions to Phase 1 or 2 question-level drills or Phase 3 timing practice rather than holding them for Phase 4 simulation.
Rotate active practice across subjects on a weekly or biweekly basis so no single subject exhausts its supply ahead of the rest—perfect weekly balance is not the goal, just parallel progress over time. A paper is spent as a simulation once you’ve seen and marked it—but individual questions with logged errors remain usable for Phase 1 or 2 reattempts weeks or months later, and knowing exactly which questions and papers remain is what makes the phase decision for each subject a deliberate choice rather than an improvised one.
Starting Your Four-Phase Past-Paper Plan
Phases 1 and 2 don’t just build content knowledge—they determine whether Phase 3 pacing data is meaningful and whether Phase 4 scores arrive early enough to drive real adjustments. Students who compress or skip those early phases can reach full mock conditions with a score and no remaining paper supply to do anything with it. That’s the real cost of missequencing: not just a hard result, but an irreversible one.
The starting action is immediate: identify which phase fits your current position in each subject, confirm with your coordinator which sessions are available, and begin with the lowest supply cost—Phase 1 single-question drills. Open a full paper before the system earns it and you’ve spent an irreplaceable diagnostic on a day when you weren’t ready to use what it told you.

