Rosh Review and Blueprint Prep ABFM Workflow: When to Follow the Algorithm, Override It and Move to Mixed Blocks

Featured image for Rosh Review and Blueprint Prep ABFM Workflow: When to Follow the Algorithm, Override It and Move to Mixed Blocks

This is an implementation guide for family-medicine candidates using Rosh Review (part of Blueprint Prep) for the ABFM one-day certification exam. It assumes you have the bank and want a decision rule for each week, not another comparison table. The principal limitation it manages: any feed that steers you toward missed material, and a predicted-pass indicator built on that behaviour, can leave low-volume blueprint domains under-practised while your headline metrics look healthy.

What Rosh Review and Blueprint Prep offer for ABFM right now

ItemCurrent state (vendor-reported, checked 19 July 2026)
ABFM coverageYes — Family Medicine Certification Exam Qbank, updated for the 2025 ABFM blueprint
BrandingRosh Review is part of Blueprint Prep
Question countapprox. 3,000 ABFM-formatted questions (vendor-reported)
Adaptive/AI featuresPerformance analytics, a predictive "likelihood of passing" indicator, custom filtered practice, tutor/test modes (vendor-reported)
Access period1-year plans; 90-day and 30-day options (verify shorter-period pricing)
Price1-year Basic approx. $549 / Standard approx. $649 / Premium approx. $748 (vendor-reported, 19 July 2026)
CME100 AMA PRA Category 1 (Standard/Premium); AAFP Prescribed credit (vendor-reported)
Pass guarantee100% Pass Guarantee — complete the bank, do not pass, refunded; score within 60 days (vendor-reported)

All figures vendor-reported on 19 July 2026. Verify before purchasing, especially the guarantee's completion condition.

The ABFM exam anchor

Three hundred single-best-answer MCQs, four sections of 75 at 95 minutes each — about six hours twenty of testing plus roughly 100 minutes of pooled breaks; a completed section is locked. The 2025 domains of care: Acute Care and Diagnosis 35%, Chronic Care Management 25%, Emergent and Urgent Care 20%, Preventive Care 15%, Foundations of Care 5%. Rosh Review's subject tags map onto these domains; when they disagree, the official weighting is the standard.

A note on pacing: why the section-lock changes your plan

The exam's four locked sections reward even pacing and punish the instinct to "save" hard items for a second pass — there is no second pass once a section closes. Build this into how you use Rosh Review's test mode: practise in section-length units of 75 items in 95 minutes, with a rule that anything still unresolved at about 90 seconds gets your best answer and a flag you will not return to within that block. The aim is to make finishing-in-time automatic, so that on the day your cognitive budget goes on the medicine rather than on clock management. Short tutor sets have their place for learning, but they quietly train the opposite habit — unlimited time per item — so ring-fence timed, section-length blocks as the exam approaches.

Baseline week: an unseen, blueprint-stratified sample first

Before you lean on the feed or the predicted-pass figure, sit a stratified unseen sample: roughly 75–100 items spread across the five domains by weighting, timed, mixed, no lookups. Record first-attempt accuracy per domain. This baseline is the honest reference the dashboard will soon paper over as repeats accumulate. Record those per-domain numbers somewhere outside the app before Rosh Review's analytics blend them into a running total; that raw map is the thing you will compare against in three weeks, and if you let the dashboard overwrite it you lose the only honest before-picture you will get.

First pass: set domain floors

Convert the blueprint into weekly floors the feed cannot override — for instance, a minimum of 30 first-attempt Emergent and Urgent Care items and 15 Foundations of Care items each week. Because Rosh Review lets you build custom sets by subject, count and mode, you can enforce these directly rather than hoping the reinforced feed raises them. The rule: follow the algorithm for efficiency, but never let it drop a domain below floor. Because Rosh Review builds custom sets by subject, count and mode, enforcing floors is mechanical — schedule the weak-domain custom set first each week, before any free-choice practice, so the minimum is banked before motivation and the feed pull you back toward what you already do well.

An error taxonomy worth keeping

Tag every miss with one cause: knowledge gap, misread stem, premature closure, guideline error, calculation error, or time-pressure error. Rosh Review's explanations are strong, which makes it tempting to read the rationale and move on; resist that. The tag, not the rationale, tells you whether next week needs more content, a pacing rule, or a process change.

Review interval: not everything deserves an immediate repeat

Knowledge gaps get a short source read and a fresh transfer question days later. Misread-stem and premature-closure errors get a process rule and spaced retesting on new stems. Guideline errors get the current source. Immediate repeats of the same item mostly manufacture recognition and inflate repeat accuracy — use them sparingly. Space the misses; measure on new items.

Reading your error-tag tally

The point of tagging every miss is the pattern the tags make over several weeks, not any single wrong answer. A tally dominated by knowledge gaps says your content is still immature and filtered drilling is the right medicine. A tally dominated by misread stems and premature closures says the opposite — you know enough, and you are losing marks to reading speed and anchoring, which more content will not fix. A cluster of guideline errors in one domain points to a specific out-of-date mental model worth correcting at source. A run of calculation errors says build a calculation floor and slow down on numerical items. And a wall of time-pressure tags is the clearest signal of all that you have left timed practice too late. Rosh Review's explanations are strong enough that it is tempting to close each miss by reading the rationale and feeling satisfied; the tally is what stops that comfortable loop and forces a change in what you actually do next week.

Mixed-block switch: the objective criteria

Move from topic-filtered practice toward timed random blocks when all three hold: every domain is at or above floor; first-attempt accuracy has stopped climbing on filtered sets; and per-item time is near 76 seconds. The predicted-pass label is not one of these criteria — it is a modelled trend, not a trigger. Filtered practice builds; mixed, timed blocks prove and train switching. A useful tell that you have switched too late rather than too early is a large gap between your filtered first-attempt accuracy and your mixed first-attempt accuracy: if you score well on single-topic sets but drop sharply on random mixed blocks, the missing skill is cross-domain switching, and only mixed practice closes it — so track both numbers separately and watch the gap.

Exit criteria that are not "I finished the bank"

You are ready to taper when coverage floors are met in all five domains, first-attempt accuracy is stable across two unseen mixed blocks, pacing holds, spaced misses stay corrected, and you have calibrated against official ABFM material. Completing all 3,000 items — the pass-guarantee condition — is a commercial threshold, not a readiness one; treat the two separately. One further exit test is worth applying: can you explain, without the options in front of you, why the right answer beats the second-best option on a sample of items? Single-best-answer questions turn on exactly that discrimination, and a candidate who can only recognise the key when the distractors are present has not yet reached the exit standard, however complete the bank looks.

Worked example: a seven-day plan

Rosh Review does content and weak-area work; iatroX supplies the unseen, timed measurement and Socratic rework its own dashboard cannot. Two-Q-bank rule, no algorithm claims.

  • Monday — Rosh Review: timed 40-item block, weakest domain; tag misses.
  • Tuesday — Rosh Review: 40-item block, second-weakest domain; explanations on knowledge-gap misses only.
  • Wednesday — iatroX: fresh, unseen, timed 50-item mixed ABFM block; no lookups.
  • Thursday — Rosh Review: custom floor set — Emergent and Urgent Care, images, calculations, Foundations of Care.
  • Friday — Rosh Review: spaced retest of last week's misses on new items.
  • Saturday — iatroX Socratic Tutor: rework 8–10 misses to the discriminating feature.
  • Sunday — rest or one timed mixed block; update the blueprint-coverage matrix.

Continue, supplement, switch or stop

Continue (learn) while filtered blocks still raise first-attempt accuracy and any domain is below floor. Supplement (retest on unseen) when the dashboard and predicted-pass look strong but nothing independent confirms them. Switch (simulate) to mixed, timed, full-section mocks once floors are met and pacing is stable. Stop only when remaining items duplicate skills already proven on unseen blocks — not because of the pass guarantee's completion clause or a rival's marketing.

Three mistakes this workflow is designed to stop

First, following the feed so faithfully that Emergent and Urgent Care and Foundations of Care stay thin until the final weeks. Second, treating the predicted-pass indicator as a green light and skipping independent unseen measurement. Third, chasing 100% completion for the guarantee at the expense of even blueprint coverage — a full bank with a weak domain is still a weak domain.

Frequently asked questions

Is Rosh Review and Blueprint Prep enough for ABFM on its own? As content, a ~3,000-item, blueprint-updated bank can carry most of your revision. As a readiness instrument it is incomplete, because coverage, repeats and the predicted-pass figure are all internal to the product. Corroborate with an unseen, timed source.

Which ABFM component does Rosh Review and Blueprint Prep not reproduce well? The four-section, section-locked, ~six-hour-twenty exam day, and the discipline of never revisiting a closed section. Filtered practice and a predicted-pass label do not train that stamina or switching; the label is a vendor model, not an ABFM outcome.

How many Rosh Review and Blueprint Prep questions should I complete per day for ABFM? No official number. For most, 40–60 timed items reviewed thoroughly per working day is sustainable, rising near the exam. Your access window and price (vendor-reported, 19 July 2026) set the runway; plan volume against that rather than a completion target.

When should I stop using Rosh Review and Blueprint Prep and move to mixed mocks? When all five domains are at floor, first-attempt accuracy has plateaued on fresh items, and pacing sits near 76 seconds. Then mixed, timed sections outperform further filtered drilling.

How should I combine Rosh Review and Blueprint Prep with iatroX without duplicating practice? Give them separate jobs: Rosh Review for content and drilling, iatroX for unseen, timed measurement and Socratic rework. Never re-attempt a seen item in the other bank; measure only on fresh questions so recognition does not contaminate the read.

Editorial notes and references

Written by Dr Kolawole Tytler, NHS GP and founder of iatroX. Last checked 19 July 2026. Vendor-reported figures (counts, prices, CME, guarantee terms) were accurate to the product pages on that date and change without notice — verify the current figure on the product page. Disclosure: iatroX operates a competing question bank; it is confined here to unseen readiness measurement and Socratic rework, jobs Rosh Review's dashboard cannot perform on itself, and is not offered as a replacement for the primary bank. Corrections via the feedback route on iatrox.com.

References: American Board of Family Medicine — One-Day Exam and 2025 Exam Blueprint (theabfm.org); Rosh Review Family Medicine Certification Exam Qbank (roshreview.com/fm/certification-exam/); iatroX ABFM bank (https://www.iatrox.com/abfm-family-medicine); "Your Q-Bank Percentage Is Not Your Exam Score" (https://www.iatrox.com/blog/qbank-percentage-not-your-exam-score); and the two-Q-bank rule (https://www.iatrox.com/blog/the-two-q-bank-rule-how-to-add-a-second-bank-without-duplicating-questions-or-destroying-calibration).

Run a fresh, timed ABFM block in iatroX →

Share this insight