Skip to content

Methodology

How Military Entrance Exams makes questions, checks them, and turns your answers into score estimates — so you can judge for yourself how much to trust each number.

Why there are no real test questions here

Real AFOQT, SIFT and ASTB-E questions are controlled test material. Using them would be test compromise — and the Air Force's own manual bars anyone who has taken the AFOQT from helping write a commercial AFOQT study guide (DAFMAN 36-2664, 4.12.6.1). Our questions are written without access to actual test content, from the published descriptions of each subtest. Official sample material (the AFOQT Information Pamphlet, OATTS, NMOTC samples) is used only to understand the format; it is never reproduced, and our validator rejects any question that resembles a published sample.

How a bank question is made

  1. Written from scratch for one subtest and one named technique, with an explanation and a reason for every wrong option.
  2. Checked automatically: five distinct options, no "all of the above", balanced answer positions, the key not habitually the longest option, no near-duplicates, and — for every math question — the key recomputed from an arithmetic expression.
  3. Solved blind by a separate reviewer who sees the question without the key; only questions where the independent answer matches the key, and that the reviewer passes for clarity and accuracy, go live.
  4. Logged: every question carries its provenance (author, batch, date, reviewer verdict, blind-solve match, automatic checks).
  5. Watched: the share of people who get each question right, and whether strong candidates miss it more than weak ones, is computed nightly; odd questions are flagged and pulled. Every question has a "Report an issue" link.

How the visual questions are generated

Table Reading, Instrument Comprehension and Block Counting (and the SIFT's Simple Drawings, Hidden Figures and Spatial Apperception) are produced by code. Each question comes from a seed, so it can be re-created exactly to grade an answer; the answer key is computed, not typed:

Before release, automated property tests generate at least 10,000 questions of each type and check every key against an independent calculation — for example Block Counting contacts are recounted with interval arithmetic, and the drawn artificial horizon is read back to confirm it shows the intended climb, dive and bank.

How score estimates work

The services do not publish norms (AFOQT), scoring formulas (SIFT, ASTB-E) or the PCSM algorithm, so every score number in the app is an estimate from our own model, and is labelled that way. The model:

  1. Your share correct on each subtest (lightly shrunk toward typical when you've answered only a few) is compared with an assumed typical performance for that subtest — the table below.
  2. A composite averages its subtests (using the official composite tables) and adjusts for the fact that subtests are correlated (we assume ρ = 0.5).
  3. The AFOQT result is shown as a 1–99 percentile; SIFT and OAR as 50 ± 10 on a 20–80 scale.
ExamSubtestAssumed typical share correctAssumed spread (SD)
AFOQTVerbal Analogies62%16 pts
AFOQTArithmetic Reasoning52%18 pts
AFOQTWord Knowledge55%17 pts
AFOQTMath Knowledge52%18 pts
AFOQTReading Comprehension60%16 pts
AFOQTSituational Judgment55%13 pts
AFOQTPhysical Science55%16 pts
AFOQTTable Reading62%18 pts
AFOQTInstrument Comprehension50%20 pts
AFOQTBlock Counting50%18 pts
AFOQTAviation Information50%18 pts
SIFTSimple Drawings60%15 pts
SIFTHidden Figures50%17 pts
SIFTArmy Aviation Information Test50%17 pts
SIFTSpatial Apperception Test55%18 pts
SIFTReading Comprehension Test60%15 pts
SIFTMath Skills Test55%17 pts
SIFTMechanical Comprehension Test55%17 pts
OAR / ASTB-EMath Skills Test55%17 pts
OAR / ASTB-EReading Comprehension Test60%15 pts
OAR / ASTB-EMechanical Comprehension Test55%17 pts
OAR / ASTB-EAviation & Nautical Information Test50%17 pts

These assumptions are ours. Use estimates to compare yourself with yourself over time. If you share your official scores after the test (optional, in the app), we use them only to recalibrate this table.

PCSM: the illustrative estimator weights the AFOQT Pilot composite, a TBAS self-rating and flying hours (capped at 41 per DAFMAN 36-2664; an older AFPC page says 60) and shows a range, never a single "your PCSM".

Sources

Found something out of date? Tell us — facts are re-checked whenever a source changes.