How the test works
AurorIQ is an adaptive cognitive assessment built on item response theory (IRT), specifically a three-parameter logistic (3PL) model. Rather than asking everyone the same fixed set of questions, the engine selects each item based on your live performance: it estimates your ability after every answer using maximum-likelihood estimation, then chooses the next question to extract the most additional information at that ability level, while making sure all five reasoning domains get even coverage across the test.
Your final score is computed from a single ability estimate across all 25 responses, converted to the familiar IQ scale (mean 100, standard deviation 15) and reported alongside a percentile and a 95% confidence range, because a single point estimate without a range overstates precision a 25-item test cannot honestly provide.
An honest note on calibration: the difficulty, discrimination, and guessing parameters behind each question are currently expert-seeded estimates, not values derived from a large set of real response data. We say this plainly because IRT scoring sounds more authoritative than it is when the underlying item parameters haven't been empirically calibrated yet. As AurorIQ collects more anonymized responses over time, we intend to re-estimate these parameters from real data and will update this page when that happens.
What this test does and doesn't measure
AurorIQ measures performance across five reasoning domains: pattern recognition, numeric reasoning, verbal reasoning, spatial reasoning, and working memory. These map loosely onto the broader-abilities tradition in cognitive science (see Research Foundations below), and your composite score reflects a weighted combination of how you performed across all five.
What it doesn't measure: creativity, emotional intelligence, practical wisdom, motivation, domain-specific expertise, or anything resembling a complete picture of a mind. No 15-minute test, clinical or otherwise, captures all of that, and we'd rather say so directly than imply otherwise with confident-sounding numbers.
Archetypes and rarity
Your cognitive archetype is determined by the shape of your profile, which domains are strongest relative to your own other domains, rather than your absolute score. We lead with this qualitative pattern rather than precise per-domain numbers because each domain is assessed with roughly five questions, which carries meaningfully more statistical uncertainty than the 25-item composite score. Treat per-domain results as directional strengths, not precise sub-scores.
Rarity percentages shown alongside your tier and archetype are model-based estimates from a simulation, not measurements from AurorIQ's actual user base yet. We'll replace these with real population statistics once enough anonymized results have been collected.
Research foundations
AurorIQ's design draws on published frameworks in psychometrics and cognitive science. The researchers and works listed below are cited because their published work informs how we think about measuring cognitive ability, not because they are affiliated with, employed by, or have endorsed AurorIQ in any way.
Important: none of the researchers or institutions named below are affiliated with AurorIQ, were consulted in building it, or endorse it. They are cited solely for the published scientific frameworks referenced throughout this page and our guides.
John Carroll
Three-stratum theory of cognitive abilities, the basis for the broader CHC framework underlying multi-domain ability models like the one AurorIQ uses.
Richard Haier
Research on the neuroscience of intelligence, including work on neural efficiency and the biological correlates of reasoning ability.
Robert Sternberg
The triarchic theory of intelligence, and an influential body of work arguing that conventional IQ tests capture only part of what intelligence means in practice.
Michal Kosinski
Research on computational approaches to psychometrics, relevant to measuring cognitive and behavioral traits from structured response data.
Lord & Novick
Foundational statistical text establishing modern item response theory, the mathematical basis for AurorIQ's adaptive scoring engine.
Embretson & Reise
A widely used applied reference for item response theory in psychological measurement, including the 3PL model AurorIQ implements.
Limitations
- Not a clinical instrument. AurorIQ cannot be used for clinical diagnosis, educational accommodation decisions, legal proceedings, or Mensa admission. Only a proctored assessment administered by a licensed psychologist using a validated instrument, such as the WAIS-IV or Stanford-Binet 5, qualifies for those purposes.
- Item parameters are not yet empirically calibrated. As noted above, our IRT parameters are expert-seeded starting estimates.
- Per-domain scores carry high uncertainty. With roughly five items per domain, treat domain results as directional, not precise.
- Rarity figures are simulated, not measured. They reflect a statistical model, not AurorIQ's real user population yet.
- Day-to-day factors affect scores. Sleep, stress, environment, and familiarity with the test format can shift results by several points in either direction.
- One test, one moment. A single 25-question session is a snapshot, not a complete or permanent profile of a mind.
Who built this
AurorIQ is built and maintained by Muhammad, a developer with a background in data-driven tools and a genuine interest in psychometrics and cognitive measurement. The project started from a simple question: can an online IQ test be both technically sound and completely honest about its limitations? Most existing tools either inflate scores to flatter users or hide their methodology behind vague claims. AurorIQ is designed to do neither.
The test engine, scoring algorithm, adaptive item selection, archetype system, and all editorial content on this site are independently produced. No third party has editorial influence over our methodology, our content, or how we report results. The methodology described on this page has been reviewed against the published sources cited above.
If you find an error, have a correction, or want to discuss any aspect of the methodology, we genuinely want to hear about it — reach out at contact@auroriq.com or visit our contact page.
Real proctored cognitive and IQ assessments, administered and interpreted by a licensed psychologist, typically range from several hundred to over a thousand dollars depending on the instrument and region. AurorIQ is free and exists for curiosity and self-reflection, not as a substitute for that process.