Under the hood is a real student-modeling engine — the same kind of math the SAT and NWEA MAP run on — built specifically for the way your child actually answers questions.
A single score hides four different stories. We read all four separately.
The skill is automatic — they're not thinking hard about it anymore.
Invisible to grade-book apps — and a real risk on a timed exam.
Not a guess — they're fluently applying the wrong rule.
They searched for a method and came up empty.
Only one of these four is really addressed by "assign more questions on the weak topic" — which is what most apps do to all four.
Because we write our own questions, every wrong option is tagged, at creation, with the exact misconception that produces it.
⅓ + ⅙ = ?
⅓ + ⅙ = ?
Same right answer, two students, two completely different next steps — because we know why each of them was wrong, not just that they were.
Learning-science research calls repetitive drilling on a weak topic "wheel-spinning" — kids who grind the same skill without real progress usually don't improve with more of the same. So every mission mixes four kinds of practice instead.
Every mission blends all four — weighted toward growth, never all one kind.
A roster of specialists, each with one job, and a few of them deliberately trying to catch each other out. Teaching a kid math well was never going to be simple, so we stopped building as if it were.
Two different AI families work every problem blind — separately, with no idea what the "right" answer is supposed to be.
A separate panel checks the wording, the reading level, and whether the hint actually helps — not just whether the math is right.
Told to believe one exact wrong idea and answer as that student would — proof a question actually catches the mistake it's built to catch.
Rewrite a hard question in simpler words and re-test it. If kids would answer differently, that's a wording problem, not a math one — fixed before it ships.
Remembers every mistake this bank has ever made and feeds the lesson back into how the next question gets written.
Checks the model's own difficulty judgment against real results from real kids on real published exams — never just its own opinion.
"You believe fractions add straight across — tops with tops, bottoms with bottoms. Answer ⅓ + ⅙ as that student would."
None of this is improvised. It's grounded in decades of published testing and learning-science research — the same item-response math the SAT runs on, the hint-partial-credit studies from Worcester Polytechnic's ASSISTments project, and the "wheel-spinning" finding that more of the same question doesn't help a stuck kid, which is why ours never just does that.
Instead of one score, the model holds a range — how sure it is about what your child knows. Twenty questions in, that range narrows fast. It never mistakes one lucky guess (or one bad day) for the whole story.