uMerit
Methodology

Where the numbers come from.

uMerit runs on public institutional data and a fixed set of rules. This page names every source, gives the count of colleges each source covers, and says what the scores can and cannot tell you.

The short version

Four numbers, four sources.

These are counts from the live database, not marketing figures. Where a number is rounded, the exact figure is next to it.

IPEDS
2,400+
colleges in the database

Every record is keyed to the institution’s federal IPEDS Unit ID. Exact count: 2,474.

Common Data Set
1,940
colleges with CDS admissions data

Admit rates, test-score ranges, GPA and class-rank distribution, filed by the college itself.

Written by us, checked automatically
23,000+
active SAT practice questions

Built to the official exam format. Some written by AI, some built from templates. Not retired official questions.

Benchmark set
697
official College Board reference questions

The reference set every generated question is measured against.

Section 1

College data: three sources, named

Three datasets sit behind every admit rate, score range, cost figure, and aid number in the product. All three are public. You can check any of them without us.

The Common Data Set

The Common Data Set is a standardized questionnaire that colleges complete once a year. Same questions, same definitions, same order at every school. A college fills in its own applicant and admit counts, its SAT and ACT ranges at the 25th and 75th percentile, the GPA distribution of its entering class, tuition and room and board, the average aid package, the share of demonstrated need it meets, and how much weight its admissions office gives each factor it considers.

Two things make it the strongest source available to a family. It comes from the college itself, published on the college’s own site under its own name. And it is standardized, so a figure from one school means the same thing at the next.

A ranking is a different kind of object. Rankings take figures like these, apply weights somebody chose, and print one position. The weights are an opinion. The Common Data Set is the input those opinions are built on, and it is what we work from.

Coverage, including the gaps

CDS admissions
Admit rates, applicant and admit counts, SAT/ACT ranges at the 25th and 75th percentile, GPA and class-rank distribution, ED and EA counts, test-optional policy
1,940 colleges
CDS costs
Tuition, room and board, books, total cost of attendance, net price by family income band
2,397 colleges
CDS financial aid
Average aid package, average need-based grant, percentage of demonstrated need met, average merit award, which aid forms are required
2,410 colleges
Total colleges in the database
Name, location, control type, campus structure, programs offered
2,474 colleges
Coverage is not 100%

The admissions section is the thinnest of the three: 1,940 of 2,474 colleges. Not every college publishes a complete Common Data Set, and some publish only part of one.

Where a college has not reported a field, the field stays empty. We do not fill it with an estimate and present the estimate as the school’s figure. Every record we store is filed under the academic year it describes, so you always know which cycle a figure is from.

IPEDS

IPEDS is the U.S. Department of Education’s reporting system. Any institution that takes part in federal student aid programs is required to report to it, so its coverage is close to complete in a way no voluntary survey can match. Every college in our database is keyed to its IPEDS Unit ID, which is how records get reconciled and how two campuses of the same university stay separate instead of merging into one. IPEDS supplies enrollment figures and the list of programs a school actually offers.

IPEDS lists more institutions than we carry. Our database holds 2,474 of them, so “every college in the country” is not a claim we make. If a school you are looking for is missing, that is why, and telling us is how it gets added.

College Scorecard

College Scorecard is the federal outcomes dataset, built in part from tax and financial-aid records. We use it for post-graduation earnings. It is the reason an earnings figure can sit next to a college at all.

Section 2

How the SAT questions are made, and how they’re checked

23,000+ active practice questions. Here is exactly what they are.

We write the questions. The College Board does not. They are not retired official questions and we will not describe them as such. Part of the bank is written by AI against the official format. The rest is built from templates that assemble each item from a rule set, so the answer key falls out of how the question was constructed. Practicing on them should feel like practicing on a real section, which is the reason to build them against the real format instead of inventing one.

Checking them is where the work is. We hold 697 official College Board reference questions in the database. Statistics measured from that set become the benchmark. The AI-written questions are scored against those benchmarks item by item as they are generated, and anything that drifts gets flagged. The bank as a whole is also compared back to the reference set: average question length, how often a passage appears, and how often the key lands on A, B, C, or D.

What the automated check looks at

Length
Question and passage length measured against the range the official reference questions actually occupy.
Answer-choice balance
Choice lengths compared to each other, because the longest option is a tell. Flagged when one option gives itself away.
Answer-key distribution
How often the key lands on A, B, C, or D across the bank, held against the official distribution.
Sentence density
Average words per sentence in the stem. A question that runs far denser than the reference questions gets flagged.
Stem phrasing
How the question is worded, compared with the phrasings the College Board actually uses.
Format integrity
Choice count, figure references that point at nothing, broken math characters, missing passages on sections that require one.
What we do not claim

Test-prep sites like the phrase “expert-verified.” We are not going to use it. No human subject expert has signed off on this bank item by item, so the honest description is automated validation against official reference questions.

The item-by-item scoring above runs on the AI-written questions. The template-built ones are checked as they are constructed and in the bank-level comparison, not by that same per-item pass.

That is what the mechanism does, and it is all we will say it does. If you hit a question that looks wrong, send it to contact@umerit.ai and we will pull it.

Section 3

What the scores are, and what they are not

The Merit Score and the Gap Analysis are arithmetic on data you can see. Neither one is a prediction.

The Merit Score reads what you enter: courses and grades, test scores, activities, awards, work, service. It scores eight dimensions against a fixed rubric. There is no AI in the number. The same inputs produce the same score every time, and you can see which dimension moved and why.

The Gap Analysis compares your profile against the reported figures for the schools on your list: score ranges at the 25th and 75th percentile, GPA and class-rank distribution, and, where the college reports it, how much weight that admissions office gives each factor. The benchmark is the college’s own filing, not our opinion of the school.

No tool can promise an admissions outcome, including this one

A score measures your profile against reported data. It is not a probability that a committee will admit you. Where the product shows an estimate, it is an estimate computed from reported figures, not a forecast of what a committee will decide.

Committees weigh things that appear in no public dataset: what the institution needs this year, the shape of the applicant pool you happen to land in, how your essay reads at 4pm on a long day, the exact phrasing a recommender chose. A tool that claimed to price all of that would be guessing and printing the guess as a number.

What a benchmark can do is tell you where you stand against the class a school actually admitted, and which gap is worth your next month. That is the claim we make.

Where AI is involved, and where it is not

AI writes essay feedback, answers in the advisor chat, and the written narrative in reports. It does not compute the Merit Score and it does not run the gap comparison. Those are deterministic code paths over your inputs and institutional data.

AI output can be wrong. Read it as a draft second opinion and verify anything that matters. More on the limits in our disclaimers.

Section 4

How often this data changes

Institutional data moves on the institution’s schedule, not ours.

The Common Data Set is an annual filing. Each college publishes its own, usually months after the admissions cycle it describes. We import a new academic year’s records once they are published, and we store them by year, so a new filing sits alongside the old one instead of overwriting it. Nothing gets quietly rewritten under a figure you already read.

So the honest answer on freshness: this data updates once per annual cycle, per college, as each college publishes. It does not update weekly, and we are not going to claim it does. IPEDS and College Scorecard follow their own federal release schedules.

The practice-question bank changes when new questions are generated and pass validation. There is no fixed cadence, and we do not publish one.

Section 5

Check any number

None of this rests on trusting us. Every source is open.

A college’s Common Data Set usually lives on its institutional research page, and a search for the school name plus “Common Data Set” will normally find the PDF. IPEDS and College Scorecard are free federal datasets, open to anyone.

If a figure in uMerit disagrees with a college’s own filing, tell us and we will fix the record. Name the school and the field, and send it to contact@umerit.ai.

How we store and protect your data, as opposed to theirs, is covered in Data Use & Security.

Questions

Ask us about any figure on this page.

We will tell you which source it came from and which year it was filed.