Skip to content

METHODOLOGY

How our ratings work

We keep three questions apart: whether people enjoy a game, what the play itself may help someone notice or practice, and how much support stands behind that meaningfulness claim.

First: is it actually a game?

Meaningful Games does not try to catalog every useful digital product. The first gate is narrower and more practical: would people still want to play this if the useful effect disappeared?

The game-first question

Would people still want to play this if the useful effect disappeared?

That question helps separate games from applications that mainly borrow points, streaks, badges, progress bars, or rewards. Those features can make a product more engaging, but they do not automatically make it a game worth recommending as play.

Meaningfulness alone is not enough for inclusion. The catalog is interested in real play value first, then in what that play may also help someone notice, practice, or understand.

What every game here is held to

Four things that are true of the whole catalog. They are yes-or-no commitments rather than points on a scale, so none of them can quietly raise a rating.

Game first

Would people still want to play this if the meaningful effect disappeared? If not, it does not belong here, however educational it is.

Honest about its subject

Every game simplifies. A game is admitted when it is open about where it simplifies, dramatises or invents, and its page states that limit in plain words. This is a gate and a caveat, never points on a scale.

Access is a fact, not a virtue

Price, platform, reading load and how much adult help a younger player needs are facts we publish and let you filter on. Being free does not make a game more meaningful, and until 2026-09-14 our own scale quietly said it did.

Independence

Games connected to the team carry a visible disclosure, never get preferential ranking, and are never added to collections by default.

One rating, one confidence badge, one outside number

We do not merge enjoyment, meaningfulness, and support into one verdict, because they answer different questions.

MEANINGFULNESS

How much is there to get out of this, beyond the fun?

One editorial rating on a five-step scale, built from three dimensions. It is a judgment about the shape of the play, not a measured effect, and it is not a verdict on whether the game is good.

Every game page shows the three dimensions behind the rating and the reasoning for each one.

EVIDENCE

Who, other than us, has checked this?

A confidence badge with three levels: published research, documented use by people who answer for the outcome, or nothing but our own reading. It qualifies the meaningfulness rating; it never raises or lowers it.

It sits on the game page rather than on cards, because it answers a question you ask after a game has caught your eye.

PLAYER RECEPTION

Do people seem to enjoy it as a game?

An external number we copy, not a rating we give: the store's own score, its original format, how many reviews it rests on, and the date we checked it.

A few pages show a clearly marked sample value until a verified external rating is found; such values never feed into editorial judgments.

The three dimensions behind the step

Each is scored out of a hundred and weighted, which is a consistency framework for making judgments comparable, not a clinical, educational, or scientific measurement scale. Accessibility and accuracy used to sit here and deliberately do not any more.

Fit

45%

How much real knowledge or skill does the play itself make you use? This is the heaviest dimension. At the top, playing is doing the real thing: writing working code, planning an orbit, constructing a proof. At the bottom, the benefit is a general soft skill staged by the game, such as communication or patience, or a subject used only as scenery.

Transfer

30%

Will what you practise here work outside the game? Code, digital logic and geometry carry over directly. Subject knowledge picked up along the way carries over in part. General habits trained in an artificial situation rarely do, and the research on far transfer says as much.

Practice

25%

Does the game make you use that skill again and again as it gets harder? Hundreds of hours of rising challenge score highest; one short playthrough scores low. Only the valuable part counts, so endless repetition of a thin skill does not earn much.

The five steps

Cards show one of five steps rather than a number out of a hundred. The steps start at 85, 70, 55 and 40 on the worksheet score, which is on every game page next to the reasoning for each dimension, but the step is what we are willing to publish as a judgment. Our bar is high on purpose: a Low or Minimal game can still be well worth playing.

Exceptional

Rare. Playing is real work in a real discipline, and it holds up over many hours.

High

A substantial skill or body of knowledge that works outside the game, practised in depth.

Moderate

Real substance with a clear limit: a real system reduced to a few levers, or genuine reasoning on invented material.

Low

Something real but narrow: a general skill in an artificial situation, or a subject you look at more than work with.

Minimal

A good game that leaves little you can use elsewhere. It is here because it is worth playing, not because it teaches.

Unrated

Not yet assessed, or not ready for a public rating.

Evidence: who, other than us, has checked

Evidence qualifies the meaningfulness rating and never changes it. A higher evidence level does not make a game better, more fun, or more meaningful; it means fewer people have to take our word for it.

Research

Research-backed

Published research has studied this game, or the exact mechanic it is built on. Where an effect has also been replicated independently, the game page says so: replication is a note inside this level rather than a level above it, because research is the strongest thing we can point at either way.

Practice

Practice-backed

No published study, but people who answer for the outcome use it: schools, clinics, museums, research laboratories. This differs from the level above in kind rather than degree. That one is papers, this one is applied use.

Editorial

Editorial only

Nobody outside our desk has checked this. It is our reading of how the game plays, and it covers most of the catalog. Saying so plainly is the point of the badge.

How a game gets reviewed

  1. Game-first eligibility: the game has to stand up as play before meaningfulness enters the conversation.
  2. Honesty check: where does it simplify, dramatise or invent, and does it admit that? A game that misrepresents its subject does not get a page.
  3. Structured editorial review: we look at what the player actually does, what the game sets out to be, and who it may suit.
  4. Meaningfulness assessment: fit, transfer and practice are each scored, weighted, and turned into one of five public steps.
  5. Evidence and source review: who, other than us, has checked the claim, recorded separately from the rating it qualifies.
  6. Publication: the page explains the judgment in plain language, including the access facts that deliberately stay out of the rating.
  7. Corrections and updates: pages are revised when a game changes, better sources appear, or an earlier judgment no longer holds.

Some pages are more fully reviewed than others. The Evidence marker and source notes are there to show how much support currently sits behind a meaningfulness claim.

Conflicts of interest

Affiliated games are disclosed and handled more cautiously than other entries. They never get preferential ranking and are never added to recommendations or editorial collections by default.

If a game is materially connected to the team, its page states that connection, and any meaningfulness assessment goes through additional independent review.

The fuller policy for affiliated projects, developer submissions, and paid placement lives on the disclosure page.

Where AI may be used

AI helps prepare and organize the work; public editorial judgments stay open to human review.

AI may help with

  • metadata preparation
  • draft classification
  • structured summaries
  • consistency checks
  • initial draft preparation

AI does not independently

  • judge historical or scientific accuracy
  • raise an Evidence level
  • resolve conflicts of interest
  • publish unsupported claims

Corrections and updates

Games change. Sources change. Editorial judgments may also change when a page is improved, a source is added, or an earlier interpretation no longer holds up.

Factual or sourcing problems can be reported through the corrections page. General editorial questions or suggestions belong on the contact page.