Correction
A version of this brief published on 8 September used the same flawed measure as the earlier Productivity and Entertainment briefs and was withdrawn the same day. This is the rebuilt version.
An App Store rating looks like a verdict. It behaves more like a monument. An app sitting at four-and-a-half stars has earned that number over years of installs and early goodwill, and it moves slowly. It says very little about how the people using the app this month actually feel.
This is the Nativerse lab reading underneath that number. For a whole category we separate two things the single star rating blurs together. The first is population truth: Apple's full ratings histogram across every rating an app has ever received, and it is almost immovable. The second is the mood of the people who write: a 90-day window of written reviews, set against what that same app's reviewers were saying a year earlier. One population, two points in time. Then the question that decides whether an app recovers. When users turn, does the developer answer?
This study covers 39 tracked Games apps on the US App Store, 39,822,331 ratings in all. Their mean lifetime rating is 4.61, and it will still read about that a year from now whatever happens next. The movement is underneath it.
The Friction Matrix
Each app sits on two forces. Left to right is the movement: how the people writing reviews now rate the app against the people who wrote about it a year or more ago. Both are written reviews, so the comparison is like for like. Top to bottom is the response: how often the developer replies. On this cohort only the movement separates the apps, so two archetypes fall out.
The only app here above 20% is Sword x Staff, and it carries no movement figure, so nothing sits in the responsive half.
Fallen against their own past. Recent reviewers rate the app below its own reviewers of a year ago.
Off the matrix: Block Out!, no written reviews earlier than its recent window; Pixel Flow!, no written reviews earlier than its recent window; Sword x Staff, no written reviews earlier than its recent window. They are counted in the reply figures but carry no movement figure.
Against their own reviewers of a year ago 9 of these apps have fallen and 9 have risen. 18 have held, and the lifetime average shows none of it. 9 apps are Ghost Ships, fallen against their own past. 27 apps are Complacent Giants, steady against their own past.
Replying and answering are different things. Across the category about 3.5% of recent reviewers get a reply, but only 1% get a substantive one. The rest repeat a templated opening. The clearest case is Sword x Staff, the busiest replier, where a large part of the answers share the same macro wording.
The movement, ranked
The same measure, app by app. Bars run left of the line where today's reviewers rate an app below its own reviewers of a year ago, and right where they rate it higher. The two n values on each label are how many written reviews the earlier figure and the recent figure each rest on, so a bar built on fifty can be read against one built on five hundred.
What it means
The lifetime star is the slowest number on the page to move, and the easiest to mistake for a signal. What moves is the mood of the people who write, measured against what that same group was saying a year earlier. The category answers about 3.5% of recent reviewers, a median of 2 days later. Strip out the templated replies and the substantive figure is about 1%.
For Games, 20 of the 39 apps studied answer nobody at all.
The matrix shows where each app sits. The deep dive explains why, app by app: the complaint archetypes behind each drop, and how the developer responded, with representative quotes.
Read the full analysisThe cohort
12 of 39 shown; the full cohort is on the deep page.
Method in brief
Lifetime ratings are population truth from Apple's histogram. Movement is measured inside written reviews only: an app's reviews from the 90-day window days ending at its latest captured review, against its own reviews from a year or more earlier. Both are people who chose to write, so the skew toward the dissatisfied sits on both sides. Re-run at 60 and 120 days, the count of apps past -0.5 is 9/8 and past +0.5 is 8/9; the 90-day figures are reported. The taxonomy is rule-based. We make no claim tying sentiment to a specific app release.
Independent research from the Nativerse lab. Population data from Apple's public ratings histogram. Movement is measured inside written reviews only: an app's reviews from the 90 days ending at its latest captured review, against its own reviews from a year or more earlier. Figures are cited, not invented.



































