Correction
A version of this brief published on 8 September used the same flawed measure as the earlier Productivity and Entertainment briefs and was withdrawn the same day. This is the rebuilt version.
An App Store rating looks like a verdict. It behaves more like a monument. An app sitting at four-and-a-half stars has earned that number over years of installs and early goodwill, and it moves slowly. It says very little about how the people using the app this month actually feel.
This is the Nativerse lab reading underneath that number. For a whole category we separate two things the single star rating blurs together. The first is population truth: Apple's full ratings histogram across every rating an app has ever received, and it is almost immovable. The second is the mood of the people who write: a 90-day window of written reviews, set against what that same app's reviewers were saying a year earlier. One population, two points in time. Then the question that decides whether an app recovers. When users turn, does the developer answer?
This study covers 39 tracked Games apps on the US App Store, 39,822,331 ratings in all. Their mean lifetime rating is 4.61, and it will still read about that a year from now whatever happens next. The movement is underneath it.
The Friction Matrix
Each app sits on two forces. Left to right is the movement: how the people writing reviews now rate the app against the people who wrote about it a year or more ago. Both are written reviews, so the comparison is like for like. Top to bottom is the response: how often the developer replies. On this cohort only the movement separates the apps, so two archetypes fall out.
The only app here above 20% is Sword x Staff, and it carries no movement figure, so nothing sits in the responsive half.
Fallen against their own past. Recent reviewers rate the app below its own reviewers of a year ago.
Off the matrix: Block Out!, no written reviews earlier than its recent window; Pixel Flow!, no written reviews earlier than its recent window; Sword x Staff, no written reviews earlier than its recent window. They are counted in the reply figures but carry no movement figure.
Against their own reviewers of a year ago 9 of these apps have fallen and 9 have risen. 18 have held, and the lifetime average shows none of it. 9 apps are Ghost Ships, fallen against their own past. 27 apps are Complacent Giants, steady against their own past.
Replying and answering are different things. Across the category about 3.5% of recent reviewers get a reply, but only 1% get a substantive one. The rest repeat a templated opening. The clearest case is Sword x Staff, the busiest replier, where a large part of the answers share the same macro wording.
The movement, ranked
The same measure, app by app. Bars run left of the line where today's reviewers rate an app below its own reviewers of a year ago, and right where they rate it higher. The two n values on each label are how many written reviews the earlier figure and the recent figure each rest on, so a bar built on fifty can be read against one built on five hundred.
The anatomy of the drop
Behind the movement are recurring complaints. We classify recent reviews with a rule-based taxonomy and name the dominant patterns. These are illustrative archetypes from a biased sample, not a verdict.
9 apps have fallen more than half a star against their own reviewers of a year ago. The 6 steepest are below.
Love and Deepspace
The Encroaching Paywall (10.1%)The Update That Broke It (7.7%)The Wall at Level 200 (6.8%)
Since the removal of Valko, half the game is glitchy. House is broken, memories lag, Journals freeze my game and won’t ever work.2★ · 2026-08 · broken_build
My phone always had enough storage space and never had any issues but for some reason now it’s lagging like crazy. It’s so slow to the point where I can’t play the game and it’s frustrating. Please fix this bug2★ · 2026-08 · broken_build
Tasty Travels
The Game in the Advert (63.8%)The Wall at Level 200 (9.8%)The Encroaching Paywall (6.7%)
This is nothing to do with the game that is advertised. The game is stupid at that. Horrible graphics and all you do is match stuff and build some sort of whatever it’s called. I hate how every single game that’s…1★ · 2026-08 · bait_ads
If you purchase gems (aka credits) and use them, the game is incredibly buggy and will not reliably deduct the correct number of credits from the account. When you contact the Support team, they expect you to have taken…1★ · 2026-08 · broken_build, progress_loss, support_void
NYT Games
The Update That Broke It (46.2%)The Encroaching Paywall (4.4%)The Wall at Level 200 (4.1%)
It’s frustrating how much worse the iPad app has gotten recently. At least it isn’t constantly crashing in the middle of my crossword puzzles anymore, but it’s so slow. It’s still so sluggish that half the time it can’t…2★ · 2026-07 · broken_build
Very disappointed with new notifications which make the product virtually unusable. I’ve tried to contact nyt but no one has fixed the bug2★ · 2026-08 · broken_build
Gardenscapes
The Game in the Advert (25.7%)The Wall at Level 200 (22.6%)The Encroaching Paywall (22.3%)
I used to like this app; but the developers have made some levels so difficult to complete. Tries to force player to spend a lot of money to continue a level. Makes playing more frustrating & annoying rather than…2★ · 2026-06 · monetization, ux_friction
This game crashes mid-game, causing you to lose your streak, your bonuses, etc. so don’t waste your money on extras. You might pay money for upgraded rewards for a certain festival or theme, only for those rewards to…1★ · 2026-08 · broken_build
Township
The Game in the Advert (39.1%)The Encroaching Paywall (13.9%)The Wall at Level 200 (12.1%)
I just want to say I HATE this new update. THE CASH THING IS ANNOYING. All they are HUNGRY FOR MONEY the pigs. Literally soo annoying after you lose a match having to pay $900 of cash! Matches don’t even give you enough…1★ · 2026-08 · broken_build
This games advertisements are nothing like the game they are just garbage clickbait. The game sucks it’s horrible worst game ever it deserves -5 stars.1★ · 2026-09 · bait_ads
Whiteout Survival
The Game in the Advert (54.6%)The Encroaching Paywall (13.2%)The Wall at Level 200 (3.8%)
Terrible game isn’t fun at all and wasn’t the game I saw in the add. Don’t waste your time downloading this garbage all reviews on here are fake bots. Developers I hope your game fails.1★ · 2026-08 · bait_ads, unfair_match
Awful. Nothing like the game play in the ad. Deceptive and frustrating. Just a ploy to get downloads. Despicable. Updated to add: the generic and duplicative reply I received in response to this review has just…1★ · 2026-08 · bait_ads
The corporate response
Developer replies are a proxy for how hard a team is fighting the friction. Across this category the reply share is about 3.5%, a median of 2 days after the review.
| App | Reply share | Median days | Templated |
|---|---|---|---|
| Sword x Staff | 64.5% | 1 | 96% |
| All in Hole | 14% | 2 | 27% |
| Tasty Travels | 10% | 2 | 98% |
| Merge Cooking® | 6.3% | 2 | 54% |
| Evony | 5.3% | 3 | 25% |
| Total Battle | 5.3% | 1 | 23% |
| Last Z | 5% | 2 | 86% |
| Last War | 4.3% | 2 | 41% |
| Kingshot | 4.1% | 2 | 89% |
| Disney Solitaire | 3.7% | 1 | 15% |
What it means
The lifetime star is the slowest number on the page to move, and the easiest to mistake for a signal. What moves is the mood of the people who write, measured against what that same group was saying a year earlier. The category answers about 3.5% of recent reviewers, a median of 2 days later. Strip out the templated replies and the substantive figure is about 1%.
For Games, 20 of the 39 apps studied answer nobody at all.
Method and limits
- Ratings and star distribution are population truth.
- Recent average and response rate are from a biased sample.
- Movement over time is measured within one population only: an app's written reviews from the 90 days ending at its latest captured review (minimum 50) against its own written reviews from at least 365 days earlier (minimum 50).
- The histogram average and the written-review average are never subtracted from one another to claim a decline. The difference between them is a standing population offset, reported as written_vs_population and nothing more.
- Taxonomy is rule-based keyword/n-gram matching (v1, heuristic); buckets can overlap and some reviews are unclassified.
- No version-tied analysis: app_version is sparse and snapshots are not version-segmented; no claim links sentiment to a release.
- The reviewer-sentiment series (where shown) is sample-based and self-selection-biased; deep-backfilled apps only.
- Developer-reply latency uses the response last-edit date as a proxy for first reply.
- Reply quality (substantive vs boilerplate) is heuristic: boilerplate = a reply sharing a templated opening 8-gram with another reply. It catches macros, not a bespoke but empty non-answer.
- Review capture runs nightly and missed some nights; for busy apps the store may hold fewer reviews than were written in the window.
- Re-run at 60 and 120 days, the count of apps past -0.5 is 9/8 and past +0.5 is 8/9; the 90-day figures are reported.
- Quotes are short illustrative excerpts selected by polarity and length, not a representative sample.
- The free/deep split is structural; no payment gating exists yet.
Grounded in prior art on app-review mining and review selection bias:
The cohort
Independent research from the Nativerse lab. Population data from Apple's public ratings histogram. Movement is measured inside written reviews only: an app's reviews from the 90 days ending at its latest captured review, against its own reviews from a year or more earlier. Figures are cited, not invented.



































