Tony Wang11 min readWhat 1,610 Reviews of the Claude App Reveal: It's the Money, Not the Model
We mined 1,610 App Store, Google Play, and Reddit posts about the Claude app. In all three, pricing and rate-limits beat quality complaints — every time.
Claude by Anthropic has spent the past month sitting in the US iOS Top-10 Free chart. In that same window, one of its own paying customers filed a proposed federal class-action lawsuit alleging Anthropic misled subscribers about how much they were actually allowed to use. We pulled 1,610 real reviews and posts — 939 from the App Store, 279 from Google Play, 392 from Reddit's r/ClaudeAI — and theme-coded every one. The result agrees with itself three times over: the loudest thing people say about Claude isn't about how smart it is. It's about the bill and the limit.
Three corpora, nobody coordinated, same answer
App Store reviewers, Google Play reviewers, and Reddit posters don't read each other's platforms. They're different populations, self-selected for different reasons, writing in different formats. So when all three independently produce the same shape, that shape is worth trusting more than any one of them alone.
We coded every review and post into eight themes — pricing, rate-limits, bugs, comparisons to rival apps, "the model got worse" complaints, safety refusals, memory/context loss, and quality praise — using a documented keyword classifier (methodology below). Combine the two money-related themes, pricing and rate-limits, into one bucket, and it's the single largest bucket in all three corpora:
That composite beats every other bucket, including the ones you'd expect to dominate an AI app's reviews:
| Theme | App Store (n=939) | Google Play (n=279) | Reddit (n=392) |
|---|---|---|---|
| Pricing + rate-limits | 231 (24.6%) | 57 (20.4%) | 108 (27.6%) |
| Comparisons to ChatGPT/Gemini | 220 (23.4%) | 29 (10.4%) | 49 (12.5%) |
| Quality praise | 158 (16.8%) | 45 (16.1%) | 24 (6.1%) |
| Bugs / reliability | 128 (13.6%) | 27 (9.7%) | 37 (9.4%) |
| Safety refusals / bans | 58 (6.2%) | 9 (3.2%) | 28 (7.1%) |
| Memory / context loss | 16 (1.7%) | 5 (1.8%) | 16 (4.1%) |
| "Got worse" / model degradation | 9 (1.0%) | 3 (1.1%) | 10 (2.6%) |
(A review can carry more than one tag, so rows don't sum to 100%. "Comparisons" is bidirectional — it catches both "way better than ChatGPT" and "ChatGPT is better than this" — so treat it as attention, not sentiment.)
The App Store corpus alone, broken out in full:
What the money complaints actually say
The pattern repeats almost word-for-word across all three platforms. From the App Store:
"Garbage can't use a paid $20 subscription to do any work" — 1★ "I hit my usage limit on Fable 5 with one prompt." — 5★ "Literally after 3 prompts, I have reached my limit." — 3★
From Google Play:
"Every time I start working on something important, I hit another limit." — 1★ "the main issue I really hate is the limit, and the super expensive subscription" — 5★ "Constantly out of tokens, will use all up for a single reply" — 1★
And from Reddit's r/ClaudeAI — where the complaint has a name (Fable is one of Anthropic's current model tiers, alongside Opus, Sonnet, and Haiku):
"I hit the limit right after asking 1 fricking question" — r/ClaudeAI "Yeah I've hit 75% of my weekly on pro." — r/ClaudeAI "$200/month is practically free" — r/ClaudeAI, in a thread arguing the opposite case
That last one isn't a typo. r/ClaudeAI is genuinely split between users who think Claude is underpriced for what it does and users who think the usage math doesn't add up — which is itself informative. Nobody in this corpus is arguing Claude can't write or reason. They're arguing about the meter.
The subreddit's own moderators back this up independently of anything we counted: r/ClaudeAI runs standing, pinned megathreads titled "Usage Limits Discussion" and "Performance and Bugs," plus a "user problem report log" that tracks report volume over time. A community builds dedicated infrastructure for the complaints it can't stop absorbing. Nobody built a megathread for "Claude's writing is boring."
The clearest confirmation this isn't just review-mining noise: on June 15, 2026, a Claude Max subscriber filed a proposed class-action lawsuit in the Northern District of California, alleging Anthropic's Max 5x ($100/month) and Max 20x ($200/month) plans deliver far less usage than advertised — the plaintiff reported that a single five-hour coding session consumed roughly 15% of his weekly allowance. Engadget, Yahoo Finance, and four other outlets corroborated it independently. This is not a Reddit rumor; it is the same pricing/rate-limit theme our review-mining surfaced, now in federal court.
A top-10 app, and the loudest number is the bill
Claude hasn't been struggling for visibility. Using our own App Store rank-tracking dataset — not a live scrape, a standing daily snapshot — Claude sat inside the US iOS Top-10 Free apps chart for most of the past month:
Chart values are Apple's overall Top Free ranking (all categories), not a category-specific chart.
A top-10 free app is not a struggling product. It's a popular one whose loudest complaint happens to be about money — which is a different, more specific story than "people don't like the AI."
Written reviews lie about how happy people are
Before drawing conclusions from any of the numbers above, one honesty check the data itself insists on: people who write reviews are not a random sample of people who use the app.
| Platform | Written-review sample | 1★ share (written) | 1★ share (true aggregate) |
|---|---|---|---|
| App Store | 939 reviews | 29.6% | 4.3% (of 201,437 ratings) |
| Google Play | 279 reviews | 25.4% | 8.0% (of 614,011 ratings) |
On both platforms, the written-review 1-star rate runs 3 to 6 times higher than the true rating aggregate. Claude's actual App Store score is 4.72 out of 201,437 ratings; its actual Google Play score is 4.53 out of 614,011. The people motivated enough to type a paragraph are angrier, on average, than the people who tapped a star and left. Every theme count in this piece describes that vocal, self-selected slice — not "how Claude users feel," which the silent majority answers far more positively.
How Claude stacks up against the other majors
Pulled live, same day, same method, across the five biggest AI assistant apps:
| App | iOS rating (score) | Android rating (score) |
|---|---|---|
| ChatGPT (OpenAI) | 8,641,725 (4.83) | 51,184,375 (4.77) |
| Grok (xAI) | 1,300,072 (4.88) | 3,490,694 (4.83) |
| Perplexity | 486,341 (4.83) | 2,021,073 (4.58) |
| Gemini (Google) | 1,991,310 (4.72) | 39,712,132 (4.58) |
| Claude (Anthropic) | 201,437 (4.72) | 614,011 (4.53) |
Two things stand out. First, Claude has by far the smallest rating base of the five on both platforms — roughly 43x fewer iOS ratings and 83x fewer Android ratings than ChatGPT — which is the real headline about Claude's consumer app: it is a small, fast-growing challenger, not yet a mass-market incumbent, whatever the chart-rank moment above suggests. Second, on Android specifically, Claude's 4.53 is the lowest score of the five — behind Gemini and Perplexity, which tie at 4.58.
The cottage industry Claude's rate limits built
The clearest sign a pain point is real: strangers build unpaid tools to manage it. Searching Google Play for Claude-adjacent apps turns up "AI Usage: Claude & Gemini" — a free, independent app (4.25★, unaffiliated with Anthropic) whose entire purpose is tracking when your Claude usage window resets, syncing that reset time to your calendar, and sending a lock-screen notification so you don't waste a session waiting.
Reddit corroborates the same anxiety from the demand side: one post title reads "I built an app to monitor your Claude usage limits in real-time"; another describes building "iPhone widgets to make Claude's rate limits easier to see." Nobody builds a widget to track how good the writing is.
What holds up, and what doesn't
What holds up: three independent, non-overlapping corpora converge on the same shape — pricing and rate-limits are the single largest complaint theme, ahead of bugs, ahead of model-quality complaints, and ahead of comparisons to rivals. A federal lawsuit filed the same month makes the same argument the reviews do. A third-party utility app exists solely to manage the pain point. That's four independent signals, not one.
What doesn't, and the caveats that matter:
- Theme classification is keyword-based, not semantic. It undercounts anyone describing a rate limit without using rate-limit words, and a review can be miscounted if it uses a theme's keywords in an unrelated sense. Treat the counts as a shape, not a census.
- Store reviews are the angriest fraction of a userbase, as the written-vs-aggregate section above shows directly — do not read "29.6% one-star" as "30% of Claude users are unhappy."
- Reddit's public feed exposes no upvote or comment-count field on posts (it's RSS-backed), so post counts above are unweighted by engagement; we used the subreddit's own auto-mod thread-size disclosures as a secondary engagement signal, not a primary one.
- Google Play's
sort=ratingparameter returned five-star reviews exclusively — a descending-score sort, not a balanced sample. All negative-review coverage on Android came from the newest and most-helpful sorts. - "Comparisons to rivals" is bidirectional and was not split into praise-of-Claude versus praise-of-a-competitor; read it as attention to the competitive question, not as evidence either way.
- This is a snapshot, not a trend line. We have 15 days of App Store rank history, not months — the "top-10 for a month" finding is real but short-window.
Sources
Methodology
Every number here was pulled through Crawlora's own APIs on July 16, 2026, not taken secondhand: App Store reviews and ratings (both sort orders, 10 pages each), Google Play reviews (newest, rating, and helpfulness sorts, 100 each) and app details, subreddit posts and search across r/ClaudeAI (top/month, top/year, hot, plus three targeted searches), thread comments on the 12 highest-signal threads, and our own App Store charts dataset for the 15-day rank history.
Corpus: 1,610 items — 939 App Store reviews, 279 Google Play reviews, and 392 Reddit posts about the Claude app, all mined via a documented keyword classifier (theme keyword lists are in the reproducible script). We publish aggregates only — no raw review text tied to individual reviewers. The aggregate dataset — theme counts, star distributions, rank history, and the competitive benchmark — is open on GitHub under CC BY 4.0.
Reproduce it yourself in the Playground; the how to scrape App Store reviews, how to scrape Google Play, and how to scrape Reddit guides walk through the exact calls, and our app review sentiment dataset has millions of pre-scraped reviews if you'd rather skip crawling entirely. For the same pattern applied to Codex and Claude Code specifically, see Codex vs Claude Code: What 1,828 Posts and Reviews Say.
Frequently asked questions
What do people complain about most in Claude app reviews?
Pricing and rate-limits, combined, are the single largest complaint theme across three independent corpora we mined in July 2026: 24.6% of 939 App Store reviews, 20.4% of 279 Google Play reviews, and 27.6% of 392 Reddit posts in r/ClaudeAI. That beats bugs, comparisons to ChatGPT and Gemini, and complaints that the model itself got worse. Bugs and reliability issues are the next-largest theme (13.6% on the App Store), followed by quality praise (16.8%).
Is Claude's app rating lower than ChatGPT's and Gemini's?
On Android, yes — Claude's 4.53 score (614,011 ratings) is the lowest of the five major AI assistant apps we compared, behind ChatGPT (4.77), Grok (4.83), and tied for second-lowest with Perplexity and Gemini (both 4.58). On iOS, Claude's 4.72 sits mid-pack, close to Gemini's 4.72 and behind ChatGPT (4.83), Grok (4.88), and Perplexity (4.83). Claude also has by far the smallest rating base of the five on both platforms — roughly 43x fewer iOS ratings and 83x fewer Android ratings than ChatGPT.
Why did Anthropic get sued over Claude's usage limits?
On June 15, 2026, a Claude Max subscriber filed a proposed federal class-action lawsuit in the Northern District of California, alleging that Anthropic's Max 5x ($100/month) and Max 20x ($200/month) plans deliver far less usage than advertised. The plaintiff said a single five-hour coding session consumed roughly 15% of his weekly allowance. The lawsuit was independently reported by Engadget, Yahoo Finance, and several other outlets, and it echoes the same pricing/rate-limit complaints that dominate our App Store, Google Play, and Reddit review mining.
Are app store reviews a reliable way to judge how happy users are?
Not on their own. Written reviews skew far more negative than the true rating aggregate: 29.6% of Claude's sampled App Store text reviews are one-star, versus 4.3% of Apple's full 201,437-rating aggregate. On Google Play the gap is 25.4% (written sample) versus 8.0% (of 614,011 ratings). People motivated enough to write a review are angrier, on average, than the silent majority who just tap a star. Read review themes as a shape of what vocal users discuss, not a survey of overall sentiment.
How was this Claude app review study done?
We pulled 939 App Store reviews (both mostRecent and mostHelpful sort orders, 10 pages each), 279 Google Play reviews (newest, rating, and helpfulness sorts, 100 each), and 392 Reddit posts from r/ClaudeAI (top/month, top/year, hot, plus three targeted searches) via Crawlora's own APIs on July 16, 2026, then theme-coded every item with a documented keyword classifier across eight categories. We also pulled top comments on the 12 highest-signal Reddit threads and 15 days of Claude's own App Store chart-rank history from our apps-charts dataset. The full reproducible script and committed aggregate dataset are linked in the post's methodology section.