What it means
Every registered business phone number on the WhatsApp Business Platform carries a quality rating. It is not a deliverability score in the email sense, and it has nothing to do with whether your messages arrive. It is a summary of how the people receiving your messages reacted to them.
The rating has three values. Green means high quality and is where a healthy number sits. Yellow means medium and is a warning that recent feedback has been worse than usual. Red means low, and it is the state that triggers consequences.
Two things make it different from the deliverability metrics teams are used to. First, it is scored per number, not per account or per template, so it lives at the same level as your Phone Number ID. Second, it is computed over a rolling recent period rather than over your lifetime history. A number with two flawless years can be red by tomorrow evening, and a number that recovers stops being punished once the bad period ages out of the window.
What Meta is actually measuring
The inputs are recipient actions. Blocking your number is the heaviest signal. Reporting the business is next, and it is worse than a block because WhatsApp asks the user why, and the reason travels with the report. The reason that damages you most is the one that says the recipient never agreed to hear from you, which is a policy statement as well as a quality one.
Things that people expect to count, but do not, are worth listing explicitly:
- Undeliverable numbers do not lower quality. They waste money and pollute your list, but they are not a recipient reacting badly.
- Volume by itself does not lower quality. Large senders with wanted content stay green.
- Read receipts are not a direct input, although a message nobody opens is a very good predictor of the block that follows.
On the positive side, engagement is a genuine counterweight. Numbers that hold real two-way conversations are much harder to push into red than numbers that only broadcast, because the ratio of good interactions to complaints stays high.
From rating to consequence
The rating is the score. The status is what Meta does about it, and the escalation runs in a predictable order.
- Connected. Normal operation. Rating green or yellow.
- Flagged. The rating dropped to red. Meta watches the number for a recovery period. You can still send, and this is the window in which you can fix things.
- Limit reduced. If quality has not recovered by the end of that period, the number's messaging limit drops a tier. This is the first consequence you can feel in revenue, because you can now start conversations with fewer unique people per day.
- Restricted. The number has hit its limit and can no longer start new conversations until the period resets.
The critical insight is that flagged is a grace period, not a punishment. Teams that treat the flag as informational and keep sending the same campaign are the ones that lose a tier. Teams that stop the offending campaign immediately usually return to connected without any lasting effect.
Templates have their own parallel quality mechanism. A specific template that collects negative feedback gets paused so it cannot be sent for a cooling-off period, and repeated pauses eventually disable it. That is a useful diagnostic: a paused template tells you precisely which piece of copy the audience objected to, which a number-level rating cannot.
Why it matters
Quality rating is the mechanism that converts marketing impatience into permanent loss of reach. Nothing about a single aggressive campaign feels dangerous at the time. Messages send, they deliver, the dashboard looks fine. The damage arrives two days later as a lower tier, and the tier is what decides how many people you can reach for the rest of the quarter.
It also compounds. A reduced tier means a slower recovery, because you have fewer conversations available to rebuild a good engagement ratio with. Getting back up a tier requires both healthy quality and enough volume to demonstrate it, and the limit you just lost is the thing constraining that volume.
Real-world examples
- The purchased list. A retailer imports twelve thousand numbers from an events partner and sends a marketing template. Within four hours the rating is red, because a large share of recipients have no memory of the partner, let alone the retailer. Two days later the tier drops. The list cost less than the reach it destroyed.
- The night send. A team schedules in their own timezone and reaches a market seven hours ahead at two in the morning. Blocks spike, not because the content was bad but because the phone buzzed at 02:00.
- The recovery that worked. A logistics company hit yellow, paused all marketing for a week, kept answering inbound support conversations, and was back to green without changing a single template. The audience was the problem, not the writing.
- The diagnostic template pause. Three templates were running. Only the win-back one was paused, which identified the segment and the message in a single signal and saved a week of guessing.
Common mistakes
- Reacting to red by rewriting the copy. If the audience did not consent, better writing does not help.
- Treating flagged as informational. It is a countdown to a tier reduction, and it is your only cheap chance to intervene.
- Rotating to a fresh number. A new number is lower tier, has no history, and will reproduce the same rating on the same list within a day.
- Not storing the history. Meta does not give you an exportable time series, so if you did not poll and record, you cannot correlate a drop with a campaign.
- Hiding the opt-out. An easy unsubscribe converts a block into a quiet opt-out. One damages quality, the other does not.
- Sending to everyone who ever bought. Recency matters. A customer from three years ago does not remember consenting.
Related concepts
- Messaging limit: the tier that a sustained red rating reduces.
- Opt-in: the single biggest lever on the signals that drive this score.
- Message template: carries its own parallel quality state and pausing.
- 24-hour customer service window: conversations inside it protect the ratio.
- Phone Number ID: the node the rating is attached to.
- Cold outreach: the practice most likely to put a number in red.
How Pinlyx handles it
Pinlyx polls the quality rating and status for every connected number, stores the history you cannot get from Meta, and plots it against your campaign timeline so a drop points at a cause rather than a mystery. A move to yellow raises an alert, a move to red can pause outbound campaigns automatically before the tier goes, and per-segment engagement is surfaced so you can stop sending to the list that is doing the damage rather than to everyone. See WhatsApp CRM for how this sits next to campaign tooling.