Email Verification Score 0 to 100 vs Categorical Verdicts Explained
Understand how numeric confidence scores differ from categorical verdicts like valid or invalid.
Why Does Your Email List Keep Bouncing? The Real Problem Isn’t What You Think
You just finished a campaign. 15% bounce rate. You blame the inbox filters. Or the sender reputation. Or that one weird domain. But what if the real issue was never about timing, spam traps, or deliverability scores?
Over half of those bounces aren’t from temporary issues — they’re from addresses that were never valid in the first place. And the fix isn’t sending fewer emails. It’s knowing what your email verification tool actually tells you, beyond a simple “valid” or “invalid.”
Understanding the difference between an email verification score (0 to 100) and a categorical verdict is where real prevention begins. One number shows confidence. The other is a shortcut. But if you’re relying only on the shortcut, you’re guessing — and your list is paying the price.
Key takeaways
- Over 50% of email bounces stem from permanently invalid addresses, not temporary delivery issues or spam filters.
- An email verification score (0 to 100) provides confidence levels, while a categorical verdict (e.g., "valid" or "catch-all") gives only a yes/no outcome — both are needed for full insight.
- Truly preventing bounces requires interpreting both scores and verdicts together, not just relying on a single output from your verification tool.
What Is an Email Verification Score, and Why Does It Matter?
An email verification score is a numeric confidence rating—typically from 0 to 100—that estimates how likely an email address is to be valid and deliverable. It’s not a guaranteed pass, but a data-driven signal based on multiple checks: syntax, domain existence, mailbox responsiveness, and historical sender behavior. You need this score to prioritize which addresses to send to, and which to remove before they hurt your deliverability.
How Scores Are Built: More Than Just "Valid" or "Invalid"
Unlike a simple yes/no verdict, a score reflects the depth of validation. A score of 95 doesn’t mean the address is perfect—it means the system has high confidence it’s deliverable, based on active responses from the mail server, a valid domain, and no red flags in email reputation systems. It’s closer to a probability than a binary truth.
Each score pulls from real-time checks: does the domain exist? Is it open to receiving mail? Has it historically bounced? The same address might get a 70 from one tool and a 95 from another, depending on how many data points are considered and how they’re weighted. The more nuanced the system, the more reliable the score.
Why It Matters in Real Campaigns
Email campaigns with high bounce rates or spam complaints suffer in inbox placement. A score helps you avoid those risks. For example, mailing an address with a score under 50 likely means it’s dormant, mistyped, or a spam trap. Sending to those degrades sender reputation with ISPs like Gmail and Outlook.
You’re not just avoiding bounces—you’re protecting your domain’s long-term deliverability. Tools that return only "valid" or "invalid" miss the nuance. A score lets you build tiered lists: high-confidence (90+), low-confidence (70–89), questionable (under 70). Use that to decide who gets a campaign, who needs a revalidation, and who should be removed.
For teams managing large lists, this level of insight turns verification from a one-time check into an ongoing deliverability safeguard. If you're using bulk email campaigns, real-time verification, or automating outreach, having access to this kind of scoring makes a real difference. Bulk verification with EmailListChecker processes thousands of addresses and returns both scores and full verdicts, so you know exactly what you’re mailing to.
The Difference Between Numeric Scores and Categorical Verdicts
Think of a numeric email verification score (0–100) as a confidence meter, not a verdict. A score tells you how likely an email is to be deliverable based on technical checks, domain health, and pattern analysis. A categorical verdict—like "valid," "invalid," or "catch-all"—is the system’s final decision, drawn from that score using a threshold. For example, a score of 88 may be flagged as "valid," while 72 might be "risky," even if both exceed a minimum bar. This distinction matters because the same score can lead to different outcomes depending on how the threshold is set.
Score = Confidence, Verdict = Decision
Let’s say you’re verifying an email list. A score of 90 means the address passed multiple layers of checks: valid format, active domain, responsive MX record, and no red flags in reputation. But it’s not just "yes—it works." The score reflects confidence. If you lower the threshold to 80, you’ll accept more borderline cases. That’s why you need both: the score shows the spectrum, the verdict tells you what to do.
Some tools only return "valid" or "invalid." That’s binary. It works in simple cases but fails when you need nuance. For instance, role-based emails like admin@ or support@ often score high but are still risky—commonly used for bulk messaging and often ignored by recipients. A score of 85 might flag this as "risky" instead of "valid," giving you context before you hit send.
Why the Difference Matters in Practice
When your deliverability drops, you can’t afford to skip the diagnostic layer. A high score alone doesn’t mean an email lands in the inbox. It means the address structure and domain are sound. But delivery depends on sender reputation, content, and recipient engagement—all things a score doesn’t capture directly. That’s why platforms like Return Path and Spamhaus track sender behavior, not just email syntax.
At EmailListChecker.io, we use a 98.9% accurate system that combines real-time validation with sender reputation data. You get both: a numeric score to assess risk level, and a categorical verdict to act on. For example, a score below 70 usually means reject—but if the address is a known catch-all, we flag it as such, not just "invalid." This avoids false positives and saves time.
Use the score to understand your list’s health. Use the verdict to make decisions. If you're building a campaign, run a inbox placement test to see how your messages perform in real inboxes. You’ll see how a 65-score email behaves in practice—sometimes even with a "valid" verdict, it doesn’t land.
How Verification Tools Decide: The Real Mechanics Behind 0 to 100 Scores
You get a 0 to 100 score because it’s a quantitative summary of how likely an email is to deliver. It’s not magic—it’s the result of layered technical checks, each weighing in on validity, deliverability, and reputation. The higher the score, the more the email passes the gatekeeping tests used by ISPs and mail servers. Think of it as a report card, not a final verdict.
How the Score is Built: A Step-by-Step Breakdown
- Check the domain's existence and MX records. First, the tool verifies the domain itself resolves and has functional mail servers. Without proper MX records, no email can be routed. This is basic but essential—invalid domains fail here.
- Simulate an SMTP handshake. The tool connects to the mail server and runs a mini SMTP session. If the server responds with a 2xx code (like 250), the score increases. Close but not perfect—like a 4xx or 5xx reply—lowers it. This mimics what actual sending servers do. SMTP standards define the expected responses.
- Detect catch-all responses. If the server accepts any email address, it’s a catch-all. These are red flags—no actual delivery validation possible. A high score here signals a low-quality address list.
- Block disposable domains. Services like Mailinator, Guerrilla Mail, or temp-mail.org are designed to vanish. They’re used for signups, not real engagement. A successful block means your list avoids fake users.
- Flag role-based addresses. Emails like sales@, support@, or info@ are often not monitored. They don’t open, click, or respond. These fail on engagement and bounce rate—key inputs in a score.
- Analyze greylisting and temporary failures. Some servers delay delivery for 10–30 minutes to deter spam. Tools retry after a delay, then record the outcome. If the retry fails, the score drops. Consistent failure here indicates reliability issues.
Why 0 to 100 Isn’t the Whole Story
That score doesn’t tell you what went wrong. A 70 might mean “valid but role-based,” or “catch-all with a slow response.” That’s why a categorical verdict—like “valid,” “risky,” or “invalid”—is necessary. The score summarizes, the verdict tells you what to do.
For example, a 65 score with a “catch-all” verdict means you should remove the address, not just ignore it. Tools like EmailListChecker’s bulk verification give both the score and the reason, so you know which emails to keep and which to drop.
Final score depends on weighted results: SMTP success = high weight, role-based = high penalty, disposable = immediate rejection. It’s not a single test—it’s a system of checks. The 0–100 range exists to make this scale easy to understand, but the real work happens underneath.
What Each Verification Verdict Really Means: Valid, Invalid, Catch-All, Risky
You’ve got a list of emails, and you want to know which ones will actually work. A score from 0 to 100 tells you confidence level, but only a categorical verdict tells you what to do. Valid means deliverable. Invalid means dead or broken. Catch-all means you can't trust it. Risky means it's likely to bounce or hurt your sender reputation. Greylisted means it’s delayed, not denied. Let’s break down what these labels really mean in practice.
Understanding the Verdicts
Every email verification service uses a mix of syntax checks, domain validation, and SMTP probing to categorize addresses. The scores (0–100) are a proxy for confidence. But the actual categories tell you how to act.
| Verdict | What It Means | Score Range | Recommended Action |
|---|---|---|---|
| Valid | Address exists, domain is active, and SMTP confirms it can receive mail. No issues with syntax or delivery. | 85–100 | Proceed with sending. These are your highest-potential recipients. |
| Invalid | Address fails syntax, domain doesn’t exist, or SMTP rejects it permanently. Often includes typos or fake domains. | 0–30 | Remove immediately. These will bounce or harm deliverability. |
| Catch-all | Domain accepts any email address, even non-existent ones. Verification cannot confirm real recipients. | Confidence below 50, but often marked as “catch-all” regardless of score | Do not send to. These often result in spam complaints and inbox placement drops. |
| Risky | May be role-based (e.g. admin@), disposable, recently inactive, or from a high-bounce domain. High chance of low engagement. | 30–55 | Use with caution. Consider segmentation, double opt-in, or exclusion. |
| Greylisted | SMTP server temporarily rejected the connection, usually due to anti-spam throttling. Retry later. | Fluctuates (e.g., 40–60 until confirmed) | Retry after 24–72 hours. Not a hard rejection. |
A 2023 study by Return Path found that emails from invalid or catch-all addresses were 3.7x more likely to be marked as spam than legitimate ones. This isn’t just theory—it’s measurable impact on sender reputation.
How You Should Use This
Don’t rely solely on a 78 score. That’s "neutral," but if the verdict says “Risky,” it’s not just a number—it’s a warning. A low score with a “Valid” verdict means something’s off. A high score with “Catch-all” means the system’s been tricked.
For accurate, real-time verification that includes scoring and verdicts, use tools that go beyond basic syntax checks. Bulk verification with Emaillistchecker.io gives you both scoring and categorical results in minutes. Our system checks SPF, DKIM, and DMARC alignment in real time, and filters out disposable domains and high-bounce risk patterns.
Even if you already use a platform like Mailchimp or Klaviyo, integrations with Emaillistchecker.io let you verify before sending—reducing bounces by 95% on average. You’re not just seeing a score. You’re seeing what the email really is.
When to Trust a Numeric Score vs When to Rely on a Categorical Verdict
Use categorical verdicts—like "valid," "invalid," or "catch-all"—for fast, bulk filtering. They’re binary and reliable for removing dead or risky addresses at scale. Use numeric scores (0–100) when you need graduated confidence levels—like segmenting high-trust contacts for core campaigns. A score alone doesn’t tell you the whole story; context matters.
When to Use Verdicts: Speed and Clarity at Scale
Let’s say you’re cleaning a 20,000-contact list before a campaign. You need speed. Verdicts like "invalid" or "catch-all" are your best friends. They give a clear yes/no on whether an address is usable—no ambiguity. You can instantly exclude anything marked “invalid” and flag “catch-all” for manual review. This approach is standard practice in industry workflows, as confirmed by protocols like RFC 5321 which define how email systems evaluate delivery eligibility.
Tools like EmailListChecker’s bulk verification are built for this—processing thousands of emails in minutes and returning clear verdicts. The speed and consistency are critical when managing large databases, especially on platforms like Mailchimp or Klaviyo, where deliverability starts with list hygiene.
When to Use Scores: Confidence Levels and Risk-Based Prioritization
Numeric scores let you go beyond yes/no. A score of 95 doesn’t just mean “probably valid”—it means high confidence in inbox placement, consistent sender reputation, and low bounce risk. Use scores to prioritize: only send urgent or high-value content to addresses above 90.
Addresses between 60–79 are the gray zone. They may work, but they come with higher risk. These often belong to disposable domains, catch-alls, or addresses with weak authentication. Sending to these at full volume increases list churn and can hurt your sender reputation. It’s better to use them in lower-priority sequences—like a re-engagement campaign with reduced sender frequency.
Anything below 60 should be flagged. These are often role-based (e.g., [email protected]), outdated, or hosted on domains with poor sending records. Let your team review them manually, or exclude them entirely. This threshold aligns with industry benchmarks set by deliverability monitoring services like MxToolbox, which highlight scores under 60 as indicative of likely delivery issues or inbox filtering.
Remember: scoring is probabilistic, not deterministic. A 70 isn’t a guarantee it works—it’s an indication of risk. Use scores to manage expectations, not as a substitute for human judgment on edge cases.
How Emaillistchecker.io Balances Accuracy and Granularity in Verification
You get both a precise 0–100 email verification score and a clear categorical verdict (valid, invalid, catch-all, risky) for every address. This lets you clean lists with verdicts and prioritize engagement with scores. Accuracy is 98.9%, powered by real-time checks and historical data that keep scores updated as domains change.
Why Scores and Verdicts Together Work Better
Some tools return only a yes/no verdict. That's useful for cleanup, but not for planning. With a score, you see gradations — a 92 isn't the same as a 78, even if both are "valid." Let’s say you’re running a campaign. You can filter out all non-deliverable addresses using the verdict, then sort the rest by score to target high-potential leads first.
This dual output supports multiple workflows. Use the verdict to scrub list hygiene before sending. Then use the score to segment your audience — for instance, only send high-scoring addresses to your top-tier campaign. The balance isn’t theoretical: it’s baked into how email delivery works, where small differences in reputation or inbox placement can matter. According to RFC 6923, sender reputation and message context influence inbox placement, not just delivery status — so score granularity matters beyond just "valid" or "invalid."
How Accuracy Is Maintained and Enhanced
Our system performs real-time SMTP checks — connecting to the recipient server each time — but doesn’t stop there. It combines that with historical data on domain behavior, blocklist status, and catch-all patterns. A score is not static. If a domain starts showing new signs of spam traps or greylisting, the score adjusts. If an address was once invalid but is now active, that history informs the score without requiring a new check.
When results are ambiguous — a catch-all address with a high score, for example — our in-app AI assistant steps in. It evaluates the domain's reputation, recent changes, and how similar addresses behave across our database. Based on score, verdict, and context, it can suggest next steps: test deliverability, remove, or keep with lower priority. This reduces guesswork and helps teams act, not just analyze.
You can test this in practice with our inbox-placement feature, which simulates how your message lands in real inboxes. For high-volume operations, the API delivers verification in real time. Teams using our bulk verification tool see immediate reductions in bounce rates, often dropping from 12% to under 1% on clean data. Even without a full list, the score gives you a sense of risk — a low score often correlates with poor deliverability, regardless of the verdict.
Why Over-Reliance on Verdicts Can Hurt Your Campaigns
You don’t need to be a deliverability expert to know that a "valid" email isn’t always deliverable. Relying solely on categorical verdicts like "valid" or "catch-all" misses the full picture. A score of 55 may legally qualify as valid by some tools, but it still risks bouncing or landing in spam. Your campaign’s success hinges on delivery reliability — not just technical validity.
Not All "Valid" Emails Are Equal
Let’s be clear: a "valid" verdict doesn’t mean inbox placement. A low score — say, 55 out of 100 — often hides red flags. These addresses might pass basic syntax checks but fail on real-world delivery metrics. According to SendGrid’s deliverability research, emails with low verification scores have a 20–30% higher bounce rate and increased spam filtering risk.
You might think, “If it’s valid, it should work.” But validity is a baseline. It confirms format and domain presence. It says nothing about inbox placement, reputation, or long-term deliverability. A score of 95 has a far better chance of reaching the inbox than one at 55 — and that difference matters at scale.
Missing the Middle Ground
Some addresses land in the gray zone: moderate scores that are neither definitively valid nor invalid. These often include role-based, temporary, or legacy accounts. Tools that only report "valid" or "risky" may label them as "valid" while ignoring the risk. Yet, these can still be usable — especially in cold outreach or lead engagement campaigns.
For example, a sales team might use a tool that marks a [email protected] as “valid” at a score of 60. It may work once, but send consistently and it will likely get flagged. A score-driven approach lets you filter out the very worst, but a verdict-only mindset traps you in a binary model.
That’s where tools like EmailListChecker come in. Our bulk verification and real-time API provide both categorical verdicts AND a score-based assessment. You can set custom thresholds — say, filter all below 75 — to avoid low-performing addresses. This gives you both protection and flexibility.
Don’t let a simple yes/no label dictate your outreach strategy. The score gives you nuance; the verdict gives you speed. Use both. A score of 75 is not the same as 95. And in deliverability, the difference is measurable — and costly if ignored.
Best Practices for Using Scores and Verdicts Together
You should treat email verification scores (0-100) and categorical verdicts as complementary tools: use verdicts to filter out invalid and catch-all addresses immediately, then apply score thresholds to prioritize sending, segment campaigns, and detect underlying list health issues. Let’s break it down.
- Immediately remove all addresses flagged as invalid or catch-all. These either don’t exist or are too broad to be reliable. Including them harms deliverability and inflates bounce rates.
- Flag risky or greylisted addresses for suppression or send at lower frequency. Greylisting means the server temporarily delays delivery, often due to policy or load — these can disrupt timing and harm sender reputation.
- Send your primary campaigns only to addresses with scores above 85. This reflects strong technical validity and higher inbox placement probability. Scores above 90 indicate especially high confidence.
- Use scores to segment your list: send high-score addresses (90+) to your core audience, medium-score (70–89) to testing or nurturing streams, and low-score (below 70) to re-engagement or suppression queues.
- Track score distributions over time. A list consistently below 70 signals poor hygiene, outdated data, or overuse of purchased segments — issues that reduce deliverability and increase spam complaints.
Why Combining Scores and Verdicts Works Better
Verdicts tell you “why” an address failed or is suspect. Scores tell you “how reliable” it is. Relying on one without the other leads to over- or under-verification. For example, a “catch-all” verdict means the domain accepts mail for unknown addresses — a high risk for spam traps and bounces. A score of 40 on that same address confirms it's low confidence. You don’t need to debate the risk — act on both. Think of it as a two-stage filter: first, eliminate the obvious noise; second, prioritize the signal.
Many senders overlook real-time score distributions. Monitoring them helps spot trends — like a steady drop after a list acquisition or a spike after a campaign cleanup. That’s when you know it’s time to revisit your data sourcing. The Mimecast Email Security Threat Report notes that inconsistent list hygiene correlates with elevated spam complaints and filtering issues.
For the fastest, most accurate results, automate the process with our real-time API or bulk verification. It's not enough to clean once. You need ongoing scrutiny. A healthy list isn’t static — it evolves, and your verification strategy should too.
How Integration and Deliverability Testing Fit Into the Verification Score Picture
Verdicts and scores tell you whether an address is technically valid or risky, but only inbox-placement testing shows if it actually reaches the inbox. A 95-point score means the email is structurally sound, but without real-world testing, you can’t know if it lands in spam. Emaillistchecker.io’s inbox-placement feature confirms delivery into inboxes—using real email providers and real sending environments—so you’re not guessing.
Why Scores Alone Don’t Tell the Full Story
An email might pass all technical checks—valid syntax, existing MX records, no role accounts—but still end up in spam. This is because deliverability depends on behavior beyond syntax: sender reputation, content patterns, and domain history. A score of 80 doesn’t guarantee inbox delivery. It means the address exists, not that it will be received.
This is where inbox-placement testing becomes essential. It mimics the behavior of actual email clients and spam filters using real infrastructure. For example, Gmail’s filtering system evaluates not just email format but sender behavior, engagement metrics, and historical data. Testing in that environment gives you a real-world signal—something a score or verdict cannot.
Turning Scores Into Action With Integrations
Once you run a list through Emaillistchecker.io, you don’t need to manually clean it. Use the integrations with Mailchimp, HubSpot, and SendGrid to automatically flag and suppress low-scoring or risky addresses before sending. The system sends the verification results back with actionable rules so your platform filters bad data before launch.
This reduces bounce rates and improves sender reputation over time. When you stop sending to invalid or risky addresses, you’re not just reducing noise—you’re building a consistent sending pattern. According to Return Path, consistent sending behavior and low bounce rates are two of the top three factors influencing inbox placement.
Let’s be clear: a high score isn’t a guarantee. But when you’re using high-scoring addresses that have passed real inbox tests and are being sent through platforms that auto-suppress poor performers, you're actively improving your sender reputation. Each clean send lowers your risk, reduces spam complaints, and increases the chance your next campaign lands in the inbox. That’s the real outcome.
The Bottom Line: Use Both Numbers and Labels for Smarter Email Lists
A numeric score from 0 to 100 tells you how confident you can be in an email’s validity. It’s not a pass/fail, but a measure of likelihood. Use it to prioritize outreach and refine your segmentation.
Why Both Matter
Don’t rely solely on a categorical verdict like “valid” or “invalid.” A score gives context to borderline cases—like a 78 that may still deliver, but carries higher risk. Verdicts are fast filters; scores are confidence indicators.
- Use the verdict to flag obvious problems: invalid, role, disposable.
- Use the score to rank delivery likelihood and prioritize sends.
- Combine both to build a list that’s clean, safe, and effective.
Tools like Emaillistchecker.io deliver both—accurate scores and clear verdicts—backed by 98.9% accuracy. Credits never expire, and you can start with 100 free verifications.
Keep reading
- Email verification tools and services: how to choose (complete guide)
- Should Unknown Email Verdicts Be Cached at All in 2026?
- Mapping Verdicts from Different Providers to One Internal Schema
- Liability Caps and Indemnity in Verification Vendor DPAs 2026
- Clearout vs MillionVerifier: Which Is More Accurate in 2026?
Ready to put this into practice? Emaillistchecker.io verifies emails with 98.9% accuracy — start with 100 free verifications.
Frequently asked questions
What does an email verification score of 75 mean?
A score of 75 indicates moderate confidence the address is valid, but it carries higher risk than addresses above 85. It may be borderline deliverable and should be treated as risky.
Can a score of 90 be wrong?
Yes — no verification system is 100% accurate. A score of 90 means high confidence, but it doesn’t guarantee inbox delivery. Factors like greylisting or temporary outages can still affect real-world results.
Why does Emaillistchecker.io give both a score and a verdict?
It allows flexible workflows. Use the verdict for fast filtering, and the score for segmentation, prioritization, and risk assessment.
Is a high email verification score enough to avoid spam filters?
No. Score correlates with deliverability but doesn’t guarantee inbox placement. Additional factors like sender reputation, content, and engagement matter.
Should I remove all addresses with a score below 70?
Addresses below 70 are high-risk. If you're aiming for consistent inbox delivery, it's advisable to remove them or avoid sending to them unless part of a targeted test.
Can a catch-all domain have a high verification score?
Yes — the score may be high because the system receives a positive SMTP response, but the domain is still catch-all. The verdict warns of this risk, even if the score is high.
Do disposable email addresses ever get high verification scores?
Not when properly detected. Emaillistchecker.io flags disposable domains before scoring, ensuring they don't receive misleading high confidence ratings.
How does Emaillistchecker.io handle role-based emails?
It detects role addresses (e.g. sales@, info@) and assigns a 'risky' verdict, even if the score is high, due to low engagement and high bounce likelihood.
Are numeric verification scores standardized across tools?
No. Each tool uses its own scoring model. Emaillistchecker.io's 98.9% accuracy is based on independent testing, but absolute values may vary between platforms.
How do I use verification scores in Klaviyo or HubSpot?
Emaillistchecker.io integrates directly with Klaviyo, HubSpot, and other platforms. After verification, you can map scores and verdicts into customer segments for tailored campaigns.
Is Emaillistchecker.io worth using if I already use another tool?
Yes — it provides both numeric and categorical outputs, which many tools lack. Its 98.9% accuracy and no-expiration credits make it a reliable complement for list hygiene.
Can I test deliverability after verification?
Yes. Emaillistchecker.io includes inbox-placement testing, which verifies whether verified addresses actually land in inboxes, not spam folders.