How to Evaluate Email Parsing Services - 95% Accuracy Is Not a Good Score

Key Takeaways

  • 95% accuracy sounds like a pass. At 2,000 emails a month it is 100 emails landing back on somebody's desk, every month, forever.
  • The billing unit decides your bill, not the plan name. With Parseur one email is 1 credit whatever its length, and a three-page PDF attachment is 3.
  • Attachments are where trials quietly fail. Send your worst real scan, not a clean sample.
  • Ask who maintains the parser the day a supplier redesigns their order confirmation. That answer is your running cost, not the sticker price.
  • "Compliant" and "certified" are two different words. Parseur is GDPR compliant, EU-hosted, and SOC 2 Type II compliant. HIPAA is in progress, not issued.

Start with the arithmetic, not the feature grid

Run 2,000 emails a month through a parser that is 95% accurate and you have not bought a 95% solution. You have bought 100 emails a month landing back on somebody's desk, every month, forever. Take it to 99% and 20 are still there. The question was never how accurate it is. The question is what it does with the ones it gets wrong, and whose Tuesday afternoon that becomes.

Most guides on how to evaluate email parsing services hand you a feature grid instead, and wish you luck. The grid is the problem. Every vendor on your shortlist supports attachments, integrations, and AI-powered extraction, so it comes back all green and you are exactly where you started.

The questions that separate these tools are duller and far more specific. What does this cost at my volume, not at the entry tier? What happens to the emails it cannot read? Who fixes it when a supplier redesigns their order confirmation? None of that fits in a checkbox, and all of it arrives in month three.

Your shortlist is also going to get longer. Archive Market Research puts the email parsing software market at roughly $2 billion, growing at a 15% compound annual rate through 2033, which buys you more vendors, more feature grids, and more green ticks. Not one of them will make the decision for you.

The ten criteria, and what to measure

The short version. Each row is a question to put to every vendor, plus the measurement that answers it.

# Criterion What to measure
1 Format variability How many distinct senders and layouts, and what breaks when one changes
2 Attachment handling Whether your worst real scan parses, not your cleanest sample
3 Exception handling The monthly count of emails a human must touch
4 Integration fit Whether the data lands in your system without a middle step
5 Volume pricing Cost per email, per page, or per credit, at your real monthly figure
6 Routing and mailboxes Whether leads, orders, and invoices can be separated cleanly
7 Line items and tables Whether a five-line order returns five rows or one
8 Maintenance burden Who updates it, and how long a new layout takes
9 Data handling Retention window, training use, and which certificates actually exist
10 Reliability Public status page, documented retries, and what happens to lost mail

1. How much do your email formats vary?

This decides which category of tool you are shopping for, so answer it before you book a single demo. One system sending one layout every time? A rule-based parser will serve you well and cost you little. Thirty senders, three marketplaces, and a web form that marketing keeps redesigning? Rules become a part-time job.

Count your distinct senders. Then ask each vendor what happens when one of them changes their layout without warning, because one of them will, usually on a Friday.

How Parseur handles it: you name and describe the fields you want, and the AI parsing engine finds them across layouts with no template to define first. A redesigned notification email usually needs no work at all. If you would rather have the tighter control of one template per layout, that engine exists too. We put the two approaches head to head if you want the long version.

2. What is inside the attachments, scans included?

This is where trials go wrong quietly. Your team tests with a clean PDF, everything works, the contract gets signed, and then the first purchase order arrives from the customer who prints it, signs it, photographs it on a desk in bad light and emails you the photo.

Ask for the format list in writing, and read it for what is missing rather than what is present. Some well-known email parsers do not handle images or scanned documents at all, and that limit tends to live on a pricing page rather than come up in a sales call.

How Parseur handles it: it extracts data from the email body and from attachments including PDFs, scans, images, CSV files, Excel, MS Word, text files, and web pages. The printed, signed, photographed document is an image, and images are on that list. Do not take our word for it though. The free tier is 20 pages a month, which is 20 chances to throw your ugliest real paperwork at it before anyone asks you for a credit card. A clean sample passes everywhere and proves nothing.

3. What happens to the emails it cannot read?

Every parser fails sometimes. The ones worth buying fail loudly.

At 2,000 emails a month, 95% accuracy leaves roughly 100 exceptions to handle by hand. What you need is not the percentage but the path: does the failed email land in a review queue, does anyone get alerted, can you fix it and reprocess it, and can anything disappear without a trace? A parser that silently drops a malformed order costs far more than one that flags it, because the bill arrives weeks later as a customer complaint.

Ask to see the failure path during the trial. Send a deliberately broken email and watch where it goes.

Gartner puts the average cost of poor data quality at $12.9 million a year per organization. Scale that down to your forty people and it is still somebody's afternoon spent unpicking an order that shipped to the wrong address, and bad data extracted automatically spreads faster than bad data typed by hand.

How Parseur handles it: it flags what it could not extract instead of guessing, and processed documents stay available for reprocessing inside your retention window.

4. Extraction is not the finish line

A parser that cannot reach your system leaves you copying and pasting from a nicer interface. Write the destination down before you compare tools: CRM, ERP, spreadsheet, database, warehouse, or your own API. Then check whether the vendor reaches it natively or quietly expects you to build the last mile yourself.

How Parseur handles it: data goes out through Zapier, Make, Power Automate, or a direct webhook, which covers thousands of destinations without anyone writing code. New to the mechanics? Start with what an email parser is, then how unstructured content becomes structured data.

5. What does it cost at your volume, not the entry tier?

Entry plans are marketing. Compare at your real monthly figure plus a year of growth, and pin down the billing unit before anything else, because per email, per document, per page, and per credit are four different bills for the same mailbox.

Put these to every vendor before Friday's budget conversation:

  • Is pricing per email, per document, per page, or per credit?
  • Do attachments bill separately from the email that carried them?
  • Does a five-page PDF cost one unit or five?
  • Are failed parses billed?
  • Do reprocessed emails bill twice?
  • Are extra mailboxes and users included?
  • What happens when you go over?

How Parseur handles it: one credit equals one page. An email counts as a single page regardless of length, so 2,000 plain emails cost 2,000 credits. A spreadsheet is one page too, so a CSV with 100 rows costs 1 credit. PDFs bill per page, so a three-page invoice costs 3 credits. Which means 2,000 emails each carrying a three-page PDF is 8,000 credits, not 2,000, and working that out before you buy is the entire reason to ask. Mailboxes are unlimited on every plan, including the free tier that starts at 20 pages a month. Take your credit figure to the pricing page and the plan picks itself.

6. Can you route different email types separately?

At any real volume you stop having one inbox and start having several. Leads, orders, invoices, and support requests want different fields, different destinations, and different people watching the exceptions.

Check whether the tool gives you separate mailboxes with their own extraction settings, or one bucket with rules bolted on. The bucket works right up until two email types collide.

How Parseur handles it: each mailbox has its own fields and its own extraction settings, and you can create as many as you need on any plan, free tier included. For what other teams route this way, we collected seven common email parser use cases.

7. A five-line order should return five rows

Order emails rarely contain one thing. They contain five SKUs, quantities, unit prices, tax, freight, and a total. A parser that returns only the first row is worse than one that fails outright, because the gap stays invisible until someone notices the missing stock. Test with a genuinely multi-line order during the trial. Single-item test emails hide this failure completely.

How Parseur handles it: it extracts tables and tabular data from both the email body and its attachments, so a five-line order returns five rows.

8. Who maintains it in six months?

The best tool is the one your operations team can run without booking engineering time. Ask who updates the parser when a sender changes their layout, how long onboarding a new format takes, whether a non-technical person can do it, and whether there is a visual editor or a rules language to learn.

That answer compounds. A parser that needs an engineer twice a month is not cheaper than one that costs more per page.

Those hours come out of a week that is already short. A 2012 McKinsey study put 20 to 35% of working time into searching and gathering information. Every hour of parser maintenance is drawn from the same pot you bought the parser to protect.

How Parseur handles it: point-and-click, self-serve, no rules to write and no template to maintain when you use the AI engine. To see the setup before signing up for anything, we walk through creating an email parser step by step.

9. What happens to your customers' data?

Order and lead emails carry customer names, addresses, phone numbers, and payment references. That makes this a procurement question, not an IT footnote.

Four things to establish in writing:

  • Where is the data hosted, and under which regulations?
  • What is the retention window, and can you set it?
  • Is your data used to train the vendor's AI models?
  • Which certifications exist, and when were they issued?

That last one matters more than it looks. "Compliant", "aligned with", and "certified" are three different claims, and vendors are not always careful about which one they are making.

How Parseur handles it: Parseur is GDPR compliant, hosted in the European Union in an ISO 27001 data center, and aligned with UK GDPR, California CCPA and CPRA, and Singapore PDPA. Customer data is never reused to train our AI models and is never sold. Retention runs from 1 day to unlimited depending on your plan, and you set it in your mailbox settings. Parseur is SOC 2 Type II compliant, and you can request the report through the trust center. HIPAA compliance work is still in progress and not yet certified. We would rather write that plainly here than let it blur in a sales call. The security page has the detail.

10. The day it goes down

Email forwarding is part of the pipeline, so an outage is not an inconvenience. It is a set of orders you never received. Ask for a public status page, a documented retry policy, and a straight answer on what happens to mail that arrives during an incident. If a vendor cannot produce a status page, you have your answer.

How Parseur handles it: up to 99.99% uptime, published on a public status page, with retry mechanisms so documents are not lost when a downstream service has a temporary disruption.

There is also a human on the other end of that incident, and support quality is a product decision rather than a cost line. A Forbes survey found that 83% of companies whose CEOs are directly involved in customer experience report higher customer satisfaction, and 58% report increased revenue. Test it yourself during the trial. Ask something hard on a Friday afternoon and time the reply.

Run your shortlist in two weeks

Choosing an email parser is not a software comparison. It is a forecast of how much manual work your team will still be doing in six months.

Take all ten questions to every vendor at once, using your own emails rather than their samples. Shortlists collapse fast once you stop reading feature grids and start counting exceptions, credits, and the number of people who have to touch the thing each month. Compare accuracy percentages instead, and you will meet the second number hiding behind 95% at the worst possible moment.

If you want the head-to-head rather than the framework, we tested and compared the leading email parsers against each other. And if one specific question is still nagging, the email parser FAQ probably covers it.

Where Parseur lands on all ten

Parseur is an AI email parser built for teams who want the exceptions to stop, not a nicer interface for handling them.

Criterion Parseur
Format variability AI engine adapts across layouts, no template required
Attachments PDFs, scans, images, CSV, Excel, Word, text, web pages
Exceptions Flags what it could not extract, reprocessing available
Integrations Zapier, Make, Power Automate, webhooks, API
Pricing unit 1 credit = 1 page. Email or spreadsheet = 1 credit. PDF = per page
Mailboxes Unlimited on every plan, free tier included
Line items Tables extracted from body and attachments
Maintenance Point-and-click, self-serve, no rules to write
Data EU-hosted, ISO 27001, GDPR compliant, never used for training. SOC 2 Type II compliant
Reliability Up to 99.99% uptime, public status page, retry mechanisms

The free tier is 20 pages a month. Start with the twenty documents you least want to explain to a salesperson.

Sign up to Parseur for Free
Try out our powerful document processing tool for free.

Last updated on

Get started

Ready to automate your
document data extraction?

Start free in minutes and see how Parseur fits into your workflow.

No model training required
Automates data entry from any document
Scales from point-and-click to API

Frequently Asked Questions

The questions buyers ask when they are past the demo and trying to work out what a parser will actually cost them, break on, and refuse to do.

Ten questions, in this order:

  1. How much do your email formats vary, and how many senders do you have?
  2. What is inside the attachments, including the scanned and photographed ones?
  3. What happens to the emails the parser cannot read?
  4. Where does the extracted data have to land?
  5. What does it cost at your real monthly volume, not at the entry tier?
  6. Can you route leads, orders, and invoices separately?
  7. Can it pull repeating line items and tables?
  8. Who maintains it when a sender changes their layout?
  9. What happens to your customers' data, and which certificates actually exist?
  10. What happens when the service has a bad day?

Accuracy alone is a bad first question. The number that decides your workload is not the accuracy rate, it is the pile of exceptions left behind. At 2,000 emails a month, 95% accuracy leaves roughly 100 of them.

With Parseur it is one credit per page, so a three-page PDF costs 3 credits and a ten-page PDF costs 10. Emails and spreadsheets are the exception: they count as a single page regardless of length, so a CSV with 100 rows still costs 1 credit. Ask every vendor this question specifically, because per-document and per-page pricing can differ by a factor of ten on the same mailbox.

Some can and many cannot, and this is the most common reason an evaluation ends in a refund request. Parseur extracts data from emails and from attachments including PDFs, scans, images, CSV, Excel, Word, text files, and web pages. The document a customer printed, signed, photographed on a desk and emailed back is an image, and images are on that list. Send that one during the trial, not a clean sample, because clean samples pass everywhere.

Parseur is GDPR compliant, EU-hosted in an ISO 27001 data center, and aligned with UK GDPR, California CCPA and CPRA, and Singapore PDPA. Parseur is SOC 2 Type II compliant, with the report available on request, while HIPAA compliance work is still in progress and not yet certified. That distinction is worth stating plainly because plenty of vendors let "compliant" and "certified" blur together in a sales call. Ask any vendor to name the certificate and the date it was issued.

With an AI parsing engine you name and describe the fields you want and the engine finds them, so a new sender layout usually needs no work at all. With a template engine you build a template per layout, which is more predictable but means someone maintains it every time a sender redesigns their notification email. The difference between the two approaches is the single biggest driver of long-term maintenance cost.

Ask for a public status page and a documented retry policy, and treat the absence of either as an answer. Parseur runs at up to 99.99% uptime with retry mechanisms so that documents are not lost during a temporary disruption, and the status page is public. Email forwarding is part of the pipeline, so a parser that silently drops messages during an incident costs you orders, not uptime percentages.

With Parseur, 2,000 plain emails cost 2,000 credits, because one email is one credit no matter how long it is. Attachments are what move the number: a three-page PDF invoice attached to each of those emails adds 3 credits per email, so the same 2,000 emails become 8,000 credits. Work out your real credit figure first, then price it. Doing it the other way round is how teams end up on the wrong plan in month two. Full rules are on the pricing page.

This is the question most evaluations skip, and it is the one that decides how much manual work you keep. At 2,000 emails a month, a parser that is 95% accurate hands you roughly 100 emails to fix by hand every month. Ask to see the failure path before you buy: is there a review queue, does it alert you, can you reprocess a failed email, and does anything get silently dropped.

Parseur extracts tables and tabular data from both the email body and its attachments, which is what order emails with multiple SKUs, quantities, and prices need. This is where basic parsers tend to fail, so test it with a real multi-line order rather than a single-item one. A parser that returns only the first row of a table is worse than useless, because the gap is invisible until someone notices the missing stock.

Not at Parseur. Customer data is never reused to train our AI models and is never sold. Retention is yours to set, anywhere from 1 day to unlimited depending on your plan, and documents are deleted automatically when the window closes. Ask every vendor the same two questions in writing, because the answer sits in the terms rather than on the homepage.

No. Parseur is self-serve and point-and-click, and the data goes out through Zapier, Make, Power Automate, or a webhook if you want it in your own stack. Any evaluation that starts with a mandatory sales call is telling you something about how the product is built.

Rule-based parsers are excellent when your emails are genuinely consistent, and brittle the moment a sender changes their layout. AI parsers cost less to maintain across varied formats and need no template building. If your emails come from more than a handful of senders, or from a marketplace that redesigns its notifications without telling you, the maintenance argument decides it. We put the two approaches head to head if you want the long version.