AI document extraction

Your paperwork is a dumpster fire.
Dump the doc. Keep the data.

XDumps reads the messy PDFs and phone-photo scans nobody wants to retype — invoices, bills of lading, receipts, rate confirmations — and hands back clean, structured data in seconds.

See it work in seconds — no signup, no install. Or grab the toolkit and run it yourself.

What you get today

SCANNED UPSIDE DOWN

What XDumps gives back

vendor_nameKeystone Freight LLC
date2026-07-08
total_amount$4,182.50
line_items7 rows, itemized

How it works

Three steps. Under thirty seconds per document. No template setup, no "training period."

STEP 01

Dump it in

Drag in a PDF or a photo of the document — crooked scans, coffee stains, and fax artifacts included. If a human can read it, XDumps can.

STEP 02

AI does the retyping

A vision-capable AI model reads the document like a person would and extracts vendor, date, totals, and every line item into a strict, validated schema.

STEP 03

Keep the data

Get clean JSON and a readable table instantly. Paste it into your TMS, spreadsheet, or accounting system — or wire the API straight into your workflow.

Built for people buried in paper

If part of someone's job is retyping documents into a system, XDumps pays for itself the first week.

🚛 Freight brokers & carriers

  • Carrier invoices → audit against the rate confirmation
  • Bills of lading → shipment records without the data-entry clerk
  • Lumper & accessorial receipts → clean cost tracking per load

🔧 Trade subcontractors

  • Supplier invoices → job-cost tracking that's actually current
  • Material receipts from 5 different supply houses → one clean ledger
  • Delivery tickets → verify what hit the jobsite vs. what was billed

📚 Bookkeepers & back offices

  • Client shoeboxes of receipts → itemized, dated, totaled
  • Vendor bills → structured rows ready for the accounting import
  • Month-end catch-up in hours, not weekends

Pricing

Start with the self-hosted toolkit today, or get on the hosted pilot where we run everything for you.

XDumps Cloud

Free first 25 docs
Nothing to install. Then pay per document — from 4¢.
  • Extract in your browser at xdumps.com/app
  • First 25 documents free — no card required
  • Then simple credit packs: 250/$19 · 1,000/$59 · 5,000/$199
  • Copy results as JSON or CSV
  • Credits never expire
  • Email-in option: forward a doc, get data back
Start free — 25 documents

Get started

Tell us what you're drowning in. We'll follow up within one business day — usually with your first documents already extracted.

FAQ

Straight answers.

Is this just OCR?

No. OCR gives you a blob of text and leaves the hard part — figuring out which text is the vendor, the total, or line 4's unit price — to you. XDumps uses a vision AI model that reads the document the way a person does and returns validated, structured fields. It handles crooked scans, photos, and layouts it has never seen before, with no templates to configure.

How accurate is it?

On clean, legible documents, field-level accuracy is very high — and every extraction shows you the result next to the source so you can verify in seconds instead of retyping for minutes. For anything the model can't find, you get an explicit null, never a guess dressed up as an answer.

What happens to my documents?

With the self-hosted toolkit, files stay on your machine and are sent only to Anthropic's API for extraction under their commercial data terms (not used for training). With the hosted pilot, documents are processed and retained only as long as you need the results.

What does the self-hosted toolkit require?

A computer with Python (free, and the setup guide walks you through installing it), plus an Anthropic API key. Typical cost per document is one to three cents. If you can install a printer, you can install this — and if you get stuck, email us.

Can it push data into my TMS / QuickBooks / spreadsheet?

Today you get clean JSON and a copy-paste-ready table; the API endpoint means anything with an import or a Zapier hook can consume it. Direct integrations are prioritized by what pilot customers ask for — tell us what you need in the form above.

Why is it called XDumps?

Because that's what your document pile is — a dump. We're not going to pretend paperwork is a "digital transformation journey." Dump the doc, keep the data, get back to work.