👤 807 total uses◯ Free: 5 uses/day • Resets in 0h 36m

Batch OCR

Upload a ZIP of PDFs or images (up to 25 MB) and get one combined markdown file with all the text.

Learn more

Batch OCR extracts text from many documents at once. Upload a ZIP containing PDFs or images (up to 25 MB total) and get back a single combined markdown file where each source file becomes its own heading. Powered by Mistral OCR, it is built for digitizing stacks of paper documents, scanned archives, and multi-file bundles in one pass.

Drop your ZIP of PDFs/images here
or click to browse — ZIP up to 25 MB. We extract up to 10 files into one markdown.

✓ Free to use — no signup, no credit card.

Small Business

Scanned invoices batch

Bulk-convert scanned invoices into searchable text records.

See input + output preview

Input

Content
A folder of 30 scanned supplier invoice PDFs that need their text extracted in bulk into searchable, copyable documents.

Output (excerpt)

Across all 30 files the tool returns extracted text per page, e.g. "Invoice #INV-2041 — Acme Supplies — Date: 2026-05-22 — Net 30 — Total Due: $1,284.00 — Item: Office chairs x4." Each scanned PDF becomes searchable, copyable text, letting you index, search, and pull totals from the whole batch at once.
Students

Archive document scan

Digitize archive pages in bulk for a research project.

See input + output preview

Input

Content
A set of 50 photographed pages from old library archive books that need to be converted to editable text for a research project.

Output (excerpt)

The tool processes all 50 photographed pages and outputs clean text per page, preserving paragraph breaks, e.g. "The expedition departed in the spring of 1887, recording detailed observations of the coastal flora..." The batch is delivered as editable text files ready to search, quote, and cite in the research project.
Freelancers

Receipt pile OCR

Process a quarter's worth of receipts in a single batch.

See input + output preview

Input

Content
A batch of 20 photos of business receipts that need their text pulled out together for a quarterly expense log.

Output (excerpt)

All 20 receipt images are processed at once, returning per-file text such as "Staples — 2026-04-09 — Printer Paper $12.99, Ink Cartridge $34.50 — Total $47.49." The extracted text lands in a consolidated list, making it easy to copy totals and merchants into a quarterly expense log in one pass.

Your Batch OCR results will appear here

You'll get clean markdown with tables, equations, and headings preserved — ready to paste or edit.

How to Use Batch OCR

  1. Put your PDFs and images into a single ZIP archive (up to 25 MB total).
  2. Upload the ZIP to Batch OCR.
  3. Run the extraction and wait while each file is processed.
  4. Download the combined markdown file with one heading per source document.

Use Cases

1

Digitize a folder of scanned paper documents in one pass

2

Convert a batch of invoices or forms into searchable text

3

Extract text from a multi-page scanned archive bundle

4

Turn a set of image-only PDFs into editable markdown

5

Build a searchable knowledge base from a stack of legacy files

Tips for Best Results

  • Scan documents at 300 DPI or higher for the cleanest text recognition.
  • Name the files inside the ZIP clearly, since those names guide the headings in the output.
  • Keep the total ZIP under 25 MB by compressing images or splitting into multiple batches.
  • Review tables and complex layouts in the markdown, as intricate formatting may need minor cleanup.

Frequently Asked Questions

What does Batch OCR do?

It reads text from multiple documents at once and merges the results into a single markdown file, with a heading for each source file so you can tell the content apart.

What do I upload and what are the limits?

Upload one ZIP archive containing your PDFs and image files, with a combined total of up to 25 MB. Compress or split larger sets to stay under the limit.

What format is the output?

You get one combined markdown (.md) file. Each original file appears under its own heading, with the extracted text in reading order beneath it.

How accurate is the text extraction?

Mistral OCR is accurate on clear printed text and preserves structure like headings and lists well. Low-resolution scans or heavy handwriting may reduce accuracy.

Does it keep tables and formatting?

It preserves structure such as headings, lists, and tables in markdown where possible, though very complex layouts may need light cleanup afterward.

Can I use the extracted text commercially?

Yes, you can use the output in your own archives, documents, and products. Free covers 5 batch runs per day with no signup; Pro is $19/month for higher volume.

What happens to the files I upload?

Your ZIP and its contents are processed only to extract the text and are then discarded. They are not stored or used to train models.

🔒
Your Privacy is Protected

We don't store your text. Processing happens in real-time and your input is discarded immediately after generating the result.

Unlock Unlimited Access

Free users: 5 uses per day | Pro users: Unlimited

Related tools

Try this agent

Student ResearchGo from topic to thesis to full paper outline in one run — including a formatted…Try this agent →

Related workflow

Podcast → Tweet ThreadUpload a podcast audio file → transcribe → ship a 7-tweet thread + hashtag pack.Run workflow →

Read more