Skip to main content

Scanned & Image-Only PDF BOQ Handling

A scanned or image-only PDF contains page images rather than selectable digital text. Quantara detects these pages, but it does not currently extract their text with Optical Character Recognition (OCR).

Who Encounters Scanned BOQs?

Teams working with legacy documents or physical tender packages regularly receive scanned or image-only PDFs.

  • Estimators receiving physical printouts
  • Contractors archiving legacy project data
  • Consultants processing third-party hardcopies
  • Subcontractors dealing with faxed or low-quality scans

The Challenge of Image-Based Documents

Unlike text-based PDFs where digital characters are stored in the file, a scanned PDF contains page images rather than selectable text.

Scanned BOQs can contain skewed or blurred pages and handwritten annotations. Quantara detects image-only pages but does not currently extract their text with OCR, so the team must use a manual transcription and review path.

What Quantara Does With Scanned PDFs Today

Quantara rasterizes PDF pages and classifies them as text-based, scanned/image-only or mixed. Scanned and image-only pages are flagged as requiring OCR; no text is invented or guessed for them.

OCR text recognition is not currently available in Quantara. Scanned BOQ content must be transcribed manually and checked against the rendered source page.

Relevant Features

Scanned/Image-Only Detection

Detects image-only pages and reports that text extraction is unavailable.

Limited

OCR Text Recognition

Automated conversion of image-based text into selectable digital data is not currently available.

Not available

Manual Review Requirement

All manually entered data requires professional review before commercial use.

Available

Handling a Legacy Scanned BOQ Today

How a team currently handles a physical tender package without OCR text extraction:

1

Scan & Upload

The physical document is scanned to PDF and uploaded.

2

Automatic Detection

Quantara rasterizes the pages and flags them as scanned/image-only, requiring OCR.

3

Manual Transcription

Since automated OCR is not yet available, the team manually transcribes quantities and descriptions from the page images.

4

Data Structuring

The transcribed data is organized into the digital BOQ hierarchy.

5

Professional Review

A qualified professional verifies every transcribed item before commercial use.

Supported Inputs

Scanned/Image-Only PDF — Detection

Limited

Detects image-only pages and reports that text extraction is unavailable.

Note: Quantara does not currently perform OCR text extraction from scanned or image-only PDFs.

Scanned/Image-Only PDF — OCR

Not available

Automated text recognition for scanned pages is not currently implemented.

Note: Scanned content requires manual transcription.

Text-based PDF

Available

Supported digital PDFs with an existing text layer.

Note: Plain paragraph text is not automatically converted into BOQ candidates; table results depend on the source layout and must be checked against the original file.

Supported Outputs

Structured Database

Available

Authorized project storage for reviewed records.

XLSX Export

Available

Export reviewed tabular data to XLSX.

Note: Generated documents are not professional approval and remain subject to project-specific review.

Current Limitations

  • Automated OCR for scanned/image-only PDFs is not currently implemented, so affected content must be transcribed manually.
  • Scanned/image-only detection does not attempt to guess or reconstruct text; it reports that OCR is required.
  • Every manually transcribed item, unit and quantity requires professional review against the source image.

Professional Disclaimer

Project information, extracted data, measurements, calculations, rates and outputs require review by the responsible construction professional before tender, procurement, contractual or construction use.

Frequently Asked Questions

What is a scanned PDF?

A scanned PDF is an image-based file whose text cannot be selected as a normal digital text layer.

Does Quantara currently OCR scanned BOQ documents?

No. Quantara detects and flags scanned or image-only pages, but OCR text extraction is not currently available, so the content requires manual transcription.

What happens after Quantara detects a scanned page?

The page is identified as image-only and kept available for review. The project team must use a manual transcription path and compare entered information with the source image.

How is manually transcribed scanned data checked?

Every item, unit and quantity entered from a scanned document should be cross-checked against the original rendered page by the responsible professional.

Can manually transcribed scanned information be exported?

After the information has been entered, reviewed and structured in a supported workflow, it can be included in supported outputs such as XLSX.

Does Quantara perform drawing takeoff?

No. Quantara does not perform automatic measurement or quantity takeoff from scanned drawings.

Related Resources

Ready to review a structured BOQ workflow?

Is Quantara right for your team?

Our automated sales advisor helps you choose your next step. OpenAI may process your message. Please leave out confidential information. Privacy

Scanned PDF BOQ Detection and OCR Status | Quantara