LATAM context
Modern LATAM banks generally deliver native PDFs from online banking. Some older formats or re-scanned printouts (when a customer prints a statement and scans it again) still need OCR for extraction.
Concrete example
A BBVA México PDF downloaded from Net Cash is native: you can select and copy the text. A PDF you received over WhatsApp after someone printed and re-scanned it is a scanned PDF, and it needs OCR.
How it shows up on your bank statement
It's the first thing worth checking before processing a batch of statements. A native PDF carries its text inside the file and reads exactly; a scanned one is an image, and every character has to be reconstructed. The same bank can deliver both formats depending on the channel: the one downloaded from online banking is usually native, and the one printed at a branch stops being native the moment you scan it.
How does finO$ handle this?
finO$ automatically detects the PDF type: native PDFs go straight to structured parsing (faster, 99%+ accuracy). Scanned ones go through OCR first (97%+ accuracy on good-quality scans).