Document OCR
Generic tax document OCR software extracts characters. Paloma extracts form data - Box 1 is distinguished from Box 12, a 1099-DIV from a 1099-B, with automatic double-checks before your team sees a number.
Off-the-shelf scanners output text soup. Tax work needs Box 2 as a number in the right field, not '$4,2Bl.OO' somewhere in a blob.
One transposed digit at entry becomes a wrong workpaper, a wrong return, and an amended filing.
Clients send sideways, shadowed kitchen-table shots - and traditional OCR simply gives up on them.
What it does
Wages, withholding, distributions, coded K-1 boxes - each field lands in its named place, not a text dump.
Phone photos are auto-righted and converted; scans of any quality get the same box-aware read.
Low-confidence fields automatically get a targeted second pass before anyone relies on them.
What can't be read confidently is flagged for human eyes - uncertainty is surfaced, not hidden in a cell.
Every extracted value displays next to the exact spot on the original form it came from - trust by inspection.
A client uploads a form.
Every box is extracted automatically.
It double-checks numbers it is not sure about.
Your team sees each number next to the real form.
Form reading is the engine inside AI tax preparation: documents arrive through collection, get read here, and feed straight into workpapers.
Common Questions
15 minutes. Upload a crumpled W-2 photo and watch every box land in the right place.
Pick a timePrefer to write first? Contact us about Data Extraction.