Bank statement OCR that keeps rows intact

I’m preparing a funds-trace on 11 accounts (about 24,300 lines) and need OCR that preserves line order and column fidelity in Bates-stamped PDFs so the export drops cleanly into IDEA/Power Query; Adobe and ABBYY keep merging carryover lines and subtotal rows, which breaks exhibit tie-outs. If you’ve had consistent results with a specific profile or tool, what settings or engines kept date/description/amount aligned for production-quality CSVs?

‌⁠‍⁠​‍​‍‌⁠‌​​‍​‍​⁠‍‍​‍​‍‌‍‌‍‌‍⁠⁠‌⁠​‍‌‍‌‌‌‍⁠‍‌⁠​⁠‌‍‍‌‌‍​⁠‌‍​‌‌‍​⁠‌‍​⁠‌‍⁠⁠‌⁠‌‌‌‍⁠‍‌⁠‌​‌‍​‌‌‍⁠‍‌⁠‌​​‍​‍​‍⁠​​‍​‍‌‍‍⁠​‍​‍​⁠‍‍​‍​‍‌‍⁠‍‌‍‌‌‌⁠‌⁠‌‌⁠⁠‌⁠‌​‌‍⁠⁠‌⁠​​‌‍‍‌‌‍​⁠​‍​‍​‍⁠​​‍​‍‌‍‍‌‌‍‌​​‍​‍​⁠‍‍​‍​‍‌‍⁠‍‌‍‌‌‌⁠‌⁠​‍​‍​‍⁠​​‍​‍‌‍‌​​‍​‍​⁠‍‍​‍​‍​⁠​‍​⁠​​​⁠​‍​⁠‌‌​⁠​‌​⁠​‍​⁠​‌​⁠‌⁠​‍​‍​‍⁠​​‍​‍‌‍‍​​‍​‍​⁠‍‍​‍​‍‌​​‍‌​​‍‌‍‌‍‌⁠​‌‌‌​‍‌⁠​‌‌‍⁠‌‌​⁠‌‌‌‌​‌‍​‌‌⁠‍‌‌‍‌​‌‍​‍‌⁠‌⁠‌‍‌​‌‌‌​​‍​‍‌⁠⁠‌​​

OCRmyPDF and Tesseract 5 with ‘tsv’ and psm 6 kept rows; flatten Bates first. https://ocrmypdf.readthedocs.io.

‌⁠‍⁠​‍​‍‌⁠‌​​‍​‍​⁠‍‍​‍​‍‌‍‌‍‌‍⁠⁠‌⁠​‍‌‍‌‌‌‍⁠‍‌⁠​⁠‌‍‍‌‌‍​⁠‌‍​‌‌‍​⁠‌‍​⁠‌‍⁠⁠‌⁠‌‌‌‍⁠‍‌⁠‌​‌‍​‌‌‍⁠‍‌⁠‌​​‍​‍​‍⁠​​‍​‍‌‍‍⁠​‍​‍​⁠‍‍​‍​‍‌⁠​‍‌‍‌‌‌⁠​​‌‍⁠​‌⁠‍‌​‍​‍​‍⁠​​‍​‍‌‍‍‌‌‍‌​​‍​‍​⁠‍‍​⁠​‌​⁠​​​⁠‍​​‍⁠​​‍​‍‌‍‌​​‍​‍​⁠‍‍​‍​‍​⁠​‍​⁠​​​⁠​‍​⁠‌‌​⁠​‌​⁠​‍​⁠​‍​⁠‌​​‍​‍​‍⁠​​‍​‍‌‍‍​​‍​‍​⁠‍‍​‍​‍‌⁠​‌​‍⁠‌‌‍‍​‌‌‌⁠‌‍‌​‌‌​‌‌‌‌‍‌‍​‍‌‌‌‌‌⁠‌‍‌‍​⁠​⁠​‌‌‍‌​‌​⁠⁠‌‍‍‌‌⁠‌⁠​‍​‍‌⁠⁠‌​​