Hello,
My use case is extracting semantic content from PDFs for downstream chunking and RAG processing. I am currently evaluating Aspose.OCR for Python via .NET (aspose-ocr-python-net==26.5.0) and ran into layout detect…...added my own reading-order correction, especially for multi-column... tables, formulas, figures/images, and other semantic elements...