Drop a file in a folder and have it processed automatically. It feels like magic, but it’s easy to set up—and there’s masses of further potential.
Fast Rust library for PDF classification and text extraction. Detects whether a PDF is text-based or scanned, extracts text with position awareness, and converts to clean Markdown — all without OCR.