extracting-with-ocr
xberg-io/xberg
Extract text from scanned PDFs, photos, screenshots, and images without a text layer using the xberg CLI. Covers OCR backends (Tesseract, PaddleOCR, VLM), language packs, force-OCR, batch runs, and performance tuning.