I've been running PaddleOCR-VL-1.5 via llama.cpp's server for OCR on book pages. It handles complex layouts, tables, and mixed text/figure pages surprisingly well. Setup: - Model: PaddleOCR-VL-1.5-GGUF + mmproj.gguf - Backend: llama-server (Vulkan on Windows) - Pipeline: layout detection → region OC