TY - RPRT TI - From Pixels to Structure: Lightweight Vision-Language Models for Document OCR and Structured JSON Extraction AU - Uddipan Basu Bir AU - Vincent Christlein AU - Andreas Maier AU - Mathias Zinnen PY - 2026 DO - 10.1007/978-3-032-36039-7_30 UR - https://arxiv.org/abs/2610.11818 ID - 2610.11818 ER -