Skip to main content
The layout model is a vision detector. It runs on a rasterized page before OCR reads anything. Output is bounding boxes and a region type, not text. It is part of convert on the Middleware API.

Configuration