Models
Layout
Page-region detection before OCR
The layout model is a vision detector. It runs on a rasterized page before OCR reads anything. Output is bounding boxes and a region type, not text.
It is part of convert on the Middleware API.
Page-region detection before OCR