Models
Pipeline
Language models for schema and extraction
Pipeline language models run after convert. They do not read page pixels. They work on the cleaned text that layout, OCR, and format converters already produced.
They power shape and extract on the Middleware API. Split stays structure-aware (sections and headings) and is not this model family.