Overview ​
This section covers capabilities related to PDF page content, page structure, and document-level presentation. Some features write or adjust PDF content (headers and footers, page objects, page organization); others read page information (text extraction, page labels, layer state).
- Headers and footers: Maintain headers, footers, and page numbers across the document; changes are written to the PDF—save or export according to your workflow.
- Text extraction: Extract body and form-related text by page segment, whole page, or rectangle, with optional character-level geometry.
- Page organizer: Insert, delete, move, copy, rotate, extract, and merge pages, and adjust page boxes and page transitions.
- Page objects: Read, hit-test, add, modify, or remove text, image, and path objects in page content streams.
- Page labels: Read page display labels and locate pages by label text.
- Layers: Read the PDF layer tree, control layer visibility, and handle layered page import scenarios.
To search by keyword and obtain hit regions only, see Text search.