Skip to content

Overview ​

This section covers capabilities related to PDF page content, page structure, and document-level presentation. Some features write or adjust PDF content (headers and footers, page objects, page organization); others read page information (text extraction, page labels, layer state).

  • Headers and footers: Maintain headers, footers, and page numbers across the document; changes are written to the PDF—save or export according to your workflow.
  • Text extraction: Extract body and form-related text by page segment, whole page, or rectangle, with optional character-level geometry.
  • Page organizer: Insert, delete, move, copy, rotate, extract, and merge pages, and adjust page boxes and page transitions.
  • Page objects: Read, hit-test, add, modify, or remove text, image, and path objects in page content streams.
  • Page labels: Read page display labels and locate pages by label text.
  • Layers: Read the PDF layer tree, control layer visibility, and handle layered page import scenarios.

To search by keyword and obtain hit regions only, see Text search.