THotPDF.ExtractLoadedPageCharacters

Returns one THPDFCharacterRecord for every logical Unicode scalar extracted from a loaded page

Delphi syntax

function ExtractLoadedPageCharacters(PageIndex: Integer;
  out ACharacters: THPDFCharacterArray): boolean;

Description

The method expands positioned glyph clusters without discarding multi-scalar /ToUnicode mappings. Every record identifies its source glyph, cluster range, ligature component index and count, source content token, marked-content identifier, transformed quadrilateral, and baseline

Ligature components share the source glyph geometry because the PDF content stream positions the glyph as one unit. GeometryExact is true when the font supplies both a glyph width and descriptor ascent and descent; otherwise HotPDF returns a conservative fallback box and sets the flag to false

See also ExtractLoadedPageGlyphs