THotPDF.ExtractLoadedPageCharacters
Returns one THPDFCharacterRecord for every logical Unicode scalar extracted from a loaded page
Delphi syntax
function ExtractLoadedPageCharacters(PageIndex: Integer; out ACharacters: THPDFCharacterArray): boolean;
Description
The method expands positioned glyph clusters without discarding multi-scalar /ToUnicode mappings. Every record identifies its source glyph, cluster range, ligature component index and count, source content token, marked-content identifier, transformed quadrilateral and baseline
Ligature components share the source glyph geometry because the PDF content stream positions the glyph as one unit. GeometryExact is true when the font supplies both a glyph width and descriptor ascent and descent; otherwise HotPDF returns a conservative fallback box and sets the flag to false
See also ExtractLoadedPageGlyphs