THotPDF.ExtractLoadedPageGlyphs Method

 

THotPDF.ExtractLoadedPageGlyphs

THotPDF

 

Top

Extracts positioned glyph clusters from a loaded page, including complete Unicode sequences, transformed quadrilaterals, baselines, and source locations

 

Delphi syntax:

function ExtractLoadedPageGlyphs(PageIndex: Integer; out AGlyphs: THPDFGlyphArray): boolean;

 

Description

Each THPDFGlyphRecord preserves the complete UnicodeSequence produced by /ToUnicode, while Unicode remains the first scalar for compatibility. ClusterStart, ClusterLength, and IsLigature identify logical-character membership. QuadX, QuadY, and the baseline endpoints are transformed through the text matrix and graphics CTM. GeometryExact is true when both the glyph width and font descriptor ascent and descent are available

Use ExtractLoadedPageCharacters when one result record per logical Unicode scalar is required

The glyphs come from the page's own content stream, so their source locations always point into it; text that the page paints through Form XObjects is not included, while ExtractLoadedPageText includes it