PdfRecordedText class
Text metadata captured from a complete, in-memory page recording.
Capture after the page content walk finishes and before drawing annotations:
PdfTextExtractor searches page content, not annotation appearances. The
recording interpreter must use collectCharOffsets: true for exact
selection geometry. Capture before serializing render commands, whose wire
format omits marked-content ids and embedded-font character offsets.
Graphics, glyph outlines and soft-mask definitions are discarded. Text in Type3 and tiling cells stays compact and expands only when runs is read. Invisible text and text inside masked source groups remain searchable, just as in a fresh extraction. The snapshot does not retain the input graph and is unaffected by subsequently clearing or extending its command lists.
Constructors
-
PdfRecordedText.capture(List<
PdfRenderCommand> commands) -
factory
Properties
- estimatedBytes → int
-
Approximate retained bytes for cache budgeting, including object/list
overhead, UTF-16 strings, glyph metadata and copied numeric buffers.
Shared cells and runs count once. This is a portable estimate, not a heap
measurement, and excludes the expanded extraction produced from runs.
final
- hashCode → int
-
The hash code for this object.
no setterinherited
- isEmpty → bool
-
no setter
-
runs
→ Iterable<
PdfTextRun> -
Positioned source text in the interpreter's original encounter order.
no setter
- runtimeType → Type
-
A representation of the runtime type of the object.
no setterinherited
Methods
-
noSuchMethod(
Invocation invocation) → dynamic -
Invoked when a nonexistent method or property is accessed.
inherited
-
toString(
) → String -
A string representation of this object.
inherited
Operators
-
operator ==(
Object other) → bool -
The equality operator.
inherited