Media Type

Words and Text

DZC Documents

Words and Text Available


Words and Text — Live from Microsoft

Secure, read-only view focused directly on the Words and Text folder.

Words-and-Text-Specific Fields

These eight values are extracted from each individual Words and Text file. They report the detected representation, measurable text volume, format-appropriate structural units, embedded resources, references, and analysis outcome. Generic facts such as filename, path, size, timestamps, and SHA-256 remain governed by File Master Primary.

Detected Format / Version

Reports the representation and version detected from the file contents or parsed container, such as a text encoding signature, Markdown, Word binary or OOXML, a font-table version, a parsed binary CMap, or a Digital Performer project format.

Text Character Count

Counts the characters recovered through format-aware decoding or text extraction. It measures actual readable content rather than file bytes and remains zero only when the parsed artifact exposes no text.

Word / Token Count

Counts word-like units in extracted prose and structured tokens in formats whose meaningful content is mappings, font records, or project data. The value is calculated from the parsed file rather than assigned by extension.

Paragraph / Record Count

Counts paragraphs for documents and Markdown, non-empty records for line-oriented text and logs, and the closest format-native record units for other supported representations.

Table / Mapping / Glyph Count

Counts the format-appropriate structural inventory: document or Markdown tables, expanded CMap mappings, font glyphs, or comparable indexed units. CMap codespace and notdef records are excluded; formats without one of these structures report a measured zero.

Embedded Resource Count

Counts resources carried inside the parsed artifact, including embedded images, objects, fonts, package parts, or project assets when the format exposes them.

Link / Reference Count

Counts hyperlinks, package relationships, external targets, font references, Markdown links, and other resolvable references detected by the format-specific analyzer.

Analysis Status

Reports the completed file-level extraction as Analyzed, Partial, or Empty. Files that cannot be parsed abort the refresh instead of receiving extension-based placeholders.