{
  "id": "keywords/pdf",
  "mentions": [
    {
      "context": "…e section of the entity page of a document that is not a note: a slide deck, a PDF, a transcript, or a note merged with one of them. It offers the file itself, …",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…des on where the tool goes after its first site. The deck is the original, its PDF the preview the build would have converted anyway, and this note its written …",
      "file": {
        "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html",
        "label": "framing/2026/2026-05-07-roadmap-outline/roadmap-outline.md"
      },
      "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 6,
      "surface": "PDF",
      "title": "Roadmap outline",
      "type": "document",
      "typeLabel": "Document"
    },
    {
      "context": "The text of an office document comes from the PDF the converter produced, never from the document's own format. A reader still opens the original file, for its …",
      "file": {
        "href": "../../specs/decisions/ingestion/single-extraction-path/index.html",
        "label": "decisions/ingestion/single-extraction-path.md"
      },
      "href": "../../specs/decisions/ingestion/single-extraction-path/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 5,
      "surface": "PDF",
      "title": "Single extraction path",
      "type": "decision",
      "typeLabel": "Decision"
    },
    {
      "context": "The page of an office document, a slide deck, a report, a spreadsheet or a PDF, alone or merged with its note. Same shell as the entity page: the tree of the sp…",
      "file": {
        "href": "../../specs/screens/pages/document-page/index.html",
        "label": "screens/pages/document-page.md"
      },
      "href": "../../specs/screens/pages/document-page/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 5,
      "surface": "PDF",
      "title": "Document page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…l, each file with its kind, \"Markdown note\", \"Presentation\", \"Text document\", \"PDF\", then \"Separate these files\", the way to contest the grouping, which leads t…",
      "file": {
        "href": "../../specs/screens/pages/entity-page/index.html",
        "label": "screens/pages/entity-page.md"
      },
      "href": "../../specs/screens/pages/entity-page/index.html#L17",
      "kind": "recognised",
      "line": 17,
      "passages": 4,
      "surface": "PDF",
      "title": "Entity page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "… and, for a transcript, its cues with their timecodes; the converter gives its PDF, cached by the fingerprint of the file, and the text of every page of that PD…",
      "file": {
        "href": "../../specs/processes/ingestion/build-pipeline/index.html",
        "label": "processes/ingestion/build-pipeline.md"
      },
      "href": "../../specs/processes/ingestion/build-pipeline/index.html#L14",
      "kind": "recognised",
      "line": 14,
      "passages": 3,
      "surface": "PDF",
      "title": "Build pipeline",
      "type": "process",
      "typeLabel": "Process"
    },
    {
      "context": "… of them. It offers the file itself, its text and, on demand, the pages of its PDF, without ever loading the viewer with the page. The document page of an offic…",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…he copy the build placed next to the page, and, when the conversion produced a PDF, \"Open the PDF\", a plain link the browser follows on its own. A rail of posit…",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L12",
      "kind": "recognised",
      "line": 12,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…ld placed next to the page, and, when the conversion produced a PDF, \"Open the PDF\", a plain link the browser follows on its own. A rail of positions comes next…",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L12",
      "kind": "recognised",
      "line": 12,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…bundle, a build of pdf.js the page never references in a script, and opens the PDF in the container: a toolbar with the counter, 3 / 12, the zoom between its tw…",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L14",
      "kind": "recognised",
      "line": 14,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…ule scripts there, the container shows a sentence saying so with a link to the PDF instead. The viewer and its worker are built only for a site that shows a PDF…",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L14",
      "kind": "recognised",
      "line": 14,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "… PDF instead. The viewer and its worker are built only for a site that shows a PDF, are announced in the build summary with their size, and change nothing in th…",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L14",
      "kind": "recognised",
      "line": 14,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…thumbnails. The text shown and searched is the text the build extracted from the PDF, cut at build.extracted_text_max_chars in all; the download gives the rest.",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L16",
      "kind": "recognised",
      "line": 16,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "Open the PDF → the PDF as the browser shows it",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L25",
      "kind": "recognised",
      "line": 25,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "Open the PDF → the PDF as the browser shows it",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L25",
      "kind": "recognised",
      "line": 25,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "Open the viewer → the pages of the PDF, page by page, with zoom and find",
      "file": {
        "href": "../../specs/viewers/document-viewer/index.html",
        "label": "viewers/document-viewer.md"
      },
      "href": "../../specs/viewers/document-viewer/index.html#L26",
      "kind": "recognised",
      "line": 26,
      "passages": 11,
      "surface": "PDF",
      "title": "Document viewer",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "Documents: office documents converted to PDF at publication and cached by fingerprint, transcripts read cue by cue with their speakers pseudonymised before anyt…",
      "file": {
        "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html",
        "label": "framing/2026/2026-05-07-roadmap-outline/roadmap-outline.md"
      },
      "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html#L13",
      "kind": "recognised",
      "line": 13,
      "passages": 6,
      "surface": "PDF",
      "title": "Roadmap outline",
      "type": "document",
      "typeLabel": "Document"
    },
    {
      "context": "…r its first site.  Six slides framing the next loops of the tool.  A deck, its PDF preview and its notes, grouped by the build into one page.  Roadmap outline ·…",
      "file": {
        "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html",
        "label": "framing/2026/2026-05-07-roadmap-outline/roadmap-outline.pdf"
      },
      "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html#L1",
      "kind": "recognised",
      "line": 1,
      "location": "p. 1",
      "passages": 6,
      "surface": "PDF",
      "title": "Roadmap outline",
      "type": "document",
      "typeLabel": "Document"
    },
    {
      "context": "…· DOCUMENTS  Decks, transcripts and their notes  Office documents converted to PDF at publication, cached by fingerprint.  Transcripts read cue by cue, speakers…",
      "file": {
        "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html",
        "label": "framing/2026/2026-05-07-roadmap-outline/roadmap-outline.pdf"
      },
      "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html#L3",
      "kind": "recognised",
      "line": 3,
      "location": "p. 3",
      "passages": 6,
      "surface": "PDF",
      "title": "Roadmap outline",
      "type": "document",
      "typeLabel": "Document"
    },
    {
      "context": "…r its first site.  Six slides framing the next loops of the tool.  A deck, its PDF preview and its notes, grouped by the build into one page.  Roadmap outline ·…",
      "file": {
        "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html",
        "label": "framing/2026/2026-05-07-roadmap-outline/roadmap-outline.pptx"
      },
      "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html#L1",
      "kind": "recognised",
      "line": 1,
      "location": "slide 1",
      "passages": 6,
      "surface": "PDF",
      "title": "Roadmap outline",
      "type": "document",
      "typeLabel": "Document"
    },
    {
      "context": "…· DOCUMENTS  Decks, transcripts and their notes  Office documents converted to PDF at publication, cached by fingerprint.  Transcripts read cue by cue, speakers…",
      "file": {
        "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html",
        "label": "framing/2026/2026-05-07-roadmap-outline/roadmap-outline.pptx"
      },
      "href": "../../briefs/framing/2026/2026-05-07-roadmap-outline/roadmap-outline/index.html#L3",
      "kind": "recognised",
      "line": 3,
      "location": "slide 3",
      "passages": 6,
      "surface": "PDF",
      "title": "Roadmap outline",
      "type": "document",
      "typeLabel": "Document"
    },
    {
      "context": "… index, the recognition and the comparison of twin resources are read from the PDF, page by page, with one PDF library; a PDF given as a source goes through the…",
      "file": {
        "href": "../../specs/decisions/ingestion/single-extraction-path/index.html",
        "label": "decisions/ingestion/single-extraction-path.md"
      },
      "href": "../../specs/decisions/ingestion/single-extraction-path/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 5,
      "surface": "PDF",
      "title": "Single extraction path",
      "type": "decision",
      "typeLabel": "Decision"
    },
    {
      "context": "…the comparison of twin resources are read from the PDF, page by page, with one PDF library; a PDF given as a source goes through the same reading, and a transcr…",
      "file": {
        "href": "../../specs/decisions/ingestion/single-extraction-path/index.html",
        "label": "decisions/ingestion/single-extraction-path.md"
      },
      "href": "../../specs/decisions/ingestion/single-extraction-path/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 5,
      "surface": "PDF",
      "title": "Single extraction path",
      "type": "decision",
      "typeLabel": "Decision"
    },
    {
      "context": "…of twin resources are read from the PDF, page by page, with one PDF library; a PDF given as a source goes through the same reading, and a transcript, which no c…",
      "file": {
        "href": "../../specs/decisions/ingestion/single-extraction-path/index.html",
        "label": "decisions/ingestion/single-extraction-path.md"
      },
      "href": "../../specs/decisions/ingestion/single-extraction-path/index.html#L8",
      "kind": "recognised",
      "line": 8,
      "passages": 5,
      "surface": "PDF",
      "title": "Single extraction path",
      "type": "decision",
      "typeLabel": "Decision"
    },
    {
      "context": "…t and one place to fix: the position of every word is a page or a slide of the PDF the reader will open in the document viewer, so a mention cites what the read…",
      "file": {
        "href": "../../specs/decisions/ingestion/single-extraction-path/index.html",
        "label": "decisions/ingestion/single-extraction-path.md"
      },
      "href": "../../specs/decisions/ingestion/single-extraction-path/index.html#L10",
      "kind": "recognised",
      "line": 10,
      "passages": 5,
      "surface": "PDF",
      "title": "Single extraction path",
      "type": "decision",
      "typeLabel": "Decision"
    },
    {
      "context": "… the repository, and a datum the corpus lacks is simply absent. A deck and the PDF preview committed next to it are read as one document: the deck is the origin…",
      "file": {
        "href": "../../specs/screens/pages/document-page/index.html",
        "label": "screens/pages/document-page.md"
      },
      "href": "../../specs/screens/pages/document-page/index.html#L14",
      "kind": "recognised",
      "line": 14,
      "passages": 5,
      "surface": "PDF",
      "title": "Document page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…w committed next to it are read as one document: the deck is the original, the PDF its preview, and a page count or a size the original does not state is read f…",
      "file": {
        "href": "../../specs/screens/pages/document-page/index.html",
        "label": "screens/pages/document-page.md"
      },
      "href": "../../specs/screens/pages/document-page/index.html#L14",
      "kind": "recognised",
      "line": 14,
      "passages": 5,
      "surface": "PDF",
      "title": "Document page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…and the strip follows the page shown. Without JavaScript the browser shows the PDF itself in the same place, first page in view. Under the rendering, two notes:…",
      "file": {
        "href": "../../specs/screens/pages/document-page/index.html",
        "label": "screens/pages/document-page.md"
      },
      "href": "../../specs/screens/pages/document-page/index.html#L16",
      "kind": "recognised",
      "line": 16,
      "passages": 5,
      "surface": "PDF",
      "title": "Document page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "… from the repository date.\" when the date was, and \"Size and page count of the PDF preview, the original stating none.\" when one of them comes from the preview;…",
      "file": {
        "href": "../../specs/screens/pages/document-page/index.html",
        "label": "screens/pages/document-page.md"
      },
      "href": "../../specs/screens/pages/document-page/index.html#L18",
      "kind": "recognised",
      "line": 18,
      "passages": 5,
      "surface": "PDF",
      "title": "Document page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "The page of a slide deck, a PDF or a transcript, or of a note merged with one, carries the document viewer after the article: the download link of each file, it…",
      "file": {
        "href": "../../specs/screens/pages/entity-page/index.html",
        "label": "screens/pages/entity-page.md"
      },
      "href": "../../specs/screens/pages/entity-page/index.html#L39",
      "kind": "recognised",
      "line": 39,
      "passages": 4,
      "surface": "PDF",
      "title": "Entity page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…ies the document viewer after the article: the download link of each file, its PDF, a rail of its pages, slides or cues, their extracted text in disclosure bloc…",
      "file": {
        "href": "../../specs/screens/pages/entity-page/index.html",
        "label": "screens/pages/entity-page.md"
      },
      "href": "../../specs/screens/pages/entity-page/index.html#L39",
      "kind": "recognised",
      "line": 39,
      "passages": 4,
      "surface": "PDF",
      "title": "Entity page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…ail of its pages, slides or cues, their extracted text in disclosure blocks, and the viewer that opens the PDF on demand. The markdown of the note is untouched.",
      "file": {
        "href": "../../specs/screens/pages/entity-page/index.html",
        "label": "screens/pages/entity-page.md"
      },
      "href": "../../specs/screens/pages/entity-page/index.html#L39",
      "kind": "recognised",
      "line": 39,
      "passages": 4,
      "surface": "PDF",
      "title": "Entity page",
      "type": "screen",
      "typeLabel": "Screen"
    },
    {
      "context": "…PDF, cached by the fingerprint of the file, and the text of every page of that PDF, the only text an office document ever contributes. A conversion that fails i…",
      "file": {
        "href": "../../specs/processes/ingestion/build-pipeline/index.html",
        "label": "processes/ingestion/build-pipeline.md"
      },
      "href": "../../specs/processes/ingestion/build-pipeline/index.html#L14",
      "kind": "recognised",
      "line": 14,
      "passages": 3,
      "surface": "PDF",
      "title": "Build pipeline",
      "type": "process",
      "typeLabel": "Process"
    },
    {
      "context": "…of every page of its documents for the search index, the original file and the PDF of each document kept for the page, a pseudonymised transcript written from i…",
      "file": {
        "href": "../../specs/processes/ingestion/build-pipeline/index.html",
        "label": "processes/ingestion/build-pipeline.md"
      },
      "href": "../../specs/processes/ingestion/build-pipeline/index.html#L28",
      "kind": "recognised",
      "line": 28,
      "passages": 3,
      "surface": "PDF",
      "title": "Build pipeline",
      "type": "process",
      "typeLabel": "Process"
    },
    {
      "context": "The transformation of an office document into a PDF at build, by a converter plugin, so that its text is extracted from that PDF and nothing else, and the conve…",
      "file": {
        "href": "../../glossary/ingestion/documents/conversion/index.html",
        "label": "ingestion/documents/conversion.md"
      },
      "href": "../../glossary/ingestion/documents/conversion/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Conversion",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "…a PDF at build, by a converter plugin, so that its text is extracted from that PDF and nothing else, and the conversion: block of the configuration that bounds …",
      "file": {
        "href": "../../glossary/ingestion/documents/conversion/index.html",
        "label": "ingestion/documents/conversion.md"
      },
      "href": "../../glossary/ingestion/documents/conversion/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Conversion",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "A slide presentation, a .pptx file that a converter turns into a PDF at build so that the site previews it and reads its text slide by slide. Its page offers th…",
      "file": {
        "href": "../../glossary/ingestion/documents/deck/index.html",
        "label": "ingestion/documents/deck.md"
      },
      "href": "../../glossary/ingestion/documents/deck/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "passages": 2,
      "surface": "PDF",
      "title": "Deck",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "… reads its text slide by slide. Its page offers the original for download, the PDF the conversion produced or the preview committed next to it, a rail of its sl…",
      "file": {
        "href": "../../glossary/ingestion/documents/deck/index.html",
        "label": "ingestion/documents/deck.md"
      },
      "href": "../../glossary/ingestion/documents/deck/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "passages": 2,
      "surface": "PDF",
      "title": "Deck",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "…e that no rule types: a markdown file left as-is, a slide deck, a Word file, a PDF, a transcript. A document is indexed and its words are recorded from its extr…",
      "file": {
        "href": "../../glossary/ingestion/documents/document/index.html",
        "label": "ingestion/documents/document.md"
      },
      "href": "../../glossary/ingestion/documents/document/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Document",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "…documents. Its page offers the file for download, shows its extracted text page by page and, when the conversion produced a PDF, opens it in a viewer on demand.",
      "file": {
        "href": "../../glossary/ingestion/documents/document/index.html",
        "label": "ingestion/documents/document.md"
      },
      "href": "../../glossary/ingestion/documents/document/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Document",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "The words of a document as the build reads them: the text of every page of the PDF the converter produced, or of the PDF given as a source, and the cues of a tr…",
      "file": {
        "href": "../../glossary/ingestion/documents/extracted-text/index.html",
        "label": "ingestion/documents/extracted-text.md"
      },
      "href": "../../glossary/ingestion/documents/extracted-text/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Extracted text",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "…eads them: the text of every page of the PDF the converter produced, or of the PDF given as a source, and the cues of a transcript with their timecodes. An offi…",
      "file": {
        "href": "../../glossary/ingestion/documents/extracted-text/index.html",
        "label": "ingestion/documents/extracted-text.md"
      },
      "href": "../../glossary/ingestion/documents/extracted-text/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Extracted text",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "The PDF form of an office document, produced by the conversion at build and cached by fingerprint, or committed next to the original by its author; either way t…",
      "file": {
        "href": "../../glossary/ingestion/documents/preview/index.html",
        "label": "ingestion/documents/preview.md"
      },
      "href": "../../glossary/ingestion/documents/preview/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Preview",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "…nd and the original stays downloadable. previews: false on a source keeps the PDFs out of the published site: the download link and the extracted text remain. A…",
      "file": {
        "href": "../../glossary/ingestion/documents/preview/index.html",
        "label": "ingestion/documents/preview.md"
      },
      "href": "../../glossary/ingestion/documents/preview/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDFs",
      "title": "Preview",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "…t notes: W-CONV-FAILED, an office document the converter could not turn into a PDF; W-CONV-SUSPECT, a converted PDF without extractable text although the docume…",
      "file": {
        "href": "../../glossary/quality/checks/document-checks/index.html",
        "label": "quality/checks/document-checks.md"
      },
      "href": "../../glossary/quality/checks/document-checks/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "passages": 2,
      "surface": "PDF",
      "title": "Document checks",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "… document the converter could not turn into a PDF; W-CONV-SUSPECT, a converted PDF without extractable text although the document is large; W-DOC-NOMD, a docume…",
      "file": {
        "href": "../../glossary/quality/checks/document-checks/index.html",
        "label": "quality/checks/document-checks.md"
      },
      "href": "../../glossary/quality/checks/document-checks/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "passages": 2,
      "surface": "PDF",
      "title": "Document checks",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "…, commit and date, and, for documents, the metadata extracted by a reader, the PDF produced by a converter and the text extracted from that PDF, page by page, o…",
      "file": {
        "href": "../../specs/objects/ingestion/resource/index.html",
        "label": "objects/ingestion/resource.md"
      },
      "href": "../../specs/objects/ingestion/resource/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Resource",
      "type": "business_object",
      "typeLabel": "Business object"
    },
    {
      "context": "… by a reader, the PDF produced by a converter and the text extracted from that PDF, page by page, or the cues of a transcript with their timecodes; an office do…",
      "file": {
        "href": "../../specs/objects/ingestion/resource/index.html",
        "label": "objects/ingestion/resource.md"
      },
      "href": "../../specs/objects/ingestion/resource/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "passages": 2,
      "surface": "PDF",
      "title": "Resource",
      "type": "business_object",
      "typeLabel": "Business object"
    },
    {
      "context": "…pecifications, the integrator of the wiki and the quality owner. The deck, its PDF preview and the transcript sit next to these minutes and the build groups the…",
      "file": {
        "href": "../../briefs/meetings/2026/2026-08-27-theme-override-model/theme-override-model/index.html",
        "label": "meetings/2026/2026-08-27-theme-override-model/theme-override-model.md"
      },
      "href": "../../briefs/meetings/2026/2026-08-27-theme-override-model/theme-override-model/index.html#L9",
      "kind": "recognised",
      "line": 9,
      "surface": "PDF",
      "title": "Theme override model",
      "type": "meeting",
      "typeLabel": "Meeting"
    },
    {
      "context": "…he plugin API that turns a document into representations the pipeline reads: a PDF today, thumbnails and text later, written under the fingerprint cache and nev…",
      "file": {
        "href": "../../glossary/ingestion/documents/converter/index.html",
        "label": "ingestion/documents/converter.md"
      },
      "href": "../../glossary/ingestion/documents/converter/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "surface": "PDF",
      "title": "Converter",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "One of the files an entity is made of: its note, a slide deck, the PDF a converter produced from it, a transcript, the contract an operation was imported from. …",
      "file": {
        "href": "../../glossary/ingestion/sources/representation/index.html",
        "label": "ingestion/sources/representation.md"
      },
      "href": "../../glossary/ingestion/sources/representation/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "surface": "PDF",
      "title": "Representation",
      "type": "term",
      "typeLabel": "Term"
    },
    {
      "context": "One of the files an entity is made of: its note, a slide deck, the PDF produced from it, a transcript, the contract an operation was imported from. A note alone…",
      "file": {
        "href": "../../specs/objects/ingestion/representation/index.html",
        "label": "objects/ingestion/representation.md"
      },
      "href": "../../specs/objects/ingestion/representation/index.html#L6",
      "kind": "recognised",
      "line": 6,
      "surface": "PDF",
      "title": "Representation",
      "type": "business_object",
      "typeLabel": "Business object"
    },
    {
      "context": "An office document could not be converted to PDF: timeout, size, corruption or missing converter. The document stays downloadable, without preview or extracted …",
      "file": {
        "href": "../../specs/rules/documents/conversion-failed/index.html",
        "label": "rules/documents/conversion-failed.rule.md"
      },
      "href": "../../specs/rules/documents/conversion-failed/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "surface": "PDF",
      "title": "Conversion failed",
      "type": "rule",
      "typeLabel": "Business rule"
    },
    {
      "context": "… other documents only. Its words are recorded from the text extracted from its PDF or its cues, but nobody wrote about it. The to-do page lists these documents;…",
      "file": {
        "href": "../../specs/rules/documents/document-without-markdown/index.html",
        "label": "rules/documents/document-without-markdown.rule.md"
      },
      "href": "../../specs/rules/documents/document-without-markdown/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "surface": "PDF",
      "title": "Document without markdown",
      "type": "rule",
      "typeLabel": "Business rule"
    },
    {
      "context": "A converted PDF holds no extractable text although the document is large; it is probably made of images. Nothing of it enters search or the occurrence scan.",
      "file": {
        "href": "../../specs/rules/documents/suspect-conversion/index.html",
        "label": "rules/documents/suspect-conversion.rule.md"
      },
      "href": "../../specs/rules/documents/suspect-conversion/index.html#L7",
      "kind": "recognised",
      "line": 7,
      "surface": "PDF",
      "title": "Suspect conversion",
      "type": "rule",
      "typeLabel": "Business rule"
    },
    {
      "context": "…d of the row when the twin-resource reconciliation merged them. A deck and the PDF preview kept next to it are one representation: the \"Deck\" tab offers the ori…",
      "file": {
        "href": "../../specs/screens/pages/meeting-page/index.html",
        "label": "screens/pages/meeting-page.md"
      },
      "href": "../../specs/screens/pages/meeting-page/index.html#L16",
      "kind": "recognised",
      "line": 16,
      "surface": "PDF",
      "title": "Meeting page",
      "type": "screen",
      "typeLabel": "Screen"
    }
  ]
}
