Five indexes run in parallel,
none waits for another.

Once a handler has sealed a passage, five independent indexers receive it simultaneously. No pipeline. No queue. Each index is complete before the first query arrives, and each is deterministic: the same corpus yields the same index, byte-identical, at any point in time.

LLX-ARCH-IDX · R1

5Parallel indexes
0Inter-index dependencies
SHA-512Every entry sealed
100%Deterministic
Passage index · 01 Verbatim, grounded, cited

Every passage is pinned to the byte it came from.

The passage index is the evidentiary foundation. Each entry carries the verbatim text, its byte-range offset in the source file, the page or sheet it sits on, and the SHA-512 seal that ties it to the sealed ingestion record. A retrieval from this index is a citation, not a summary.

PI 01

Verbatim extraction

The handler passes exact text to the indexer — no paraphrase, no normalisation, no loss. The passage in the index is the passage in the file.

PI 02

Byte-range coordinates

Every passage is indexed with its start byte, end byte, page number and positional order within the page. A query returns the address, not just the text.

PI 03

SHA-512 anchor

The passage entry carries the SHA-512 hash of its byte range. The hash matches the sealed ingestion record. Tampering with one breaks the other.

PI 04

Cited retrieval

Every retrieval from the passage index returns a citation: file path, file hash, page, byte-range, timestamp of ingestion. Nothing is returned without its provenance.

Metadata index · 02 Every attribute the document carries

Metadata is indexed as a first-class finding.

The metadata index records every attribute the handler extracts from a file: title, author, revision, date, classification, document number, subject, and any custom field the file format exposes. These attributes are indexed separately from the passage text so they can be queried, filtered and cross-referenced without re-reading the file.

MI 01

Structural attributes

Title, author, revision number, creation date, last-modified date, document number and classification — extracted from the file's internal structure, not from its name.

MI 02

Format-specific fields

A drawing carries sheet number, scale, issue date and discipline. A schedule carries baseline dates and activity IDs. The handler extracts what the format exposes.

MI 03

Indexed for cross-reference

Metadata is queryable across the corpus. Every document authored by a party, every revision issued after a date, every drawing in a discipline — returned as a set, not a search result list.

MI 04

Sealed with the passage

Metadata entries carry the same SHA-512 seal as passage entries. A document whose metadata has been altered does not match its ingestion record.

Spatial index · 03 Position is evidence

Where a passage sits on the page is part of the finding.

The spatial index maps every passage to its exact coordinates within its document: column, row, bounding box and reading order. On drawings, it maps title blocks, revision clouds and annotation zones. The IoU spatial merger reconciles passages that span format boundaries — a table cell that crosses a column break, an annotation that overlays a block of text — and presents them as a single addressable unit.

SI 01

Bounding box per passage

Every passage is stored with its bounding box: x1, y1, x2, y2 in normalised page coordinates. The box is queryable. Find all text in the right margin, the title block, the header — by position.

SI 02

Reading order preserved

The spatial index records the reading order the handler determined from the file structure. Column-aware, table-aware, footnote-aware. The sequence is the sequence the author intended.

SI 03

IoU spatial merger

Intersections over Union: passages from different extraction passes that overlap in space are merged into a single spatial record. No duplicate, no gap. The merger is deterministic.

SI 04

Drawing zone indexing

On engineering drawings, the spatial index maps title blocks, revision schedules, general notes and zone grids as named addressable regions. A query can target zone A3 on sheet 12 directly.

XATR index · 04 Cross-attribute temporal relations

XATR maps when documents speak to each other across time.

XATR — Cross-Attribute Temporal Relation — is the chronological index. It records every date attribute found in the corpus, resolves conflicts between stated dates and transmission dates, and builds a temporal map of the corpus: what existed when, what was superseded by what, and where the timeline breaks. XATR is the index that answers delay questions.

XATR 01

Date attribute resolution

Every document date is recorded: creation date, issue date, revision date, transmission date, received date, programme date. Where they conflict, XATR records all of them and flags the discrepancy.

XATR 02

Supersession chain

XATR tracks which revision supersedes which, which drawing replaces which, which instruction cancels which. The chain is built from metadata and cross-reference analysis — not from filenames.

XATR 03

Chronological ordering

The corpus is ordered by every date type simultaneously. A query for events between two dates returns documents sorted by the date type that matters to the question: issue, receipt, or programme.

XATR 04

Delay surface

XATR exposes the delay surface: the gap between when a document was dated and when it was transmitted, between what the programme assumed and what was actually issued. The gap is indexed, not calculated.

KUNZU index · 05 Knowledge unit — the semantic cornerstone

KUNZU binds what documents say to what they mean.

KUNZU — Knowledge Unit — is the semantic index. It extracts entities, obligations, quantities and relations from passage text and indexes them as structured facts. A KUNZU entry is not a passage: it is a claim extracted from a passage, linked to its source, and typed by its semantic class. KUNZU is what makes the engine answer questions about the record, not just return passages from it.

KZ 01

Entity extraction

Parties, locations, items, dates and quantities are extracted from passage text and indexed as named entities. Every entity entry links back to the passage it came from and carries its seal.

KZ 02

Obligation and claim indexing

Obligations, representations, warranties and claims are identified and indexed by type. A query for all obligations issued by a party returns the set, each with its source passage and date.

KZ 03

Relation graph

KUNZU builds a directed graph of relations between entities: party-to-party, document-to-document, obligation-to-response. The graph is queryable. Cycles, breaks and contradictions surface as graph anomalies.

KZ 04

Deterministic semantic linter

The Advanced Semantic Linter runs over the KUNZU graph after each ingestion cycle. It proves relations, flags contradictions and scores consistency — without a model, without inference, without probabilistic output. The result is deterministic.

Determinism · 06 The cornerstone property

The same corpus yields the same indexes, always.

All five indexes are deterministic. Given the same corpus at the same version, the indexer produces byte-identical output. No randomness. No model sampling. No approximation. The index is a function of the files, not of the moment it was computed. This is the property that makes PARALLAX RC® findings reproducible and forensically defensible.

DT 01

No model in the index path

No large language model participates in any of the five indexing passes. Models are used only at the answering layer, on passages already retrieved and sealed. The index is not learned — it is computed.

DT 02

Version-pinned

Every index entry records the version of the handler and indexer that produced it. A re-ingestion with the same version produces the same entry. A version upgrade is explicit and audited.

DT 03

Byte-identical at T+96 months

The same query on the same sealed corpus returns the same passages, with the same coordinates and the same seals, ninety-six months after ingestion. Long after a cloud endpoint would have drifted.

DT 04

The forensic guarantee

Determinism is not a performance property. It is the forensic guarantee. A finding can be reproduced by any party with access to the corpus and the engine version. The reproducibility is the proof.

Five indexes. One deterministic record.

Indexation is not search pre-computation. It is the analytical act. Every chronology, every register, every finding drawn from PARALLAX RC® rests on what these five indexes hold and what the custody record proves.

Request access