Invarra

Research foundation

Papers behind invariant-risk auditing.

Invarra publishes the measurement principles behind its audit work while keeping operational corpus construction, scoring logic, and client protocols private. Each paper includes a short summary, rendered math, an embedded PDF, and a downloadable copy.

Paper 01

June 28, 2026

The Latent Invariance Principle

An epistemic constraint on measurement under indirect observation

Short summaryOpen

Many important targets cannot be observed directly. Intent, belief, understanding, legal scope, risk state, and conceptual mastery are usually seen through representations: prompts, documents, tests, interfaces, symptoms, surveys, or sensor channels.

LIP states that when the target is latent, correct behavior under one representation is not enough. Stability across meaning-preserving representational variation is the admissible evidence that a system is tracking the latent phenomenon rather than the surface form.

The principle is not a model, algorithm, learning rule, or theory of truth. It is a measurement-validity constraint for settings where the thing being measured is only visible through representation.

Representation process

r=g(Φ,c,ϵ)r = g(\Phi, c, \epsilon)

An observable representation is modeled as a function of the latent phenomenon, the representation channel, and residual variation.

Observed behavior

B(r)B(r)

The evaluator observes behavior under a representation, while the measurement question concerns behavior with respect to the latent phenomenon.

Public boundary

The paper gives the research principle and measurement argument. It does not disclose operational audit construction, private corpus methods, scorer configuration, thresholds, or client-specific protocols.

Full paper

Open the embedded PDF here, or use the direct links.

Open viewer

Paper 02

June 28, 2026

Canonical Semantic Realization

A measurement framework for controlled semantic variation

Short summaryOpen

CSR is a measurement framework for semantic systems: natural-language inputs, policy descriptions, clinical notes, legal documents, system logs, instructions, survey items, and other artifacts whose meaning is not determined solely by surface form.

CSR separates three layers that are often collapsed into one row: the canonical semantic unit, the realization through language or presentation form, and the observed outcome from the system.

The framework treats canonical meaning as the experimental unit and controlled realizations as repeated measurements. This makes semantic brittleness, uncertainty, and representation sensitivity measurable without claiming automatic truth, correctness, or normative resolution.

Realization relation

p=π(s,c)p = \pi(s,c)

An observable realization is modeled as an expression of a canonical semantic unit under a representation condition.

Semantic preservation

p1semp2for valid realizations p1,p2E(s)p_1 \equiv_{\text{sem}} p_2 \quad \text{for valid realizations } p_1,p_2 \in E(s)

Valid realizations preserve the meaning-bearing commitments of the same semantic unit.

Public boundary

This page follows the arXiv-safe framing. It explains the measurement layers and validity constraint, but does not publish transformation libraries, admission procedures, provenance schemas, scoring logic, or operational corpus construction.

Full paper

Open the embedded PDF here, or use the direct links.

Open viewer

License

Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International.