All notable changes to ContextSafe are documented here. The format is based on Keep a Changelog; the project has no tagged release yet, so everything to date lives under Unreleased.
-
Dependency-update automation (SEC-14):
.github/dependabot.ymlcovering the two places this repository pins — theuvlock and the SHA-pinned GitHub Actions — weekly, with a seven-day cooldown on both ecosystems (SEC-26 asks for 72 hours) and Python updates grouped into one pull request. The repository previously had neither adependabot.ymlnor arenovate.json, so no advisory against a locked dependency could open a pull request, nothing kept the action pins current, and OpenSSF Scorecard'sDependency-Update-Toolcheck scored 0 by construction. -
B-039 slice:
tests/test_privacy_canaries.py, the near-miss, log, and crash-output half of the canary suite RG-12 gates on. It pins the privacy boundary in both directions — approved codes that resemble identifiers must not be false positives, values one character from acceptable must fail closed with a named code — and records three identifier-shaped values the pattern scan does not catch (a date outside its 19xx/20xx window, a dotted date, a seven-digit local number) as blind spots for the independent security review rather than as accepted behavior; the synthetic-namespace grammar is what bounds them. It adds a structural log canary (no module importsloggingor prints, and no accepted or rejected command emits a record), a crash canary (an unexpected failure after the boundary read carries neither evidence content nor the caller's source path, and a CLI rejection prints a structured error rather than a traceback), an index canary (raw bytes stay in the content-addressed object; the queryable SQLite index carries hashes, tokens, and provenance only), and a matrix property that no rejection echoes the value that triggered it. No detector, schema, or runtime behavior changes. B-039 is not closed: pattern tuning is a security-owned decision whose independent review has not happened, FHIR/HL7/LIS sources do not exist (B-023–B-025), and the diagnostics, support bundle, and local logs RG-12 also covers are B-046. -
B-021 slice:
tests/test_determinism.py, the three-run reproducibility evidence R-10 and RG-15 ask for and the process half of status-algebra invariant 10. Each shipped command runs three times in fresh interpreters under different time zones, locales, hash seeds, UTF-8 modes, working directories, and input directories, and must produce byte-identical exit codes, stdout, stderr, and--outputartifacts. Every artifact must be one canonical UTF-8 JSON line with one terminal newline and no carriage return, the referenceevaluatedocument has a pinned SHA-256, no absolute input path or environment value may reach an artifact, a caller-declaredclaimed_generated_atmust move the envelope without movingpayload_sha256, and a fail-closed rejection must emit the same stderr bytes and error code every run. A CI matrix (ubuntu-24.04,macos-15,windows-2025) reproduces the pinned digest, and a monkeypatched test pins the documented fail-closed rejection on platforms without descriptor-relative no-follow open — Windows among them, wherepack validate,plan validate, andevidence preflighttherefore cannot run. This is byte-reproducibility evidence only: packaging and fresh-install evidence remain B-045, and B-021 stays open pending normalization (B-019/B-026) and signing (B-035). -
docs/PUBLICATION-READINESS.md: a gate-by-gate audit of whether this repository could ever be made public, with evidence. Gate 0 is the IP/inventions-agreement question created by the repository's creation date falling during prior employment, which only the maintainer's attorney can clear; Gate 1 is a dual-use and misuse assessment specific to this project — a tool that reports where transgender and nonbinary identity data is lost also reports where it is retained — including what the threat model already covers, four things it does not, and what must be decided before B-010. The verdict is technically ready pending IP clearance, not ready to publish. No tag, release, visibility change, or history rewrite accompanies it. -
CLI:
contextsafe evidence preflightnow accepts--output, matchingpack validate,plan validate, andevaluate. Previously the only way to obtain the boundary-check result was stdout, so combining--quietwithevidence preflightsilently discarded the command's only output and left nothing but the exit code.--outputwrites the same non-sensitive result document (boundary_check_status, hashes, declared scope —PreflightResultnever carries evidence content) that would otherwise print; it does not change what the command reads, copies, indexes, or logs. -
B-033 slice:
schemas/contextsafe-receipt-v0.1.schema.json, the published contract for the receipt document and its deterministic payload — the pre-1.0 shape of the receipt schema required bydocs/04-ARCHITECTURE.mdsection 8. The contract closes every object (additionalProperties: false), pins the unsigned envelope constants so a signing layer cannot relabel these documents in place, keeps the payload claim-minimal by rejecting timestamp, signature, reviewer, run-environment, and semantic-value fields, pins the mandated limitation set as a closed ordered list so a stripped, reworded, reordered, or padded disclosure fails validation (F-030) and the payload carries no unbounded free-text channel, and publishes closed status, reason, checkpoint, and concept enums. Tests enforce schema/runtime agreement on the reference document, theevaluate --outputartifact, and every Hypothesis-generated bundle; a companion test asserts that every file inschemas/is a valid, self-consistent Draft 2020-12 contract. Outcome reasons are now the typedOutcomeReasonenum, so an unpublished reason string cannot reach a receipt without a schema change. Receipt bytes are unchanged. -
B-027 slice: Hypothesis-based property tests seeding the documented property layer (
docs/09-TEST-AND-EVALUATION.mdsection 2) for the machine-checkable status-algebra invariants — no pass without exactly one affirmative evidence match, not-applicable only from a predeclared rule, fail-closed cross-concept rejection, order-independent byte-identical receipts, and value-minimized receipts that never echo generated semantic values. Invariants needing pack lifecycle, review signatures, HTML, or signature verification remain untested because those components do not exist yet. -
B-020 slice: every CLI command accepts
--quiet(suppress the stdout success payload; exit codes,--outputfiles, and stderr JSON errors unchanged) and--no-color(an explicit pin of the always-plain contract — output never contains ANSI escape sequences), and exit codes are documented and stable:0success,2fail-closed contract rejection,64command-line usage error (previously argparse's default2, which collided with contract rejections). -
V1 planning corpus (
docs/00–16): PRD, service design, architecture, data and evidence model, security/privacy threat model, governance, test strategy, operations, roadmap, backlog, risk register, and release checklist. -
Iteration 1: strict versioned case and observation-set schemas; separately typed GI, RSG, SPCU, name-to-use, and pronoun values; fail-closed cross-concept rejection; pure exact-match evaluator (missing/ambiguous evidence is indeterminate); deterministic value-minimized JSON receipts; offline
validateandevaluateCLI commands with a synthetic reference fixture. -
Iteration 2: strict pack envelope, deterministic unsigned compiler with semantic component hashes and lifecycle/withdrawal checks; strict engagement and execution-plan contracts with fail-closed non-production attestations, host allowlisting, and hash pinning.
-
Iteration 3: canonical JSON evidence boundary envelope with field allowlist, namespace pins, PHI canaries, and direct-identifier checks; read-only
evidence preflight; recoverable two-pass persistence into a SHA-256 object store with an update/delete-protected SQLite index. -
Iteration 4 (B-021 slice): receipt payload/envelope separation.
contextsafe evaluatenow emits a receipt document instead of the bare iteration-1 receipt — the byte-identical deterministic payload pluspayload_sha256over the payload only, and an untrusted envelope with caller-declaredclaimed_generated_at(optional canonical whole-second UTC, via--claimed-generated-at),signature_status: not_signed, andtrusted_time: false. Timestamps and signatures stay outside the deterministic payload (P0-14); no signing or trusted-time path exists. -
Standards-conformance baseline (2026-07-16 sweep): LICENSE (Apache-2.0), SECURITY.md, CONTRIBUTING.md, CITATION.cff, CHANGELOG, pre-commit config, Semgrep/gitleaks/pip-audit security workflow, tag-triggered release workflow, ADR log seed (existing ADRs relocated from
docs/decisions/todocs/adr/), docs/I18N.md declaration, and a README Standards Conformance table.
- The Semgrep SAST gate (SEC-07) had been red on
mainfor every one of its fourteen runs since 2026-07-17, on four blocking findings against the two evidence-index header PRAGMAs inevidence_store.py. SQLite does not accept bound parameters in a PRAGMA, so the statements are now rendered once at module scope from their integer constants with the:dconversion — which can emit only digits and an optional sign — and_publish_new_databaseexecutes those constants instead of building a string at the call site. A new test pins the exact rendered text of both statements, requires each to matchPRAGMA [a-z_]+ = -?\d+, and asserts that SQLite rejects the parameterized form. No waiver,.semgrepignore, or# nosemgrepwas added; the registry auto config now reports 0 findings over 72 targets. Store bytes and the on-disk index header are unchanged. - Command output is written as UTF-8 bytes instead of through a text stream. Text-mode writes translate the terminal newline into the platform line separator and encode with the platform's preferred encoding, so the same receipt would have left a POSIX host and a Windows host with different bytes and different file digests — the cross-platform nondeterminism R-10 names and RG-15 gates. Artifact and payload content is unchanged on POSIX hosts.
CITATION.cffno longer advertises a release that was never cut. It carrieddate-released: 2026-07-17whilegit tag -lis empty and no GitHub release exists. CFF treatsversionanddate-releasedas optional; both return when a release is actually tagged.- The README no longer points readers at
../STANDARDS, a path that exists only in the author's local checkout and names a repository a reader cannot open. The standards are now described rather than linked; the conformance table is unchanged. .gitignorecovers.hypothesis/, which was previously ignored only by the nested ignore file Hypothesis generates for itself.
- The Semgrep SAST gate (SEC-07) reported a green check on every pull request
while scanning nothing.
semgrep ciresolves a diff baseline on apull_requestevent by runninggit fetch origin --force --depth=1 <head-sha>; this repository is private and the checkout setspersist-credentials: false, so the fetch failed, Semgrep aborted before scanning, and its default--suppress-errorsturned the aborted run into exit 0. The job logs carry the scan-environment banner and the fetch error with no scan summary — no rule count, no target count, no findings line. A HIGH finding introduced by a pull request would have passed the gate. Replaced withsemgrep scan --config auto --error --strict, which needs no baseline and no credential, runs the identical full scan on push, pull request, and schedule, and fails on an analysis error so a scan that cannot run can no longer report success. See ADR 0004.