32 KiB
AGENTS.md - Research Stack Operating Contract
This file is the first stop for coding agents working in this repository.
⛔ READ-ONLY NOTICE — SilverSight Has Replaced This Repo
The Research Stack is a READ-ONLY archive. SilverSight is the clean-slate successor. All new formal work MUST go to the SilverSight repository.
| Action | Allowed? |
|---|---|
| Add new Lean modules to Research Stack | NO → use SilverSight repo |
| Add new Python shims to Research Stack | NO → use SilverSight repo |
| Edit existing Research Stack files (bug fixes, doc updates) | YES |
Run lake build on existing modules |
YES |
| Run DB sync scripts | YES |
| Add new AGENTS.md rules to Research Stack | NO → add to SilverSight |
SilverSight repository: https://github.com/allaunthefox/SilverSight
SilverSight local clone: /home/allaun/SilverSight (or wherever you clone it)
SilverSight formal modules: formal/SilverSight/
SilverSight AGENTS.md: AGENTS.md in SilverSight repo
Enforcement: Any agent that adds new files to the Research Stack instead of the SilverSight repository is violating this contract.
Local Hermes Deployment (2026-06-20)
Hermes Agent v0.14.0 is the primary chat/dashboard gateway, serving the fully-local Hermes-3 model on qfox-1's RTX 4070. OpenClaw is decommissioned.
- Hermes dashboard on qfox-1 (CachyOS):
http://100.88.57.96:9119- User-level systemd unit:
~/.config/systemd/user/hermes-dashboard.service - Run with
--host 0.0.0.0 --port 9119 --insecure --no-open --skip-build(--insecuredisables the OAuth auth gate for tailnet-only use). - Web UI built once to
/home/allaun/.hermes/hermes-agent/hermes_cli/web_dist. - Public access:
https://chat.researchstack.infovia Caddy on racknerd (reverse_proxy http://100.88.57.96:9119). - Note:
100.85.244.73isnixos-steamdeck-1, not qfox-1.
- User-level systemd unit:
- Local inference on qfox-1: Official Ollama (v0.30.10) serving
hermes3:latestbuilt from/var/lib/gemma-models/Hermes-3-Llama-3.1-8B-Q4_K_M.gguf.- Endpoint:
http://127.0.0.1:11434/v1 - Systemd unit:
ollama-hermes3.service - GPU offload verified:
ollama psreports100% GPUand logs showoffloaded 33/33 layers to GPU.
- Endpoint:
- Hermes config:
~/.hermes/config.yamlpoints at the local Ollama:model: default: "hermes3:latest" provider: "openai" base_url: "http://127.0.0.1:11434/v1" api_key: "sk-local"
Common failure modes
500 Internal Errorfromchat.researchstack.infoafter a model swap: the~/.hermes/config.yamlstill references an old model name. Verifymodel.defaultmatches the live/v1/modelslisting andproviderisllamacpp(orollama/lmstudio/vllm) — not the legacyproviders.<name>:schema, which is silently ignored.- Browser shows a 500 cached from an earlier failed deploy: hard refresh (Ctrl+Shift+R) or open an incognito window.
Headroom Context Compression (2026-06-19)
Headroom (v0.26.0) compresses tool outputs, logs, RAG chunks, and conversation history before they reach the LLM — 60–95% fewer tokens, same answers. Installed via pipx on the workstation.
- MCP server installed for Claude Code and Codex (headroom_compress,
headroom_retrieve, headroom_stats). See
~/.claude/config. - Anthropic/Codex proxy:
headroom proxyon port 8787. Route through it withANTHROPIC_BASE_URL=http://127.0.0.1:8787for automatic compression. - Kimi Code proxy:
headroom-kimi-proxy.serviceon port 8789. Routes Kimi Code through Headroom using the OpenAI-compatible Kimi endpoint. Started via systemd user unit.~/.kimi/config.tomluses the direct Kimi endpoint by default; to enable compression, run:KIMI_BASE_URL=http://127.0.0.1:8789 kimi. - Memory:
headroom proxy --memory --learnenables persistent memory and automatic failure-pattern mining → AGENTS.md updates. - Project config lives in
.headroom/at the repo root. headroom learn --project . --applymines past Claude Code/Codex session failures and writes corrections to AGENTS.md. RequiresclaudeCLI available for the analysis LLM.
Neon proxy (2026-06-19)
Headroom proxy also runs on neon-64gb (netcup ARM64, NixOS) for remote agent sessions:
- Tailscale endpoint:
http://100.92.88.64:8787(also reachable ashttp://neon-64gb:8787via MagicDNS) - Startup:
/home/allaun/.headroom/headroom-proxy-start.sh(wrapsLD_LIBRARY_PATHfor NixOS libstdc++ compatibility) - Installed via: Python venv at
~/headroom-venv/ - Config:
--memory --learn --port 8787 --host 0.0.0.0 - Route through it:
ANTHROPIC_BASE_URL=http://100.92.88.64:8787 claude
Common failure modes
headroom learnanalysis fails withclaude CLI not found: the analysis LLM depends on theclaudebinary being on PATH. Install Claude Code or set--modelto an available provider (e.g.--model gpt-4o).pipx installfails withCERTIFICATE_VERIFY_FAILED: SSL inspection proxy. Install Rust first so maturin doesn't need to download it.Proxy dependencies not installedon NixOS:libstdc++.so.6missing. Installnix profile install nixpkgs#stdenv.cc.cc.liband setLD_LIBRARY_PATHin the startup script.
Repository Extraction Notice (2026-06-02)
The codebase has been split into three repositories:
distributed-compute-fabric (generic compute infrastructure)
- VCN/LUPINE hardware acceleration
- Hardware abstractions (Virtio, SPIR-V, QEMU)
- Spatial hash indexing (vectorless graph database)
- LyteNyte visualization dashboards
- Ray actors and hermes orchestrator
- Kubernetes manifests and monitoring
- Tailscale networking
- gRPC inference proxy (optional)
research-compute-fabric (research-specific algorithms)
- Braid eigensolid compression
- PIST/RRC classification
- Flexure/joint analysis
- Lean formalization bridges
- Research probes (Hutter Prize, Tang Nano, etc.)
Research Stack (this repo — formalization and documentation)
- Lean/Semantics formalization
- Hardware bring-up documentation
- Citation references and provenance
- Research artifacts and design docs
See respective repositories for components. Shared utilities have been duplicated to maintain independence between repositories.
Ground Rules
- Use
/home/allaun/Research Stackas the active host checkout and git root. When theresearch-stackdev container is available, run Lean builds, WGSL shader work, and research execution throughpodman exec research-stackfrom the container-local checkout at/home/researcher/repo/. Treat the host checkout and container checkout as synchronized views of the same repo; verify paths before copying artifacts between them. - The dev container is
research-stack(podman,research-stack-otom:latest). - Research artifacts (design docs, experimental Lean files, exploration output)
must be written to container-local paths under
/home/researcher/research/. These are not in the git tree and do not pollute the working tree. Usepodman cpto extract research artifacts when they are ready for promotion to production. - Read the nearest nested
AGENTS.mdbefore editing a subtree. - Preserve user work. The working tree is often intentionally dirty; do not revert, delete, or stage unrelated files.
- Prefer repo-native tools and receipt generators over ad hoc summaries.
- Treat Lean as the source of truth for formal or hardware-adjacent claims.
- Load the
lean-proofskill before any Lean work. It enforces the proof quality contract: no bare sorries, no tautologies, Q16_16 compliance, and#evalwitness requirements. Trigger patterns auto-load it on.lean,lake build,sorry,Q16_16,theorem,lemma, etc. - No
Floatin compute paths.ofFloatis only permitted at the external boundary (JSON parsing, sensor data). All core/main computation usesQ16_16.ofNat,Q16_16.ofRatio, orQ16_16.ofRawInt. This applies to Lean, Python, and Verilog. The HiGHS boundary is the only exception (it requires float inputs, but the result is immediately converted to Q16_16). - Keep claims bounded: a receipt proves only the gate it actually checks.
- Secrets are runtime-only. Use environment variables such as
OLLAMA_API_KEYorDEEPSEEK_API_KEY; never paste, print, or commit literal provider keys. - For repo-stabilization tasks, finish with a clean
git status --branch --short --untracked-files=alland an emptygit clean -nddry run before claiming the tree is stable.
Core Surfaces
- Lean/Semantics:
0-Core-Formalism/lean/Semantics/— Compiler surface includesSemantics.SieveLemmas,Semantics.InteractionGraphSidon,Semantics.PIST.Spectral, andSemantics.PIST.Classify(8604 jobs, 0 errors). Thepist-classify-traceexecutable is the sole authority for offline proof-trace shape/tactic classification; Python shims are pure I/O. - Infrastructure shims and probes:
4-Infrastructure/shim/ - Hardware bring-up:
4-Infrastructure/hardware/ - Documentation and wiki surfaces:
6-Documentation/ - Virtio-Net DMA Compute Spec:
6-Documentation/docs/specs/virtio_net_compute_fabric_spec.md - Entropy exploration → RRC certification pipeline:
4-Infrastructure/shim/geometric_entropy_explorer.py— Exploration-phase candidate generator4-Infrastructure/shim/candidate_certification_bridge.py— Bridge to LeanCandidates.lean0-Core-Formalism/lean/Semantics/Semantics/RRC/EntropyCandidates/— Lean-certifiableBraidStatefixturesshared-data/data/stack_solidification/candidates/— Candidate JSON + batch manifests- DP-RRC spec:
6-Documentation/docs/specs/DP_RRC_RECEIPT_ENCODING_SPEC.md
- Citation reference map:
CITATION.cff(external sources are provenance and terminology references unless a Lean theorem or receipt explicitly promotes a bounded claim) - Stack receipts:
shared-data/data/stack_solidification/ - Promoted review receipts:
shared-data/artifacts/deepseek_review/ - Canonical Ollama/DeepSeek review emitter:
5-Applications/tools-scripts/llm/ollama_deepseek_review_emitter.py - CAD harness:
5-Applications/text-to-cad/ - Historical scoped staging maps:
6-Documentation/docs/stack_solidification_staging_manifest_2026-05-09.mdand6-Documentation/docs/stack_solidification_staging_manifest_2026-05-10.md
Verification Expectations
- For Lean changes, run the narrow target first, then the broader
lake buildwhen feasible. - For Python shims, run
python3 -m py_compileon touched files. - For JSON receipts, run
python3 -m json.toolor a repo-native receipt parser. - For promoted Ollama/DeepSeek review receipts, run
python3 5-Applications/tools-scripts/llm/ollama_deepseek_review_emitter.py --verify-onlysoanswer_sha256is checked against the answer file after write. - For hardware claims, distinguish software witness, bitstream presence, SRAM load, flash persistence, UART beacon, and live hardware receipt.
- Before committing, run
git diff --cached --checkand a staged secret scan for touched source/receipt files. The repository credential hook is a final gate, not a substitute for local review. - For root CAD setup changes, JSON-parse
package.json,.vscode/settings.json, and.vscode/tasks.json, then use the pinned Python/CAD commands documented in5-Applications/text-to-cad/AGENTS.md.
Post-Interaction Workflow (mandatory after every agent session)
After every interaction with the user that changes code, Lean modules, shim scripts, receipts, or architecture decisions, the agent MUST run the following steps before claiming the session is complete:
1. Update AGENTS.md files
Update the nearest scoped AGENTS.md for every subtree touched:
| Subtree touched | AGENTS.md to update |
|---|---|
0-Core-Formalism/lean/Semantics/ |
0-Core-Formalism/lean/Semantics/AGENTS.md |
4-Infrastructure/ |
4-Infrastructure/AGENTS.md |
| Root or multi-subtree | AGENTS.md (this file) |
Minimum required updates per touched subtree:
- Blessed surface table — add/remove/update module rows if Compiler roots changed
- Build baseline — update job count and commit hash if
lake buildwas run - Architecture section — update if data-flow or boundary rules changed
- Pending proof work — add new
TODO(lean-port)stubs; remove resolved ones - Quarantine table — add quarantined files; remove revived ones
2. Verify the build
cd 0-Core-Formalism/lean/Semantics && lake build Compiler
If Lean files were touched, also run the full build:
cd 0-Core-Formalism/lean/Semantics && lake build
3. Commit
Stage only explicitly touched files (never git add .).
Commit message format:
<type>(<scope>): <summary>
<body — what changed and why>
Build: <N> jobs, 0 errors (lake build)
Types: feat, fix, chore, docs, refactor.
Scope examples: lean, rrc, avm-isa, infra, corpus250.
4. Check tree cleanliness
git status --branch --short --untracked-files=all
Untracked files that are not generated artifacts should either be staged or noted as intentionally dirty. Do not claim the tree is stable if there are unexpected modifications.
5. Programming choice flow (check before writing any new code)
Before writing or placing any new logic, run through this decision tree in order. Stop at the first rule that applies.
New logic needed?
│
├── Does it make an admissibility, routing, alignment, or gating decision?
│ └── YES → Write it in Lean. No Python equivalent allowed.
│ File: Semantics/RRC/Emit.lean or a new Semantics.* module.
│
├── Does it mint, stamp, or emit a top-level receipt or JSON bundle?
│ └── YES → It belongs in Semantics.AVMIsa.Emit ONLY.
│ AVMIsa.Emit is the sole output boundary. No other module
│ may call leanBuildReceipt and produce a top-level JSON.
│
├── Does it classify rows, run an alignment gate, or compute scores?
│ └── YES → Lean (Semantics.RRC.Emit or a new Semantics.RRC.* module).
│ Python may call it via #eval / lake exe but may not replicate it.
│
├── Does it supply raw input features (equation text, route_hint, domain_type,
│ equation_id hashing, weak_axes count)?
│ └── YES → Python shim is acceptable. The shim must:
│ (a) produce only a List FixtureRow (or equivalent data structure)
│ (b) carry no admissibility logic — every field it sets is "raw"
│ (c) be regenerable from source (document the regeneration command)
│ (d) live in 4-Infrastructure/shim/ with a clear TODO(lean-port) if
│ any logic in it could eventually move to Lean
│
├── Does it use floating-point arithmetic in a compute path?
│ └── YES → STOP. Use Q16_16.ofNat / Q16_16.ofRatio / Q16_16.ofInt instead.
│ ofFloat is only permitted at the external boundary (JSON parsing,
│ sensor input) and must be immediately bracketed.
│
├── Does it advance promotion status (e.g. set promotion = "promoted")?
│ └── YES → STOP. Promotion is always not_promoted until a Lean gate
│ explicitly passes. Never advance it in shim space or by hand.
│
└── Is it pure I/O (read JSON, write JSONL, call subprocess, format output)?
└── YES → Python shim is fine. Keep it in 4-Infrastructure/shim/.
Receipt-writing Python must still route output through
AVMIsa.Emit (Lean stamps; Python only formats/stores).
Summary rule: Lean owns all decisions. Python owns all I/O. If you find yourself writing decision logic in Python, stop and port it to Lean.
What "every interaction" means
This workflow triggers whenever the user message results in:
- Any file edit (Lean, Python, TOML, JSON, shell, Rust, etc.)
- Any
lake buildrun - Any architectural decision that changes the data-flow or boundary rules
- Any new
TODO(lean-port)or quarantine boundary
It does NOT trigger for pure read-only exploration, explanation, or questions that result in no file changes.
Do Not Sweep
Avoid broad cleanup or staging commands such as:
git add .
git add 0-Core-Formalism 4-Infrastructure 6-Documentation shared-data
git checkout -- .
git clean -fdx
Use explicit file lists from the relevant staging manifest.
shared-data/ is ignored by default because most of it is generated or
offloaded. Promote only specific, durable receipts with git add -f -- <path>,
and keep empty/failed model outputs out of Git unless they are themselves the
evidence under review.
Git Remote Hygiene
- The active branch may not have an upstream. Inspect with
git rev-parse --abbrev-ref --symbolic-full-name @{u}before assuming push state. - For GitHub sync, prefer the
githubremote and verify the remote head after push:
git fetch github <branch>
git rev-list --left-right --count FETCH_HEAD...HEAD
git push -u github <branch>
git ls-remote --heads github <branch>
- Dependabot banners printed by GitHub after push may be stale relative to the live alert API. Treat the push result and remote-head hash separately from dependency-alert remediation.
Pending Proof Work (2026-06-22)
| Theorem | File | Status |
|---|---|---|
cleanMerge_preservesGap |
GraphRank.lean:209 | sorry with TODO(lean-port) — computational kernel (mergeCheck_all) verified all 256×256 byte pairs; Q16_16↔byte bridge for 8-bin signatures pending |
Q16_16 Unification Migration
The codebase has 6 different Q16_16 type definitions causing fragmentation. A migration
is in progress to unify all consumers to the canonical Semantics.FixedPoint.Q16_16.
Migration Status (Completed)
- ✅
ManifoldStructures.lean- Migrated (type-only usage) - ✅
Errors.lean- Migrated (function-using, bridge-based approach) - ✅
SubstrateProfile.lean- Migrated - ✅
BoundaryDynamics.lean- Migrated - ✅
CausalGeometry.lean- Migrated - ✅
CosmicStructure.lean- Migrated - ✅
CriticalityDynamics.lean- Migrated - ✅
ExoticSpacetime.lean- Migrated - ✅
MagnetoPlasma.lean- Migrated - ✅
ManifoldPotential.lean- Migrated - ✅
MultiBodyField.lean- Migrated - ✅
SpikingDynamics.lean- Migrated - ✅
PhysicsScalarBridge.lean- Created bridge with full function forwarding
Migration Approach
For files using PhysicsScalar.Q16_16:
-
Type-only files (simple): Replace imports and open statements
-import Semantics.PhysicsScalar +import Semantics.FixedPoint -open Semantics.PhysicsScalar +open Semantics.FixedPoint -
Function-using files (requires bridge): Use
Semantics.PhysicsScalarBridgeimport Semantics.PhysicsScalarBridge -- Replace PhysicsScalar.Q16_16.gt with PhysicsScalarBridge.gt -- Replace constants with PhysicsScalarBridge.one, .two, .three, etc.
Available Bridge Functions
toFixedPoint : PhysicsScalar.Q16_16 → FixedPoint.Q16_16fromFixedPoint : FixedPoint.Q16_16 → PhysicsScalar.Q16_16add,mul,sub- bridged arithmeticgt,le- bridged comparisonszero,one,two,three,four,half,quarter- constants
Migration Guide
See LEAD_Q16_16_MIGRATION_GUIDE.md for detailed microsteps.
Glossary (before reading further)
These terms appear throughout all AGENTS.md files and the codebase:
- Sidon label — an address from a set where all pairwise sums are unique. Powers of 2 (1,2,4,8,16,32,64,128) are the canonical Sidon set for 8 strands. Sidon slack = address budget − max label used (encodes capacity headroom).
- BraidStorm — the 8-strand braid topology used by the eigensolid compressor. Strands cross pairwise; each crossing merges phase and produces a residual.
- eigensolid — the converged, stable state of a braid crossing loop.
Detected when
crossStep(s) = s. The DC baseline in the TNT BraidCarrier model. - scar — a FAMM failure record in
ene.scars. Stores scar_pressure, failure_mode, and optional coarsening_agent for remediation path. Scar absence (∅) is a positive receipt dimension. - Yang-Baxter — the braid relation
βij βjk βij = βjk βij βjkthat defines braid-order invariance. Represented inBraidedFieldPaths.lean; operational braid-action proofs must state their exact evaluation model. - Anti-BraidStorm — adversarial dual that tests Yang-Baxter invariance and receipt aliasing as a validation layer for the compressor.
- MORE FAMM — Memory-Optimized Recursive Entropy Fractal Aggregate Memory Model. Capability-based memory isolation for the runtime stack.
- TSM — Topological S3C Manifold. Thermodynamic safety monitor with Builder/Warden/Judge clock domains.
- GCL — Genetic Code Language. Self-improving evolutionary program representation.
- AVM — Adaptive Virtual Machine. Universal bridge between math languages and
Python bytecode. Core ISA is live in
Semantics.AVMIsa.*(Types, Value, Instr, State, Step, Run). AVM is the sole output boundary for RRC receipts:AVMIsa.Emitstamps all top-level receipt JSON;RRC.EmitandRRC.Corpus250feed it as classifier and raw-feature supplier respectively. - enwik9 — the Hutter Prize 1GB Wikipedia XML corpus, used as the canonical end-to-end test vector for the hierarchical compressor.
- receipt — a machine-readable attestation record stored in
ene.receipts. Receipt dimensions (C, σ, k, ε_seq, t, ∅_scars) together form the encoding of the compressed state. Every gate generates a receipt; a receipt proves only the gate it actually checks.
Compression First Principles
- Zero-gap-timing-spacing-silence IS signal. No byte, no space, no absent row is noise. The compressor encodes everything; the decompressor must reconstruct everything, including the gaps, because the gaps ARE the compression.
- The receipt IS the compressed state — not metadata around it. Receipt dimensions (crossing matrix C, Sidon slack σ, step count k, residual series ε_seq, write timing t, scar absence ∅) together form the encoding. Invertibility of this receipt is the definition of lossless compression.
- Every compute substrate denies being a CPU: GPU (WGSL/wgpu), ASIC (SHA-256), FPGA (Verilog), PCIe/DMA (bus mastering), storage (NVMe/BTRFS), blitter (6502 OISC), hydraulic (pipes). All compute identically because Q16_16 integer arithmetic is deterministic across all of them. No substrate is privileged.
- The database IS the test field. enwik9 (1GB) and 259 TiddlyWiki tiddlers live
in the same
ene.packagestable. The compressor does not distinguish between them. The schema is the substrate. - Two distinct Lean theorems are required for every compressor:
eigensolid_convergence— the braid crossing loop stabilizesreceipt_invertible— given the receipt, the original state is reconstructible within bounded error, including all gap/timing/absence dimensions
- Float (
ofFloat) is forbidden in compute paths.Q16_16.ofNatandQ16_16.ofRatioare the canonical constructors.ofFloatis only permitted at the external boundary (parsing JSON, reading sensor data) and must be immediately bracketed.
Legacy Recovery Trigger
The phrase RECOVER LEGACY INFORMATION is the explicit retrieval trigger
for archived or quarantined concepts. Treat this as a user-controlled cold
archive request, not permission to revive an old branch wholesale.
Accepted trigger forms:
RECOVER LEGACY INFORMATION: <path, commit, concept, or artifact>
Recover Legacy Information: <path, commit, concept, or artifact>
recover from cornfield: <path, commit, concept, or artifact>
When this trigger appears:
- Inspect the requested legacy source first with read-only commands such as
git show,git log, or targeted file reads. - Recover only the named file, concept, commit slice, or receipt requested.
- Modernize the recovered material onto the current clean branch before committing it.
- Never merge, reset to, or base new work on a legacy/cornfield branch unless the user explicitly asks for that exact branch operation.
- Preserve the legacy branch as retrievable archive state.
Current cornfield ref:
backup/distilled-with-vcd-history-2026-05-11
Nested Contracts
- Strict Lean/docs contract:
6-Documentation/docs/AGENTS.md - Lean module-local contract:
0-Core-Formalism/lean/Semantics/AGENTS.md - Infrastructure contract:
4-Infrastructure/AGENTS.md - CAD harness contract:
5-Applications/text-to-cad/AGENTS.md - QC flagger contract:
scripts/qc-flag/AGENTS.md - Lean expert agent contract:
shared-data/artifacts/lean_expert_agent/AGENTS.md
Clean PIST predictions pipeline (no Rust authority)
Canonical IDs + dedup
Join key is invariant ID: equation_id := invariant_receipt.object_id (format rrc_eq_<hex>).
The 278 source file contains duplicate object_id groups; predictions artifacts must be
deduped by equation_id.
Preserve auditability by attaching provenance:
summary.total_source_records = 278summary.unique_equation_ids = 250- each prediction row includes
source_records : Array {equation_record_id, name}for all source rows sharing that invariant.
Deterministic representative selection
When multiple compiled records share the same equation_id, select a representative
deterministically:
representative := min(records, key = equation_record.equation_id) (lexicographic)
Never "last wins".
Prediction artifact (v1, matrix-only)
Generate: shared-data/rrc_pist_predictions_250_v1.json
Constraints:
schema: "rrc_pist_predictions_250_v1"claim_boundary: "matrix-only;no-classifier;no-lean-spectral"proxy_pred: null,exact_pred: null(until a classifier surface is defined)- include:
global_vocab_hashmatrix_schema: "token_strand_adjacency_8x8_v1"matrix_hash(sha256 of canonical row-major JSON, no whitespace)matrix_8x8(Int counts)
Generation rules (matrix_schema v1)
- global vocab = all unique tokens across corpus, sorted
strand(token) = vocab_index % 8- matrix is 8×8 adjacency of token bigrams in original order, projected to strands:
M[strand(t_i)][strand(t_{i+1})] += 1
Merge path into Lean corpus
pist_matrix_builder.py
→ writes rrc_pist_predictions_250_v1.json (dedup by invariant id)
→ build_corpus250.py reads it and merges by equation_id = rrc_eq_<hex>
→ regenerates Semantics/RRC/Corpus250.lean with pistProxyLabel/pistExactLabel
populated when present
→ emit250.json alignment gate becomes non-missing_prediction only when labels exist.
Required validations (every change)
python3 -m py_compileon touched shim scriptspython3 -m json.toolon generated JSON- reproducibility check: two consecutive runs must produce identical file SHA256
lake build(full workspace) must stay green
NOTE on counts
Prediction artifact row count may be 250 while source record count is 278. This is correct: one prediction per invariant equation id, with provenance for all source records.
Fable (Claude Code) Session Artifacts — 2026-06-10
A previous Fable session created three Claude Code skills in ~/.claude/skills/:
| Skill | Content | Current status |
|---|---|---|
lean-formal-build |
Build workflow for singer-theorem-lean + Semantics | Current |
fabric-k3s-ops |
k3s cluster operations | Current |
nixos-image-build |
NixOS qcow2 image build (netcup + openstack) | Current |
ContextStream workspace mappings in ~/.contextstream/mappings.json were updated
to point at the singer submodule's new permanent path.
Nutbreaker prompts were written to 6-Documentation/docs/:
NUTBREAKER_SSMS_PROMPT.md, NUTBREAKER_ADJUGATE_MATRIX_PROMPT.md,
NUTBREAKER_BRAID_SPHERION_BRIDGE_PROMPT.md, NUTBREAKER_BURGERS_BRIDGE_PROMPT.md.
For sessions that need these skills, reference ~/.claude/skills/<name>/SKILL.md.
Meta-Solid Finding (2026-06-19)
A ContextStream node was written (e967f515-3af9-46c9-9fc8-e5c766a6c4fc, type=fact) documenting the meta-solid topological triple point and the abelian→nonabelian mixing transition. Key points:
- Two abelian strand species forced to mix become nonabelian (braid group rep dim > 1)
- At mixing fraction x = 1/7 ≈ 0.14, three phase projections become simultaneously exact:
- Trace closure (global average) → Gas
- Local YB isotopy (Reidemeister moves) → Liquid
- Braid class (Alexander polynomial, crossStep fixed point) → Solid
- The meta-solid is when all three quotients are exact — no information lost in projection
- The 1/7 threshold = one complete Sidon doubling step (1 of 7 doublings 2→128) consumed by the size spread
- Maps directly onto the ~14% terminal polydispersity from the hard-sphere consensus literature (Kofke+, Fasolo+Sollich)
- Canonical receipts:
rrc_photonic_stress_test_final_receipt.json,rrc_bosonic_tensor_final_receipt.json - Shim:
rrc_bosonic_tensor_network.py(quimb-based, bypasses Perceval 256-mode FockState cap) - ContextStream node query:
search(mode="keyword", query="meta-solid") - Granular-superconductor analogy tested and negative (2026-06-19): The hypothesis that a universal reduced field H*/Hc₂ ≈ 1/7 exists across granular superconductors was tested via Consensus search. No paper reports such a universal ratio; H* is always microstructure-dependent, varies by orders of magnitude, and is never normalized to bulk Hc₂. The 1/7 threshold remains grounded in Sidon doubling combinatorics and hard-sphere polydispersity only. See CITATION.cff for the 25 vortex-glass/granular references reviewed.
Finsler-Randers QAP Benchmark (2026-06-21)
benchmark_finsler_qap.py formulates Finsler-Randers directed routing as QAP (Quadratic Assignment
Problem via TSP-MTZ) and benchmarks against the QUBO subset-selection approach at n=8,12,24,48.
Results
| n | β-aniso | QUBO-SA (sym) | QUBO-card (K=n/2) | QAP-LP relaxed | QAP-MIP (path) | QUBO-card NN | TSP-on-subset |
|---|---|---|---|---|---|---|---|
| 8 | 0.650 | 0.0 | 20.95 (K=4) | 21.10 | 15.17 (0.07s) | 3.96 | 6.47 |
| 12 | 1.041 | 0.0 | 55.18 (K=6) | 27.30 | 19.45 (0.30s) | 5.11 | 8.31 |
| 24 | 1.127 | 0.0 | 289.21 (K=12) | 38.78 | 24.53 (2.54s) | 10.68 | 12.54 |
| 48 | 1.127 | 0.0 | 1294.73 (K=24) | 91.57 | 37.90 (13.08s) | 21.57 | 20.85 |
Key Findings
- QAP-MIP (MTZ TSP) scales well — n=48 solves to feasibility in 13s; n=24 in 2.5s
- QAP-LP is a loose bound — always 2-3x higher than QAP-MIP path cost (assignment relaxation ignores subtour elimination)
- QUBO is structurally degenerate for all-positive Q_ij — unconstrained QUBO always selects 0 nodes; cardinality constraint (K=n/2) gives meaningful subsets but the symmetric energy is fundamentally different from directed path cost
- 2-phase strategy viable: select K via QUBO-card, then route via TSP-on-subset. The combined cost at K=n/2 is roughly half the full TSP cost at n
- NN-tour vs TSP-on-subset: greedy nearest-neighbor is within 3% of optimal TSP at n=48 but 30-60% off at n=8,12
- Solver times: QAP-MIP grows as O(n³) empirically (0.07s n=8 → 0.30s n=12 → 2.54s n=24 → 13.08s n=48)
Relevant Files
4-Infrastructure/shim/benchmark_finsler_qap.py— benchmark harness with MTZ-TSP, QUBO-MIP-card, QAP-LP, and cross-evaluation4-Infrastructure/shim/qubo_highs.py— QUBO→MIP bridge andsolve_route_lp(TSP assignment relaxation)4-Infrastructure/shim/qaoa_adapter.py—FinslerMetric,finsler_metric_to_qubo,geodesic_assignment/tmp/finsler_benchmark_v3.json— n=8,12,24 results (structued JSON)/tmp/finsler_benchmark_n48.json— n=48 results
When to Use ContextStream Search:
✅ Project is indexed and fresh ✅ Looking for code by meaning/concept ✅ Need semantic understanding