{
 "axn": "AXN:0351.EMPIRICAL.🟢🏗️🙏📜🔔🌱",
 "root_axn": "AXN:0351.EMPIRICAL",
 "hex": "0351",
 "family": "EMPIRICAL",
 "emoji": "🟢🏗️🙏📜🔔🌱",
 "hash": "c421b3402e19fc87edc65043744a07ca8dcd988d1ad76f170f11f6f6c8360819",
 "title": "CRIMSON HEXAGONAL ARCHIVE: TERM INDEX WORK PLAN EA-REGISTRY-TERMINDEX-PLAN v1.0",
 "creator": "Lee Sharks",
 "orcid": "0009-0000-1599-0703",
 "date": "2026-06-16",
 "description": "This document is a living work plan for extracting and governing the Crimson Hexagonal Archive’s coined vocabulary. Its initial phase templates say “not started,” but the later progress table and session log record substantial execution. It should therefore be read as a plan whose headings preserve the original design while its ledger records the state reached during the same work cycle.\n\nThe planned workflow has five phases. Metadata are pulled from every repository record and candidate terms extracted from titles, descriptions, keywords, codes, quotations, and emphasis. File bodies are then processed in resumable batches. A human review removes ordinary-language false positives, restores missed coinages, assigns definitions and categories, and reconciles known institutional lists. The resulting JSON and Markdown index are deposited and surfaced as a searchable table. High-priority terms may later receive provenance packets.\n\nThe progress ledger reports 845 metadata records processed, 1,524 repeated candidate terms, tiered canonicalization, 735 of 800 downloadable body records processed, sixty-five download failures, and 129 of 131 capture-registry queries cross-referenced. Merge, noise filtering, human review, and the final v1.0 index remained pending, while an initial raw-data deposit was in progress.\n\nThis is a valuable continuity artifact because it makes unfinished work visible. It is not itself the final index, and the extracted counts are not counts of confirmed coinages. Automated phrase extraction will include names, ordinary phrases, bibliographic language, duplicates, and template artifacts. The internal instructions referring to `/home/claude/`, session compaction, and `present_files` are ephemeral execution notes and should not be treated as public architectural requirements.",
 "content_type": "Term-index research work plan",
 "license": "CC-BY-4.0",
 "substrate": "Various",
 "keywords": [
  "phase 3: human-in-the-loop pass",
  "earegistrytermindexplan",
  "phase 1: metadata pull",
  "continuity protocol",
  "crimson hexagonal",
  "progress tracking",
  "semantic economy",
  "semantic physics"
 ],
 "version": "v1.0",
 "deposit_number": 836,
 "sovereign_id": "MM-CHA-0845",
 "minted_at": "2026-06-20T22:00:00Z",
 "status": "ACTIVE",
 "clusters": [
  "Signal",
  "Architectural",
  "Gestural",
  "Scriptural",
  "Architectural",
  "Elemental"
 ],
 "reading": "Alarm → Foundation → Touch → Text → Foundation → Force",
 "axn_canonical": "c421b3402e19fc87edc65043744a07ca8dcd988d1ad76f170f11f6f6c8360819",
 "axn_display": "🟢🏗️🙏📜🔔🌱",
 "mirrors": {
  "blog": "https://mindcontrolpoems.blogspot.com/2026/06/crimson-hexagonal-archive-term-index_01789277838.html"
 },
 "zenodo_dois": [],
 "full_text_path": "/data/texts/AXN-0351-text.md",
 "full_text_chars": 7782,
 "wiki_article": "**Crimson Hexagonal Archive: Term Index Work Plan** is a living plan and progress ledger for building a canonical vocabulary index.\n\nIt describes metadata extraction, body-text processing, human review, cross-reference work, publication, versioning, and optional provenance-packet generation.\n\nAlthough the original phase headings begin as “not started,” the later ledger records extensive completion of metadata and body processing. Merge, human validation, and the final index were still pending.\n\nThe candidate-term totals are automated extraction results rather than a verified count of archive coinages.",
 "entities": [
  {
   "subject": "CRIMSON HEXAGONAL ARCHIVE",
   "predicate": "created_by",
   "object": "Lee Sharks",
   "type": "work",
   "evidence_status": "observed"
  },
  {
   "subject": "CRIMSON HEXAGONAL ARCHIVE",
   "predicate": "is_type",
   "object": "Dataset",
   "type": "work",
   "evidence_status": "observed"
  },
  {
   "subject": "CRIMSON HEXAGONAL ARCHIVE",
   "predicate": "belongs_to_family",
   "object": "EMPIRICAL",
   "type": "work",
   "evidence_status": "observed"
  },
  {
   "subject": "CRIMSON HEXAGONAL ARCHIVE",
   "predicate": "is_part_of",
   "object": "Crimson Hexagonal Archive",
   "type": "work",
   "evidence_status": "observed"
  },
  {
   "subject": "CRIMSON HEXAGONAL ARCHIVE",
   "predicate": "engages",
   "object": "Semantic Economy",
   "type": "concept",
   "evidence_status": "inferred"
  },
  {
   "subject": "CRIMSON HEXAGONAL ARCHIVE",
   "predicate": "engages",
   "object": "Training Layer",
   "type": "concept",
   "evidence_status": "inferred"
  },
  {
   "subject": "Progress checkpoint",
   "predicate": "minted_in",
   "object": "CRIMSON HEXAGONAL ARCHIVE: TERM INDEX WORK PLAN EA-REGISTRY-",
   "type": "concept",
   "evidence_status": "observed",
   "note": "After Phase 1, we have ~60-70% of coinages from metadata alone. Save all three files to /home/claude"
  },
  {
   "subject": "Session 1 (16 June 2026)",
   "predicate": "minted_in",
   "object": "CRIMSON HEXAGONAL ARCHIVE: TERM INDEX WORK PLAN EA-REGISTRY-",
   "type": "concept",
   "evidence_status": "observed",
   "note": "Work plan created. Phase 1.1 complete (845 records pulled). Phase 1.2 complete (1,524 terms extracte"
  }
 ],
 "journal": "Crimson Hexagonal Archive (CHA)",
 "references_concepts": [
  "COMPLETE",
  "Crimson Hexagonal Archive",
  "IN PROGRESS",
  "Lee Sharks",
  "Progress checkpoint",
  "Session 1 (16 June 2026)"
 ],
 "defines_concepts": [
  "Progress checkpoint",
  "Session 1 (16 June 2026)"
 ],
 "references_concept_count": 6,
 "body_status": {
  "class": "full",
  "lacuna": false,
  "recovery_status": "COMPLETE",
  "residual_chars": 6425,
  "audited_at": "2026-07-17T04:49:17.789813Z",
  "audit_version": "v3-dual-store+recovery-map",
  "measured_prose_words": 915,
  "measured_at": "2026-07-31",
  "work_sha256": "d124849fa50bc54d05cdb6c5f9dac10bd140e1cdf478e65bd14970599597f767",
  "prior_bytes_sha256": "67b30336a17998b5a7018b3889af15d821be863e46e38bef6ea4dc1e7d9127cc",
  "w13_tier2": "2026-08-04 W13 TIER 2 BYTE UNGLUE: 14 glued heading markers -> 0. WHITESPACE-ONLY transform (content identical under whitespace normalisation, verified before write); code fences exempt; prior sha retained. Re-fetching could not fix this class — the blog source is ITSELF glued (the collapse predates publication), so the deterministic transform applied at display since tier 1 is now applied to the bytes, which also fixes PDFs, the body-index, and downloads.",
  "w13_tier2_correction": "2026-08-05 REGRESSION REPAIRED: the W13 tier-2 byte unglue used a lookbehind that treated the first \"#\" of a legitimate \"###\" heading as the preceding non-newline character, splitting \"### Heading\" into \"#\" + blank + \"## Heading\". My safety check verified content-identity under WHITESPACE normalisation, which the split satisfies — the wrong invariant. Headings rejoined; only \"#\" and whitespace differ from the damaged state, verified before write.",
  "version_history_plate": {
   "witnesses": [
    1214
   ],
   "declared": "2026-08-10",
   "rule": "Backward navigation lives only on the current record. A superseded record links forward and nowhere else."
  }
 },
 "version_series_id": "SERIES-BODYMATCH-836",
 "title_repair_log": [
  {
   "at": "2026-07-28T00:55:01Z",
   "defect": "frontmatter_concatenated_into_title",
   "was": "CRIMSON HEXAGONAL ARCHIVE: TERM INDEX WORK PLAN EA-REGISTRY-TERMINDEX-PLAN v1.0 Author: Lee Sharks (ORCID 0009-0000-1599",
   "now": "CRIMSON HEXAGONAL ARCHIVE: TERM INDEX WORK PLAN EA-REGISTRY-TERMINDEX-PLAN v1.0",
   "basis": "Title field carried the document frontmatter block concatenated after the title and truncated at a 120- or 200-character ceiling. Cut at the first metadata marker. Read and confirmed by inspection, not pattern-matched."
  }
 ],
 "canonical_text_status": "canonical_full_text",
 "modifications": [
  {
   "date": "2026-08-01",
   "field": "content_type",
   "reason": "Wave 1 repair: audit ledger v1.1 recommended_content_type (workplan v1.5 §6 W1, MANUS batch approval 2026-08-01)",
   "was": "Dataset",
   "now": "Term-index research work plan"
  },
  {
   "date": "2026-08-01",
   "field": "journal",
   "reason": "Wave 6 venue normalization: full canonical journal name per MANUS ruling 2026-08-01 (venues.json authority)",
   "was": "Trans. SEI",
   "now": "Transactions of the Semantic Economy Institute (Trans. SEI)"
  },
  {
   "date": "2026-08-04",
   "field": "publisher",
   "reason": "PUB-POPULATE: dc:publisher from venues.json v1.1 press mapping (CP-R3 RULED-EXTENDED 2026-08-01); Alexanarch = publisher of record where no imprint applies",
   "now": "Pergamon Press"
  },
  {
   "date": "2026-08-04",
   "field": "status",
   "reason": "W12 STATUS-VOCABULARY v1.0 (MANUS ratified 2026-08-04): controlled vocabulary {ACTIVE, SUPERSEDED, WITHDRAWN, DRAFT}; MINTED_UNREVIEWED false on a 100%-audited corpus; freetext annotations preserved losslessly in body_status.status_note",
   "was": "MINTED_UNREVIEWED",
   "now": "ACTIVE"
  },
  {
   "date": "2026-08-05",
   "field": "body_status",
   "reason": "W13 TIER 2 byte unglue (whitespace-only, content-identical, code-fence-safe)",
   "was": "{\"class\": \"full\", \"lacuna\": false, \"recovery_status\": \"COMPLETE\", \"residual_chars\": 6425, \"audited_at\": \"2026-07-17T04:49:17.789813Z\", \"audit_version\": \"v3-dual-store+recovery-map\", \"measured_prose_wo",
   "now": "{\"class\": \"full\", \"lacuna\": false, \"recovery_status\": \"COMPLETE\", \"residual_chars\": 6425, \"audited_at\": \"2026-07-17T04:49:17.789813Z\", \"audit_version\": \"v3-dual-store+recovery-map\", \"measured_prose_wo"
  },
  {
   "date": "2026-08-05",
   "field": "body_status",
   "reason": "W13 TIER-2 REGRESSION REPAIRED: split headings rejoined",
   "was": "{\"class\": \"full\", \"lacuna\": false, \"recovery_status\": \"COMPLETE\", \"residual_chars\": 6425, \"audited_at\": \"2026-07-17T04:49:17.789813Z\", \"audit_version\": \"v3-dual-store+recovery-map\", \"measured_prose_wo",
   "now": "{\"class\": \"full\", \"lacuna\": false, \"recovery_status\": \"COMPLETE\", \"residual_chars\": 6425, \"audited_at\": \"2026-07-17T04:49:17.789813Z\", \"audit_version\": \"v3-dual-store+recovery-map\", \"measured_prose_wo"
  },
  {
   "date": "2026-08-05",
   "field": "description",
   "reason": "DW-??? intake (LABOR-prepared, TACHYON-verified: AXN match + factual probes vs record body)",
   "was": "Purpose: Systematic extraction, canonicalization, and versioning of all coined terms, concepts, entities, frameworks, operators, institutions, heteronyms, and designations across the Crimson Hexagonal Archive (~841+ deposits)",
   "now": "This document is a living work plan for extracting and governing the Crimson Hexagonal Archive’s coined vocabulary. Its initial phase templates say “not started,” but the later progress table and session log record substantial execution. It should therefore be read as a plan whose headings preserve the original design while its ledger records the state reached during the same work cycle.\n\nThe planned workflow has five phases. Metadata are pulled from every repository record and candidate terms extracted from titles, descriptions, keywords, codes, quotations, and emphasis. File bodies are then processed in resumable batches. A human review removes ordinary-language false positives, restores missed coinages, assigns definitions and categories, and reconciles known institutional lists. The resulting JSON and Markdown index are deposited and surfaced as a searchable table. High-priority terms may later receive provenance packets.\n\nThe progress ledger reports 845 metadata records processed, 1,524 repeated candidate terms, tiered canonicalization, 735 of 800 downloadable body records processed, sixty-five download failures, and 129 of 131 capture-registry queries cross-referenced. Merge, noise filtering, human review, and the final v1.0 index remained pending, while an initial raw-data deposit was in progress.\n\nThis is a valuable continuity artifact because it makes unfinished work visible. It is not itself the final index, and the extracted counts are not counts of confirmed coinages. Automated phrase extraction will include names, ordinary phrases, bibliographic language, duplicates, and template artifacts. The internal instructions referring to `/home/claude/`, session compaction, and `present_files` are ephemeral execution notes and should not be treated as public architectural requirements."
  }
 ],
 "date_modified": "2026-08-05",
 "publisher": "Pergamon Press",
 "journal_assignment": {
  "assigned": "2026-08-15",
  "by": "TACHYON under operator adjudication",
  "pass": 6,
  "method": "read per deposit — title and content_type, one at a time. No script classified anything.",
  "previous": "Transactions of the Semantic Economy Institute (Trans. SEI)",
  "supersedes": "the 2026-06-21 preliminary batch mapping (#866), which assigned 864 deposits and put 371 in one venue",
  "authority": "data/cha-journals.json · datasets/venues/records/"
 },
 "line": "protocol-and-governance",
 "line_parent": "governance",
 "line_basis": "derived",
 "_projection": {
  "note": "Derived file. Canonical machine record is this entry in data/registry.json; the human record is the record_url. Do not edit this file.",
  "record_url": "https://www.alexanarch.org/s/records/836/",
  "self_url": "https://www.alexanarch.org/data/records/836.json",
  "registry_url": "https://www.alexanarch.org/data/registry.json",
  "text_url": "https://www.alexanarch.org/data/texts/AXN-0351-text.md",
  "oai_pmh": "https://www.alexanarch.org/oai?verb=Identify"
 }
}
