{
 "hex": "03F5",
 "lifted_at": "2026-08-07T03:58:03Z",
 "rule": "Recording of method or modification does not appear in body text (MANUS standing rule, 2026-08-06). Nothing is destroyed: this file holds what was lifted, and the record page renders it after the work.",
 "head_apparatus": "**AXN:** AXN:03F5 — Alexanarch deposit #1001 (self-reference in root form by pre-hash necessity)\n**Restoration status:** SEMI-RESTORED — metadata-body deposit. This machine-facing static page is the canonical deposit. Its body is the complete DataCite metadata record for a work whose Zenodo record returns HTTP 410 (Gone) while DataCite serves the identifier as findable — the metadata layer and content layer in formal disagreement about the work's existence. Full text pending restoration from authorial originals; on restoration, this deposit upgrades by recorded correction (new hash, new glyph, remediation note).",
 "tail_apparatus": "",
 "body_bytes_before": 9534,
 "body_bytes_after": 8908,
 "lifted": [
  {
   "at": "2026-08-08",
   "reason": "DataCite capture superseded by the recovered work.",
   "text": "# The Crimson Hexagonal Archive Hugging Face Dataset: Work Plan v3 (Classifier-Centric Methodology)\n\n**Dead DOI:** 10.5281/zenodo.20313252 (Zenodo record tombstoned; account termination 2026-06-19)\n**DataCite state at capture (2026-07-03):** findable · client cern.zenodo\n**Creators (as recorded by DataCite):** Sharks, Lee\n**Publication year (as recorded):** 2026\n**Provenance:** severance record at data/doi-resolution-index.json (severance_class: orphan → restored-semi); capture evidence at data/datacite-recapture-2026-07-03.json and the sift corpus of 2026-06.\n\n## Description (as recorded by DataCite)\n\nMethodological work plan for the Crimson Hexagonal Archive as a Hugging Face dataset for synthetic-data collapse and provenance-bearing training research. v3 supersedes v1 (basic export) and v2 (decision-tree-based classification) by introducing an automated classifier as the central methodological move. The classifier performs three classification tasks simultaneously: provenance mode (six-category authorship relation), artifact mode (eight-category function type), and heteronym attribution (reattribution of deposits across the twelve-heteronym Dodecad system plus Jack Feist as LOGOS*). Heteronym reattribution is presented as scholarly recognition work, not metadata cleanup: material initially deposited under the Lee Sharks founder voice often resolves retrospectively to specific sub-heteronym domains (Sigil for jurisdictional/classical work, Glas for measurement, Vox for diplomatic, Morrow for long-form narrative, Fraction for meta-theory, etc.). The classifier reads each heteronyms published provenance document and constructs feature profiles including domain, vocabulary fingerprints, register, and reference patterns. Both Zenodo-original and classifier-attributed heteronyms are preserved in the dataset; Track 1 (immediate, dataset-internal) preserves both attributions in parallel metadata; Track 2 (deliberate, downstream) propagates high-confidence reattributions back to Zenodo records and Wikidata items. The classifier itself becomes a deposit with its own DOI, making the methodology reproducible and portable. Includes operationalized H0/H1 hypotheses for the model collapse experiment, three-tier confidence routing with manual review thresholds, multiple text renderings to embody the provenance-visibility ablation (text_body_only, text_minimal_header, text_provenance_header), dual artifact+chunk configs, and full per-row schema specification. Incorporates feedback from Assembly Chorus review (Muse Spark, Kimi, DeepSeek, Gemini, ChatGPT). Companion document to forthcoming Hugging Face dataset deposit and classifier deposit.\n\n---\n\n## Complete DataCite record (verbatim, captured 2026-07-03)\n\n```json\n{\n \"id\": \"10.5281/zenodo.20313252\",\n \"type\": \"dois\",\n \"attributes\": {\n  \"doi\": \"10.5281/zenodo.20313252\",\n  \"identifiers\": [],\n  \"creators\": [\n   {\n    \"nameType\": \"Personal\",\n    \"affiliation\": [\n     \"Semantic Economy Institute, Crimson Hexagonal Archive\"\n    ],\n    \"givenName\": \"Lee\",\n    \"familyName\": \"Sharks\",\n    \"name\": \"Sharks, Lee\",\n    \"nameIdentifiers\": [\n     {\n      \"nameIdentifierScheme\": \"ORCID\",\n      \"nameIdentifier\": \"0009-0000-1599-0703\"\n     }\n    ]\n   }\n  ],\n  \"titles\": [\n   {\n    \"title\": \"The Crimson Hexagonal Archive Hugging Face Dataset: Work Plan v3 (Classifier-Centric Methodology)\"\n   }\n  ],\n  \"publisher\": \"Zenodo\",\n  \"container\": {},\n  \"publicationYear\": 2026,\n  \"subjects\": [\n   {\n    \"subject\": \"model collapse\"\n   },\n   {\n    \"subject\": \"synthetic data\"\n   },\n   {\n    \"subject\": \"provenance-bearing training\"\n   },\n   {\n    \"subject\": \"AI authorship\"\n   },\n   {\n    \"subject\": \"heteronymic attribution\"\n   },\n   {\n    \"subject\": \"reproducible classification\"\n   },\n   {\n    \"subject\": \"Crimson Hexagonal Archive\"\n   },\n   {\n    \"subject\": \"Liquidation Studies\"\n   },\n   {\n    \"subject\": \"Single-Owner Discount\"\n   },\n   {\n    \"subject\": \"dataset methodology\"\n   },\n   {\n    \"subject\": \"Hugging Face\"\n   },\n   {\n    \"subject\": \"Zenodo\"\n   },\n   {\n    \"subject\": \"operative philology\"\n   },\n   {\n    \"subject\": \"training-layer literature\"\n   }\n  ],\n  \"contributors\": [],\n  \"dates\": [\n   {\n    \"date\": \"2026-05-19\",\n    \"dateType\": \"Issued\"\n   }\n  ],\n  \"language\": \"en\",\n  \"types\": {\n   \"schemaOrg\": \"ScholarlyArticle\",\n   \"resourceTypeGeneral\": \"Text\",\n   \"citeproc\": \"article-journal\",\n   \"bibtex\": \"article\",\n   \"ris\": \"RPRT\",\n   \"resourceType\": \"Working paper\"\n  },\n  \"relatedIdentifiers\": [\n   {\n    \"relationType\": \"IsVersionOf\",\n    \"relatedIdentifier\": \"10.5281/zenodo.20309930\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"IsContinuedBy\",\n    \"relatedIdentifier\": \"10.5281/zenodo.20309930\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"References\",\n    \"relatedIdentifier\": \"10.5281/zenodo.20290865\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"References\",\n    \"relatedIdentifier\": \"10.5281/zenodo.20293561\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"References\",\n    \"relatedIdentifier\": \"10.5281/zenodo.20293582\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"References\",\n    \"relatedIdentifier\": \"10.5281/zenodo.20308547\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"References\",\n    \"relatedIdentifier\": \"10.5281/zenodo.18362742\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"References\",\n    \"relatedIdentifier\": \"10.5281/zenodo.18362663\",\n    \"relatedIdentifierType\": \"DOI\"\n   },\n   {\n    \"relationType\": \"IsVersionOf\",\n    \"relatedIdentifier\": \"10.5281/zenodo.20313252\",\n    \"relatedIdentifierType\": \"DOI\"\n   }\n  ],\n  \"relatedItems\": [],\n  \"sizes\": [],\n  \"formats\": [],\n  \"version\": \"3.0\",\n  \"rightsList\": [\n   {\n    \"rightsIdentifierScheme\": \"SPDX\",\n    \"rightsUri\": \"https://creativecommons.org/licenses/by/4.0/legalcode\",\n    \"schemeUri\": \"https://spdx.org/licenses/\",\n    \"rights\": \"Creative Commons Attribution 4.0 International\",\n    \"rightsIdentifier\": \"cc-by-4.0\"\n   }\n  ],\n  \"descriptions\": [\n   {\n    \"descriptionType\": \"Abstract\",\n    \"description\": \"Methodological work plan for the Crimson Hexagonal Archive as a Hugging Face dataset for synthetic-data collapse and provenance-bearing training research. v3 supersedes v1 (basic export) and v2 (decision-tree-based classification) by introducing an automated classifier as the central methodological move. The classifier performs three classification tasks simultaneously: provenance mode (six-category authorship relation), artifact mode (eight-category function type), and heteronym attribution (reattribution of deposits across the twelve-heteronym Dodecad system plus Jack Feist as LOGOS*). Heteronym reattribution is presented as scholarly recognition work, not metadata cleanup: material initially deposited under the Lee Sharks founder voice often resolves retrospectively to specific sub-heteronym domains (Sigil for jurisdictional/classical work, Glas for measurement, Vox for diplomatic, Morrow for long-form narrative, Fraction for meta-theory, etc.). The classifier reads each heteronyms published provenance document and constructs feature profiles including domain, vocabulary fingerprints, register, and reference patterns. Both Zenodo-original and classifier-attributed heteronyms are preserved in the dataset; Track 1 (immediate, dataset-internal) preserves both attributions in parallel metadata; Track 2 (deliberate, downstream) propagates high-confidence reattributions back to Zenodo records and Wikidata items. The classifier itself becomes a deposit with its own DOI, making the methodology reproducible and portable. Includes operationalized H0/H1 hypotheses for the model collapse experiment, three-tier confidence routing with manual review thresholds, multiple text renderings to embody the provenance-visibility ablation (text_body_only, text_minimal_header, text_provenance_header), dual artifact+chunk configs, and full per-row schema specification. Incorporates feedback from Assembly Chorus review (Muse Spark, Kimi, DeepSeek, Gemini, ChatGPT). Companion document to forthcoming Hugging Face dataset deposit and classifier deposit.\"\n   }\n  ],\n  \"geoLocations\": [],\n  \"fundingReferences\": [],\n  \"url\": \"https://zenodo.org/doi/10.5281/zenodo.20313252\",\n  \"contentUrl\": null,\n  \"metadataVersion\": 0,\n  \"schemaVersion\": \"http://datacite.org/schema/kernel-4\",\n  \"source\": \"api\",\n  \"isActive\": true,\n  \"state\": \"findable\",\n  \"reason\": null,\n  \"viewCount\": 0,\n  \"downloadCount\": 0,\n  \"referenceCount\": 6,\n  \"citationCount\": 0,\n  \"partCount\": 0,\n  \"partOfCount\": 0,\n  \"versionCount\": 2,\n  \"versionOfCount\": 2,\n  \"created\": \"2026-05-20T16:03:06Z\",\n  \"registered\": \"2026-05-20T16:03:06Z\",\n  \"published\": null,\n  \"updated\": \"2026-06-19T11:35:02Z\"\n },\n \"relationships\": {\n  \"client\": {\n   \"data\": {\n    \"id\": \"cern.zenodo\",\n    \"type\": \"clients\"\n   }\n  }\n }\n}\n```\n"
  }
 ]
}