{
  "schema": "org-writing@v1",
  "slug": "label-drift",
  "kg": {
    "id": "org:writing:label-drift",
    "type": "brick",
    "graph": "/kg.json"
  },
  "title": "Label drift",
  "subtitle": "A method's vocabulary spreads faster than its skill, until the word marks a stance rather than a capability",
  "abstract": "Why informed and inspired multiply around every successful method, what the honest version of that suffix looks like, and the one question that tells a claim from a stance. The canonical treatment of label drift.",
  "kind": "brick",
  "topics": [
    "Tacit knowledge"
  ],
  "courseMemberships": [
    {
      "course": "org:courses:tacit",
      "topic": "Tacit knowledge",
      "wall": "org:walls:practice",
      "position": 3,
      "total": 6
    }
  ],
  "publishedAt": "2026-08-03T00:00:00.000Z",
  "updatedAt": "2026-08-11T00:00:00.000Z",
  "version": 3,
  "guidelinesVersion": 15,
  "brief": {
    "problem": {
      "text": "The words a method invents are free to adopt while the practice behind them costs years, so informed and inspired variants proliferate far ahead of any fidelity, and the vocabulary ends up describing an intention rather than a competence.",
      "claims": [
        "IFS-informed practice proliferates without full certification",
        "many widely used AI safety benchmarks correlate highly with general model capabilities"
      ]
    },
    "mechanism": {
      "text": "Adoption of a label costs a sentence and adoption of the skill costs an apprenticeship, so the two diffuse at wildly different speeds until the word stops predicting anything about what happens in the room.",
      "claims": [
        "self-determined by the schools themselves"
      ]
    },
    "move": {
      "text": "Ask what would have to be true for the label to be false, and if nothing would be, treat the word as a statement of intent and go looking for the evidence separately.",
      "claims": []
    }
  },
  "sources": [
    {
      "repo": "mnstry-strategy",
      "path": "docs/20-business/10-discovery/segment-deep-research.md"
    }
  ],
  "canonicalPath": "/writing/label-drift/",
  "body": "Every method that works acquires a suffix. Internal Family Systems certification takes years of practice and consultation after the training, and meanwhile IFS-informed practice proliferates without it, which our research record flags as the label drift risk in exactly those words. Trauma-informed made the same journey further and faster, from a specific set of clinical commitments to a phrase that appears on funding applications, job descriptions, and organizational values pages. The critique has caught up with it. A 2025 paper in Culture, Medicine and Psychiatry asks in its title whether the wide reach of the trauma-informed model exceeds its narrow grasp, and practitioners writing from inside the field describe the term collapsing into a box to tick, satisfied by a two-hour training and a policy statement.\n\nThe asymmetry is the whole mechanism. Adopting a method's vocabulary costs one sentence and adopting its judgment costs an apprenticeship, so the two diffuse at speeds that differ by orders of magnitude, and past some ratio the word stops predicting anything about what happens in the room. Nothing dishonest has to occur for this to run. Most people using the label learned something real, meant it sincerely, and simply had no way to know what the remaining distance was, because the part they did not get is the part nobody could have written down for them.\n\nThere is an honest version of the suffix, and it is worth studying because it shows what candor about the remainder looks like in practice. The North American Reggio Emilia Alliance states plainly that the municipality of Reggio Emilia approves no certifications in the approach, that there are no Reggio Emilia schools outside the city itself, and that when schools elsewhere call themselves Reggio-inspired the phrase is self-determined by the schools themselves, with vast differences in what it means. That is a label doing the one thing a drifting label never does. It declares its own uncertainty in the same breath it invites use. A parent reading it knows the word is a direction of travel and not a guarantee, which is precisely the information that trauma-informed no longer carries.\n\nWe apply this to our own field before it feels comfortable, because safety is the word most exposed to it right now. Ren and colleagues, in a NeurIPS 2024 paper they titled Safetywashing, measured widely used AI safety benchmarks against general capability and training compute and found many of them highly correlated, which means a lab can improve its safety scores by making a bigger model and report the result as safety progress. The vocabulary of the field is doing what IFS-informed did, spreading across products at conversational cost while the underlying discipline spreads at the speed of people who actually know how to do it. We use the word too. Every restraint claim on this site is a sentence anybody could write, which is why the ones that matter here are attached to structural facts a reader can check rather than to the adjective.\n\nThe kinship is with the metric proxy, one level up. There a countable stand-in replaced the quality of a bond with a quantity of it, and once the number existed, maintaining the number substituted for relating. Here the stand-in is a word rather than a number, and the substitution is the same shape: once the label exists, holding the label substitutes for holding the skill, and the market rewards the label because the label is the part it can see. Goodhart's law does not require arithmetic. A proxy will do.\n\nSo the audit is one question, asked of any label including ours. What would have to be true for this word to be false here? If there is an answer, hours of supervision, a consultation requirement, a measurement someone else could take, then the word is a claim and it can be checked. If nothing whatsoever would make it false, it is a statement of intent, and intent is worth something, but it is not the thing the word is being read as. Ask it out loud and you find the practitioners who have gone the distance, because they are the ones who answer it with relief.",
  "apparatus": {
    "note": "The human-facing essay is deliberately practical; this apparatus carries the full references, evidence-graded claims, article-local concepts, and research context behind it. Canonical concept definitions come from the concept registry.",
    "references": [
      {
        "id": "org:references:label-drift:r01",
        "author": "North American Reggio Emilia Alliance",
        "work": "General FAQs",
        "relevance": "The honest suffix. The alliance states that no certifications in the approach are approved by the municipality of Reggio Emilia, that there are no Reggio Emilia schools outside the city, and that Reggio-inspired language is self-determined by schools, with vast differences in what it means in practice."
      },
      {
        "id": "org:references:label-drift:r02",
        "author": "Vojtech Pisl, Sanne te Meerman, Allen Frances, Laura Batstra",
        "work": "Does the Wide Reach of the Trauma-informed Model Exceed its Narrow Grasp? (Culture, Medicine, and Psychiatry 49)",
        "year": 2025,
        "relevance": "The peer-reviewed critique of the most widely diffused informed label, examining the discursive practices that present the broad trauma model as uncontested and the risks that follow."
      },
      {
        "id": "org:references:label-drift:r03",
        "author": "Richard Ren, Steven Basart, Adam Khoja, Alice Gatti, Long Phan, Xuwang Yin, Mantas Mazeika, Alexander Pan, Gabriel Mukobi, Ryan H. Kim, Stephen Fitz, Dan Hendrycks",
        "work": "Safetywashing, Do AI Safety Benchmarks Actually Measure Safety Progress? (NeurIPS 2024, Datasets and Benchmarks Track)",
        "year": 2024,
        "relevance": "The same drift inside our own field, measured. Many widely used safety benchmarks correlate highly with general capabilities and training compute, so capability gains can be reported as safety progress."
      },
      {
        "id": "org:references:label-drift:r04",
        "author": "MNSTRY research record",
        "work": "Methodology licensing analysis, internal",
        "year": 2026,
        "relevance": "The source of the IFS-informed observation and of the phrase label drift itself. Internal market research; only the vocabulary observation is used, none of the surrounding segment or prospect material."
      }
    ],
    "claims": [
      {
        "id": "org:claims:label-drift:c01",
        "claim": "Our research record reports that IFS-informed practice proliferates without full certification fidelity, and names this the label drift risk.",
        "basis": "The internal licensing analysis, which observes the pattern against IFS certification requirements. An observation from a single internal synthesis rather than a measured prevalence.",
        "confidence": "directional",
        "sources": []
      },
      {
        "id": "org:claims:label-drift:c02",
        "claim": "The spread of the trauma-informed model has drawn peer-reviewed critique that its reach exceeds its evidential grasp, alongside practitioner accounts of the term collapsing into a box to tick.",
        "basis": "Pisl, te Meerman, Frances and Batstra 2025 for the academic critique, verified to title, authors, journal and thesis; the practitioner accounts are commentary rather than measurement, and the combined claim is graded down to their level.",
        "confidence": "directional",
        "sources": []
      },
      {
        "id": "org:claims:label-drift:c03",
        "claim": "The Safetywashing paper found that many widely used AI safety benchmarks correlate highly with general model capabilities and training compute, so capability improvements can be presented as safety progress.",
        "basis": "Ren et al., NeurIPS 2024, abstract and stated findings, checked at the paper's own listing during authoring.",
        "confidence": "verified",
        "sources": []
      },
      {
        "id": "org:claims:label-drift:c04",
        "claim": "The North American Reggio Emilia Alliance states that no certifications in the approach are approved, that there are no Reggio Emilia schools outside the city, and that Reggio-inspired language is self-determined by the schools themselves, with vast resulting differences.",
        "basis": "Fetched directly from the alliance's own FAQ during authoring.",
        "confidence": "verified",
        "sources": []
      }
    ],
    "concepts": [
      {
        "id": "org:concepts:label-drift",
        "name": "Label drift",
        "definition": "The condition in which a method's vocabulary diffuses far ahead of its practice. Adopting the word costs a sentence and adopting the judgment costs an apprenticeship, so the label and the skill spread at speeds differing by orders of magnitude, and past some ratio the word stops predicting anything about what happens in the room. The informed and inspired suffixes are its usual carriers. Drift does not require bad faith; it requires only that the part which failed to transfer is the part nobody could have written down. The honest form of the same suffix declares its own uncertainty, as the Reggio Emilia alliance does in stating that Reggio-inspired is self-determined by the schools using it. The diagnostic is one question asked of any label, including our own safety vocabulary: what would have to be true for this word to be false here?",
        "provenance": "canonical"
      },
      {
        "id": "org:concepts:tacit-remainder",
        "name": "Tacit remainder",
        "definition": "What is left of a practice after the best possible writing-down. Codification carries everything a method can state about itself, the sequence, the vocabulary, the diagnostic categories, and leaves behind the situational judgment that decides when the stated thing applies, because that judgment was never in sentence form even in its originator's head. The remainder is not diffuse: it sits in four dimensions, decision trees (what to do when the person does not do the expected thing), micro-timing (when to probe and when to wait), interpretive frameworks (what a behavior means inside the method's worldview), and feedback loops (how a practitioner learns they are drifting). Polanyi's observation that we can know more than we can tell is its oldest statement. Its practical consequence is that codification does not transmit expertise at some percentage of fidelity; it clears the ground around the remainder so that whoever is forming the next practitioner can see exactly what still has to be learned by proximity.",
        "provenance": "canonical"
      }
    ],
    "researchContext": "The agent-facing body for \"Label drift\": the case set, the grading, and the\nreason our own field is in it.\n\n## The self-application\n\nThe brick applies the argument to safety vocabulary, including ours, because a\npiece about labels that exempted its author would be demonstrating the failure\nit describes. The Safetywashing result is the strongest available evidence\nthat the word can move independently of the thing in a field with far better\ninstrumentation than coaching has. Our own restraint claims are held to the\nstated test: attached to structural facts a reader can check, or treated as\nintent.\n\n## Kinship rather than duplication\n\nThe metric proxy brick owns the numeric version of this mechanism, where a\ncountable stand-in replaces the quality of a bond and maintaining the number\nsubstitutes for relating. This brick owns the vocabulary version and cites the\nkinship in one paragraph rather than re-arguing it. Goodhart's law appears in\nboth because it is the same law; the brick's contribution is that the proxy\ndoes not have to be arithmetic.\n\n## Grading\n\nTwo claims are verified from primary pages, Reggio and Safetywashing. The\ntrauma-informed claim is graded directional because it bundles a peer-reviewed\ncritique with practitioner commentary, and the weaker half sets the grade. The\nIFS-informed observation is directional as an internal synthesis.\n\n## De-identification\n\nThe internal source is segment research. Nothing about segments, prospects, or\npricing appears here; the only material taken is the vocabulary observation\nand the term label drift."
  },
  "contract": "https://mnstry.org/contracts/org/org-writing.v1.schema.json",
  "releaseHash": "d0395003512439635eab8349b39233b72b8614ca4f1116fd27bb25adaa77da06",
  "versions": [
    {
      "version": 3,
      "cutAt": "2026-08-11",
      "note": "Tacit remainder demotion release: reparented to The honest instrument and Tacit Knowledge course projection updated.",
      "visibility": "published",
      "path": "/writing/label-drift/",
      "contentHash": "sha256:2aae0b463b648f64",
      "releaseHash": "d0395003512439635eab8349b39233b72b8614ca4f1116fd27bb25adaa77da06"
    },
    {
      "version": 2,
      "cutAt": "2026-08-03",
      "note": "Remove authorial should language without changing the claim.",
      "visibility": "published",
      "path": "/writing/label-drift/v/2/",
      "contentHash": "sha256:2aae0b463b648f64",
      "releaseHash": "4352d6e5c544f2a7e9ca18a6dd000a553a3b807ec079ef3d03efc3f7c3219e21"
    },
    {
      "version": 1,
      "cutAt": "2026-08-03",
      "note": "Initial publication, tacit wave",
      "visibility": "published",
      "path": "/writing/label-drift/v/1/",
      "contentHash": "sha256:e0fb957886201154"
    }
  ]
}