Library / Semantic Integration Engineering Principles Framework
Jump to passage
In this reading

Link to current text

Published source confirmed at last check

Source changed 2026-10-02 23:06:08 UTC · snapshot created 2026-10-03 01:38:24 UTC · last check 2026-10-03 03:00:06 UTC

SIE.4 - Judge Cross-Source Correspondences and Their Permitted Uses

Type: Method pattern Status: Eternal alpha Normativity: Normative method guidance within SIE; examples are constructed and non-normative.

Primary working result: a QualifiedCorrespondenceSet@Use whose rows identify exact source-local endpoints, relation or incompatibility, orientation, bounded use, permitted loss, justification and evidence, source versions, counterexamples, confidence where meaningful, and accepted, rejected, unresolved, or incompatible disposition.

SIE.4:1 - Problem Frame

Use this when two or more governed sources contain concepts, types, relations, fields, codes, model elements, or claims that appear to match and a receiving use needs their relation made explicit. Typical triggers are a shared label, a crosswalk, an automated matcher result, a spreadsheet of “equivalences”, or a model-to-model mapping whose semantic premise has never been judged.

The primary EntityOfConcern is one use-qualified cross-source correspondence row. The first move is to recover both endpoint senses and ask which direct relation, difference, or incompatibility is actually supported. The first result is a set that preserves accepted, rejected, unresolved, and incompatible rows rather than forcing every candidate into an equivalence.

The payoff is a defensible semantic premise for identity work, claim composition, executable mappings, and validation. Do not use SIE.4 to construct a missing semantic model, decide cross-source identity, define transformation code, or choose the authoritative value. A local relation inside one governed source remains with its source owner unless the cross-source use makes it an integration premise.

SIE.4:2 - Problem

Correspondence candidates are cheap. Labels can be normalized, lexical similarity scored, hierarchies compared, and model structures aligned. Yet “similar”, “related”, “narrower”, “same field name”, and “safe to substitute here” answer different questions. A row can be a true relation and still lose distinctions that the receiving use requires.

When candidate generation and qualification collapse, downstream mapping code inherits an unstated ontology and loss policy. Counterexamples are discarded as data-quality defects, source versions disappear, and a reviewer cannot tell whether the relation itself or only its use is unsupported.

SIE.4:3 - Forces

ForceTension
AutomationMatchers can surface many candidates, while their scores do not establish relation truth or permitted use.
ReuseA general crosswalk is attractive, while relation adequacy can change with receiver, grain, interval, and accepted loss.
PrecisionExact relation kinds prevent overclaim, while sources may not support a complete formal characterization.
CoveragePressure favors filling every row, while unresolved and incompatible rows can be the most decision-useful result.
DirectionTransformation needs an orientation, while many conceptual relations are not symmetric or safely reversible.
VersioningStable mapping identifiers aid reuse, while endpoint meaning can change across source editions and profiles.

SIE.4:4 - Solution

Separate candidate discovery, direct relation judgment, and bounded-use qualification. For each row, identify exact source-local endpoints, state the relation or incompatibility and orientation, justify it with evidence and counterexamples, then decide whether that relation may support the named use under a stated loss boundary. Keep the direct relation and its use qualification independently inspectable.

SIE.4:4.1 - Pattern-Use Unfolding

  1. Take the question from the contract. State the receiving answer and why this endpoint relation can change it.
  2. Bind exact endpoints. Reference the SIE.2 source rows, schemes, definitions, editions, profiles, effectivity, units, codes, and local extensions. Do not map labels without their senses.
  3. Generate candidates without promotion. Lexical, structural, instance, expert, historical, or matcher evidence may nominate a row. Record the candidate source and score where useful, but do not treat nomination as acceptance.
  4. State the direct relation or difference. Use the narrowest warranted relation vocabulary. State direction and whether the inverse, symmetry, or transitivity is actually supported. If no positive relation is established, choose rejected, unresolved, or incompatible with reason.
  5. Test examples and counterexamples. Include instances or cases that should satisfy the relation and cases that distinguish endpoints, including version, unit, code, granularity, lifecycle, and authority differences.
  6. Qualify the receiving use. State which substitution, comparison, join, navigation, or transformation the row permits, what semantic loss it introduces, and which uses remain forbidden.
  7. Preserve evidence and authority. Record justification, evidence source, reviewer or domain authority where required, uncertainty or confidence, and the owner of any unresolved domain judgment.
  8. Assign a row disposition. Use accepted-for-use, accepted-with-loss, rejected, unresolved, or incompatible. A row can state a true broader/narrower relation yet be rejected for this use.
  9. State dependencies and tests. Name whether SIE.5, SIE.6, or SIE.7 consumes the row, plus a test or observation that reopens it.

SIE.4:4.2 - Record the Result

Correspondence positionRequired content
row identity and usestable row identifier, receiver/use, consuming result
source endpointsource, scheme/model, exact sense, edition/profile/effectivity, identifier
target endpointsource, scheme/model, exact sense, edition/profile/effectivity, identifier
direct relationrelation or explicit difference/incompatibility, orientation, inverse/symmetry/transitivity qualification
use qualificationpermitted operation, accepted loss, forbidden reuse, conditions
warrantcandidate origin, justification, evidence, authority, uncertainty/confidence where meaningful
challengepositive examples, counterexamples, negative/unlike cases
disposition and continuationaccepted-for-use, accepted-with-loss, rejected, unresolved, incompatible; downstream reference and reopen condition

SIE.4:4.3 - What Changes in Practice

The team stops treating a mapping spreadsheet as a bag of equalities. A matcher result becomes a candidate; a domain judgment becomes a relation warrant; and a receiving-use decision states what the relation may safely support. Unresolved and incompatible rows remain visible to the interface and tests.

SIE.4:5 - Archetypal Grounding - Product Feature and Inspection Characteristic

In the AP242/QIF application, the lexical candidate feature ↔ characteristic is too broad. The source inventory distinguishes an AP242 product-definition feature, a QIF inspection characteristic, the plan relation that applies a characteristic to a feature, and a measured result concerning that characteristic.

Candidate rowDirect relation judgmentUse qualification and disposition
AP242 feature F-17 ↔ QIF characteristic C-17not identity and not class equivalence; the QIF plan asserts that C-17 concerns the identified product feature under configuration/effectivity R/Eaccepted-for-use as a directed inspection-characteristic-concerns-feature relation only after SIE.5 qualifies the feature endpoint; no reverse substitution
AP242 feature type hole ↔ QIF characteristic type diametera diameter characteristic can characterize some holes, but the types are neither equal nor simply broader/narroweraccepted-for-use for navigation from a selected plan row with a qualified feature endpoint to its linked hole feature and diameter characteristic. Return their separate source identifiers and types with the plan relation; forbidden as a global type crosswalk
AP242 feature F-22 ↔ QIF characteristic C-99the candidate arose from a shared local label; effectivity and plan evidence do not support the relationrejected; the interface returns unmatched rather than joining
legacy feature class ↔ current QIF characteristic classone local extension lacks an authoritative definitionunresolved and stopped for dependent mapping; return to the extension owner

The set records endpoint editions and counterexamples such as a non-dimensional inspection characteristic and a feature absent at effectivity E. It establishes semantic premises; it does not decide that identifiers denote the same entity or specify executable joins.

SIE.4:6 - Bias-Annotation

LensLikely driftRepair
GovernanceA matcher or integrator silently becomes the domain relation authority.Record the warrant, decision scope, and required domain authority for each accepted row.
ArchitectureAvailable mapping technology limits the relation vocabulary.State the semantic relation first; choose its carrier later.
Ontology/EpistemologySimilarity, correlation, broader/narrower relation, and equivalence collapse.Use the narrowest supported relation and preserve uncertainty and counterexamples.
PragmaticsEvery possible endpoint pair is reviewed.Restrict rows to relations that can change the contract or a dependent result.
DidacticsDirection arrows are read as a mandatory processing order.Explain orientation as relation or transformation meaning, not lifecycle sequence.

SIE.4:7 - Conformance Checklist

  • Every row serves a named receiving use and consuming result.
  • Both endpoints reference exact source-local senses and applicable editions/profiles.
  • Candidate generation is distinguishable from relation judgment.
  • The direct relation, difference, or incompatibility uses the narrowest warranted kind and states orientation.
  • Inverse, symmetry, and transitivity are not inferred without support.
  • Positive examples and discriminating counterexamples are present where load-bearing.
  • Permitted operation, semantic loss, conditions, and forbidden reuse are explicit.
  • Evidence, uncertainty/confidence, and decision authority are inspectable.
  • Rejected, unresolved, and incompatible rows remain available to downstream interfaces and tests.
  • The set claims neither cross-source identity nor executable transformation by implication.

SIE.4:8 - Common Anti-Patterns and How to Avoid Them

Anti-patternRepair
Same label, therefore same meaningRecover endpoint senses and test counterexamples.
Matcher score as truthTreat the score as candidate evidence; obtain a relation warrant and use decision.
Every relation is sameAsDistinguish identity, equivalence, broader/narrower, related, concerns, transforms-to, difference, and incompatibility.
One global crosswalkQualify rows by use, source editions, loss, and forbidden reuse.
Symmetric spreadsheet mappingState orientation and test whether the inverse is supported.
Delete failed rowsRetain rejected, unresolved, and incompatible dispositions for receiver and validation behavior.

SIE.4:9 - Consequences

The correspondence set gives mapping and interface work explicit semantic premises and preserves honest negative results. It makes version and loss dependencies visible and allows a narrow use to proceed while other rows remain unresolved.

The cost is row-level judgment and counterexample work. Automated matchers remain useful for candidate discovery but cannot close the relation or its receiving-use qualification alone.

SIE.4:10 - Rationale

A direct relation and permission to use that relation answer different questions. Keeping them separate allows a true relation to be rejected for a loss-sensitive use and a qualified narrow relation to support one operation without being promoted globally. This is the smallest structure that protects both semantic truth and practical action.

SIE.4:11 - SoTA-Echoing

The best-known line combines explicit relation vocabularies, mapping metadata, alignment evaluation, and direct Bridge truth. The serious default is a two-column crosswalk plus confidence score. Its defect is that it hides endpoint senses, relation algebra, bounded-use loss, and counterexamples. SIE.4 adapts current mapping practice into a set whose negative dispositions and use qualification are first-class.

Source lineAdopt, adapt, or rejectRole and limit
Current FPF F.9adoptGoverns each direct cross-context Bridge and separate bounded-use claim; a package row must not replace Bridge truth.
SKOS ReferenceadaptSupplies distinct mapping relation forms and cautions through their semantics; SKOS does not prove a candidate relation or authorize substitution.
SSSOM 1.0adaptContributes inspectable subjects, predicates, objects, provenance, justification, confidence, and source versions; metadata completeness is not relation truth.
OAEI 2025 resultsadaptDemonstrate task- and track-dependent alignment evaluation; benchmark performance does not decide a local row.
label-only or score-only crosswalkreject as sufficientIt can nominate candidates but cannot establish endpoint senses, relation, permitted loss, or authority.

Reopen when an endpoint meaning or edition changes, a counterexample defeats the relation or use qualification, the receiver’s loss boundary changes, or a better relation vocabulary materially changes action.

SIE.4:12 - Relations

  • SIE.2 supplies exact endpoint senses, editions, profiles, authority limits, and model gaps.
  • SIE.3 is a conditional external return when available models cannot express a required endpoint distinction.
  • SIE.5 consumes rows that expose a load-bearing cross-source identity question.
  • SIE.6 uses correspondences to decide which source claims are comparable or composable without erasing scope or conflict.
  • SIE.7 consumes only accepted rows under their stated use and loss conditions; it adds executable transformation semantics.
  • SIE.10 tests correspondence truth and use qualification independently from mapping execution.

SIE.4:End

Referenced in the corpus

44 literal mentions in other sections. Read their context to establish the relation.