Skip to content

docs: specialize org governance for OCR evidence and model lifecycle - #31

Merged
Teakowa merged 1 commit into
mainfrom
docs/issue-27-ocr-governance
Sep 29, 2026
Merged

Teakowa merged 1 commit into
mainfrom
docs/issue-27-ocr-governance

Conversation

@e54-bot

@e54-bot e54-bot commented Sep 29, 2026

Copy link
Copy Markdown
Contributor

Summary

  • AGENTS.md applies the organization testing, verification, engineering-quality, and entropy policies explicitly, keeping OCRKit's stricter recognition/model specialization local.
  • Expected truth: recognition expected values require authority independent from the model/parser under test (reviewed labels, reproducible screenshot content, accepted layout/API contract, provenance-tracked regression); stable model output is not its own oracle and labels are not rewritten to make a new model pass; aggregate accuracy and fixture/test counts are evidence summaries; model versions, layout/fixture counts, and dataset cardinality are not durable contracts.
  • Model/evaluation boundary: training evidence and the decisive held-out evaluation set stay disjoint (the versioned source-level split in training/README.md is the enforcement mechanism); mutable release pointers/channels are never fixture truth.
  • Mechanism admission: fallback models, routing layers, and multimodal paths require the same measured recognition gap as training a new model.
  • Entropy: OCRKit-specific cleanup targets (duplicate preprocess paths, obsolete model fallbacks, stale layout compatibility, duplicated parser normalization, ownerless generated artifacts, migration adapters) with a consumer-mapping requirement before removing recognition paths; uncertainty/status reporting, validation, privacy boundaries, artifact verification, and relied-on compatibility stay.

Test plan

Fixes #27

Route shared testing/verification/engineering-quality/entropy policy to the organization owner and keep OCRKit's stricter local rules: expected truth independent from the model under test, disjoint training vs decisive evaluation evidence, mutable release channels are not fixture truth, dynamic model/layout/fixture counts are not durable contracts, no production OCR surface added for test access, measured-gap admission for model complexity, and consumer mapping before entropy removal.

Fixes #27
@Teakowa
Teakowa merged commit b2ff593 into main Sep 29, 2026
1 check passed
@Teakowa
Teakowa deleted the docs/issue-27-ocr-governance branch September 29, 2026 18:13
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

Specialize org governance for OCR evidence, evaluation, and model lifecycle

2 participants