provenance truth pass: resolve prior-art citations, correct pre-production/compression claims - #2
Open
keithbinkly wants to merge 2 commits into
Open
provenance truth pass: resolve prior-art citations, correct pre-production/compression claims#2keithbinkly wants to merge 2 commits into
keithbinkly wants to merge 2 commits into
Conversation
The live meta-context page's vendor matrix cites spec/prior-art.md as its 'full detail' source for all rows, but 6 tools had no entry there: Euno, Kilo Code, Cassis, WHOOP snowflake-semantic-tools, and Ramp Research. Added each as a descriptive entry (what it stores, how it delivers), sourced from data-centered's six-axis-context-tool-comparison DRAFT v2 and the dbt Summit 2026 context-tools research report. No scoring against this schema's five layers, no L1-L5 labels applied to the added tools, no competitor framing. Power BI (prep-for-AI) intentionally excluded: the six-axis draft itself flags it as part of an unlinked 'Set A citation gap' pending a citation pass against current vendor docs — no primary source to cite. Ref: dc-tracks/dbt-summit-econ-open provenance-adjudication-2026-07-30.md R4.
Three factual overclaims/errors caught by a provenance adjudication: - spec/known-gaps.md Gap 1 'Raised by: Production deployment...' was wrong — the source (an external adopter, unnamed pending their own consent to be identified) explicitly described pre-production, design-and-development experience, not production metrics. Corrected the framing without naming the adopter. - CHANGELOG.md 0.1.0 'validated in a production financial-services deployment' overclaimed; reframed to pilot/design-phase language. - eval/results.md and CHANGELOG.md both stated a '5.6x compression' distillation ratio that contradicts the eval's own word-count table (4,743 / 1,157 words = 4.1x, which is what the live page already states). Corrected both instances to 4.1x. Ref: dc-tracks/dbt-summit-econ-open provenance-adjudication-2026-07-30.md R5.
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
Two frozen adjudication rulings (R4, R5) from a provenance sweep of the live meta-context pages, applied to this repo. Ruling source:
dc-tracks/dbt-summit-econ-openprovenance-adjudication-2026-07-30.md.R4 — vendor matrix citation now resolves
The live meta-context page's vendor comparison matrix cites
spec/prior-art.mdas its "full detail" source for every row, but 6 of 17 rows had no entry there. Added:AGENTS.md, custom rules, deprecated Memory Bank), context bound to repo paths not data objectssnowflake-semantic-tools— open-source CLI compiling git-versioned YAML to Snowflake Semantic Views; structured fields + verified-query exemplars, business rules as a free-text instruction stringEach entry describes what the tool stores and how it delivers context — no scoring against this schema's five layers, no L1–L5 labels applied to the added tools, no competitor language, no ✓/✗ marks. Each carries its source verification date (2026-06-26 or 2026-07-01, per the source documents).
Power BI (prep-for-AI) is intentionally skipped. The six-axis comparison draft (the only source naming it) itself flags Power BI as part of an unlinked "Set A citation gap" — pending a citation pass against current vendor docs, with no primary source fetched. Inventing a defensible entry from an admittedly-unverified row would repeat exactly the kind of provenance failure this adjudication exists to fix, so that row is left out rather than filled with guessed detail.
Sources used (read-only, not modified):
analytics-workspace/data-centered/content/articles/drafts/six-axis-context-tool-comparison-DRAFT-v2.md(the v2 draft, per instruction — not the non-v2 version) and.../content/research/dbt-summit-2026/context-tools-research-report.md.R5 — three factual corrections
(a)
spec/known-gaps.mdGap 1 "Raised by" line**Raised by:** Production deployment in regulated public-transport domain**Raised by:** A pre-production field report from a regulated public-transport deployment (design-and-development phase, not production metrics)(b)
CHANGELOG.md0.1.0 entry...validated in a production financial-services deployment....piloted in a financial-services deployment during its design phase.(c)
eval/results.md+CHANGELOG.md— 5.6× → 4.1× compressionGPT-5.6model-name references ineval/design-v2.md, left untouched).Gates (pasted)
1. Validator test suite (from the
validator/directory):2. Grep — zero remaining overclaim strings:
3. All 6 tools present in
spec/prior-art.md(Power BI skipped, reason above):4.
git diff --stat(only intended files):Test plan
git diff --stattouches only the 4 intended filesCorrections to public claims; Keith merges.
https://claude.ai/code/session_017uymU7VUQ2ngpuASgzt3t4