Intake
Candidate name, primary URL, row ID, and topic enter from the source matrix. The intake topic is preserved rather than inferred by the judge.
The catalog reflects the current CEI source pipeline and its drk_0805 rubric. Retrieval failures and uncertainty remain visible instead of being converted into low scores, and every published row carries assessment provenance.
Candidate name, primary URL, row ID, and topic enter from the source matrix. The intake topic is preserved rather than inferred by the judge.
The pipeline fetches and normalizes page or PDF text, extracts dates and metadata, and writes a crawl-quality report beside the exact evidence used for review.
Five eligibility gates test retrieval, AI focus, governance relevance, attribution, and whether the source is still in force. A failed gate prevents scoring.
Gate passers receive 0–2 scores for actionability, authority, and currency. Judged claims include evidence rationales; currency is computed from observed dates.
Perspective, coverage topic, lifecycle stage, layer, track, and granularity are attached as non-scored labels, then the deterministic outcome rule is applied.
G1 is computed from retrieval evidence. G2–G5 are grounded judgments over fetched text. When a required supporting quote is missing or invalid, the verdict becomes unresolved for review rather than an automatic rejection.
Does the URL resolve to enough full text to assess?
Is AI or an algorithmic system the source’s primary subject?
Does it address rules, oversight, rights, risk, safety, ethics with consequences, or policy?
Is a named author or issuing body responsible for the source?
Is this the current version, with no positive evidence of repeal, withdrawal, or replacement?
Only sources that pass all five gates are scored. If one criterion is unknown, the pipeline may pro-rate the available scores; at least two criteria must be scored before a tier can be assigned. Unknown evidence is never silently treated as zero.
Distinguishes description or analysis from general direction and from a specific obligation, control, or duty.
Rates the status of this document and its issuer—not the authority of sources it merely cites.
Uses the newest observed publication or modification date against a freshness window selected by source perspective. Missing dates remain unknown, never guessed.
Scores determine a tier only after eligibility is established. Classification labels describe the source but cannot move its tier. The browser selects Core and Supporting by default; the complete ledger also exposes Context only, Unresolved, and Rejected rows.
Each completed pipeline run is an immutable bundle containing the normalized crawl evidence, quality report, gate judgments, scores, assessment CSV and JSON, rejection log, model usage, and a manifest. The manifest records input and rubric hashes, model configuration, options, timestamps, source-control state, and outcome counts. Per-source revisit dates keep review work explicit.
Automated judgments support source triage; they do not replace subject-matter, legal, or editorial review. Unresolved rows and pipeline flags are intentionally retained so reviewers can see where evidence or model agreement was insufficient.