feat(mrl): add benchmark flags and sealed canary fixture - #295
feat(mrl): add benchmark flags and sealed canary fixture#295TheHalfMoon wants to merge 33 commits into
Conversation
|
Important
This repository does not receive automatic reviews because it has fewer than 10 stars. ⚙️ Run configurationConfiguration used: defaults Review profile: CHILL Plan: Pro Plus Run ID: Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out. Comment |
|
ⓘ Qodo reviews are paused because your trial has ended. Ask your workspace admin to add credits to resume reviews. Manage billing |
|
@coderabbitai review Please perform a fresh exact-head review of PR #295 at |
|
The author of this PR is on the CodeRabbit Free Plan. In order to use the Chat feature, please upgrade the PR author to CodeRabbit Pro by visiting https://app.coderabbit.ai/settings/billing. |
|
@cubic-dev-ai review this PR Perform a fresh direct semantic review of exact PR #295 head |
@TheHalfMoon You've already used all 20 free PR reviews this month. Please upgrade your plan to keep running reviews: https://www.cubic.dev/settings?tab=subscription |
|
Exact-head qualification blocker recheck — August 29, 2026 The exact candidate head remains Fresh direct retries were performed again on the same exact head:
These results are external pre-execution hosted-runner blocker evidence only. They are not Python, Ruff, mypy, pytest/coverage, MRL machine-state, |
Exact-head qualification retry — attempt 4A fresh direct retry was performed on the unchanged exact head
The candidate remains These attempt-4 results are external pre-execution hosted-runner blocker evidence only. They are not Python/Ruff/mypy/pytest/coverage/machine-state/ |
|
@coderabbitai review Perform a fresh exact-head review of PR #295 at |
|
The author of this PR is on the CodeRabbit Free Plan. In order to use the Chat feature, please upgrade the PR author to CodeRabbit Pro by visiting https://app.coderabbit.ai/settings/billing. |
|
@coderabbitai full review |
|
|
@coderabbitai full review |
|
|
@coderabbitai full review Please perform a fresh exact-head review of PR #295 at |
|
The author of this PR is on the CodeRabbit Free Plan. In order to use the Chat feature, please upgrade the PR author to CodeRabbit Pro by visiting https://app.coderabbit.ai/settings/billing. |
|
Superseded by canonical main. The MRL-0604 benchmark-derived-generation hardening and MRL-0606 sealed temporal-canary fixture workflow carried here are present on current |
Summary
Implement MRL-0604 benchmark-derived-generation flags and MRL-0606 R2-compatible sealed temporal-canary fixture workflow while preserving the evidence-only MRL boundary.
MRL-0604 — Benchmark-derived generation flags
NOT_BENCHMARK_DERIVED,BENCHMARK_DERIVED, andINDETERMINATEclassificationsMRL-0606 — Sealed temporal-canary fixture workflow
sealed=True,fixture_only=True,can_enter_training=False,can_enter_search=False, andcan_authorize=FalseUpstream reconciliation chain
Manual semantic review has established an upstream MRL-0601 construction-identity defect. PR #305 now carries the isolated MRL-0601 fix. PR #299 must first be reconciled against canonical #305 so MRL-0602/MRL-0603 validate the original construction-bound lineage. This PR must then be reconciled against the final canonical #299 implementation.
Therefore the current head is explicitly pre-reconciliation. Its qualification/review evidence cannot authorize merge and becomes stale after the required upstream reconciliation chain:
Canonical base
bf92dd2977d24aa597d2442decabc215f7bd3dbfCurrent exact pre-reconciliation head
b6c0fb2e3ef35bc19011451e1e7e0151f13f8a67The current scope is exactly four intended files and prior live compare showed
behind_by=0.Current exact-head qualification blocker
Fresh automatic workflows on this pre-reconciliation head terminate before any workflow step executes:
3326902072199144046903:failure,steps=null99144047058:failure,steps=null3326902072499144046831:failure,steps=nullThese are external pre-execution hosted-runner blocker results only. They are not Ruff, format, strict mypy, pytest/coverage, MRL machine-state,
medscale check, or CodeQL-analysis results and do not authorize merge.A separate repository-level security qualification blocker remains: prior CodeQL execution reached SARIF upload and GitHub reported code scanning is not enabled for this private repository. Connected GitHub tooling exposes no repository-security or Actions billing/budget mutation for removing these blockers.
No human reviewer has been requested or contacted. No paid review usage is enabled.
Boundary
This candidate records metadata and runs only deterministic in-memory fixture evaluation. It performs no benchmark/corpus read, benchmark-derived generation, real model/provider/network/GPU work, training, promotion, deployment, release, or clinical action. Item-level fixture values are not emitted and the canary remains prohibited from training/search reuse.
MRL-0607 remains separately dependent on canonical MRL-0606 and is not eligible while this PR is unqualified/unmerged.
Fresh exact-head Python 3.11/3.12 CI, Ruff lint/format, strict mypy, full pytest/coverage, MRL machine-state drift/manual-edit gate,
medscale check, CodeQL/security qualification, exact intended scope,behind_by=0, mergeability, and zero unresolved material review findings/threads are required before guarded merge with the then-currentexpected_head_sha.No force-push, rebase, destructive history rewrite, real-asset access, provider spend, or governance bypass is used.