fix(rectification): bump scoring identity to scoring-10 for functional profile v2; dated contract by generation (BUG-1181)
Functional roles feed the *_functional_*_auxiliary rules, so57782aeachanges candidate scores for identical input (memoization fixture 12:00: 8.6274 -> 8.4977). Per the "scoring semantics change => bump ALGORITHM_VERSION" precedent (scoring-7 -> 8 -> 9), the identity moves to scoring-10; policy v3, input contract v5 and Skill versions are unchanged, history is not relabeled. Five frontend sites and one SQL guard tested `=== "...scoring-9"` for the dated candidate-window contract; they now use isDatedScoringAlgorithmVersion / a generation regex (>= 9). Migration 20261002010000 only recreates validate_dated_rectification_candidate (one-line guard change). Memoization golden v2 written by the test's own write_golden; v1 (scoring-8) frozen by sha256. Real-engine scoring-10 cross-midnight golden added. Research records re-frozen per ERR-110 (label functional_v2_2026_10_02) and scripts/functional_benefics.py added to the frozen production identity (ERR-114:57782aeachanged scores without tripping it). Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01N4f2nya58RoRu4yEmJgRGE
This commit is contained in:
co-authored by
Claude Opus 5.5
parent
57782aea8a
commit
3733b9787b
@@ -12,6 +12,7 @@ from tests.test_reported_offset_research import ( # noqa: F401
|
||||
from tests.test_sealed_holdout_contract_freshness import ( # noqa: F401
|
||||
test_changed_frozen_identity_fails_before_any_replay,
|
||||
test_contract_tracks_actual_current_scorer_and_dataset_audit,
|
||||
test_functional_roles_are_part_of_the_production_identity,
|
||||
test_fixed_protocol_rerun_is_auditable_but_never_independent_blind,
|
||||
test_frozen_record_matches_dataset_scorer_and_evaluator_bytes,
|
||||
test_extended_identity_drift_rejected_even_when_legacy_hash_unchanged,
|
||||
|
||||
Reference in New Issue
Block a user