Scientific objective
Build the evidence required for a Prob4D-centered paper claim:
Partial gauge information is admitted only when it supports a declared physical query; complete-prior uncertainty is retained in unobserved directions; unsupported precision is rejected; otherwise the exact physical fallback is preserved.
The implementation is on science/query-aware-observability-v1. This issue tracks promotion beyond controlled mechanism evidence.
Stage 1 — implementation and analytic control
Stage 2 — DLO-like controlled sweep
Freeze before running:
Required baselines:
Required outputs:
Stage 3 — fresh-provider qualification
Coordinate with #333 and #49.
No target access is authorized merely because the adapter executes.
Stage 4 — held-out registration
Before opening target outputs:
Stage 5 — end-to-end evidence
Evaluate on identical held-out factors/queries:
Primary evidence bar:
Claim boundary
Until Stages 3–5 pass, this work supports only a controlled method claim. It must not be described as fresh real-provider competence, a new benchmark result, or established downstream BayesianPhysTwin improvement.
Scientific objective
Build the evidence required for a Prob4D-centered paper claim:
The implementation is on
science/query-aware-observability-v1. This issue tracks promotion beyond controlled mechanism evidence.Stage 1 — implementation and analytic control
factor-supported.fallback-unresolved.prior-bounded.fallback-invalid-updatedespite sub-tolerance reported uncertainty.Stage 2 — DLO-like controlled sweep
Freeze before running:
Required baselines:
Required outputs:
Stage 3 — fresh-provider qualification
Coordinate with #333 and #49.
No target access is authorized merely because the adapter executes.
Stage 4 — held-out registration
Before opening target outputs:
prior-boundedis permitted;FlorianPfaff/BayesianPhysTwin-Paper.Stage 5 — end-to-end evidence
Evaluate on identical held-out factors/queries:
Primary evidence bar:
Claim boundary
Until Stages 3–5 pass, this work supports only a controlled method claim. It must not be described as fresh real-provider competence, a new benchmark result, or established downstream BayesianPhysTwin improvement.