Interpreting the evidenceWhat stays fixed in the test.
Persistent and Decode–Inject consume the same frozen donor code and share the frozen query encoder. Prediction horizon and consuming interface are controlled explicitly.
What this setting establishes.
Direct persistent context can outperform an accurate decoded-parameter bottleneck. In the reported h32 comparison, this occurs despite decoder R² above .996.
Scope of the evidenceD-Clean is an analytically integrated state environment with no native camera. Different baseline and downstream populations remain separate; these are prediction comparisons, not closed-loop control.
Full result recordMeasurements and comparisons.
Reported results are kept with their own populations and conditions. Development, held-out and sealed comparisons remain separate.
01 / Study Table 8Adapted methods and native rollout references — C. D-Clean: common-reader H32
D-Clean report population; matched-reader H32 has 3 sources × 3 readers
Raw-state MSE ↓
BBold marks the SPRII matched-reader error in this common-reader comparison.
- SPRIIFixed Wrong and Fixed Zero change the fitted reader’s input; Fresh Null is trained separately.
Common-reader raw-state H32 errors use three sources and three readers. Fresh Null is fitted separately; Fixed Wrong and Fixed Zero change only the Matched reader’s input. Unevaluated interventions remain unavailable.
02 / Study Table 8Adapted methods and native rollout references
D-Clean report population; matched-reader H32 has 3 sources × 3 readers
Raw-state MSE ↓
BBold distinguishes direct prediction and free rollout; numerical errors are not pooled.
The dashed rule separates target horizon and rollout protocol.
Selected TDS source decoders are evaluated on the D-Clean report population. Direct h16 prediction and free rollout h32 use different prediction procedures and horizons; neither is the common-reader H32 comparison.
03 / Study Table 13D-Clean source acquisition and donor routes
200 one-shot sealed-test systems; 3 source seeds
Drag probe R²; source-specific learned-embedding h16 errors
BBold identifies strong drag accessibility under the two relation-trained alignment recipes.
The dashed rule separates the supervised-parameter reference.
- AlignSelf and donor errors use each source’s own learned-embedding targets. They are not a common physical-unit leaderboard.
Means ± sample SD across three sources on 200 one-shot sealed-test systems. Self and donor errors use each source’s learned-embedding target. Correct uses another same-system rollout; Gap is shuffled minus correct. Native has no donor route.
04 / Study Table 28Direct context versus decoded physical parameters — A. Absolute reader and intervention error
200 validation systems; 3 sources × 3 readers; same donor history; h32 primary
Raw-state MSE ↓; Decode-minus-Persistent gap; decoder R²
BBold compares direct and decoded context at the prespecified h32 endpoint.
- 32Oracle receives true parameters; Fixed Wrong is an input intervention on the Persistent reader, not a fifth fitted arm.
The four fitted readers use the same donor histories and frozen query encoder; h32 is primary. Fixed Wrong substitutes the Persistent reader’s context without fitting another reader. Oracle supplies true parameters to a learned reader, rather than an optimal-risk bound.
05 / Study Table 28Direct context versus decoded physical parameters — B. Paired contrasts and intervention uncertainty
200 validation systems; 3 sources × 3 readers; same donor history; h32 primary
Raw-state MSE ↓; Decode-minus-Persistent gap; decoder R²
BBold marks the h32 Decode-minus-Persistent contrast and its paired interval.
- 32Positive D−P favors direct context. Contrast values use the stated 10⁻⁴ scale.
D−P is Decode minus Persistent; positive favors direct context. Gaps and intervals use the stated 10⁻⁴ scale. Intervals are pointwise paired-system 95% intervals, conditional on the fitted grid; h32 is the primary endpoint.
06 / Study Table 28Direct context versus decoded physical parameters — C. Independent frozen-decoder validation
200 validation systems; 3 sources × 3 readers; same donor history; h32 primary
Raw-state MSE ↓; Decode-minus-Persistent gap; decoder R²
BBold marks independently validated decoder R² for all three samplings.
- 0No sampling is selected as best. Accurate decoding does not determine the most useful downstream interface.
The frozen drag decoder is independently validated on 200 systems with 32 windows each, without fitting on validation data. Sampling seeds are shown separately. Accurate factor decoding does not establish which downstream interface minimizes task error.