Offline robot forecasting across input routes — A. Self-context
25 held-out tasks; task-macro means over 3 fitted sources
Normalized force/torque MSE ↓; TCP error in cm ↓
| Condition | FT MSE h1 | FT MSE h4 | FT MSE h16 | TCP cm h4 | TCP cm h16 |
|---|---|---|---|---|---|
| Native | 0.6346 | 1.1175 | 1.4971 | 1.9621 | 4.0632 |
| Structure | 0.6199 | 1.0993 | 1.4985 | 1.9444 | 3.9645 |
| Align–Indep | 0.6633 | 1.1316 | 1.5031 | 2.2618 | 4.4288 |
| Align–Random | 0.6232 | 1.1064 | 1.5098 | 2.0454 | 4.2568 |
| Cross–SameEp | 0.6145 | 1.0898 | 1.4948 | 1.9628 | 3.9652 |
| Cross–Indep | 0.6230 | 1.0930 | 1.4974 | 1.9015 | 3.9022 |
| Cross–Random | 0.6307 | 1.0961 | 1.4950 | 1.9476 | 3.9505 |
| A+C–SameEp | 0.6407 | 1.1202 | 1.5003 | 2.0976 | 4.3710 |
| A+C–Indep | 0.6602 | 1.1296 | 1.5046 | 2.2683 | 4.4347 |
| A+C–Random | 0.6225 | 1.1005 | 1.5068 | 2.0359 | 4.2575 |
| LD–Native | 0.5909 | 1.0893 | 1.4893 | 1.7296 | 3.8873 |
| LD–Cross–Indep | 0.5760 | 1.0695 | 1.4880 | 1.5979 | 3.7484 |
| LD–A+C–SameEp | 0.6162 | 1.1177 | 1.5045 | 2.0324 | 4.3018 |
| LD–A+C–Random | 0.6090 | 1.1224 | 1.5054 | 1.8664 | 4.1913 |
Bold compares Structure and independent-episode Cross on self-context TCP forecasts.
The dashed rule separates full-input and LowDim models.
- Cross–IndepOther conditions and outputs remain visible. Compare each input-eligible condition on its own target scale.
Task-macro means average three fitted sources on 25 held-out tasks. Force/torque uses normalized MSE; TCP uses cm. Self-context results are distinct from matched-donor forecasting, the 24-task specificity bank, and the additional 25-task intervention bank. These are offline forecasts.