Skip to content
Interactive demos/The 684 Measurements
Home ↗
← Back to slide 17The transfer study: 684 measurements
← All interactives
The 684 Measurements
Every trained model on every benchmark. Pick an architecture and encoder state: the grid is peak PCK; the plot beside it is transfer against a distance for one target, a stratum, or all nine, with the Spearman ρ.
hover a point for its source
Rows are training sources, columns are target benchmarks, cells are peak PCK@5% (the best checkpoint). Use the target buttons, or click a column header, to choose what the plot shows; "All 9 targets" overlays every context with ρ computed per target and averaged, which is how the paper pools it. Coverage ρ is computed within a (target, architecture, encoder-state) context, so no benchmark or model effect can inflate it; the pooled values reproduce the slide's table. Dashes mark models that were not trained (RAFT was not trained on SPair; FlowFormer-frozen skips it).