Experiment Comparison
Comparison presents metric and coverage differences, dataset and split identity, feature and threshold changes, warnings, and error overlap. It does not automatically select a winner.
Experiments without completed summaries or compatible meaning fail with an unsupported-comparison result. Statistical significance is never implied by a displayed difference.